alignment

14 projects share this GitHub topic

alignment — aeneas ★2.9kalignmentrlhf-book — ★2.2kWFGY — ★1.8kAwesome-Knowledge-Distillation-of-LLMs — ★1.3kDataDreamer — ★1.1kgangealing — ★1kSimPO — ★956Awesome-LLM-in-Social-Science — ★636CipherChat — ★628deita — ★599awesome-image-alignment-and-stitching — ★462awesome-alignment-of-diffusion-models — ★431AlignProp — ★324VADER — ★315rlhf-book★ 2.2kWFGY★ 1.8kAwesome-Knowledge-Distil…★ 1.3kDataDreamer★ 1.1kgangealing★ 1kSimPO★ 956Awesome-LLM-in-Social-Sc…★ 636CipherChat★ 628deita★ 599awesome-image-alignment-…★ 462awesome-alignment-of-dif…★ 431AlignProp★ 324VADER★ 315

Lines connect members that are measurably related to each other. Dot size reflects stars.

🧬 Members
aeneas
aeneas is a Python/C library and a set of tools to automagically synchronize audio and text (aka forced…
★ 2.9k
rlhf-book
Textbook on reinforcement learning from human feedback
★ 2.2k
WFGY
WFGY is heading toward WFGY 5.0 Polaris Protocol, a major open-source release for AI reasoning, RAG, agents,…
★ 1.8k
Awesome-Knowledge-Distillation-of-LLMs
This repository collects papers for "A Survey on Knowledge Distillation of Large Language Models". We break…
★ 1.3k
DataDreamer
DataDreamer: Prompt. Generate Synthetic Data. Train & Align Models.   🤖💤
★ 1.1k
gangealing
Official PyTorch Implementation of "GAN-Supervised Dense Visual Alignment" (CVPR 2022 Oral, Best Paper…
★ 1k
SimPO
[NeurIPS 2024] SimPO: Simple Preference Optimization with a Reference-Free Reward
★ 956
Awesome-LLM-in-Social-Science
Awesome papers involving LLMs in Social Science.
★ 636
CipherChat
A framework to evaluate the generalization capability of safety alignment for LLMs
★ 628
deita
Deita: Data-Efficient Instruction Tuning for Alignment [ICLR2024]
★ 599
awesome-image-alignment-and-stitching
A curated list of awesome resources for image alignment and stitching ...
★ 462
awesome-alignment-of-diffusion-models
[ACM Computing Surveys] The collection of awesome papers on alignment of diffusion models.
★ 431
AlignProp
AlignProp uses direct reward backpropogation for the alignment of large-scale text-to-image diffusion models.…
★ 324
VADER
Video Diffusion Alignment via Reward Gradients. We improve a variety of video diffusion models such as…
★ 315
🔗 Related families

Measured from GitHub topics shared by both projects, weighted by how rare each topic is.