A curated reading list for large language model (LLM) alignment. Take a look at our new survey "Large Language Model Alignment: A Survey" for more details!
☆81Sep 28, 2023Updated 2 years ago
Alternatives and similar repositories for llm-alignment-survey
Users that are interested in llm-alignment-survey are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A Bilingual Role Evaluation Benchmark for Large Language Models☆44Jan 9, 2024Updated 2 years ago
- Implement of our TKDE paper: Hyperbolic Graph Learning for Social Recommendation☆13Jun 3, 2024Updated 2 years ago
- Official implementation for "ALI-Agent: Assessing LLMs'Alignment with Human Values via Agent-based Evaluation"☆21Jan 31, 2026Updated 7 months ago
- a brief repo about paper research☆15Sep 4, 2024Updated 2 years ago
- ☆12Apr 25, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The implement of papar "Enhanced Graph Learning for Collaborative Filtering via Mutual Information Maximization"☆18Aug 11, 2021Updated 5 years ago
- ☆10Jan 28, 2024Updated 2 years ago
- An implementation of SEAL: Safety-Enhanced Aligned LLM fine-tuning via bilevel data selection.☆24Feb 20, 2025Updated last year
- Code for "When LLM Meets DRL: Advancing Jailbreaking Efficiency via DRL-guided Search" (NeurIPS 2024)☆18Oct 22, 2024Updated last year
- Tensorflow implementation of our SIGIR 2023 accepted paper "Generative-Contrastive Graph Learning for Recommendation"☆32Aug 26, 2024Updated 2 years ago
- This is the oficial repository for "Safer-Instruct: Aligning Language Models with Automated Preference Data"☆17Feb 22, 2024Updated 2 years ago
- Aligning Large Language Models with Human: A Survey☆738Sep 11, 2023Updated 2 years ago
- Pytorch code for "Learning Guidance Rewards with Trajectory-space Smoothing" (NeurIPS 2020)☆12Jul 7, 2021Updated 5 years ago
- Repository for the paper "Cognitive Mirage: A Review of Hallucinations in Large Language Models"☆49Oct 21, 2023Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Puzzle Generator; Einstein's Riddle, Zebra Puzzle and Blood Donation Puzzle Solver. For non-commercial use only!☆20Mar 4, 2023Updated 3 years ago
- ☆20Sep 23, 2018Updated 7 years ago
- Multi-agent Social Simulation + Efficient, Effective, and Stable alternative of RLHF. Code for the paper "Training Socially Aligned Langu…☆357Jun 18, 2023Updated 3 years ago
- A framework to train language models to learn invariant representations.☆14Jan 24, 2022Updated 4 years ago
- [NeurIPS 2024 Oral] Aligner: Efficient Alignment by Learning to Correct☆196Jan 16, 2025Updated last year
- Source codes for "Structure-Aware Abstractive Conversation Summarization via Discourse and Action Graphs"☆65Oct 27, 2023Updated 2 years ago
- The code for NeurIPS 2023 paper DSR☆14Oct 8, 2023Updated 2 years ago
- The official code release for Don't Trust: Verify -- Grounding LLM Quantitative Reasoning with Autoformalization☆39Mar 9, 2025Updated last year
- Pytorch implementation of 'Commonsense Knowledge Aware Conversation Generation with Graph Attention'☆26Jul 25, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- NTK scaled version of ALiBi position encoding in Transformer.☆69Aug 16, 2023Updated 3 years ago
- Rewarded soups official implementation☆66Sep 27, 2023Updated 2 years ago
- ☆13Feb 17, 2025Updated last year
- ☆10Mar 19, 2024Updated 2 years ago
- Paper List for In-context Learning 🌷☆191Feb 21, 2024Updated 2 years ago
- Code for ACL22 short Paper "Hierarchical Curriculum Learning for AMR Parsing"☆13Jun 1, 2022Updated 4 years ago
- ☆17Dec 12, 2020Updated 5 years ago
- [ACL 2024] A Survey of Chain of Thought Reasoning: Advances, Frontiers and Future☆501Jan 16, 2025Updated last year
- ☆41Jun 7, 2023Updated 3 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Flames is a highly adversarial benchmark in Chinese for LLM's harmlessness evaluation developed by Shanghai AI Lab and Fudan NLP Group.☆68May 21, 2024Updated 2 years ago
- A Quasi-Wasserstein Loss for Learning Graph Neural Networks (QW loss)☆10May 20, 2024Updated 2 years ago
- ☆24Dec 2, 2023Updated 2 years ago
- Learnable Global Pooling Layers Based on Regularized Optimal Transport (ROT)☆16Mar 17, 2024Updated 2 years ago
- EMNLP'23 survey: a curation of awesome papers and resources on refreshing large language models (LLMs) without expensive retraining.☆135Dec 12, 2023Updated 2 years ago
- Secrets of RLHF in Large Language Models Part I: PPO☆1,430Mar 3, 2024Updated 2 years ago
- Training Neural Networks Without Gradients: A Scalable ADMM Approach python implement☆14Jun 20, 2017Updated 9 years ago