Feeling confused about super alignment? Here is a reading list
☆43Jan 9, 2024Updated 2 years ago
Alternatives and similar repositories for about-super-alignment
Users that are interested in about-super-alignment are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆317Jul 16, 2024Updated 2 years ago
- Explore what LLMs are really leanring over SFT☆28Mar 30, 2024Updated 2 years ago
- [ICML 2023] "Data Efficient Neural Scaling Law via Model Reusing" by Peihao Wang, Rameswar Panda, Zhangyang Wang☆14Jan 4, 2024Updated 2 years ago
- Fine-Tuning Pre-trained Transformers into Decaying Fast Weights☆20Oct 9, 2022Updated 3 years ago
- ☆14Jun 20, 2022Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆13Apr 2, 2024Updated 2 years ago
- [ACL 2024]Official GitHub repo for OlympiadBench: A Challenging Benchmark for Promoting AGI with Olympiad-Level Bilingual Multimodal Scie…☆195Jun 8, 2025Updated last year
- Implementing a Commercial Level Face Payment System with Yolo+Arcface (yolo + arcface 实现 商用级别人脸支付系统 )☆11May 18, 2026Updated 2 months ago
- Tinkering RL☆28Updated this week
- 大语言模型训练和服务调研☆37Aug 4, 2023Updated 3 years ago
- A lightweight post-training framework for LLMs and VLMs. 51 algorithms, 38 verified models. Scales with DeepSpeed, vLLM, and Ray.☆20Aug 4, 2026Updated last week
- Teaching Models to Express Their Uncertainty in Words☆39May 26, 2022Updated 4 years ago
- Crafting Adversarial Examples for Neural Machine Translation☆10Apr 7, 2023Updated 3 years ago
- [EMNLP 2023 Industry Track] A simple prompting approach that enables the LLMs to run inference in batches.☆76Mar 8, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The implementation for our paper, "Improving Simultaneous Machine Translation with Monolingual Data," accepted to AAAI 2023. 🎉☆12Jul 19, 2023Updated 3 years ago
- [ICML 2025] M-STAR (Multimodal Self-Evolving TrAining for Reasoning) Project. Diving into Self-Evolving Training for Multimodal Reasoning☆75Jul 13, 2025Updated last year
- Tools for formatting WMT hypothesis and test sets in XML☆27Apr 18, 2025Updated last year
- [NeurIPS'24] Official code for *🎯DART-Math: Difficulty-Aware Rejection Tuning for Mathematical Problem-Solving*☆121Dec 10, 2024Updated last year
- ☆13Oct 18, 2023Updated 2 years ago
- A curated list of papers & resources linked to concept learning☆13Aug 9, 2023Updated 3 years ago
- [ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective☆21Jun 13, 2025Updated last year
- LaTeX Drawing☆18Dec 22, 2025Updated 7 months ago
- xVerify: Efficient Answer Verifier for Reasoning Model Evaluations☆148Nov 13, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official code of "Exploring the Personality Traits of LLMs through Latent Features Steering"☆18Jan 30, 2025Updated last year
- Official code for the paper: "Multi-User Large Language Model Agents"☆27May 11, 2026Updated 3 months ago
- the instructions and demonstrations for building a formal logical reasoning capable GLM☆54Sep 3, 2024Updated last year
- [TACL 2025] RepreGuard: Detecting LLM-Generated Text by Revealing Hidden Representation Patterns☆16Jan 27, 2026Updated 6 months ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated 2 years ago
- Implementation of the model: "Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models" in PyTorch☆29Aug 3, 2026Updated last week
- WorldSense benchmark for grounded reasoning in language models☆25Nov 28, 2023Updated 2 years ago
- Source codes for paper "BlockFFN: Towards End-Side Acceleration-Friendly Mixture-of-Experts with Chunk-Level Activation Sparsity".☆19Jan 10, 2026Updated 7 months ago
- Awesome papers on Language-Model-as-a-Service (LMaaS)☆545May 14, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆19Jun 13, 2023Updated 3 years ago
- PyTorch implementation of experiments in the paper Aligning Language Models with Human Preferences via a Bayesian Approach☆32Nov 6, 2023Updated 2 years ago
- Separable Diffusion Model Unlearning☆13Jan 29, 2025Updated last year
- GAU-alpha-pytorch☆20May 11, 2022Updated 4 years ago
- Paper collections of methods that using language to interact with environment, including interact with real world, simulated world or WWW…☆128Jul 26, 2023Updated 3 years ago
- Implementation of Influence Function approximations for differently sized ML models, using PyTorch☆18Sep 15, 2023Updated 2 years ago
- Paper list of "The Life Cycle of Knowledge in Big Language Models: A Survey"☆58Aug 24, 2023Updated 2 years ago