☆15Feb 10, 2026Updated 7 months ago
Alternatives and similar repositories for FreeLM
Users that are interested in FreeLM are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆73Mar 3, 2026Updated 6 months ago
- ☆13Jan 22, 2025Updated last year
- ☆18Aug 19, 2024Updated 2 years ago
- ☆20Nov 3, 2024Updated last year
- [ICML 2026] E-mem: Multi-Agent Based Episodic Context Reconstruction for LLM Agent Memory☆31May 3, 2026Updated 4 months ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆16Jul 23, 2024Updated 2 years ago
- SELF-GUIDE: Better Task-Specific Instruction Following via Self-Synthetic Finetuning. COLM 2024 Accepted Paper☆34May 29, 2024Updated 2 years ago
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆17Apr 16, 2026Updated 5 months ago
- Repository for the Q-Filters method (https://arxiv.org/pdf/2503.02812)☆35Mar 7, 2025Updated last year
- ☆22May 23, 2025Updated last year
- ☆16Sep 3, 2026Updated 2 weeks ago
- ☆50Feb 4, 2026Updated 7 months ago
- A list of papers about data quality in Large Language Models (LLMs)☆27Dec 14, 2023Updated 2 years ago
- ☆16Feb 21, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICML 2026] InnoEval: On Research Idea Evaluation as a Knowledge-Grounded, Multi-Perspective Reasoning Problem☆30Jun 21, 2026Updated 2 months ago
- Does Socialization Emerge in AI Agent Society? A Case Study of Moltbook☆18Feb 17, 2026Updated 7 months ago
- ☆17Nov 6, 2025Updated 10 months ago
- Socratic-Zero is a fully autonomous framework that generates high-quality training data for mathematical reasoning☆37Oct 26, 2025Updated 10 months ago
- ☆28Mar 14, 2026Updated 6 months ago
- ☆15Oct 20, 2024Updated last year
- HIP/ROCm fork optimized for AMD RDNA2 (gfx1030) with PrismML Q1_0_G128 1-bit quant support, RotorQuant, TurboQuant, EAGLE3 and P-EAGLE sp…☆25Jun 23, 2026Updated 2 months ago
- Try HopWeaver: The first automatic synthesis framework based on any corpora, with quality approaching manual annotation.☆28Apr 7, 2026Updated 5 months ago
- [ICLR 2026] Thinking on the Fly: Test-Time Reasoning Enhancement via Latent Thought Policy Optimization☆32Mar 6, 2026Updated 6 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 6 months ago
- FinVault: Benchmarking Financial Agent Safety in Execution-Grounded Environments☆19Jun 4, 2026Updated 3 months ago
- A highly contextualized retrieval system integrating Large Language Models (LLMs), embeddings, and a dynamic agent-driven framework. Supp…☆27May 21, 2026Updated 3 months ago
- Enhanced Search-R1 Implementation: Improved Compatibility and Modern Framework Integration☆33Dec 8, 2025Updated 9 months ago
- Codebase for Instruction Following without Instruction Tuning☆36Sep 24, 2024Updated last year
- 30 tok/s for 20B MoE on 8 GB VRAM. Flat throughput to 32K context. Native MXFP4 + GGUF Q4_K/Q5_K/Q6_K via ggml CUDA kernels — zero dequan…☆22Apr 7, 2026Updated 5 months ago
- [NeurIPS 2024 poster] Cross-model Control: Improving Multiple Large Language Models in One-time Training☆15Oct 25, 2024Updated last year
- ☆12Feb 14, 2024Updated 2 years ago
- [ACL 2026 Main] Official Repo for Paper "Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Ali…☆16Jul 1, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [EMNLP'25] AutoSDT is a fully automatic pipeline to collect data-driven scientific coding tasks to train co-scientist models.☆22Aug 11, 2025Updated last year
- Official code for the paper "Why Do Self-Supervised Models Transfer? Investigating the Impact of Invariance on Downstream Tasks".☆16Dec 7, 2021Updated 4 years ago
- Vstream - Video Analytics pipeline with Hardware based accelerations (dev - stage)☆10Feb 2, 2024Updated 2 years ago
- Materials for "Multi-property Steering of Large Language Models with Dynamic Activation Composition"☆14Nov 22, 2024Updated last year
- An implementation of DecorrelatedBN by tensorflow☆13Jun 30, 2022Updated 4 years ago
- 𝐀 𝐅𝐨𝐫𝐤 𝐨𝐟 𝐂𝐡𝐢𝐜𝐚𝐠𝐨𝟗𝟓 𝐦𝐚𝐝𝐞 𝐭𝐨 𝐥𝐨𝐨𝐤 𝐥𝐢𝐤𝐞 𝐖𝐢𝐧𝐝𝐨𝐰𝐬 𝟐𝟎𝟎𝟎.☆22Sep 1, 2026Updated 2 weeks ago
- This is a meta-model distilled from LLMs for information extraction. This is an intermediate checkpoint that can be well-transferred to a…☆30Feb 23, 2025Updated last year