☆45Jan 30, 2026Updated 6 months ago
Alternatives and similar repositories for ConceptMoE
Users that are interested in ConceptMoE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official implementation of “ReGuLaR: Variational Latent Reasoning Guided by Rendered Chain-of-Thought”☆54Feb 2, 2026Updated 6 months ago
- ☆17May 25, 2026Updated 2 months ago
- ☆17Feb 4, 2026Updated 6 months ago
- On demand communication☆34Apr 16, 2026Updated 3 months ago
- Source code for GreaTer ICLR 2025 - Gradient Over Reasoning makes Smaller Language Models Strong Prompt Optimizers☆36Apr 18, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Official Implementation of ConceptLM.☆23Mar 18, 2026Updated 4 months ago
- FinVault: Benchmarking Financial Agent Safety in Execution-Grounded Environments☆16Jun 4, 2026Updated 2 months ago
- ☆35Updated this week
- MDM-Prime-v2: Binary Encoding and Index Shuffling Enable Scaling of Diffusion Language Models☆27May 23, 2026Updated 2 months ago
- When and How Much to Imagine: Adaptive Test-Time Scaling with World Models for Visual Spatial Reasoning☆18Jun 2, 2026Updated 2 months ago
- Post-Trained MoE Can Skip Half Experts via Self-Distillation☆38May 19, 2026Updated 2 months ago
- Neural Lexicon Reader: Reduce Pronunciation Errors in End-to-end TTS by Leveraging External Textual Knowledge☆21Jul 25, 2022Updated 4 years ago
- ☆17Jun 25, 2025Updated last year
- ☆19May 12, 2026Updated 2 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ICLR 2026] Harder Is Better: Boosting Mathematical Reasoning via Difficulty-Aware GRPO and Multi-Aspect Question Reformulation☆128May 17, 2026Updated 2 months ago
- ProAct is a framework designed to enable Large Language Model (LLM) agents to perform accurate, multi-turn lookahead reasoning in interac…☆18Feb 11, 2026Updated 5 months ago
- Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents☆31Apr 16, 2026Updated 3 months ago
- Source code to accompany research paper on training multi token prediction language models using self-distillation.☆39Feb 21, 2026Updated 5 months ago
- Code and models for the paper: Hybrid Linear Attention Done Right: Efficient Distillation and Effective Architectures for Extremely Long …☆40Apr 9, 2026Updated 3 months ago
- mamba2-jax: A pure JAX/Flax implementation of Mamba-2 for language modeling and time series forecasting.☆17Jun 23, 2026Updated last month
- This repository contains the code and data for the paper "Chaining the Evidence: Robust Reinforcement Learning for Deep Search Agents wit…☆72Apr 8, 2026Updated 3 months ago
- Homepage for paper “MeKi : Memory-based Expert Knowledge Injection for Efficient LLM Scaling”☆29Mar 5, 2026Updated 4 months ago
- A proof-of-concept implementation of Titans: models mixing long-term, short-term and persistent memories☆24Apr 9, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A benchmark for evaluating contextual agents on realistic multimodal personal-computer environments with profiling and factual-retention …☆31Apr 2, 2026Updated 4 months ago
- Template for creating audio encoders compatible with X-ARES☆19Feb 11, 2026Updated 5 months ago
- Compress and Attend Transformers (CATs) 😸☆23Jul 13, 2026Updated 3 weeks ago
- [ICLR-2026] Official Implementation of our paper "THOR: Tool-Integrated Hierarchical Optimization via RL for Mathematical Reasoning".☆33Feb 26, 2026Updated 5 months ago
- Skill-Targeted Adaptive Training☆25Mar 12, 2026Updated 4 months ago
- Official code and dataset for our paper: RefineBench: Evaluating Refinement Capability of Language Models via Checklists☆17Dec 1, 2025Updated 8 months ago
- ☆50Aug 5, 2025Updated 11 months ago
- ABC-Bench is a benchmark for Agentic Backend Coding. It evaluates whether code agents can explore real repositories, edit code, configure…☆33Jan 20, 2026Updated 6 months ago
- ☆11Nov 8, 2022Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for paper "FineRS: Fine-grained Reasoning and Segmentation of Small Objects with Reinforcement Learning" Neurips2025.☆15Jan 29, 2026Updated 6 months ago
- ☆36Apr 18, 2024Updated 2 years ago
- ☆22Jan 29, 2026Updated 6 months ago
- [ICLR 2026 🔥] Dr.LLM: Dynamic Layer Routing in LLMs☆57Apr 24, 2026Updated 3 months ago
- Visual Programming for Text-to-Image Generation and Evaluation (NeurIPS 2023)☆57Jul 25, 2023Updated 3 years ago
- Network representation learning on drug-target-side effects-indication graphs for side effect prediction☆13Feb 4, 2020Updated 6 years ago
- The official code of TACL 2022, "Break, Perturb, Build: Automatic Perturbation of Reasoning Paths Through Question Decomposition".☆12Oct 18, 2021Updated 4 years ago