Kinetics: Rethinking Test-Time Scaling Laws
β87Jul 11, 2025Updated last year
Alternatives and similar repositories for Kinetics
Users that are interested in Kinetics are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NAACL'25 π SAC Award] Official code for "Advancing MoE Efficiency: A Collaboration-Constrained Routing (C2R) Strategy for Better Expertβ¦β16Feb 4, 2025Updated last year
- β65Jun 12, 2025Updated last year
- Vortex: Programmable Sparse Attention for Agents as Algorithm Designersβ67Jun 24, 2026Updated 3 weeks ago
- [ICML 2026] Less Is More: Training-Free Sparse Attention with Global Locality for Efficient Reasoningβ34Sep 12, 2025Updated 10 months ago
- β16Jun 15, 2026Updated last month
- Managed Kubernetes at scale on DigitalOcean β’ AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- β81Jun 8, 2026Updated last month
- Official repository for Parallax (Parameterized Local Linear Attention)β65Jul 7, 2026Updated last week
- β56Jul 7, 2025Updated last year
- β77May 29, 2026Updated last month
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"β17Jun 30, 2025Updated last year
- Compact and Agent-Native MoE Training Systemβ288Updated this week
- [EMNLP 25] An effective and interpretable weight-editing method for mitigating overly short reasoning in LLMs, and a mechanistic study unβ¦β19Dec 17, 2025Updated 7 months ago
- A bunch of kernels that might make stuff slower πβ91Updated this week
- An efficient implementation of the NSA (Native Sparse Attention) kernelβ133Jun 24, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ICML 2024] Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inferenceβ400Jul 10, 2025Updated last year
- β34Oct 13, 2025Updated 9 months ago
- Official code for the paper "HEXA-MoE: Efficient and Heterogeneous-Aware MoE Acceleration with Zero Computation Redundancy"β15Mar 6, 2025Updated last year
- β88Jun 16, 2025Updated last year
- β16May 27, 2026Updated last month
- The official implementation of HybridNorm: Towards Stable and Efficient Transformer Training via Hybrid Normalizationβ19Mar 7, 2025Updated last year
- Simple & Scalable Pretraining for Neural Architecture Researchβ336Mar 31, 2026Updated 3 months ago
- Ring is a reasoning MoE LLM provided and open-sourced by InclusionAI, derived from Ling.β109Aug 5, 2025Updated 11 months ago
- β36Mar 12, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer β’ AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- β135May 29, 2025Updated last year
- β14Oct 3, 2024Updated last year
- ArcherCodeR is an open-source initiative enhancing code reasoning in large language models through scalable, rule-governed reinforcement β¦β44Aug 6, 2025Updated 11 months ago
- π₯ A minimal training framework for scaling FLA modelsβ403Apr 22, 2026Updated 2 months ago
- VibeRL is a Reinforcement Learning framework built essentially through vibe coding with Kimi K2.β17Jul 13, 2026Updated last week
- [ICMLβ25] Official code for paper "Occult: Optimizing Collaborative Communication across Experts for Accelerated Parallel MoE Training anβ¦β13Apr 17, 2025Updated last year
- Code for "Reasoning to Learn from Latent Thoughts"β133Mar 28, 2025Updated last year
- [ICML 2025] XAttention: Block Sparse Attention with Antidiagonal Scoringβ280Jul 6, 2025Updated last year
- Implementation of "RaanA: A Fast, Flexible, and Data-Efficient Post-Training Quantization Algorithm"β17Apr 11, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Physics of Language Models: Part 4.2, Canon Layers at Scale where Synthetic Pretraining Resonates in Realityβ356May 20, 2026Updated 2 months ago
- β114Aug 26, 2024Updated last year
- Dataflow-Oriented Reinforcement Learning for (Multi-)Agentic LLMsβ95Updated this week
- [ICML2024 Spotlight] Fine-Tuning Pre-trained Large Language Models Sparselyβ24Jun 26, 2024Updated 2 years ago
- [ICLR 26] The official code repository for the paper "Mirage or Method? How ModelβTask Alignment Induces Divergent RL Conclusions".β18Feb 9, 2026Updated 5 months ago
- Official code repo for paper "Great Memory, Shallow Reasoning: Limits of kNN-LMs"β24Apr 30, 2025Updated last year
- Long-RL: Scaling RL to Long Sequences (NeurIPS 2025)β726Sep 24, 2025Updated 9 months ago