Benchmarking long-horizon chain-of-thought reasoning.
☆43Apr 20, 2026Updated 4 months ago
Alternatives and similar repositories for longcot
Users that are interested in longcot are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Storing the LongCoT-mini results for RLM(GPT-5.2)☆20Apr 26, 2026Updated 3 months ago
- Implementations of Stable Contrastive RL☆22Apr 13, 2025Updated last year
- ☆19Apr 28, 2024Updated 2 years ago
- An MLX implementation of Meta AI's ESM-2 protein language model☆16Aug 16, 2025Updated last year
- code for ACL2024-main: BatchEval: Towards Human-like Text Evaluation☆19May 20, 2024Updated 2 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆26Oct 29, 2025Updated 9 months ago
- Scripts for fine-tuning an HPC Code LLM☆17Jul 19, 2024Updated 2 years ago
- This code accompanies the paper "Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration."☆39Jul 11, 2025Updated last year
- A benchmark to measure AI progress on unsolved research problems in mathematics.☆32Updated this week
- ☆32Mar 11, 2026Updated 5 months ago
- Production focused Self-harnessed LM runtime (RLM) that allows the LM to call its sub-lm with DSPy signatures. Define your inputs, output…☆428Aug 5, 2026Updated 2 weeks ago
- code for EACL2024-main:Generative Dense Retrieval: Memory Can Be a Burden☆32Jan 19, 2024Updated 2 years ago
- ☆13Mar 5, 2025Updated last year
- Code for MatGPTQ: Accurate and Efficient Post-Training Matryoshka Quantization☆23Feb 18, 2026Updated 6 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [XLLM@ACL2025] Official Code for "Less is More: Enhancing Structured Multi-Agent Reasoning via Quality-Guided Distillation"☆22Jul 29, 2025Updated last year
- Public repository for the Remote Labor Index (RLI)☆75Nov 3, 2025Updated 9 months ago
- Codes and files for the paper Are Emergent Abilities in Large Language Models just In-Context Learning☆33Jan 9, 2025Updated last year
- High performance hybrid Monte Carlo simulation☆10Jun 29, 2026Updated last month
- ☆14Nov 1, 2023Updated 2 years ago
- Minimal Transformer base in JAX. A single backbone for language modelling, diffusion, classification, etc...☆16May 28, 2025Updated last year
- ☆15Jan 14, 2026Updated 7 months ago
- A self-extending MCP server☆22Apr 10, 2026Updated 4 months ago
- Giza++☆12May 12, 2015Updated 11 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆11Oct 31, 2021Updated 4 years ago
- Code for the paper "Data Feedback Loops: Model-driven Amplification of Dataset Biases"☆18Sep 9, 2022Updated 3 years ago
- This repository is the official implementation of the Hybrid Self-Attention NEAT algorithm. It contains the code to reproduce the results…☆15Jun 19, 2023Updated 3 years ago
- Repo of "Seeing the Whole Elephant: A Benchmark for Failure Attribution in LLM-based Multi-Agent Systems" (ACL 2026)☆19Apr 27, 2026Updated 3 months ago
- CoEvolve: Training LLM Agents via Agent-Data Mutual Evolution☆23Apr 27, 2026Updated 3 months ago
- ☆34Sep 10, 2025Updated 11 months ago
- Overlooked Factors in Concept-based Explanations: Dataset Choice, Concept Learnability, and Human Capability (CVPR 2023)☆10Mar 14, 2023Updated 3 years ago
- Agent Skill to help convert transformer LLMs to mlx-lm☆51Jun 16, 2026Updated 2 months ago
- A framework for steering MoE models by detecting and controlling behavior-linked experts.☆36Sep 12, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A challenging aggregation benchmark for long-context models☆52Feb 22, 2026Updated 6 months ago
- Challenges for general-purpose web-browsing AI agents☆68Jun 2, 2025Updated last year
- UQ: Assessing Language Models on Unsolved Questions☆30Aug 26, 2025Updated 11 months ago
- My collection of dotfiles☆14Apr 22, 2026Updated 4 months ago
- ☆27Jun 7, 2026Updated 2 months ago
- Evaluation repository of wikipedia index with Dria☆10Mar 14, 2024Updated 2 years ago
- A very hacky set of functions for getting plotly to do what I want when doing mech interp research, designed to be compatible with PyTorc…☆15Jun 16, 2023Updated 3 years ago