☆56Jun 25, 2026Updated 2 months ago
Alternatives and similar repositories for futuresim
Users that are interested in futuresim are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Codebase from our first release.☆63Feb 17, 2026Updated 6 months ago
- ☆21Apr 3, 2026Updated 5 months ago
- ☆57Mar 18, 2026Updated 5 months ago
- Official code and dataset for our paper: RefineBench: Evaluating Refinement Capability of Language Models via Checklists☆17Dec 1, 2025Updated 9 months ago
- Official Implementation for the paper "Discovering Knowledge Deficiencies of Language Models on Massive Knowledge Base"☆26Sep 2, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆33Mar 18, 2026Updated 5 months ago
- FrontierSmith, a new system that uses AI to synthesize open-ended coding problems at scale☆52May 30, 2026Updated 3 months ago
- ☆27Jun 22, 2026Updated 2 months ago
- Trust Region Preference Approximation: A simple and stable reinforcement learning algorithm for LLM reasoning☆14Jun 28, 2025Updated last year
- Repo for Paper: Discovering Interpretable Algorithms by Decompiling Transformers to RASP☆16May 25, 2026Updated 3 months ago
- Code for 'Answer Matching Outperforms Multiple Choice for Language Model Evaluation' paper☆19Jul 4, 2025Updated last year
- [COLM-LLA 2026] The official implementation for paper "AdaExplore: Failure-Driven Adaptation and Diversity-Preserving Search for Efficien…☆24Aug 23, 2026Updated 3 weeks ago
- This repository contains the code and data for the paper "Chaining the Evidence: Robust Reinforcement Learning for Deep Search Agents wit…☆73Apr 8, 2026Updated 5 months ago
- ☆33Apr 29, 2026Updated 4 months ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- Decomposing and measuring evaluation awareness in existing benchmarks and our proposed EvalAwareBench.☆19Jun 1, 2026Updated 3 months ago
- Storing the LongCoT-mini results for RLM(GPT-5.2)☆20Apr 26, 2026Updated 4 months ago
- ☆115Updated this week
- Repository for the paper: "TiC-LM: A Web-Scale Benchmark for Time-Continual LLM Pretraining" ACL Oral 2025☆24Updated this week
- [ECCV'24 Oral] PiTe: Pixel-Temporal Alignment for Large Video-Language Model☆17Feb 13, 2025Updated last year
- Head Vis Public Release☆41May 4, 2026Updated 4 months ago
- Source code for the collaborative reasoner research project at Meta FAIR.☆116Mar 26, 2026Updated 5 months ago
- TEVO: evolve LM motifs cheaply, then validate them in downstream train.py loops.☆19Apr 18, 2026Updated 4 months ago
- Landing page for MIB: A Mechanistic Interpretability Benchmark☆26Aug 15, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- A drop-in replacement for the standard Categorical Cross-Entropy (CCE) loss that significantly improves OOD and Calibration performance w…☆53Apr 6, 2026Updated 5 months ago
- Source code of "Leaky Thoughts: Large Reasoning Models Are Not Private Thinkers" EMNLP 2025☆17Jan 12, 2026Updated 8 months ago
- CS194-196 Course Project☆14Feb 20, 2025Updated last year
- ☆79Mar 6, 2025Updated last year
- ☆80Apr 26, 2026Updated 4 months ago
- [WWW 2026 Oral] MoE-CL:Self-Evolving LLMs via Continual Instruction Tuning☆22Dec 1, 2025Updated 9 months ago
- ☆21Apr 21, 2026Updated 4 months ago
- An unbounded n-gram language model on Tiny Shakespeare☆22Jan 21, 2026Updated 7 months ago
- Official implementation of the ΔBelief-RL method.☆31Feb 28, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- SHARPIE: Shared Human-AI Reinforcement Learning Platform for Interactive Experiments☆27Aug 3, 2026Updated last month
- The official code repo of paper "Chasing the Tail: Effective Rubric-based Reward Modeling for Large Language Model Post-Training"☆32Feb 20, 2026Updated 6 months ago
- EdgeBench: Unveiling scaling laws of learning from real-world environments☆443Updated this week
- DASH: Detection and Assessment of Systematic Hallucinations of VLMs☆16Jul 2, 2025Updated last year
- The code for "VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by VIdeo SpatioTemporal Augmentation" [CVPR2025]☆20Feb 27, 2025Updated last year
- ADAG: Transluce's MLP neuron-level circuit tracing library☆37Apr 10, 2026Updated 5 months ago
- ALMA (Automated meta-Learning of Memory designs for Agentic systems) is a framework that meta-learns memory designs to replace human-engi…☆294Apr 8, 2026Updated 5 months ago