☆22Jan 2, 2026Updated 7 months ago
Alternatives and similar repositories for QuestA
Users that are interested in QuestA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21Mar 22, 2026Updated 4 months ago
- [ICLR 2026] Official Implementation of ProxyThinker: Test-Time Guidance through Small Visual Reasoners.☆22Sep 24, 2025Updated 10 months ago
- ☆31Feb 7, 2025Updated last year
- ☆25Dec 13, 2024Updated last year
- [ACM MM25] LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models☆24Mar 29, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- This is an official repository for "Performance Scaling via Optimal Transport: Enabling Data Selection from Partially Revealed Sources" (…☆14Oct 26, 2023Updated 2 years ago
- ☆25Feb 12, 2026Updated 5 months ago
- [COLM 2026] Resa: Transparent Reasoning Models via SAEs☆49Sep 23, 2025Updated 10 months ago
- The official repo for LIFT: Language-Image Alignment with Fixed Text Encoders☆43Jun 10, 2025Updated last year
- Scaf-GRPO: Scaffolded Group Relative Policy Optimization for Enhancing LLM Reasoning☆25Feb 8, 2026Updated 6 months ago
- The official data and code for EMNLP 2023 main conference paper: CRT-QA: A Dataset of Complex Reasoning Question Answering over Tabular D…☆13May 19, 2025Updated last year
- [NeurIPS 2025] ChartMuseum: Testing Visual Reasoning Capabilities of Large Vision-Language Models☆24Apr 20, 2026Updated 3 months ago
- Code for ProTrix: Building Models for Planning and Reasoning over Tables with Sentence Context☆17Nov 15, 2024Updated last year
- (ICLR 2026) Unveiling Super Experts in Mixture-of-Experts Large Language Models☆44Sep 25, 2025Updated 10 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- {DeepL, Google, WMT-Best, davinci-003, turbo, gpt-4} × {En-De, En-Cs, En-Ru, En-Zh, De-Fr, En-Ja, Uk-En, Uk-Cs, En-Hr, En-Ha, En-Is}☆14Jun 18, 2023Updated 3 years ago
- Official code implementation for the ACL 2025 paper: 'CoT-based Synthesizer: Enhancing LLM Performance through Answer Synthesis'☆32May 19, 2025Updated last year
- ☆15Dec 15, 2025Updated 7 months ago
- ☆15Apr 11, 2024Updated 2 years ago
- ☆15Nov 19, 2025Updated 8 months ago
- This is the official implementation of the ICML 2023 paper - Can Forward Gradient Match Backpropagation ?☆13May 31, 2023Updated 3 years ago
- ☆27Jun 10, 2025Updated last year
- Command helper for slurm system. Act as if you are on compute node.☆16Feb 1, 2025Updated last year
- Post-Trained MoE Can Skip Half Experts via Self-Distillation☆39May 19, 2026Updated 2 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Source code for "Taming GANs with Lookahead–Minmax", ICLR 2021.☆15Mar 28, 2021Updated 5 years ago
- Code for "ATTA: Anomaly-aware Test-Time Adaptation for Out-of-Distribution Detection in Segmentation" (NeurIPS 23)☆16Apr 12, 2024Updated 2 years ago
- FastCuRL: Curriculum Reinforcement Learning with Stage-wise Context Scaling for Efficient LLM Reasoning (EMNLP 2025)☆61Oct 10, 2025Updated 10 months ago
- [ACL 25] SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities☆30Apr 2, 2025Updated last year
- A Physiology-Informed ECG Delineation Algorithm Based on Peak Prominence☆17Jan 4, 2026Updated 7 months ago
- ☆35Jan 20, 2026Updated 6 months ago
- ☆30Nov 9, 2025Updated 9 months ago
- ☆16Jul 23, 2024Updated 2 years ago
- Testing paligemma2 finetuning on reasoning dataset☆18Dec 28, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Code for the paper "Stack Attention: Improving the Ability of Transformers to Model Hierarchical Patterns"☆18Mar 15, 2024Updated 2 years ago
- Codes about our paper "Semi-supervised medical image segmentation via hard positives oriented contrastive learning"☆12Jan 13, 2025Updated last year
- ThetaEvolve: Test-time Learning on Open Problems, enabling RL training on AlphaEvolve/OpenEvolve and emphasizing scaling test-time comput…☆173Feb 27, 2026Updated 5 months ago
- Implementations of the algorithms described in the paper: On the Convergence Theory for Hessian-Free Bilevel Algorithms.☆11Nov 1, 2024Updated last year
- [ICML 2026 Spotlight] Code for miXed Discrete Diffusion Language Model☆29Mar 16, 2026Updated 4 months ago
- [ICLR 2026] Rectifying LLM Thought From Lens of Optimization☆15Dec 5, 2025Updated 8 months ago
- Optimization algorithm which fits a ResNet to CIFAR-10 5x faster than SGD / Adam (with terrible generalization)☆14Oct 20, 2023Updated 2 years ago