The official implementation of "ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering"
☆74Jun 21, 2025Updated last year
Alternatives and similar repositories for ML-Agent
Users that are interested in ML-Agent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The official implementation of "ML-Master: Towards AI-for-AI via Integration of Exploration and Reasoning"☆452Mar 29, 2026Updated 6 months ago
- ☆108Oct 30, 2025Updated 11 months ago
- ☆27Feb 27, 2026Updated 7 months ago
- A unified evaluation toolkit and leaderboard for rigorously assessing the scientific intelligence of large language and vision–language m…☆87Aug 30, 2026Updated last month
- Inspired by Crunchbase, ShortStack is the CfA database for public software, products and government entities☆14Oct 10, 2011Updated 15 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆18Feb 20, 2024Updated 2 years ago
- [Under Progress] Code & Data for the AAAI 2020 Paper "Likelihood Ratios and Generative Classifiers For Unsupervised OOD Detection In Task…☆10Jul 25, 2024Updated 2 years ago
- [ACL 2023] Few-shot Reranking for Multi-hop QA via Language Model Prompting☆27Oct 19, 2025Updated 11 months ago
- [NeurIPS 2025 D&B Track] MLR-Bench: Evaluating AI Agents on Open-Ended Machine Learning Research☆36Sep 8, 2026Updated last month
- ☆16Jun 4, 2025Updated last year
- ☆16Mar 6, 2025Updated last year
- [ICML-25] AutoML-Agent: A Multi-Agent LLM Framework for Full-Pipeline AutoML☆171Jul 11, 2025Updated last year
- ☆90Sep 11, 2024Updated 2 years ago
- ☆246Jul 25, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Offline probe-based evaluation harness for hermes-agent's ContextCompressor. Methodology adapted from Factory's Dec 2025 'Evaluating Comp…☆20May 16, 2026Updated 4 months ago
- [ICLR 2026] SR-Scientist: Scientific Equation Discovery With Agentic AI☆57Jan 27, 2026Updated 8 months ago
- Github repository for "Internalizing World Models via Self-Play Finetuning for Agentic RL" (EMNLP Findings 2026)☆38Nov 1, 2025Updated 11 months ago
- Instant Graph Neural Networks for Dynamic Graphs☆11Dec 28, 2022Updated 3 years ago
- Dataflow-LoopAI is an intelligent system with self-optimization capabilities that automatically detects and evaluates generation deficien…☆29Sep 29, 2026Updated last week
- Multi-Agent System Powered by LLMs for End-to-end Multimodal ML Automation☆308Sep 1, 2026Updated last month
- MLE-bench is a benchmark for measuring how well AI agents perform at machine learning engineering☆1,769Apr 24, 2026Updated 5 months ago
- R1V, trained with AI feedback, answers open-ended visual questions.☆14Apr 12, 2025Updated last year
- AIRA-dojo: a framework for developing and evaluating AI research agents☆174Apr 14, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- A Simple Active-and-Adaptive Baseline for Cross-Domain 3D Semantic Segmentation☆13Dec 22, 2022Updated 3 years ago
- This is a joint project between Helmholtz Imaging (located at DKFZ) and Lin Yang and Otmar Schmid (Helmholtz Munich).☆13Nov 6, 2024Updated last year
- ☆43Jun 4, 2026Updated 4 months ago
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 5 months ago
- Official PyTorch Implementation of ZSLViT (CVPR'24)☆18Jul 9, 2024Updated 2 years ago
- Code for Paper 'DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards'☆18May 21, 2026Updated 4 months ago
- Official implementation of X-Master, a general-purpose tool-augmented reasoning agent.☆323Oct 18, 2025Updated 11 months ago
- A comrephensive collection of learning from rewards in the post-training and test-time scaling of LLMs, with a focus on both reward model…☆74Jun 13, 2025Updated last year
- [ICLR 2026] Efficient Agent Training for Computer Use☆145Sep 5, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICLR 25 Oral] RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style☆86Jul 18, 2025Updated last year
- [ACL 2025] DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues☆26Jul 10, 2025Updated last year
- Radiology Language Evaluations☆11Nov 17, 2023Updated 2 years ago
- Data and codes for MetroGAN☆16Dec 23, 2024Updated last year
- "Open-source toolkit (Python Library, Registry API, CLI) for secure, decentralized AI agent interoperability using A2A/MCP."☆20May 10, 2025Updated last year
- The code of Dynamic Graph Learning Based on Hierarchical Memory for Origin-Destination Demand Prediction☆15Apr 29, 2022Updated 4 years ago
- [NeurIPS 2024] OlympicArena: Benchmarking Multi-discipline Cognitive Reasoning for Superintelligent AI☆106Mar 6, 2025Updated last year