☆29Jul 3, 2026Updated 3 weeks ago
Alternatives and similar repositories for ProjectEval
Users that are interested in ProjectEval are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆28Mar 11, 2026Updated 4 months ago
- ☆18Jun 25, 2026Updated last month
- A Comprehensive Benchmark for Software Development.☆138May 30, 2024Updated 2 years ago
- ☆10Jun 27, 2024Updated 2 years ago
- Code for "Possibility Before Utility: Learning And Using Hierarchical Affordances" (ICLR 2022)☆14Mar 14, 2022Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆10Jul 14, 2024Updated 2 years ago
- A mobile GUI search engine using a vision-language model☆15May 5, 2025Updated last year
- SearchEyes: Towards Frontier Multimodal Deep Search Intelligence via Search World Simulation. A typed knowledge graph unifies data synthe…☆21Jul 8, 2026Updated 3 weeks ago
- [ICLR 2026 Oral] RAIN-Merging☆15Mar 9, 2026Updated 4 months ago
- Include ML DL RL, knowledge and code☆12Feb 12, 2023Updated 3 years ago
- A Manually-Annotated Code Generation Benchmark Aligned with Real-World Code Repositories☆41Sep 4, 2024Updated last year
- Code for ACL 2022 main conference paper "Modeling Dual Read/Write Paths for Simultaneous Machine Translation"☆12Mar 31, 2022Updated 4 years ago
- A Differentially-Private Random Decision Forest using Smooth Sensitivity☆11Oct 26, 2016Updated 9 years ago
- Code for EMNLP 2022 main conference paper "Information-Transport-based Policy for Simultaneous Translation"☆13Nov 3, 2022Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ISSTA'24] A Large-Scale Dataset Capable of Enhancing the Prowess of Large Language Models for Program Testing☆12Jan 7, 2025Updated last year
- Pytorch implementation of [Feudal Net](https://arxiv.org/abs/1703.01161). ([Tensorflow version](https://github.com/dmakian/feudal_networ…☆18Jun 25, 2019Updated 7 years ago
- ☆23Jul 4, 2026Updated 3 weeks ago
- Dataset from Tip of the Tongue Known-Item Retrieval (2021) paper.☆12Nov 4, 2021Updated 4 years ago
- A Benchmark for Evaluating Safety and Trustworthiness in Web Agents for Enterprise Scenarios☆25Mar 12, 2026Updated 4 months ago
- mcp wrapper for openai built-in tools☆12Mar 13, 2025Updated last year
- Simple setup for personal dotfiles☆11Jul 4, 2026Updated 3 weeks ago
- [ECCV 2026] Official implementation of "TIR-Bench: A Comprehensive Benchmark for Agentic Thinking-with-Images Reasoning"☆25Feb 8, 2026Updated 5 months ago
- Clober Solidity Library☆10Jun 9, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [NeurIPS 2024] Code and Data Repo for Paper "Embedding Trajectory for Out-of-Distribution Detection in Mathematical Reasoning"☆28May 28, 2024Updated 2 years ago
- ☆22Dec 20, 2019Updated 6 years ago
- ☆12Mar 18, 2024Updated 2 years ago
- Project OCELoT: an Open, Collaborative Evaluation Leaderboard of Translations☆23Jul 11, 2026Updated 2 weeks ago
- ☆17Nov 19, 2025Updated 8 months ago
- A curated list of personalized Language model / Large language model (continually updated)☆10Nov 17, 2023Updated 2 years ago
- Prototype for a game testing framework using AI methods☆10Feb 25, 2023Updated 3 years ago
- ☆34Jul 13, 2026Updated 2 weeks ago
- LLMs interview notes and answers:该仓库主要记录大模型(LLMs)算法工程师相关的面试题和参考答案☆17Oct 16, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- A collection of practical code generation tasks and tests from open source projects. Complementary to HumanEval by OpenAI.☆24Jan 28, 2023Updated 3 years ago
- ☆27Feb 3, 2026Updated 5 months ago
- Repo to reproduce the First-Explore paper results☆39May 6, 2026Updated 2 months ago
- Website for Learning from "Big Code"☆30Jun 19, 2021Updated 5 years ago
- The authenticated storage benchmark platform includes implementations like OpenEthereum's MPT, RainBlock's MPT, LMPTs, and LVMT.☆16Jul 6, 2023Updated 3 years ago
- (CVPR 2026 Highlight) Official repository for Scone (Subject-driven COmposition and DistinctioN Enhancement) model, supporting subject co…☆32Apr 9, 2026Updated 3 months ago
- Tools and prompt templates used to build and evaluate SWE-rebench-v2 tasks for the paper.☆72Mar 12, 2026Updated 4 months ago