☆44May 16, 2026Updated 4 months ago
Alternatives and similar repositories for harbor-datasets
Users that are interested in harbor-datasets are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- An agent for auditing repositories of traces for violations of safety properties. Automatically finds cheating (task-level gaming and har…☆15Jun 6, 2026Updated 4 months ago
- ☆24Jun 18, 2026Updated 3 months ago
- Framework for evaluating and improving agents☆5,951Updated this week
- ☆425Apr 30, 2026Updated 5 months ago
- SWE-Marathon: an ultra long-horizon SWE benchmark☆165Sep 25, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆11Nov 16, 2019Updated 6 years ago
- Multi-agent synthetic data generation pipeline capable of generating and validating long horizon terminal/coding tasks for RL training☆74Jul 28, 2025Updated last year
- 💻 SETA: Scaling Environments for Terminal Agents - Environments☆147Feb 16, 2026Updated 7 months ago
- A curated list of awesome Harbor ecosystem projects☆55May 29, 2026Updated 4 months ago
- [NeurIPS 2026 Main] Automate the build, execution and test of software repositories across programming languages and operating systems.☆205Oct 1, 2026Updated last week
- Measuring and evolving with the frontier of agent work☆863Updated this week
- Mixture Density Network Demo in Pytorch☆14Sep 17, 2018Updated 8 years ago
- MUFFIN: Curating Multi-Faceted Instructions for Improving Instruction-Following☆16Oct 31, 2024Updated last year
- Simple WebSockets API☆10Nov 18, 2021Updated 4 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Python package for extractive NLP using the OpenAI API☆17Aug 28, 2024Updated 2 years ago
- ☆15Jan 20, 2026Updated 8 months ago
- ☆12Mar 3, 2022Updated 4 years ago
- [ICLR2026🔥Oral] SwingArena: Competitive Programming Arena for Long-context GitHub Issue Solving☆15Feb 26, 2026Updated 7 months ago
- ☆10Dec 17, 2020Updated 5 years ago
- A benchmark for LLMs on complicated tasks in the terminal☆2,600Jul 11, 2026Updated 2 months ago
- Run SWE-bench evaluations remotely☆83Aug 14, 2025Updated last year
- 登录脚本☆12Nov 4, 2022Updated 3 years ago
- ☆13Jan 21, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Hasura GraphQL Engine on Render☆15Aug 28, 2023Updated 3 years ago
- QuestBench: Can LLMs ask the right question to acquire information in reasoning tasks?☆44Jun 30, 2026Updated 3 months ago
- Convert GitHub PRs into Harbor tasks☆89Jul 13, 2026Updated 2 months ago
- ☆84Jun 25, 2026Updated 3 months ago
- GRPO training code which scales to 32xH100s for long horizon terminal/coding tasks. Base agent is now the top Qwen3 agent on Stanford's T…☆413Aug 24, 2025Updated last year
- Fast Topological Clustering with Wasserstein Distance (ICLR 2022)☆12Jun 24, 2022Updated 4 years ago
- Python SDK for Weaver.☆17Updated this week
- Finetuning Stable Diffusion from Diffusers☆11Mar 11, 2024Updated 2 years ago
- Inferring and Leveraging Parts from Object Shape for Improving Semantic Image Synthesis (CVPR 2023)☆18Dec 13, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- We introduce OpenStory++, a large-scale open-domain dataset focusing on enabling MLLMs to perform storytelling generation tasks.☆18Aug 30, 2024Updated 2 years ago
- Demo showing how to sync data with ElectricSQL from Postgres to Cloudflare's Workers KV☆17Aug 20, 2024Updated 2 years ago
- Benchmarking execution environments ability to prevent reward hacking in agent evals.☆18Sep 15, 2026Updated 3 weeks ago
- Original Implementation of Improving Domain-Adapted Sentiment Classification by Deep Adversarial Mutual Learning publicized in AAAI-2020☆19May 29, 2020Updated 6 years ago
- Open ChatGLM Eyes to See the World☆13Mar 30, 2023Updated 3 years ago
- 可以成功Lora微调的Qwen-VL模型☆16Oct 27, 2023Updated 2 years ago
- Source code and additional results for GLOD issues☆12Jan 19, 2023Updated 3 years ago