☆17May 7, 2026Updated 3 months ago
Alternatives and similar repositories for AgenticSZZ
Users that are interested in AgenticSZZ are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Compare chatbots pairwise via multi‑round evaluations for SE tasks.☆15Apr 4, 2026Updated 4 months ago
- A self-evolving coding agent in Rust: the smallest possible implementation that actually works.☆15Jul 10, 2026Updated last month
- AutoResearch official style beginner tutorial, from 0 to 1☆19Aug 11, 2026Updated 2 weeks ago
- ☆12Jun 19, 2026Updated 2 months ago
- PocketFlow from 0 to 1 | 100 行代码构建所有 LLM 应用 | 首个 PocketFlow 交互式教程 | 光速掌握智能体开发实战☆34Aug 14, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- You install it. Claude drives the CLI tool. At first it might seem like too much. Eventually, nothing less will make sense.☆26Jun 2, 2026Updated 2 months ago
- Meta-specification framework for AI Agents to generate Spec-driven X toolkits automatically.☆51Nov 22, 2025Updated 9 months ago
- A curated list of awesome autonomous researcher frameworks☆153Updated this week
- A curated list of awesome open source libraries to deploy, monitor, version and scale agentic applications and systems☆166Aug 1, 2026Updated 3 weeks ago
- A curated list of awesome leaderboard-oriented resources for AI domain☆381Updated this week
- Evaluation tools for Retrieval-augmented Generation (RAG) methods.☆171Nov 18, 2024Updated last year
- ☆18Apr 15, 2024Updated 2 years ago
- HFCommunity offers an offline up-to-date relational database built from the data available at the Hugging Face Hub, providing queriable d…☆16Oct 14, 2024Updated last year
- Program Transformation Tool for Java Methods☆10Sep 16, 2022Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- MODIT: On Multi-Modal Learning of Editing Source Code.☆20Apr 24, 2021Updated 5 years ago
- ☆20Mar 6, 2023Updated 3 years ago
- SZZ Algorithm To Detect Fault-Inducing Commits☆51Oct 24, 2023Updated 2 years ago
- A BPMN.js extension to improve working with SpiffWorkflow - the python BPMN engine.☆31Jul 14, 2026Updated last month
- CoditT5: Pretraining for Source Code and Natural Language Editing☆29Jan 16, 2025Updated last year
- ☆36May 25, 2023Updated 3 years ago
- Sample project which shows how to implement a secured AngularJS/Spring-Boot application secured by Keycloak.☆38May 17, 2016Updated 10 years ago
- ☆43May 9, 2024Updated 2 years ago
- Code for paper: "Executing Arithmetic: Fine-Tuning Large Language Models as Turing Machines"☆10Oct 11, 2024Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- 《机器学习理论导引》(宝箱书)的证明、案例、概念补充与参考文献讲解。☆1,719Aug 19, 2026Updated last week
- Codev-Bench (Code Development Benchmark), a fine-grained, real-world, repository-level, and developer-centric evaluation framework. Codev…☆49Nov 6, 2024Updated last year
- LLM benchmarks☆13Feb 22, 2024Updated 2 years ago
- 中文金融大模型测评基准,六大类二十五任务、等级化评价,国内模型获得A级☆10May 6, 2024Updated 2 years ago
- Knowledge Graph based Question Answering benchmark.☆10Feb 1, 2020Updated 6 years ago
- ☆49Nov 19, 2025Updated 9 months ago
- Replication Package for "Natural Attack for Pre-trained Models of Code", ICSE 2022☆52May 31, 2026Updated 2 months ago
- ☆11Nov 5, 2024Updated last year
- LGEB: Benchmark of Language Generation Evaluation☆16Oct 21, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Website for release of TellMeWhy dataset for why question answering☆14Nov 11, 2022Updated 3 years ago
- Code for our project CROWN (Conversational Passage Ranking by Reasoning over Word Networks)☆10Jan 11, 2024Updated 2 years ago
- [LREC-Coling 2024] PECC: Problem Extraction and Coding Challenges☆14May 30, 2024Updated 2 years ago
- Math-aware QA system☆18May 8, 2026Updated 3 months ago
- RACE is a multi-dimensional benchmark for code generation that focuses on Readability, mAintainability, Correctness, and Efficiency.☆14Oct 12, 2024Updated last year
- A framework for evaluating the effectiveness of chain-of-thought reasoning in language models.☆19Feb 6, 2025Updated last year
- Demo scripts for HPS Dataset (http://virtualhumans.mpi-inf.mpg.de/hps/)☆11Mar 10, 2025Updated last year