BigCodeBench: Benchmarking Code Generation with Diverse Function Calls and Complex Instructions
☆26Aug 8, 2024Updated 2 years ago
Alternatives and similar repositories for bigcodebench-annotation
Users that are interested in bigcodebench-annotation are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- BigCodeArena: Unveiling More Reliable Human Preferences in Code Generation via Execution☆61Oct 13, 2025Updated 10 months ago
- SWE Arena☆37Jul 6, 2025Updated last year
- Astraios: Parameter-Efficient Instruction Tuning Code Language Models☆63Apr 10, 2024Updated 2 years ago
- [ICLR'25] BigCodeBench: Benchmarking Code Generation Towards AGI☆519Jan 3, 2026Updated 7 months ago
- Making code edting up to 7.7x faster using multi-layer speculation☆23Feb 20, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆12Sep 8, 2023Updated 2 years ago
- Actor-Sharer-Learner training framework for off-policy DRL algorithms☆22Dec 29, 2024Updated last year
- This repository contains the code and released models for the paper Segmenting Text and Learning Their Rewards for Improved RLHF in Langu…☆19Jan 8, 2025Updated last year
- CREATE Environment for long-horizon physics-puzzle tasks with diverse tools☆18Nov 22, 2022Updated 3 years ago
- Scaling Data-Constrained Language Models☆347Jun 28, 2025Updated last year
- Optimizing bit-level Jaccard Index and Population Counts for large-scale quantized Vector Search via Harley-Seal CSA and Lookup Tables☆22May 18, 2025Updated last year
- ☆16Mar 24, 2023Updated 3 years ago
- This is for EMNLP 2024 Paper: AppBench: Planning of Multiple APIs from Various APPs for Complex User Instruction☆16Nov 4, 2024Updated last year
- ☆21Jan 17, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆11Oct 11, 2023Updated 2 years ago
- SciFe: Scala Framework for Efficient Generation of Data Structures with Invariants☆15Mar 15, 2024Updated 2 years ago
- Official codebase for "The Generalization Gap in Offline Reinforcement Learning" accepted to ICLR 2024☆29Apr 8, 2026Updated 4 months ago
- ☆13Jul 8, 2023Updated 3 years ago
- Collect simple coverage information in memory.☆11Oct 6, 2022Updated 3 years ago
- Replication package for evaluation of code generation metrics☆17Nov 24, 2025Updated 9 months ago
- Fast and Precise On-the-fly Patch Validation for All☆10Feb 24, 2023Updated 3 years ago
- Modular Single-file Reinfocement Learning Algorithms Library☆38May 16, 2023Updated 3 years ago
- An implementation of Wang et al.'s Signed Network Embedding in Social Media in PyTorch☆12Dec 24, 2017Updated 8 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- A correct Scheme interpreter derived from the R5RS spec's formal semantics, written in Haskell.☆23Jun 4, 2026Updated 2 months ago
- ☆42Mar 26, 2025Updated last year
- OpenAI gym environments for goal-conditioned and language-conditioned reinforcement learning☆14Jan 27, 2026Updated 7 months ago
- D3PE (Deep Data-Driven Policy Evaluation) aims to evaluation a large set of candidate policies from a fixed dataset to select best ones.☆10Jun 2, 2022Updated 4 years ago
- Official implementation of "Continual Learning by Modeling Intra-Class Variation" (MOCA). [TMLR 2023]☆16Mar 3, 2023Updated 3 years ago
- Balancing the Picture: Debiasing Vision-Language Datasets with Synthetic Contrast Sets☆12May 25, 2023Updated 3 years ago
- [ACL 2023] VSTAR is a multimodal dialogue dataset with scene and topic transition information☆16Oct 27, 2024Updated last year
- [ICML 2023] Variational Curriculum Reinforcement Learning for Unsupervised Discovery of Skills☆12Jul 15, 2023Updated 3 years ago
- Implementation of the model: "Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models" in PyTorch☆29Aug 16, 2026Updated 2 weeks ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Link Prediction with Signed Latent Factors in Signed Social Networks (SIGKDD 2019)☆15Oct 24, 2021Updated 4 years ago
- A beginner-friendly repository on Deep Reinforcement Learning (RL), written in PyTorch.☆27Mar 19, 2026Updated 5 months ago
- Source code for student lectures on dependent type theory.☆12Jun 9, 2025Updated last year
- code for the paper Offline Prioritized Experience Replay☆12Jun 13, 2023Updated 3 years ago
- Accelerated replay buffers in JAX☆47Sep 17, 2022Updated 3 years ago
- InfraLinker is an open source and modern data center infrastructure assets management solution designed to make managing and tracking you…☆11Oct 22, 2025Updated 10 months ago
- ☆14Jun 21, 2016Updated 10 years ago