Code and data for the paper "Can Large Language Models Understand Real-World Complex Instructions?"(AAAI2024)
☆51Apr 19, 2024Updated 2 years ago
Alternatives and similar repositories for CELLO
Users that are interested in CELLO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Benchmarking Complex Instruction-Following with Multiple Constraints Composition (NeurIPS 2024 Datasets and Benchmarks Track)☆103Feb 20, 2025Updated last year
- [ACL 2024] FollowBench: A Multi-level Fine-grained Constraints Following Benchmark for Large Language Models☆121Jun 12, 2025Updated last year
- Resources for our ACL 2023 paper: Distilling Script Knowledge from Large Language Models for Constrained Language Planning☆36Aug 19, 2023Updated 3 years ago
- ☆18Feb 29, 2024Updated 2 years ago
- Understanding Why and How Instruction Tuning Changes Pre-trained Models☆25Mar 18, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆24Oct 14, 2024Updated last year
- CFBench: A Comprehensive Constraints-Following Benchmark for LLMs☆57Aug 26, 2024Updated 2 years ago
- ☆62Aug 22, 2024Updated 2 years ago
- Official implementation of the paper "From Complex to Simple: Enhancing Multi-Constraint Complex Instruction Following Ability of Large L…☆55Jun 24, 2024Updated 2 years ago
- ☆17Nov 3, 2024Updated last year
- Collection of papers for scalable automated alignment.☆90Oct 22, 2024Updated last year
- [ACL 2023 Findings] What In-Context Learning “Learns” In-Context: Disentangling Task Recognition and Task Learning☆21Jul 9, 2023Updated 3 years ago
- The evaluation code for MultiIF multi-turn and multi-lingual instruction following☆63Oct 29, 2024Updated last year
- ☆10Sep 13, 2022Updated 4 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Source code of "Reasons to Reject? Aligning Language Models with Judgments"☆58Feb 29, 2024Updated 2 years ago
- Based on EEG signal, the differential entropy feature is extracted, and the convolution neural network based on time domain network and a…☆12Sep 21, 2022Updated 4 years ago
- ☆11Nov 23, 2024Updated last year
- ☆34Jan 11, 2024Updated 2 years ago
- Code and models for EMNLP 2024 paper "WPO: Enhancing RLHF with Weighted Preference Optimization"☆41Sep 24, 2024Updated last year
- 北京工业大学 嵌入式系统的4个实践项目以及综合项目☆12Apr 26, 2023Updated 3 years ago
- A Recipe for Building LLM Reasoners to Solve Complex Instructions☆32Oct 9, 2025Updated 11 months ago
- GPU-accelerated PyTorch implementation of Zero-shot User Intent Detection via Capsule Neural Networks☆16Apr 16, 2019Updated 7 years ago
- [ACL 2024 Findings] CriticBench: Benchmarking LLMs for Critique-Correct Reasoning☆31Mar 5, 2024Updated 2 years ago
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- ☆338Jul 25, 2024Updated 2 years ago
- A novel jailbreak attack unveiling an overlooked attack surface inherently in the chain-of-thought reasoning trajectory of LLMs☆22Apr 3, 2026Updated 5 months ago
- StrategyLLM: Large Language Models as Strategy Generators, Executors, Optimizers, and Evaluators for Problem Solving☆22Dec 11, 2024Updated last year
- [ICLR 2024] This is the official implementation for the paper: "Beyond imitation: Leveraging fine-grained quality signals for alignment"☆11May 5, 2024Updated 2 years ago
- ☆148Jul 1, 2024Updated 2 years ago
- Official Repo of "CIBench: Evaluation of LLMs as Code Interpreter "☆15Jul 19, 2024Updated 2 years ago
- [AAAI'25] CharacterBench: Benchmarking Character Customization of Large Language Models☆23Aug 1, 2025Updated last year
- Reformatted Alignment☆112Sep 23, 2024Updated 2 years ago
- Code for "ConsistentChat: Building Skeleton-Guided Consistent Multi-Turn Dialogues for Large Language Models from Scratch", where dataset…☆17Sep 8, 2025Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- DialogueCSE: Dialogue-based Contrastive Learning of Sentence Embeddings☆19Nov 24, 2021Updated 4 years ago
- A framework for evolving and testing question-answering datasets with various models.☆26Feb 28, 2024Updated 2 years ago
- ☆15Aug 4, 2021Updated 5 years ago
- A custom line wrap layout ,support set max lines.(自定义流式布局,支持设置最大行数)☆10Apr 13, 2018Updated 8 years ago
- ☆18May 17, 2022Updated 4 years ago
- [EMNLP 2023] Plan, Verify and Switch: Integrated Reasoning with Diverse X-of-Thoughts☆28Nov 4, 2023Updated 2 years ago
- ☆98Dec 5, 2023Updated 2 years ago