☆126Aug 13, 2024Updated 2 years ago
Alternatives and similar repositories for Sibyl-System
Users that are interested in Sibyl-System are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- new optimizer☆20Aug 4, 2024Updated 2 years ago
- A Bilingual Role Evaluation Benchmark for Large Language Models☆44Jan 9, 2024Updated 2 years ago
- Synthetic Data Generation for Evaluation☆16Feb 21, 2025Updated last year
- Code for paper "Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System"☆73Nov 14, 2024Updated last year
- Beating the GAIA benchmark with Transformers Agents. 🚀☆153Feb 19, 2025Updated last year
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆19Mar 10, 2025Updated last year
- ☆13Mar 28, 2024Updated 2 years ago
- [ICLR 2025] Automated Design of Agentic Systems☆1,639Jan 28, 2025Updated last year
- The codebase for our EMNLP24 paper: Multimodal Self-Instruct: Synthetic Abstract Image and Visual Reasoning Instruction Using Language Mo…☆85Jan 27, 2025Updated last year
- Virtual focus group with custom personas, product details, and final analysis created with AutoGen, Ollama/Llama3, and Streamlit.☆45Jul 14, 2024Updated 2 years ago
- ☆15Apr 26, 2025Updated last year
- 🔧 Compare how Agent systems perform on several benchmarks. 📊🚀☆102Aug 4, 2025Updated last year
- Multi-Granularity LLM Debugger [ICSE2026]☆101Jul 6, 2025Updated last year
- JAX port of FLUX.1 models using flax.nnx☆23Sep 28, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [NLPCC'23] ZeroGen: Zero-shot Multimodal Controllable Text Generation with Multiple Oracles PyTorch Implementation☆14Oct 7, 2023Updated 2 years ago
- [ACL 2024] AutoAct: Automatic Agent Learning from Scratch for QA via Self-Planning☆239Jan 13, 2025Updated last year
- The official implementation of the EMNLP 2023 paper "Paraphrase Types for Generation and Detection"☆12Oct 20, 2024Updated last year
- An Efficent BPE Algorithm Faster then Hugging Face Tokenizer's Implementation☆13Sep 9, 2024Updated 2 years ago
- Official code, models, and dataset for "Evolution Fine-Tuning (EFT): Learning to Discover Across 371 Optimization Tasks"☆28Jun 30, 2026Updated 2 months ago
- Code for experiments on self-prediction as a way to measure introspection in LLMs☆17Dec 10, 2024Updated last year
- SWE-Debate: Competitive Multi-Agent Debate for Software Issue Resolution [ICSE 2026]☆34Nov 11, 2025Updated 10 months ago
- ☆27Sep 11, 2024Updated 2 years ago
- The official Github repository for paper "R^2AG: Incorporating Retrieval Information into Retrieval Augmented Generation" (EMNLP 2024 Fin…☆41Dec 6, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- The official repo for "LLoCo: Learning Long Contexts Offline"☆118Jun 15, 2024Updated 2 years ago
- ☆14Dec 15, 2025Updated 9 months ago
- This is the code of a agentic rag method with dynamic workflow.☆15Jan 22, 2026Updated 7 months ago
- ☆19Aug 7, 2024Updated 2 years ago
- The code implementation of MAGDi: Structured Distillation of Multi-Agent Interaction Graphs Improves Reasoning in Smaller Language Models…☆40Feb 5, 2024Updated 2 years ago
- ☆26Jun 25, 2021Updated 5 years ago
- Official implementation of Vector-ICL: In-context Learning with Continuous Vector Representations (ICLR 2025)☆25Jun 2, 2025Updated last year
- a bunch of rubrics I made in different format and structure for llm judge and other use cases☆16Sep 22, 2025Updated 11 months ago
- Official implementation of Self-Taught Agentic Long Context Understanding (ACL 2025).☆14Sep 22, 2025Updated 11 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Open source code of the paper: "OmniEval: An Omnidirectional and Automatic RAG Evaluation Benchmark in Financial Domain"☆86Dec 20, 2024Updated last year
- Enable Next-sentence Prediction for Large Language Models with Faster Speed, Higher Accuracy and Longer Context☆42Aug 16, 2024Updated 2 years ago
- m&ms: A Benchmark to Evaluate Tool-Use for multi-step multi-modal tasks☆46Sep 26, 2024Updated last year
- [TMLR 2024] Official implementation of "Sight Beyond Text: Multi-Modal Training Enhances LLMs in Truthfulness and Ethics"☆20Sep 15, 2023Updated 3 years ago
- Code for the paper 🌳 Tree Search for Language Model Agents☆224Jul 25, 2024Updated 2 years ago
- [EMNLP 2024] Tree of Problems: Improving structured problem solving with compositionality☆20Mar 4, 2025Updated last year
- ☆37Jun 5, 2025Updated last year