☆24Jul 1, 2024Updated 2 years ago
Alternatives and similar repositories for HaluAgent
Users that are interested in HaluAgent are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A method for evaluating the high-level coherence of machine-generated texts. Identifies high-level coherence issues in transformer-based …☆12Mar 18, 2023Updated 3 years ago
- ☆13Jan 22, 2025Updated last year
- Code repository for the paper on "Predicting the Performance of Black-Box LLMs through Self-Queries".☆12Jan 9, 2025Updated last year
- ☆13Aug 26, 2024Updated 2 years ago
- MinPrompt: Graph-based Minimal Prompt Data Augmentation for Few-shot Question Answering☆14May 3, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- Examples for the Spartan HPC cluster.☆10Sep 2, 2019Updated 7 years ago
- ☆10Nov 28, 2023Updated 2 years ago
- Code and data release of the paper Enhancing LLM Complex Problem-Solving with Hybrid Thinking and Dynamic Workflows☆15Oct 4, 2024Updated last year
- codes for "Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models"☆13Feb 10, 2025Updated last year
- explainable-machine-translation-metrics☆12Jul 15, 2022Updated 4 years ago
- [ICLR 2026] Rectifying LLM Thought From Lens of Optimization☆14Dec 5, 2025Updated 9 months ago
- An AI regulatory assistant to pre-check your documentation before FDA or MDR submission.☆14Jul 31, 2024Updated 2 years ago
- Control LLM☆23Apr 6, 2025Updated last year
- Is Neuron Coverage a Meaningful Measure for Testing Deep Neural Networks? (FSE 2020)☆10Sep 23, 2021Updated 5 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Companion code to https://arxiv.org/abs/2402.15491☆22Sep 18, 2025Updated last year
- Official code for the paper Improving Language Plasticity via Pretraining with Active Forgetting, NeurIPS 2023☆21Mar 12, 2026Updated 6 months ago
- Code for ICML2020 "Sequence Generation with Mixed Representations"☆12Jun 27, 2020Updated 6 years ago
- A dataset of pitch curves for music performance assessment☆10Jun 5, 2023Updated 3 years ago
- ☆14May 25, 2026Updated 3 months ago
- Convert pretrained RoBerta models to various long-document transformer models☆11Apr 5, 2022Updated 4 years ago
- Resources for paper "DialSummEval: Revisiting summarization evaluation for dialogues"☆14Jul 22, 2025Updated last year
- 免费的AI视频生成nonebot插件,支持文生视频和图文生视频☆10May 7, 2025Updated last year
- Accompanying code for our EMNLP 2017 publication "Bringing Structure into Summaries: Crowdsourcing a Benchmark Corpus of Concept Maps"☆13Dec 5, 2017Updated 8 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Implementation code for ACL2024:Advancing Parameter Efficiency in Fine-tuning via Representation Editing☆15Apr 20, 2024Updated 2 years ago
- Yet another coding assistant powered by LLM.☆16Sep 11, 2024Updated 2 years ago
- ☆14Apr 9, 2026Updated 5 months ago
- ☆24Jun 13, 2023Updated 3 years ago
- FrugalScore is an approach to learn a fixed, low cost version of any expensive NLG metric, while retaining most of its original performan…☆16Sep 21, 2022Updated 4 years ago
- Koishi's Day 2024 Paper (NeurIPS 2024): An advanced persona-driven role-playing system with global faithfulness quantification and optimi…☆13Oct 19, 2025Updated 11 months ago
- The Infibench variant of bigcode-evaluation-harness --- a framework for the evaluation of autoregressive code generation language models.☆14Oct 19, 2024Updated last year
- ☆12Nov 14, 2024Updated last year
- MCP as a Judge is a behavioral MCP that strengthens AI coding assistants by requiring explicit LLM evaluations☆17Dec 15, 2025Updated 9 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Outline to Story: Fine-grained Controllable Story Generation from Cascaded Events☆18Jun 16, 2022Updated 4 years ago
- Modified LLaVA framework for MOSS2, and makes MOSS2 a multimodal model.☆13Sep 19, 2024Updated 2 years ago
- The official code repo and data hub of top_nsigma sampling strategy for LLMs.☆28Feb 11, 2025Updated last year
- Open-source evaluation toolkit of large vision-language models (LVLMs), support ~100 VLMs, 30+ benchmarks☆15Feb 17, 2025Updated last year
- ☆15Apr 14, 2025Updated last year
- [preprint] Think Longer to Explore Deeper: Learn to Explore In-Context via Length-Incentivized Reinforcement Learning☆19Feb 18, 2026Updated 7 months ago
- ☆16Oct 23, 2023Updated 2 years ago