This project is an attempt to create a common metric to test LLM's for progress in eliminating hallucinations which is the most serious current problem in widespread adoption of LLM's for many real purposes.
☆222Apr 6, 2023Updated 3 years ago
Alternatives and similar repositories for haltt4llm
Users that are interested in haltt4llm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- RWKV is a RNN with transformer-level LLM performance. It can be directly trained like a GPT (parallelizable). So it's combining the best …☆10Nov 3, 2023Updated 2 years ago
- Codes for NAACL 2021 paper 'Noisy Self-Knowledge Distillation for Text Summarization'☆24Jul 27, 2021Updated 5 years ago
- Token-level Reference-free Hallucination Detection☆97Jul 25, 2023Updated 3 years ago
- Trying to deconstruct RWKV in understandable terms☆14May 6, 2023Updated 3 years ago
- Alpaca dataset from Stanford, cleaned and curated☆1,606Mar 7, 2026Updated 5 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Dataset, metrics, and models for TACL 2023 paper MACSUM: Controllable Summarization with Mixed Attributes.☆34Jul 25, 2023Updated 3 years ago
- This is the repository of HaluEval, a large-scale hallucination evaluation benchmark for Large Language Models.☆595Feb 12, 2024Updated 2 years ago
- Node.js implementation binding for the RWKV.cpp module☆22Aug 2, 2023Updated 3 years ago
- Github repository for "FELM: Benchmarking Factuality Evaluation of Large Language Models" (NeurIPS 2023)☆65Dec 25, 2023Updated 2 years ago
- Easily deploy your rwkv model☆19May 5, 2023Updated 3 years ago
- A collection of modular datasets generated by GPT-4, General-Instruct - Roleplay-Instruct - Code-Instruct - and Toolformer☆1,666Sep 15, 2023Updated 2 years ago
- This repository contains code for cleaning your training data of benchmark data to help combat data snooping.☆28Apr 21, 2023Updated 3 years ago
- 4 bits quantization of LLaMA using GPTQ☆3,071Jul 13, 2024Updated 2 years ago
- [ICLR24] The open-source repo of THU-KEG's KoLA benchmark.☆57Sep 28, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆15Feb 21, 2024Updated 2 years ago
- Companion repository to "Prompt Compression and Contrastive Conditioning for Controllability and Toxicity Reduction in Language Models"☆14May 31, 2023Updated 3 years ago
- Let ChatGPT teach your own chatbot in hours with a single GPU!☆3,152Mar 17, 2024Updated 2 years ago
- Code for "Tracing Knowledge in Language Models Back to the Training Data"☆40Dec 27, 2022Updated 3 years ago
- contrastive decoding☆206Nov 14, 2022Updated 3 years ago
- ☆455Oct 15, 2023Updated 2 years ago
- ☆10Dec 30, 2021Updated 4 years ago
- A simple Google Search Engine Crawler.☆22Feb 16, 2024Updated 2 years ago
- Training a reward model for RLHF using RWKV.☆15Jun 5, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 💡Light Bulb is a tool to help you label, train, test and deploy machine learning models without any coding.☆25Feb 15, 2023Updated 3 years ago
- This repository contains code to quantitatively evaluate instruction-tuned models such as Alpaca and Flan-T5 on held-out tasks.☆552Mar 10, 2024Updated 2 years ago
- Reverse Instructions to generate instruction tuning data with corpus examples☆214Mar 5, 2024Updated 2 years ago
- ☆1,514May 12, 2023Updated 3 years ago
- Code and data for "Retrieval Enhanced Model for Commonsense Generation" (ACL-IJCNLP 2021).☆29Dec 31, 2021Updated 4 years ago
- Repo for ICML23 "Why do Nearest Neighbor Language Models Work?"☆59Jan 12, 2023Updated 3 years ago
- "FiD-ICL: A Fusion-in-Decoder Approach for Efficient In-Context Learning" (ACL 2023)☆15Jul 24, 2023Updated 3 years ago
- Fine-tuning RWKV-World model☆26Jun 6, 2023Updated 3 years ago
- Instruction Tuning with GPT-4☆4,334Jun 11, 2023Updated 3 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆534Dec 1, 2023Updated 2 years ago
- A school for camelids☆1,198May 1, 2023Updated 3 years ago
- UNISUMM: Unified Few-shot Summarization with Multi-Task Pre-Training and Prefix-Tuning☆61Jun 12, 2023Updated 3 years ago
- ☆51Jun 29, 2023Updated 3 years ago
- [NAACL 2022] GlobEnc: Quantifying Global Token Attribution by Incorporating the Whole Encoder Layer in Transformers☆24May 16, 2023Updated 3 years ago
- Benchmarking large language models' complex reasoning ability with chain-of-thought prompting☆2,774Aug 4, 2024Updated 2 years ago
- [ICLR 2024] Fine-tuning LLaMA to follow Instructions within 1 Hour and 1.2M Parameters☆5,912Mar 14, 2024Updated 2 years ago