Banishing LLM Hallucinations Requires Rethinking Generalization
☆277Jul 15, 2024Updated 2 years ago
Alternatives and similar repositories for Lamini-Memory-Tuning
Users that are interested in Lamini-Memory-Tuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆32Jul 8, 2024Updated 2 years ago
- The official repo for "LLoCo: Learning Long Contexts Offline"☆118Jun 15, 2024Updated 2 years ago
- Official code for "MAmmoTH2: Scaling Instructions from the Web" [NeurIPS 2024]☆146Oct 27, 2024Updated last year
- This repository contains various RAG patterns implemented from scratch☆19Dec 12, 2025Updated 8 months ago
- Truth Forest: Toward Multi-Scale Truthfulness in Large Language Models through Intervention without Tuning☆47Dec 19, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆1,033Dec 17, 2024Updated last year
- TextGrad: Automatic ''Differentiation'' via Text -- using large language models to backpropagate textual gradients. Published in Nature.☆3,707Jul 25, 2025Updated last year
- ☆27Jul 9, 2024Updated 2 years ago
- Q-GaLore: Quantized GaLore with INT4 Projection and Layer-Adaptive Low-Rank Gradients.☆206Jul 17, 2024Updated 2 years ago
- Deploy your agentic worfklows to production☆2,068Apr 6, 2026Updated 4 months ago
- ☆147Jul 19, 2024Updated 2 years ago
- Code for the arXiv preprint "The Unreasonable Effectiveness of Easy Training Data"☆48Jan 17, 2024Updated 2 years ago
- ☆131Oct 1, 2024Updated last year
- ☆729Aug 4, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Distilabel is a framework for synthetic data and AI feedback for engineers who need fast, reliable and scalable pipelines based on verifi…☆3,383Updated this week
- Creating Generative AI Apps which work☆17Apr 14, 2025Updated last year
- Robust recipes to align language models with human and AI preferences☆5,672May 26, 2026Updated 3 months ago
- Layer-Condensed KV cache w/ 10 times larger batch size, fewer params and less computation. Dramatic speed up with better task performance…☆156Apr 7, 2025Updated last year
- A tool to assist in the interpretation of learned features in sparse autoencoders (in particular the four SAE's trained by Joseph Bloom o…☆19Oct 4, 2024Updated last year
- Accelerate your Hugging Face Transformers 7.6-9x. Native to Hugging Face and PyTorch.☆684Aug 22, 2024Updated 2 years ago
- Memory layers use a trainable key-value lookup mechanism to add extra parameters to a model without increasing FLOPs. Conceptually, spars…☆379Dec 12, 2024Updated last year
- AdalFlow: The library to build & auto-optimize LLM applications.☆4,212May 29, 2026Updated 3 months ago
- A RAG that can scale 🧑🏻💻☆11May 28, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ACL 2024] Progressive LLaMA with Block Expansion.☆513May 20, 2024Updated 2 years ago
- S-LoRA: Serving Thousands of Concurrent LoRA Adapters☆1,923Jan 21, 2024Updated 2 years ago
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.☆106Jul 19, 2025Updated last year
- ☆56Jun 23, 2026Updated 2 months ago
- Mixing Language Models with Self-Verification and Meta-Verification☆118Dec 12, 2024Updated last year
- Generative Representational Instruction Tuning☆699Jun 25, 2025Updated last year
- Implementation of VisionLLaMA from the paper: "VisionLLaMA: A Unified LLaMA Interface for Vision Tasks" in PyTorch and Zeta☆15Nov 11, 2024Updated last year
- Simple Graph Memory for AI applications☆105Feb 23, 2026Updated 6 months ago
- Implementation of paper Data Engineering for Scaling Language Models to 128K Context☆500Mar 19, 2024Updated 2 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- The Open Source Memory Layer For Autonomous Agents☆2,641Oct 22, 2024Updated last year
- assign color hues to a collection of text fragments based on embeddings☆20Jun 15, 2024Updated 2 years ago
- Official repo for the paper PHUDGE: Phi-3 as Scalable Judge. Evaluate your LLMs with or without custom rubric, reference answer, absolute…☆53Jul 10, 2024Updated 2 years ago
- RAGElo is a set of tools that helps you selecting the best RAG-based LLM agents by using an Elo ranker☆130Aug 12, 2026Updated 2 weeks ago
- A library for advanced large language model reasoning☆2,342Jun 10, 2025Updated last year
- [ICLR 2024] Efficient Streaming Language Models with Attention Sinks☆7,268Jul 11, 2024Updated 2 years ago
- look how they massacred my boy☆63Oct 16, 2024Updated last year