Banishing LLM Hallucinations Requires Rethinking Generalization
☆277Jul 15, 2024Updated 2 years ago
Alternatives and similar repositories for Lamini-Memory-Tuning
Users that are interested in Lamini-Memory-Tuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆32Jul 8, 2024Updated 2 years ago
- The official repo for "LLoCo: Learning Long Contexts Offline"☆118Jun 15, 2024Updated 2 years ago
- Official code for "MAmmoTH2: Scaling Instructions from the Web" [NeurIPS 2024]☆146Oct 27, 2024Updated last year
- This repository contains various RAG patterns implemented from scratch☆19Dec 12, 2025Updated 9 months ago
- Truth Forest: Toward Multi-Scale Truthfulness in Large Language Models through Intervention without Tuning☆47Dec 19, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆1,033Dec 17, 2024Updated last year
- TextGrad: Automatic ''Differentiation'' via Text -- using large language models to backpropagate textual gradients. Published in Nature.☆3,737Jul 25, 2025Updated last year
- ☆27Jul 9, 2024Updated 2 years ago
- Q-GaLore: Quantized GaLore with INT4 Projection and Layer-Adaptive Low-Rank Gradients.☆206Jul 17, 2024Updated 2 years ago
- Deploy your agentic worfklows to production☆2,068Apr 6, 2026Updated 5 months ago
- ☆147Sep 11, 2026Updated last week
- Code for the arXiv preprint "The Unreasonable Effectiveness of Easy Training Data"☆48Jan 17, 2024Updated 2 years ago
- ☆132Oct 1, 2024Updated last year
- ☆731Aug 4, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Distilabel is a framework for synthetic data and AI feedback for engineers who need fast, reliable and scalable pipelines based on verifi…☆3,395Updated this week
- Creating Generative AI Apps which work☆17Apr 14, 2025Updated last year
- Robust recipes to align language models with human and AI preferences☆5,680Updated this week
- Layer-Condensed KV cache w/ 10 times larger batch size, fewer params and less computation. Dramatic speed up with better task performance…☆155Apr 7, 2025Updated last year
- A tool to assist in the interpretation of learned features in sparse autoencoders (in particular the four SAE's trained by Joseph Bloom o…☆19Oct 4, 2024Updated last year
- Accelerate your Hugging Face Transformers 7.6-9x. Native to Hugging Face and PyTorch.☆684Aug 22, 2024Updated 2 years ago
- Memory layers use a trainable key-value lookup mechanism to add extra parameters to a model without increasing FLOPs. Conceptually, spars…☆386Dec 12, 2024Updated last year
- AdalFlow: The library to build & auto-optimize LLM applications.☆4,213May 29, 2026Updated 3 months ago
- A RAG that can scale 🧑🏻💻☆11May 28, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆60Sep 13, 2024Updated 2 years ago
- [ACL 2024] Progressive LLaMA with Block Expansion.☆514May 20, 2024Updated 2 years ago
- S-LoRA: Serving Thousands of Concurrent LoRA Adapters☆1,925Jan 21, 2024Updated 2 years ago
- j1-micro (1.7B) & j1-nano (600M) are absurdly tiny but mighty reward models.☆108Jul 19, 2025Updated last year
- ☆56Jun 23, 2026Updated 2 months ago
- Mixing Language Models with Self-Verification and Meta-Verification☆118Dec 12, 2024Updated last year
- Generative Representational Instruction Tuning☆701Jun 25, 2025Updated last year
- Implementation of VisionLLaMA from the paper: "VisionLLaMA: A Unified LLaMA Interface for Vision Tasks" in PyTorch and Zeta☆15Nov 11, 2024Updated last year
- Simple Graph Memory for AI applications☆105Feb 23, 2026Updated 6 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Implementation of paper Data Engineering for Scaling Language Models to 128K Context☆499Mar 19, 2024Updated 2 years ago
- The Open Source Memory Layer For Autonomous Agents☆2,648Oct 22, 2024Updated last year
- ☆13Dec 12, 2025Updated 9 months ago
- assign color hues to a collection of text fragments based on embeddings☆20Jun 15, 2024Updated 2 years ago
- Official repo for the paper PHUDGE: Phi-3 as Scalable Judge. Evaluate your LLMs with or without custom rubric, reference answer, absolute…☆53Jul 10, 2024Updated 2 years ago
- RAGElo is a set of tools that helps you selecting the best RAG-based LLM agents by using an Elo ranker☆131Sep 10, 2026Updated last week
- A library for advanced large language model reasoning☆2,348Jun 10, 2025Updated last year