SAIL: Search Augmented Instruction Learning
☆160Jul 22, 2025Updated last year
Alternatives and similar repositories for SAIL
Users that are interested in SAIL are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- LangCode - Improving alignment and reasoning of large language models (LLMs) with natural language embedded program (NLEP).☆51Sep 22, 2023Updated 2 years ago
- Leveraging passage embeddings for efficient listwise reranking with large language models.☆51Dec 7, 2024Updated last year
- Synthetic QA generation for long documents.☆16Jul 22, 2022Updated 4 years ago
- Code repository for supporting the paper "Atlas Few-shot Learning with Retrieval Augmented Language Models",(https//arxiv.org/abs/2208.03…☆562Jul 2, 2026Updated 2 months ago
- DialOp: Decision-oriented dialogue environments for collaborative language agents☆114Nov 15, 2024Updated last year
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- SILO Language Models code repository☆83Feb 23, 2024Updated 2 years ago
- Implementation of the paper: "Making Retrieval-Augmented Language Models Robust to Irrelevant Context"☆76Aug 6, 2024Updated 2 years ago
- Code for Arxiv 2023: Improving Language Model Negociation with Self-Play and In-Context Learning from AI Feedback☆210May 24, 2023Updated 3 years ago
- Entailment self-training☆27May 30, 2023Updated 3 years ago
- This repository contains the code for our paper "Augmenting Black-box LLMs with Medical Textbooks for Clinical Question Answering" [EMNLP…☆14Oct 8, 2024Updated last year
- Code for the ICML 2025 paper "SelfCite Self-Supervised Alignment for Context Attribution in Large Language Models"☆30Mar 12, 2026Updated 5 months ago
- ☆295Dec 20, 2023Updated 2 years ago
- ☆13Jun 4, 2024Updated 2 years ago
- Salesforce open-source LLMs with 8k sequence length.☆726Jun 2, 2026Updated 3 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Synthesizing realistic and diverse text-datasets from augmented LLMs☆19Apr 4, 2026Updated 5 months ago
- BESA is a differentiable weight pruning technique for large language models.☆17Mar 4, 2024Updated 2 years ago
- [Preprint] Learning to Filter Context for Retrieval-Augmented Generaton☆199Apr 6, 2024Updated 2 years ago
- ☆25Dec 13, 2024Updated last year
- DSIR large-scale data selection framework for language model training☆276Apr 7, 2024Updated 2 years ago
- [EMNLP 2023] Enabling Large Language Models to Generate Text with Citations. Paper: https://arxiv.org/abs/2305.14627☆528Oct 9, 2024Updated last year
- LaMini-LM: A Diverse Herd of Distilled Models from Large-Scale Instructions☆824May 6, 2023Updated 3 years ago
- Reverse Instructions to generate instruction tuning data with corpus examples☆214Mar 5, 2024Updated 2 years ago
- Code for NAACL 2025 paper "AdaCAD: Adaptively Decoding to Balance Conflicts between Contextual and Parametric Knowledge"☆17Mar 2, 2026Updated 6 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Grade-School Math with Irrelevant Context (GSM-IC) benchmark is an arithmetic reasoning dataset built upon GSM8K, by adding irrelevant se…☆67Feb 13, 2023Updated 3 years ago
- Fusion-in-Decoder☆595Oct 4, 2023Updated 2 years ago
- An integration of Qdrant ANN vector database backend with Haystack☆46Updated this week
- 🦅🔗 Building FlyteGPT on Flyte with LangChain☆30Jan 23, 2024Updated 2 years ago
- [EMNLP 2022] Training Language Models with Memory Augmentation https://arxiv.org/abs/2205.12674☆192Jun 14, 2023Updated 3 years ago
- Codes for ICML 2023 Learning Dynamic Query Combinations for Transformer-based Object Detection and Segmentation☆38Sep 12, 2023Updated 2 years ago
- Benchmarks for Business Document Foundation Models☆10Apr 4, 2024Updated 2 years ago
- HyPe: Better Pre-trained Language Model Fine-tuning with Hidden Representation Perturbation [ACL 2023]☆14Jul 11, 2023Updated 3 years ago
- ModuleFormer is a MoE-based architecture that includes two different types of experts: stick-breaking attention heads and feedforward exp…☆225Sep 18, 2025Updated 11 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 🍀 Official pytorch implementation of "D2ADA: Dynamic Density-aware Active Domain Adaptation for Semantic Segmentation. Wu et al. ECCV 20…☆25Feb 2, 2023Updated 3 years ago
- GisPy: A Tool for Measuring Gist Inference Score in Text https://aclanthology.org/2022.wnu-1.5/☆13Jul 1, 2024Updated 2 years ago
- A dataset of LLM-generated chain-of-thought steps annotated with mistake location.☆90Aug 10, 2024Updated 2 years ago
- ☆46Apr 19, 2024Updated 2 years ago
- Is In-Context Learning Sufficient for Instruction Following in LLMs? [ICLR 2025]☆34Jan 23, 2025Updated last year
- [ACL'25] Mosaic-IT: Cost-Free Compositional Data Synthesis for Instruction Tuning☆20Sep 27, 2025Updated 11 months ago
- Test-time-training on nearest neighbors for large language models☆50Apr 18, 2024Updated 2 years ago