[ACL 2024 (Oral)] A Prospector of Long-Dependency Data for Large Language Models
β61Jul 23, 2024Updated last year
Alternatives and similar repositories for ProLong
Users that are interested in ProLong are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- β31Sep 12, 2025Updated 10 months ago
- (ACL 2025) π₯π₯π₯Code for "Empowering Multimodal Large Language Models with Evol-Instruct"β22May 15, 2025Updated last year
- Marathon: A Multiple-choice Long Context Evaluation Benchmark for Large Language Models.β10May 16, 2024Updated 2 years ago
- [ACL 2026] Repository of IPBenchβ23Apr 6, 2026Updated 3 months ago
- LongAttn οΌSelecting Long-context Training Data via Token-level Attentionβ15Jul 16, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [COLING 2024 (Oral)] PromISe:Releasing the Capabilities of LLMs with Prompt Introspective Searchβ23Aug 26, 2024Updated last year
- β47Nov 25, 2024Updated last year
- Ruler: A Model-Agnostic Method to Control Generated Length for Large Language Modelsβ41Sep 30, 2024Updated last year
- Official repository of the AAAI'2022 paper "GALAXY: A Generative Pre-trained Model for Task-Oriented Dialog with Semi-Supervised Learningβ¦β108Jul 15, 2022Updated 4 years ago
- SWE-Flow: Synthesizing Software Engineering Data in a Test-Driven Mannerβ40Jun 29, 2025Updated last year
- [EMNLP'24] LongHeads: Multi-Head Attention is Secretly a Long Context Processorβ32Apr 8, 2024Updated 2 years ago
- β39Apr 6, 2026Updated 3 months ago
- A toolkit for modeling and simulation of cloud-native applications.β16Aug 4, 2025Updated 11 months ago
- β28Oct 28, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways β’ AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- β63Oct 29, 2024Updated last year
- This repo contains code and data for ICLR 2025 paper MIA-Bench: Towards Better Instruction Following Evaluation of Multimodal LLMsβ38Mar 9, 2025Updated last year
- The this is the official implementation of "DAPE: Data-Adaptive Positional Encoding for Length Extrapolation"β41Oct 11, 2024Updated last year
- Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervisionβ19Apr 1, 2025Updated last year
- [ACL 2025 (Findings)] DEMO: Reframing Dialogue Interaction with Fine-grained Element Modelingβ22Dec 16, 2024Updated last year
- Aioli: A unified optimization framework for language model data mixingβ33Jan 17, 2025Updated last year
- β19Oct 14, 2024Updated last year
- (NIPS 2025) OpenOmni: Official implementation of Advancing Open-Source Omnimodal Large Language Models with Progressive Multimodal Alignβ¦β142May 9, 2026Updated 2 months ago
- Efficient retrieval head analysis with triton flash attention that supports topK probabilityβ13Jun 15, 2024Updated 2 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICLR 2026] Adaptive Social Learning via Mode Policy Optimization for Language Agentsβ51Feb 2, 2026Updated 5 months ago
- [ACL 24 Findings] Implementation of Resonance RoPE and the PosGen synthetic dataset.β24Mar 5, 2024Updated 2 years ago
- [EMNLP 2024 (Oral)] Leave No Document Behind: Benchmarking Long-Context LLMs with Extended Multi-Doc QAβ155Dec 22, 2025Updated 7 months ago
- π° Must-read papers on KV Cache Compression (constantly updating π€).β726Apr 15, 2026Updated 3 months ago
- Codebase for Instruction Following without Instruction Tuningβ36Sep 24, 2024Updated last year
- LongRecipe: Recipe for Efficient Long Context Generalization in Large Language Modelsβ79Oct 16, 2024Updated last year
- A Chinese-focused PyTorch framework for exploring Attention Residuals in Qwen3-style causal LMs, with baseline, Block AttnRes, Full AttnRβ¦β19May 3, 2026Updated 2 months ago
- Official implementation of GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agentsβ252May 5, 2025Updated last year
- Official repository of the EMNLP'2020 paper "Amalgamating Knowledge from Two Teachers for Task-oriented Dialogue System with Adversarial β¦β16Dec 9, 2021Updated 4 years ago
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- β11Aug 20, 2025Updated 11 months ago
- open-source code for paper: Retrieval Head Mechanistically Explains Long-Context Factualityβ241Aug 2, 2024Updated last year
- β76Jun 10, 2025Updated last year
- [ACL'24 Oral] Analysing The Impact of Sequence Composition on Language Model Pre-Trainingβ24Aug 18, 2024Updated last year
- [NeurIPS 2024] Fast Best-of-N Decoding via Speculative Rejectionβ56Oct 29, 2024Updated last year
- CoT-Valve: Length-Compressible Chain-of-Thought Tuningβ91Feb 14, 2025Updated last year
- [arxiv: 2604.14142] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Spaceβ17Apr 16, 2026Updated 3 months ago