An unofficial pytorch implementation of 'Efficient Infinite Context Transformers with Infini-attention'
☆56Aug 19, 2024Updated last year
Alternatives and similar repositories for infini-attention
Users that are interested in infini-attention are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementation of Infini-Transformer from "Leave No Context Behind: Efficient Infinite Context Transformers with Infini-attention…☆300May 4, 2024Updated 2 years ago
- Efficient Infinite Context Transformers with Infini-attention Pytorch Implementation + QwenMoE Implementation + Training Script + 1M cont…☆96May 9, 2024Updated 2 years ago
- ☆13Sep 12, 2024Updated last year
- Compute WER and SER for speech recognition evaluation☆26Jun 6, 2026Updated last week
- Data Benchmarking☆25May 24, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Implementation of the LDP module block in PyTorch and Zeta from the paper: "MobileVLM: A Fast, Strong and Open Vision Language Assistant …☆15Mar 11, 2024Updated 2 years ago
- code for promptCSE, emnlp 2022☆11Apr 10, 2023Updated 3 years ago
- Community Open Source Implementation of GPT4o in PyTorch☆32Jun 6, 2026Updated last week
- Implementation Code for "LLM-based Medical Assistant Personalization with Short- and Long-Term Memory Coordination"☆14May 17, 2026Updated last month
- A byte-level decoder architecture that matches the performance of tokenized Transformers.☆68Apr 24, 2024Updated 2 years ago
- Implementation of "PaLM2-VAdapter:" from the multi-modal model paper: "PaLM2-VAdapter: Progressively Aligned Language Model Makes a Stron…☆17Nov 11, 2024Updated last year
- Implementation of the paper "Variable Bitrate Residual Vector Quantization for Audio Coding"☆11Apr 10, 2025Updated last year
- This is a simple torch implementation of the high performance Multi-Query Attention☆16Aug 23, 2023Updated 2 years ago
- T5Voice is a lightweight PyTorch implementation of T5-based text-to-speech synthesis, supporting both streaming and non-streaming speech …☆28Nov 7, 2025Updated 7 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Official codebase for permutation self-consistency.☆19Feb 11, 2024Updated 2 years ago
- A codebase for data crawling and preprocessing for TTS and ASR systems training.☆23Updated this week
- EvaByte: Efficient Byte-level Language Models at Scale☆119Apr 22, 2025Updated last year
- An implementation of the base GPT-3 Model architecture from the paper by OPENAI "Language Models are Few-Shot Learners"☆22Jun 29, 2024Updated last year
- A simple and minimal open source implementation of "Introducing LFM2: The Fastest On-Device Foundation Models on the Market" from Liquid …☆26Jun 8, 2026Updated last week
- Community Repo for Nowledge Labs Products☆101Updated this week
- Code and data for the paper "Understanding Hidden Context in Preference Learning: Consequences for RLHF"☆27Aug 21, 2024Updated last year
- ☆28Feb 10, 2026Updated 4 months ago
- Tensorflow implementation of the paper "Fast Compressive Sensing Using Generative Model with Structed Latent Variables"☆10Apr 7, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Text-to-text alignment algorithm for speech recognition error analysis.☆30Apr 6, 2026Updated 2 months ago
- Rust derive macros for automating the boring stuff.☆14Aug 3, 2025Updated 10 months ago
- Recursive Self-Aggregation evals on ARC-AGI☆36Jan 26, 2026Updated 4 months ago
- Pytorch Implementation of the paper: "Learning to (Learn at Test Time): RNNs with Expressive Hidden States"☆25Jun 8, 2026Updated last week
- Implementation of the model: "Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models" in PyTorch☆28Jun 8, 2026Updated last week
- Merging Generated and Retrieved Knowledge for Open-Domain QA (EMNLP 2023)☆21Oct 8, 2023Updated 2 years ago
- Bleeding edge low level Rust binding for GGML☆17Jun 26, 2024Updated last year
- Block Transformer: Global-to-Local Language Modeling for Fast Inference (NeurIPS 2024)☆166Apr 13, 2025Updated last year
- Gemma 2B with 10M context length using Infini-attention.☆933May 12, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆15Oct 31, 2023Updated 2 years ago
- ☆11Dec 26, 2018Updated 7 years ago
- Personalized Graph-based Retrieval for LLMs Benchmark☆34Feb 16, 2025Updated last year
- Mandarin Chinese audio datasets aligned with Montreal Forced Aligner☆19Aug 13, 2024Updated last year
- Kanban board made with TailwindCSS☆11Jun 10, 2021Updated 5 years ago
- Pytorch Implementation of the Model from "MIRASOL3B: A MULTIMODAL AUTOREGRESSIVE MODEL FOR TIME-ALIGNED AND CONTEXTUAL MODALITIES"☆26Jan 27, 2025Updated last year
- Placeholder☆10Jul 17, 2023Updated 2 years ago