Implementation of Qformer from BLIP2 in Zeta Lego blocks.
☆51Nov 11, 2024Updated last year
Alternatives and similar repositories for qformer
Users that are interested in qformer are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- a suite of finetuned LLMs for atomically precise function calling 🧪☆16Sep 7, 2026Updated last week
- Implementation of the model "Hedgehog" from the paper: "The Hedgehog & the Porcupine: Expressive Linear Attentions with Softmax Mimicry"☆15Mar 11, 2024Updated 2 years ago
- ☆11May 9, 2023Updated 3 years ago
- The open source implementation of the cross attention mechanism from the paper: "JOINTLY TRAINING LARGE AUTOREGRESSIVE MULTIMODAL MODELS"☆37Mar 11, 2024Updated 2 years ago
- The Swarm Ecosystem☆31Aug 1, 2024Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Implementation of the model: "(MC-ViT)" from the paper: "Memory Consolidation Enables Long-Context Video Understanding"☆27Aug 28, 2026Updated 3 weeks ago
- A CUDA kernel for NHWC GroupNorm for PyTorch☆26Nov 15, 2024Updated last year
- Implementation of the LDP module block in PyTorch and Zeta from the paper: "MobileVLM: A Fast, Strong and Open Vision Language Assistant …☆15Mar 11, 2024Updated 2 years ago
- Towards a general language-audio model for computational paralinguistic tasks☆31Dec 14, 2024Updated last year
- Pytorch Implementation of Deepmind's SIMA: "Scaling Instructable Agents Across Many Simulated Worlds"☆35Jun 17, 2024Updated 2 years ago
- Implementation of SoundtStream from the paper: "SoundStream: An End-to-End Neural Audio Codec"☆13Jan 27, 2025Updated last year
- Repository of the IJCV'26 & WACV'24 paper☆35Apr 27, 2026Updated 4 months ago
- Implementation of "PaLM2-VAdapter:" from the multi-modal model paper: "PaLM2-VAdapter: Progressively Aligned Language Model Makes a Stron…☆17Nov 11, 2024Updated last year
- Implementation of the Pairformer model used in AlphaFold 3☆14Aug 28, 2026Updated 3 weeks ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official Implementation of GLAP - General Language Audio Pretraining☆76May 14, 2026Updated 4 months ago
- Implementation of the transformer from the paper: "Real-World Humanoid Locomotion with Reinforcement Learning"☆64Aug 29, 2026Updated 3 weeks ago
- Community Implementation of the paper: "Multi-Head Mixture-of-Experts" In PyTorch☆31Aug 29, 2026Updated 3 weeks ago
- Llama-Mimi is a speech language model that uses a unified tokenizer (Mimi) and a single Transformer decoder (Llama) to jointly model sequ…☆31Sep 20, 2025Updated last year
- SciKnowEval: Evaluating Multi-level Scientific Knowledge of Large Language Models☆32Jul 13, 2025Updated last year
- Deploy your autonomous agents to production grade environments with 99% Uptime Guarantee, Infinite Scalability, and self-healing.☆57Aug 12, 2026Updated last month
- Generative Expressive Conversational Speech Synthesis (Accepted by MM'2024)☆61Nov 1, 2024Updated last year
- An unofficial implementation of "UniCATS: A Unified Context-Aware Text-to-Speech Framework with Contextual VQ-Diffusion and Vocoding".☆26Nov 4, 2023Updated 2 years ago
- [NeurIPS 2023] Bootstrapping Vision-Language Learning with Decoupled Language Pre-training☆26Dec 5, 2023Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of VisionLLaMA from the paper: "VisionLLaMA: A Unified LLaMA Interface for Vision Tasks" in PyTorch and Zeta☆15Nov 11, 2024Updated last year
- ☆15Jul 19, 2026Updated 2 months ago
- Compute WER and SER for speech recognition evaluation☆28Jun 6, 2026Updated 3 months ago
- Official respository for ReasonGen-R1☆74Jun 23, 2025Updated last year
- LMM for VQA, tcsvt version☆10Jul 19, 2024Updated 2 years ago
- Simple Implementation of a Transformer in the new framework MLX by Apple☆19Nov 18, 2024Updated last year
- ☆11Apr 17, 2025Updated last year
- ALMOは拡張Markdownパーサ・静的サイトジェネレータです。WebAssemblyを使ってブラウザ上で完結する実行環境を提供し、サーバを必要としないサンプルコ ードの実行環境やジャッジシステムを提供するページの構築を可能にします。☆17Apr 14, 2026Updated 5 months ago
- JATTS: A modern, research-oriented Japanese Text-to-speech Open-sourced Toolkit☆44Mar 13, 2026Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆40Aug 26, 2025Updated last year
- A static deobfuscator for JavaScript Malware☆14May 6, 2020Updated 6 years ago
- ☆11Mar 18, 2025Updated last year
- A single-layer, streaming codec model providing SOTA audio quality and discrete tokens designed for superior downstream modelability.☆129Jun 4, 2025Updated last year
- Speech-To-Text forced-alignment Speech processing Universal PERformance Benchmark☆38May 7, 2025Updated last year
- This is the pytorch\DGL implementation of the AMIGO paper.☆10Feb 6, 2024Updated 2 years ago
- ☆16Aug 28, 2024Updated 2 years ago