SimplifiedTransformer simplifies transformer block without affecting training. Skip connections, projection parameters, sequential sub-blocks, and normalization layers are removed. Experimental results confirm similar training speed and performance.
β15Jul 20, 2026Updated this week
Alternatives and similar repositories for SimplifiedTransformers
Users that are interested in SimplifiedTransformers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- a suite of finetuned LLMs for atomically precise function calling π§ͺβ16Updated this week
- Implementation of a Hierarchical Mamba as described in the paper: "Hierarchical State Space Models for Continuous Sequence-to-Sequence Moβ¦β16Nov 11, 2024Updated last year
- Implementation of the model "Hedgehog" from the paper: "The Hedgehog & the Porcupine: Expressive Linear Attentions with Softmax Mimicry"β16Mar 11, 2024Updated 2 years ago
- Multi-threading, Concurrency, Asynchrony, and various Execution Methods implemented in a Rust backend for bleeding edge performance.β20Nov 11, 2024Updated last year
- Implementation of the model: "(MC-ViT)" from the paper: "Memory Consolidation Enables Long-Context Video Understanding"β27Updated this week
- Simple, predictable pricing with DigitalOcean hosting β’ AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- The Swarm Ecosystemβ29Aug 1, 2024Updated last year
- Pytorch Implementation of Deepmind's SIMA: "Scaling Instructable Agents Across Many Simulated Worlds"β35Jun 17, 2024Updated 2 years ago
- This is the accompanying repository to the paper - Automatic Estimation of Singing Voice Musical Dynamicsβ16Oct 28, 2024Updated last year
- Deploy your autonomous agents to production grade environments with 99% Uptime Guarantee, Infinite Scalability, and self-healing.β54Jul 13, 2026Updated last week
- The open source implementation of the cross attention mechanism from the paper: "JOINTLY TRAINING LARGE AUTOREGRESSIVE MULTIMODAL MODELS"β37Mar 11, 2024Updated 2 years ago
- SinGlow is a part of my Singing voice synthesis system. It can extract features of sound, particularly songs and musics. Then we can use β¦β11Oct 9, 2021Updated 4 years ago
- β15Apr 16, 2026Updated 3 months ago
- β16Mar 31, 2025Updated last year
- β19Feb 2, 2023Updated 3 years ago
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β15Updated this week
- Vocal Remover using Deep Neural Networksβ21Dec 31, 2024Updated last year
- Bilingual Singing Voice Synthesisβ18Mar 25, 2024Updated 2 years ago
- [ICMR 2025] Official Repository for The Paper, Let Network Decide What to Learn: Symbolic Music Understanding Model Based on Large-scale β¦β18Aug 17, 2025Updated 11 months ago
- Implementation of Qformer from BLIP2 in Zeta Lego blocks.β51Nov 11, 2024Updated last year
- An AR+AR TTS attempt.β18Jan 13, 2025Updated last year
- Mutiband version of HIFIGANβ19Nov 6, 2020Updated 5 years ago
- β17Dec 18, 2023Updated 2 years ago
- Public female English corpus used for Project AIβ€dolβ16Dec 28, 2025Updated 6 months ago
- 1-Click AI Models by DigitalOcean Gradient β’ AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICME 2024 oral] Official Repository for The Paper, PianoBART: Symbolic Piano Music Understanding and Generating with Large-Scale Pre-Traβ¦β23Aug 17, 2025Updated 11 months ago
- Just another FastSpeech 2 but cleaner code :)β29Jun 28, 2024Updated 2 years ago
- On-device real-time RAG App built using Jina Reader, Mediapipe, Gemma 2b IT LLM.β15Apr 15, 2024Updated 2 years ago
- Singing Voice Synthesis System based on Sinsyβ23Mar 3, 2020Updated 6 years ago
- β16Dec 11, 2023Updated 2 years ago
- The implementation of Paper: Compose Yourself: Average-Velocity Flow Matching for One-Step Speech Enhancement.β16Sep 23, 2025Updated 9 months ago
- Cog wrapper for FalconsAi / nsfw_image_detectionβ19Aug 6, 2025Updated 11 months ago
- Speech-To-Text forced-alignment Speech processing Universal PERformance Benchmarkβ39May 7, 2025Updated last year
- β18Dec 18, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer β’ AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ChatGPT Clone with React! (Next.js 13, Firebase, Tailwind CSS, TypeScript, API endpoints in Next.js, ChatGPT models, Dynamic page routingβ¦β13Mar 4, 2023Updated 3 years ago
- [Pattern Recognition 2024] Semantic-Aware Frame-Event Fusion based Pattern Recognition via Large Vision-Language Models, Dong Li, Jiandonβ¦β18Jan 18, 2025Updated last year
- SPUQ: Perturbation-Based Uncertainty Quantification for Large Language Modelsβ17Jun 24, 2024Updated 2 years ago
- Music generationβ26May 2, 2024Updated 2 years ago
- Community Implementation of the paper: "Multi-Head Mixture-of-Experts" In PyTorchβ30Updated this week
- A swarm of LLM agents that will help you test, document, and productionize your code!β19Updated this week
- Modified Score-Entropy-Discrete-Diffusion to do a character level ml model and integrate with Oxenβ22Apr 26, 2024Updated 2 years ago