Adaptation of titans-pytorch to llama models on HF
☆25Mar 6, 2025Updated last year
Alternatives and similar repositories for llama-titans
Users that are interested in llama-titans are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official repo of paper LM2☆49Feb 13, 2025Updated last year
- ☆13Apr 15, 2024Updated 2 years ago
- AI Based "Happiness Optimizer"☆12Oct 20, 2024Updated last year
- Official repository for distributing ECG-Reasoning-Benchmark dataset☆18Jul 30, 2026Updated last month
- A research-oriented training and evaluation framework for ECG-Language Models (ELMs)☆18Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Efficient PScan implementation in PyTorch☆17Jan 2, 2024Updated 2 years ago
- Python implementation of the methods in Meulemans et al. 2020 - A Theoretical Framework For Target Propagation☆31Oct 31, 2024Updated last year
- 😊 TPTT: Transforming Pretrained Transformers into Titans☆65Jun 7, 2026Updated 3 months ago
- [NeurIPS 2024] Implementation of paper - D-LLM: A Token Adaptive Computing Resource Allocation Strategy for Large Language Models☆25Apr 9, 2025Updated last year
- Integrates Imbue's Cost Aware pareto-Region Bayesian Search (CARBS) with Weights and Biases (WanDB)☆12Mar 17, 2025Updated last year
- Official Code Repository for the paper "Key-value memory in the brain"☆34Feb 25, 2025Updated last year
- Code implementation for paper "On the Efficacy of Small Self-Supervised Contrastive Models without Distillation Signals".☆17Dec 15, 2021Updated 4 years ago
- ☆14Oct 30, 2024Updated last year
- A large database of artificial neural network statistics during training☆15Dec 8, 2020Updated 5 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- This is a simple automated license plate detector developed in C++ via OpenCV.☆11Sep 26, 2020Updated 5 years ago
- Official PyTorch implementation of "Hierarchical Entity-centric Reinforcement Learning with Factored Subgoal Diffusion", Haramati et al.,…☆16Feb 23, 2026Updated 6 months ago
- HGRN2: Gated Linear RNNs with State Expansion☆59Aug 20, 2024Updated 2 years ago
- Sample data associated with the Aurora-BP study☆44Mar 18, 2026Updated 5 months ago
- ☆11Oct 11, 2023Updated 2 years ago
- ☆54Jul 18, 2024Updated 2 years ago
- The repository for HKU ENGG1340 Group Project (24/25 Semester 2).☆11Jun 22, 2025Updated last year
- [ICML 2025] Code for "R2-T2: Re-Routing in Test-Time for Multimodal Mixture-of-Experts"☆19Mar 10, 2025Updated last year
- 🍔 Chen’s Private Cuisine Menu☆10Jan 4, 2026Updated 8 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆116Mar 12, 2024Updated 2 years ago
- Official repo of dataset-decomposition paper [NeurIPS 2024]☆21Jan 8, 2025Updated last year
- A terminal text editor written in MoonBit☆10Apr 7, 2025Updated last year
- Implementation and explorations into Blackbox Gradient Sensing (BGS), an evolutionary strategies approach proposed in a Google Deepmind p…☆20Apr 17, 2026Updated 4 months ago
- Multi-agent Reinforcement Learning game using Advantage Actor Critic (A2C) algorithm☆14Sep 26, 2023Updated 2 years ago
- [ICLR 2025] "Training LMs on Synthetic Edit Sequences Improves Code Synthesis" (Piterbarg, Pinto, Fergus)☆19Feb 11, 2025Updated last year
- ☆12Mar 20, 2025Updated last year
- A sleek, customizable interface for managing LLMs with responsive design and easy agent personalization.☆19Aug 30, 2024Updated 2 years ago
- Official implementation of "NinA: Normalizing Flows in Action. Training VLA Models with Normalizing Flows"☆18Sep 22, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Accompanying repo for the paper - High-speed Autonomous Racing using Trajectory-aided Deep Reinforcement Learning☆18Jan 17, 2024Updated 2 years ago
- Code from CS152 lectures☆14Apr 13, 2026Updated 4 months ago
- [ICLR 2023] "Sparse MoE as the New Dropout: Scaling Dense and Self-Slimmable Transformers" by Tianlong Chen*, Zhenyu Zhang*, Ajay Jaiswal…☆56Feb 28, 2023Updated 3 years ago
- ROSA+: RWKV's ROSA implementation with fallback statistical predictor☆36Oct 13, 2025Updated 10 months ago
- BrainProp: How the brain can implement reward-based error backpropagation☆17Dec 8, 2022Updated 3 years ago
- [ICLR'25] ApolloMoE: Efficiently Democratizing Medical LLMs for 50 Languages via a Mixture of Language Family Experts☆53Nov 20, 2024Updated last year
- An upgraded version of Kantumruy.☆13Jan 13, 2024Updated 2 years ago