☆90May 21, 2025Updated last year
Alternatives and similar repositories for Hunyuan-TurboS
Users that are interested in Hunyuan-TurboS are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- C^3-Bench: The Things Real Disturbing LLM based Agent in Multi-Tasking☆38Mar 1, 2026Updated 4 months ago
- [EMNLP2022] Source code for Neural Machine Translation with Contrastive Translation Memories☆12Feb 15, 2023Updated 3 years ago
- Pixels, Patterns, but no Poetry: To See the World like Humans☆18Aug 11, 2025Updated 11 months ago
- Flash Attention in 300-500 lines of CUDA/C++☆39Aug 22, 2025Updated 10 months ago
- ☆10Nov 29, 2022Updated 3 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- Code for ACL 2023 Oral Paper: ManagerTower: Aggregating the Insights of Uni-Modal Experts for Vision-Language Representation Learning☆12Aug 23, 2025Updated 10 months ago
- VisualChatGPT☆16Mar 10, 2023Updated 3 years ago
- Source codes for paper "BlockFFN: Towards End-Side Acceleration-Friendly Mixture-of-Experts with Chunk-Level Activation Sparsity".☆19Jan 10, 2026Updated 6 months ago
- List of papers on Hallucination in LMM☆10Nov 29, 2023Updated 2 years ago
- Official Repository for Paper "BaichuanSEED: Sharing the Potential of ExtensivE Data Collection and Deduplication by Introducing a Compet…☆18Aug 28, 2024Updated last year
- [EMNLP 2022] Official implementation of Transnormer in our EMNLP 2022 paper - The Devil in Linear Transformer☆65Jul 30, 2023Updated 2 years ago
- Fast LLM Training CodeBase With dynamic strategy choosing [Deepspeed+Megatron+FlashAttention+CudaFusionKernel+Compiler];☆41Jan 4, 2024Updated 2 years ago
- The Newton-Muon optimizer☆30Jun 5, 2026Updated last month
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [SIGGRAPH ASIA 2024] Frankenstein: Generating Semantic-Compositional 3D Scenes in One Tri-Plane☆19Nov 25, 2024Updated last year
- Official github repo for "Fractional Reasoning via Latent Steering Vectors Improves Inference Time Compute"☆17Jun 30, 2025Updated last year
- ☆1,583Dec 6, 2024Updated last year
- Simple & Scalable Pretraining for Neural Architecture Research☆336Mar 31, 2026Updated 3 months ago
- openmmlab models visualization☆15Jan 22, 2023Updated 3 years ago
- Our 2nd-gen LMM☆34May 22, 2024Updated 2 years ago
- MoBA: Mixture of Block Attention for Long-Context LLMs☆2,150Apr 3, 2025Updated last year
- Triton implement of bi-directional (non-causal) linear attention☆78Mar 1, 2026Updated 4 months ago
- Fast and Modularized CFG-focused Models☆23Nov 8, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆810Jun 9, 2025Updated last year
- the pytorch implementation of SubCenterArcface and sphereface2. And i add the prove of easy_margin part of Arcface in the codes.☆12Dec 1, 2021Updated 4 years ago
- NumPy+Jax with named axes and an uncompromising attitude☆23Mar 4, 2025Updated last year
- Implementation of the model: "Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models" in PyTorch☆28Jul 13, 2026Updated last week
- PyTorch reimplementation of REALM and ORQA☆22Feb 3, 2022Updated 4 years ago
- ☆80Jun 20, 2025Updated last year
- [TOG 2024] BlockFusion: Expandable 3D Scene Generation using Latent Tri-plane Extrapolation☆16Jun 14, 2024Updated 2 years ago
- DeltaProduct is a new linear recurrent neural network architecture that uses products of generalized Householder matrices as state-transi…☆15Oct 13, 2025Updated 9 months ago
- 在线学习网站 教师端+学生端 (课件资源上传下载删除、教学团队、班级管理、学生管理、考勤、作业提交批改评分、讨论区、找回密码)☆11Feb 16, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆16Jan 24, 2018Updated 8 years ago
- ☆15Feb 24, 2023Updated 3 years ago
- HunyuanVideo-I2V: A Customizable Image-to-Video Model based on HunyuanVideo☆1,831Apr 7, 2026Updated 3 months ago
- ☆126Feb 19, 2026Updated 5 months ago
- Official code for the NeurIPS25 paper "RAT: Bridging RNN Efficiencyand Attention Accuracy in Language Modeling" (https://arxiv.org/abs/25…☆26Dec 10, 2025Updated 7 months ago
- Official code repo for paper "Great Memory, Shallow Reasoning: Limits of kNN-LMs"☆24Apr 30, 2025Updated last year
- Gradio chat interface for FastMLX☆12Sep 22, 2024Updated last year