MaineCoon: Pursuing a Real-Time Audio-Visual Social World Model — technical report & project links. 🌐 https://mainecoon.tech/
☆121Jun 22, 2026Updated 3 months ago
Alternatives and similar repositories for MaineCoon
Users that are interested in MaineCoon are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [TPAMI] The official implementation of our paper "Improved and Accelerated Text-to-Image Generation with Collect, Reflect, and Refine".☆30Mar 8, 2026Updated 6 months ago
- [ICLR2025] IV-Mixed Sampler: Leveraging Image Diffusion Models for Enhanced Video Synthesis☆39Feb 17, 2025Updated last year
- Bag of Design Choices for Inference of High-Resolution Masked Generative Transformer☆16Nov 21, 2024Updated last year
- The official code of "Mano: Restriking Manifold Optimization for LLM Training".☆25Jun 1, 2026Updated 4 months ago
- [ICLR2026] The official code of "Weak-to-Strong Diffusion with Reflection".☆59Jan 28, 2026Updated 8 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [ICML 2026] The official code for our work "LIVEditor-14B: Lightning Unified Video Editor via In-Context Sparse Attention".☆40May 15, 2026Updated 4 months ago
- Official Pytorch implementation of AvatarForcing: One-Step Streaming Talking Avatars via Local-Future Sliding-Window Denoising☆81May 9, 2026Updated 4 months ago
- [ICLR2025] The code of Z-Sampling, proposed in our paper "Zigzag Diffusion Sampling: Diffusion Models Can Self-Improve via Self-Reflectio…☆104May 20, 2026Updated 4 months ago
- Official Implementation of SwitchCraft: Training-Free Multi-Event Video Generation with Attention Controls [CVPR 2026]☆23Mar 2, 2026Updated 7 months ago
- Code for "OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation"☆157Jun 18, 2026Updated 3 months ago
- Official implementation of "TransNormal: Dense Visual Semantics for Diffusion-based Transparent Object Normal Estimation" (ICML 2026). Si…☆23Sep 9, 2026Updated 3 weeks ago
- [NeurIPS 2026] Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers☆44Jul 1, 2026Updated 3 months ago
- [ICML 2026] Official codebase for "Causal Forcing: Autoregressive Diffusion Distillation Done Right for High-Quality Real-Time Interactiv…☆985Sep 24, 2026Updated last week
- Modular and Automated Scene Generation for Policy Learning and Evaluation☆404Aug 27, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- official repository of GeoDiff4D☆24Mar 6, 2026Updated 6 months ago
- The official code of Yume☆683Jan 14, 2026Updated 8 months ago
- Long Video Gen Infrastructure☆2,649Updated this week
- A 3B-active-parameter native unified multimodal model for image and video understanding, generation, and editing.☆1,349Jul 14, 2026Updated 2 months ago
- Official codebase for "One-Forcing: Towards Stable One-Step Autoregressive Video Generation"☆88Sep 24, 2026Updated last week
- UE5-based Data Engine used in OmniX☆115Jul 12, 2026Updated 2 months ago
- DreamX-World: A General-Purpose Interactive World Model☆778Jul 23, 2026Updated 2 months ago
- Helios: Real Real-Time Long Video Generation Model☆2,182Aug 24, 2026Updated last month
- [ECCV 2026] LiveEdit: Towards Real-Time Diffusion-Based Streaming Video Editing☆180Aug 10, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ICML 2026 Spotlight] RelaxFlow: Text-Driven Amodal 3D Generation☆24May 21, 2026Updated 4 months ago
- 猜测妙鸭实现证件照生成的pipeline☆14Sep 5, 2023Updated 3 years ago
- A Minimal and Elegant Framework & Tutorial for Real-Time Interactive World Models☆846Sep 10, 2026Updated 3 weeks ago
- [NeurIPS 2026] Official Code of NAVA: Native Audio-Visual Alignment for Generation.☆226Jun 30, 2026Updated 3 months ago
- [ICLR'26] Topology-Preserved Auto-regressive Mesh Generation in the Manner of Weaving Silk☆120Mar 2, 2026Updated 7 months ago
- codes for RFSR: Improving ISR Diffusion Models via Reward Feedback Learning☆18Dec 8, 2024Updated last year
- [ECCV 2026 Oral] Official implementation of "OmniForcing: Unleashing Real-time Joint Audio-Visual Generation"[arXiv:2603.11647]. OmniForc…☆197Jul 23, 2026Updated 2 months ago
- Flow Map OPD for AnyStep Video Diffusion☆435Aug 14, 2026Updated last month
- Official repo for paper "Echo-Infinity: Learnable Evolving Memory for Real-Time Infinite Video Generation"☆114Jun 4, 2026Updated 3 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ECCV 2026] Official implementation of "MemRoPE: Training-Free Infinite Video Generation via Evolving Memory Tokens"☆59Jun 24, 2026Updated 3 months ago
- [ICML 2026] World-R1: Reinforcing 3D Constraints for Text-to-Video Generation☆427Jun 3, 2026Updated 4 months ago
- (CVPR 2026) Sampling Algorithm for paper "Ani3DHuman: Photorealistic 3D Human Animation with Self-guided Stochastic Sampling"☆24Jun 3, 2026Updated 3 months ago
- [ICML26] AVGen-Bench is a task-driven benchmark for multi-granular evaluation of Text-to-Audio-Video (T2AV) generation.☆31Jul 2, 2026Updated 3 months ago
- Official Repo for Self-Forcing++ High Quality Long Video Generation☆272Oct 13, 2025Updated 11 months ago
- [3DV 2026] FastMesh: Efficient Artistic Mesh Generation via Component Decoupling☆141Nov 11, 2025Updated 10 months ago
- ☆2,116Apr 11, 2026Updated 5 months ago