[ICML 2025] Repository for M3-JEPA: Multimodal Alignment via Multi-gate MoE based on the Joint-Predictive Embedding Architecture
☆36Mar 13, 2026Updated 6 months ago
Alternatives and similar repositories for M3-JEPA
Users that are interested in M3-JEPA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This repo contains the code for the paper "Intuitive physics understanding emerges fromself-supervised pretraining on natural videos"☆271Jun 3, 2026Updated 4 months ago
- A framework for training, evaluating, and comparing state-of-the-art zero-shot reinforcement learning methods. It allows reproducing the …☆70Dec 22, 2025Updated 9 months ago
- Assignments from 16-825 Learning for 3D Vision at Carnegie Mellon University☆13Apr 5, 2023Updated 3 years ago
- First neural GPT aligned with text and speech. Welcome to join us to make better foundation model in neural modality.☆14Oct 30, 2024Updated last year
- ☆16Sep 25, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Code of ICLR 2025 paper "DynaPrompt: Dynamic Test-Time Prompt Tuning"☆22Jan 29, 2025Updated last year
- 📷 Python package and CLI utility to create photo mosaics - now with GPU support☆20Mar 6, 2026Updated 7 months ago
- Efficiently discovering algorithms via LLMs with evolutionary search and reinforcement learning.☆17Apr 22, 2025Updated last year
- Official codebase for "Gazeformer: Scalable, Effective and Fast Prediction of Goal-Directed Human Attention" (CVPR 2023)☆47Jul 18, 2026Updated 2 months ago
- Official Code Repository for the paper "Continuous Diffusion Model for Language Modeling" (NeurIPS 2025).☆75Sep 25, 2025Updated last year
- Mixture of Cognitive Reasoners: Modular Reasoning with Brain-Like Specialization☆46Feb 7, 2026Updated 8 months ago
- ☆10Nov 1, 2024Updated last year
- TiNO-Edit: Timestep and Noise Optimization for Robust Diffusion-Based Image Editing (CVPR 2024)☆44Sep 12, 2025Updated last year
- Official code repository for Med-TTT.☆19Jun 30, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- [EMNLP 2025 Findings] 3D-Aware Vision-Language Models Fine-Tuning with Geometric Distillation☆40Jun 12, 2025Updated last year
- The Official PyTorch Implementation of OTSeg: Multi-prompt Sinkhorn Attention for Zero-Shot Semantic Segmentation☆35Jul 6, 2024Updated 2 years ago
- Reinforcing LLM Reasoning through Self-Training and Value-Guided Decoding☆19May 6, 2026Updated 5 months ago
- iLLaVA: An Image is Worth Fewer Than 1/3 Input Tokens in Large Multimodal Models (ICLR2026)☆23Jun 24, 2026Updated 3 months ago
- Mariana Pro colorscheme from Sublime Text ported to Vim☆10Aug 2, 2025Updated last year
- Official implementation for "MOCHA: Real-Time Motion Characterization via Context Matching" [SIGGRAPH Asia 2023]☆25May 5, 2026Updated 5 months ago
- 🤖 AI Assistant fine-tuned to provide support for coding and design questions based on the latest trends in the industry.☆17Jan 14, 2024Updated 2 years ago
- Python script to obtain dynamic functional connectivity metrics, after using a sliding window approach, statistical analyses to test for …☆12Sep 10, 2024Updated 2 years ago
- Code for "General-Purpose Brain Foundation Models for Time-Series Neuroimaging Data"☆15Dec 14, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆17Jul 19, 2026Updated 2 months ago
- Text-conditioned image-to-video generation based on diffusion models.☆55Jun 13, 2024Updated 2 years ago
- [ICML 2025] Graph4MM: Weaving Multimodal Learning with Structural Information☆28Jul 9, 2025Updated last year
- We introduce VLM-Mamba, the first Vision-Language Model built entirely on State Space Models (SSMs), specifically leveraging the Mamba ar…☆16Jan 5, 2026Updated 9 months ago
- Python library to compute functional connectivity measures from EEG☆12Oct 14, 2023Updated 2 years ago
- 🔥 [ICLR 2025] Official PyTorch Model "Visual Haystacks: A Vision-Centric Needle-In-A-Haystack Benchmark"☆26Feb 9, 2025Updated last year
- Auditing agents for fine-tuning safety☆22Oct 21, 2025Updated 11 months ago
- go + tauri + vue.js 做的一款通信软件,基于websocket实现即时聊天☆11Aug 8, 2023Updated 3 years ago
- Embedding and readout for simple multi-categorical and gaussian continuous☆20Jul 5, 2026Updated 3 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆155Aug 27, 2025Updated last year
- Towards Robust and Relible Multimodal Misinformation Recognition with Incomplete Modality☆15May 11, 2026Updated 4 months ago
- ☆21Sep 20, 2025Updated last year
- This is the Pytorch version of the VisualBack prop algorithm on the VGG16 model. Theoretical description of the algorithm can be found fo…☆11Feb 7, 2019Updated 7 years ago
- This is an official implementation for "Learning a Cross-Modality Anomaly Detector for Remote Sensing Imagery“ (TIP 2024))☆19Jul 24, 2025Updated last year
- Official Repo for Tuning-Free Noise Rectification for High Fidelity Image-to-Video Generation☆30Mar 29, 2024Updated 2 years ago
- Tiny LLM chatbot powered by Qwen2.5:0.5B-Instruct. It can run locally on a laptop. Depoloyment on Google Cloud is also supported.☆22Apr 1, 2025Updated last year