[EMNLP 2025] Representation Potentials of Foundation Models for Multimodal Alignment: A Survey
☆33Feb 3, 2026Updated 6 months ago
Alternatives and similar repositories for Representation-Alignment-Survey
Users that are interested in Representation-Alignment-Survey are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025] Scale-Free Graph-Language Models☆20Feb 22, 2025Updated last year
- [NeurIPS 2023] Latent Graph Inference with Limited Supervision☆33Feb 1, 2024Updated 2 years ago
- [NeurIPS 2025] The Indra Representation Hypothesis for Multimodal Alignment☆31Feb 3, 2026Updated 6 months ago
- A curated list of resources on on-policy distillation☆25Apr 13, 2026Updated 4 months ago
- [ICLR 2026] Official code for "Ref-Adv: Exploring MLLM Visual Reasoning in Referring Expression Tasks"☆28Mar 2, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Material for the course of "Mathematics of Transformer"☆24Aug 3, 2025Updated last year
- [NeurIPS2024] Official code for (IMA) Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs☆23Oct 15, 2024Updated last year
- Tracking the latest and greatest research papers on diffusion large language models.☆32Mar 13, 2026Updated 5 months ago
- A Codebook-Driven Approach for Low-Light Image Enhancement☆30Apr 14, 2026Updated 4 months ago
- [ICDM 2022] Making Reconstruction-based Method Great Again for Video Anomaly Detection (PyTorch)☆40Mar 25, 2024Updated 2 years ago
- ThinkJEPA: Empowering Latent World Models with Large Vision-Language Reasoning Model☆53Jul 19, 2026Updated last month
- Python logging package for easy reproducible experimenting in research☆42Jul 29, 2025Updated last year
- ☆24Nov 4, 2025Updated 9 months ago
- A survey on MM-LLMs for long video understanding: From Seconds to Hours: Reviewing MultiModal Large Language Models on Comprehensive Long…☆25Sep 12, 2025Updated 11 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [CVPR 2025] DiscoVLA: Discrepancy Reduction in Vision, Language, and Alignment for Parameter-Efficient Video-Text Retrieval☆22Jun 23, 2025Updated last year
- [ICLR'23] Trainability Preserving Neural Pruning (PyTorch)☆34May 21, 2023Updated 3 years ago
- ☆12Apr 19, 2024Updated 2 years ago
- ☆26May 29, 2024Updated 2 years ago
- Official implementation of Bayes Conditional Distribution Estimation for Knowledge Distillation Based on Conditional Mutual Information☆12Sep 28, 2023Updated 2 years ago
- [IEEE TIP] Offical implementation for the work "BadCM: Invisible Backdoor Attack against Cross-Modal Learning".☆14Aug 30, 2024Updated 2 years ago
- [ICLR 2026] Official repository of "Beyond Fixed: Training-Free Variable-Length Denoising for Diffusion Large Language Models"☆174Feb 16, 2026Updated 6 months ago
- Non-linear Motion Estimation for Video Frame Interpolation using Space-time Convolutions☆20Jun 23, 2022Updated 4 years ago
- [ICLR 2025] See What You Are Told: Visual Attention Sink in Large Multimodal Models☆119Feb 16, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [TKDE 2024, CIKM 2022] SLA²P: Self-supervised Anomaly Detection with Adversarial Perturbation.☆39Dec 26, 2024Updated last year
- Source code of ICML'22 paper: FEDformer: Frequency Enhanced Decomposed Transformer for Long-term Series Forecasting☆10Jun 10, 2022Updated 4 years ago
- TETCI: Coarse-to-Fine Low-light Image Enhancement with Light Restoration and Color Refinement☆24Nov 23, 2023Updated 2 years ago
- ☆58Apr 4, 2025Updated last year
- TRISTAN: TRI's Situation and Trajectory Anticipation Networks☆14Jun 8, 2026Updated 2 months ago
- ☆15Dec 12, 2024Updated last year
- ☆42Jun 9, 2025Updated last year
- [ECCV 2024] Learning Video Context as Interleaved Multimodal Sequences☆46Mar 11, 2025Updated last year
- ☆86Nov 5, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- Code for generating adversarial color-shifted images☆20Nov 11, 2019Updated 6 years ago
- a survey on deep research☆48Sep 9, 2025Updated 11 months ago
- ☆29May 13, 2025Updated last year
- Code for MERL's ECCV 2022 paper on Cross-Modal Knowledge Transfer Without Task-Relevant Source Data☆11Jul 19, 2022Updated 4 years ago
- Envision: Benchmarking Unified Understanding & Generation for Causal World Process Insights☆32Jan 9, 2026Updated 7 months ago
- An implementation of several unsupervised object discovery models (Slot Attention, SLATE, GNM) in PyTorch with pre-trained models.☆14May 26, 2025Updated last year
- A list of all papers related to anomaly detection in NeurIPS 2020.☆10Jan 13, 2021Updated 5 years ago