[CVPR2025] Official code for Lost in Translation Found in Context
☆24Jan 14, 2026Updated 6 months ago
Alternatives and similar repositories for LiTFiC
Users that are interested in LiTFiC are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆39Jul 9, 2025Updated last year
- [ICLR2024] Sign2GPT: Leveraging Large Language Models for Gloss-Free Sign Language Translation☆41Jun 30, 2025Updated last year
- [ICLR'25] Official Implement of "Uni-Sign: Toward Unified Sign Language Understanding at Scale"☆166Feb 11, 2026Updated 5 months ago
- PyTorch implementation of "Multi-modality Associative Bridging through Memory: Speech Sound Recollected from Face Video" (ICCV2021)☆22Apr 11, 2022Updated 4 years ago
- Official code for Metric learning for user-defined keyword spotting☆40Feb 21, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆16Jun 2, 2025Updated last year
- Official code for the paper "Scaling Multilingual Visual Speech Recognition"☆20Aug 15, 2025Updated 11 months ago
- Self-supervised video pretraining for sign language translation.☆41May 19, 2026Updated 2 months ago
- [EMNLP 2024] IFCap: Image-like Retrieval and Frequency-based Entity Filtering for Zero-shot Captioning☆15May 13, 2025Updated last year
- UCPM: Uncertainty-Guided Cross-Modal Retrieval with Partially Mismatched Pairs (TIP 2025, pytorch code)☆25Apr 16, 2026Updated 3 months ago
- ☆13Apr 12, 2026Updated 3 months ago
- A Simplied Framework of GAN Inversion☆16Mar 18, 2024Updated 2 years ago
- Efficient Training for Multilingual Visual Speech Recognition: Pre-training with Discretized Visual Speech Representation (ACM MM 2024)☆20Mar 17, 2025Updated last year
- ☆36Jan 20, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆38May 28, 2025Updated last year
- [EMNLP 2024] Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality☆24Oct 8, 2024Updated last year
- [ICASSP 2024] Official code for Slowfast Network for Continuous Sign Language Recognition☆66Jul 4, 2025Updated last year
- Personalized Lip Reading: Adapting to Your Unique Lip Movements with Vision and Language (AAAI 2025)☆24Jun 29, 2026Updated last month
- ☆28Feb 23, 2026Updated 5 months ago
- 将训练好的人脸分类器模型文件转换为.pb格式,促进工程应用。☆11Jan 1, 2020Updated 6 years ago
- A paper list of partially relevant video retrieval☆42Jul 21, 2026Updated last week
- The speaker-labeled information of LRW dataset, which is the outcome of the paper "Speaker-adaptive Lip Reading with User-dependent Paddi…☆10Oct 12, 2023Updated 2 years ago
- Official source code for the paper "Tailored Design of Audio-Visual Speech Recognition Models using Branchformers"☆15Feb 24, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Code for "Weakly-supervised Fingerspelling Recognition in British Sign Language Videos", BMVC 2022.☆12Jun 22, 2023Updated 3 years ago
- Official Implementation of the work "Audio Mamba: Bidirectional State Space Model for Audio Representation Learning"☆173Nov 24, 2024Updated last year
- Montreal Forced Aligner for Vietnamese☆15Oct 23, 2023Updated 2 years ago
- ☆17Oct 1, 2024Updated last year
- [ICASSP2025] Official code for VoiceDiT: Dual-Condition Diffusion Transformer for Environment-Aware Speech Synthesis☆52Apr 9, 2025Updated last year
- Official Implementation of Video-MA2MBA☆12Dec 3, 2024Updated last year
- Official PyTorch implementation for "Zero-AVSR: Zero-Shot Audio-Visual Speech Recognition with LLMs by Learning Language-Agnostic Speech …☆37May 11, 2025Updated last year
- ☆10Nov 27, 2024Updated last year
- 大连理工大学编译原理课程设计☆10Jan 1, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Code for Learning to Learn Language from Narrated Video☆33Oct 3, 2023Updated 2 years ago
- Pytorch implementation of "Towards Practical and Efficient Image-to-Speech Captioning with Vision-Language Pre-training and Multi-modal T…☆12Apr 29, 2026Updated 3 months ago
- [TASLP 2024] Textless Unit-to-Unit training for Many-to-Many Multilingual Speech-to-Speech Translation☆31Sep 6, 2024Updated last year
- ☆19May 6, 2024Updated 2 years ago
- Official implementation for "Signs as Tokens: A Retrieval-Enhanced Multilingual Sign Language Generator" [ICCV 2025]☆47Feb 21, 2026Updated 5 months ago
- Code for the paper "Hyperbolic Image-Text Representations", Desai et al, ICML 2023☆204Aug 23, 2023Updated 2 years ago
- Efficient Uncertainty Estimation for LiDAR-based 3D Object Detection☆10Nov 8, 2022Updated 3 years ago