Official Implementation for paper "Pretraining A Large Language Model using Distributed GPUs: A Memory-Efficient Decentralized Paradigm"
☆23May 8, 2026Updated 3 months ago
Alternatives and similar repositories for SPES
Users that are interested in SPES are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP 2025 Oral] Official codebase for Seeing More, Saying More: Lightweight Language Experts are Dynamic Video Token Compressors.☆18Sep 7, 2025Updated 11 months ago
- Official Repo for the VideoVerse☆15Mar 29, 2026Updated 5 months ago
- (ECCV2026) Dual Distribution Estimation for Zero-shot Noisy Test-Time Adaptation with VLMs☆16Aug 1, 2026Updated last month
- [ECCV 2026] Official code repository for "Self-transcendence: Is External Feature Guidance Indispensable for Accelerating Diffusion Trans…☆38Aug 25, 2026Updated last week
- Weighted Reverse Convolution for Feature Upsampling☆24May 24, 2026Updated 3 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICML 2026] Official PyTorch implementation of paper “CoCoEdit: Content-Consistent Image Editing via Region Regularized Reinforcement Lea…☆26Jun 14, 2026Updated 2 months ago
- DepthMaster: Unified Monocular Depth Estimation for Perspective and Panoramic Images☆26Jun 13, 2026Updated 2 months ago
- A simple and effective feature extractor for untrimmed videos☆13Sep 1, 2022Updated 4 years ago
- Photo3D: Advancing Photorealistic 3D Generation through Structure‑Aligned Detail Enhancement☆22Mar 18, 2026Updated 5 months ago
- [ICML 2026] The offical code of Diversity-Preserved Distribution Matching Distillation for Fast Visual Synthesis☆91Jun 2, 2026Updated 3 months ago
- GGT-100K: Generative Ground Truth for Generalizable Real-World Image Restoration☆71Jun 1, 2026Updated 3 months ago
- LongVALE: Vision-Audio-Language-Event Benchmark Towards Time-Aware Omni-Modal Perception of Long Videos. (CVPR 2025))☆62Jun 9, 2025Updated last year
- ECCV2024, LAPT: Label-driven Automated Prompt Tuning for OOD Detection with Vision-Language Models☆18Aug 9, 2024Updated 2 years ago
- [ECCV2024] Reflective Instruction Tuning: Mitigating Hallucinations in Large Vision-Language Models☆20Jul 17, 2024Updated 2 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆32Apr 29, 2026Updated 4 months ago
- ☆12Jul 18, 2024Updated 2 years ago
- Official codebase for LiveVLN: Breaking the Stop-and-Go Loop in Vision-Language Navigation☆23Apr 22, 2026Updated 4 months ago
- Official repository for the paper "MICo-150K: A Comprehensive Dataset for Multi-Image Composition".☆104Apr 21, 2026Updated 4 months ago
- ☆43May 9, 2026Updated 3 months ago
- [NeurlPS' 25] InstructRestore: Region-Customized Image Restoration with Human Instructions☆53Oct 23, 2025Updated 10 months ago
- TokLIP: Marry Visual Tokens to CLIP for Multimodal Comprehension and Generation☆236Aug 18, 2025Updated last year
- [CVPR 2026 Highlight] VideoITG: Multimodal Video Understanding with Instructed Temporal Grounding☆128Apr 17, 2026Updated 4 months ago
- [CVPR2026] BinaryAttention: One-Bit QK-Attention for Vision and Diffusion Transformers☆43Mar 17, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- (CVPR2026 Oral) ANTS: Adaptive Negative Textual Space Shaping for OOD Detection via Test-Time MLLM Understanding and Reasoning☆93Aug 19, 2026Updated 2 weeks ago
- [CVPR 2026] TimeLens: Rethinking Video Temporal Grounding with Multimodal LLMs☆174Jul 21, 2026Updated last month
- [ICLR 2026] Many-for-Many: Unify the Training of Multiple Video and Image Generation and Manipulation Tasks☆32Feb 5, 2026Updated 6 months ago
- [ECCV'24] A novel weakly supervised framework for 3D object detection from 2D bounding boxes. It can easily extend to novel scenarios and…☆36Jul 26, 2024Updated 2 years ago
- Official repository for the paper "TIIF-Bench: How Does Your T2I Model Follow Your Instructions?".☆120Jun 26, 2026Updated 2 months ago
- CheXOne: A Reasoning-Enabled Vision–Language Foundation Model for Chest X-ray Interpretation☆42Apr 12, 2026Updated 4 months ago
- Official PyTorch implementation of paper “InsViE-1M: Effective Instruction-based Video Editing with Elaborate Dataset Construction”☆35Apr 3, 2026Updated 5 months ago
- Official codes for Polyline Path Masked Attention for Vision Transformer☆17Jun 3, 2026Updated 3 months ago
- Use Hermes Agent as the control plane for local coding agents like Codex, Kimi Code, Claude Code, OpenCode, and Gemini CLI.☆36Jul 22, 2026Updated last month
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Winner solution to Generic Event Boundary Captioning task in LOVEU Challenge (CVPR 2023 workshop)☆29Jan 1, 2024Updated 2 years ago
- Official code for our CVPR 2025 paper: "Toward Generalized Image Quality Assessment: Relaxing the Perfect Reference Quality Assumption"☆68Sep 15, 2025Updated 11 months ago
- Official code for our Paper "SSL: A Self-similarity Loss for Improving Generative Image Super-resolution" in ACMMM 2024☆51Jun 6, 2026Updated 2 months ago
- Transformer: PyTorch Implementation of "Attention Is All You Need"☆15Dec 13, 2023Updated 2 years ago
- Official PyTorch codes for "Open Vocabulary 3D Scene Understanding via Geometry Guided Self-Distillation", ECCV2024☆31Jul 19, 2024Updated 2 years ago
- [ICLR 2026] - One2Scene☆51May 25, 2026Updated 3 months ago
- [NeurIPS 2025] DP²O-SR: Direct Perceptual Preference Optimization for Real-World Image Super-Resolution☆86Dec 20, 2025Updated 8 months ago