Curate, Annotate, and Manage Your Data in LightlyStudio.
☆860Jul 25, 2026Updated this week
Alternatives and similar repositories for lightly-studio
Users that are interested in lightly-studio are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- All-in-one training for vision models (YOLO, ViTs, RT-DETR, DINOv3): pretraining, fine-tuning, distillation.☆1,626Updated this week
- [EMNLP 2024] IFCap: Image-like Retrieval and Frequency-based Entity Filtering for Zero-shot Captioning☆15May 13, 2025Updated last year
- ☆15Dec 31, 2024Updated last year
- Refine high-quality datasets and visual AI models☆10,943Updated this week
- [CVPR2026 Findings] VHS: Verifier on Hidden States, an efficient inference-time scaling verification framework for DiT-based image genera…☆16Mar 25, 2026Updated 4 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- 🚀 Lightning-fast computer vision models. Fine-tune SOTA models with just a few lines of code. Ready for cloud ☁️ and edge 📱 deployment.…☆352Dec 11, 2025Updated 7 months ago
- [CVPR 2024] DyBluRF: Dynamic Neural Radiance Fields from Blurry Monocular Video☆22Jun 14, 2024Updated 2 years ago
- [ICCV'25] When Large Vision-Language Model Meets Large Remote Sensing Imagery: Coarse-to-Fine Text-Guided Token Pruning☆52Feb 16, 2026Updated 5 months ago
- Learning to Mask and Permute Visual Tokens for Vision Transformer Pre-Training☆16Jul 1, 2025Updated last year
- Landsat-Bench: Datasets and Benchmarks for Landsat Foundation Models☆20Jun 18, 2025Updated last year
- 👻 kwami.io | A 3D Interactive AI Companion Library for creating engaging AI companions with visual (blob), audio, and AI speech capabili…☆45Jun 19, 2026Updated last month
- [CVPR 2026] Offical implementation of PhysGen: Physically Grounded 3D Shape Generation for Industrial Design☆30Apr 20, 2026Updated 3 months ago
- Official implementation for "Exploring Vision Transformers for 3D Human Motion-Language Models with Motion Patches" (CVPR 2024)☆32Jul 4, 2024Updated 2 years ago
- PyTorch utilities to handle the KITTI Vision Benchmark Suite☆11Jul 21, 2022Updated 4 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- A free and opensource yolov8, yolo11 and yolo26 all in one training tool that automates file structure and yaml files, auto labeling with…☆41Feb 26, 2026Updated 5 months ago
- [IROS 2024 Oral Pitch] PyTorch Implementation of "Dual-Branch Graph Transformer Network for 3D Human Mesh Reconstruction from Video"☆15Jul 19, 2024Updated 2 years ago
- [CVPR 2024] S-DyRF: Reference-Based Stylized Radiance Fields for Dynamic Scenes☆13Jun 1, 2024Updated 2 years ago
- HumanNOVA: Photorealistic, Universal and Rapid 3D Human Avatar Modeling from a Single Image (CVPR 2026 Highlight)☆34Jun 1, 2026Updated last month
- [ICCV'25 oral] Official Code for "LoftUp: Learning a Coordinate-Based Feature Upsampler for Vision Foundation Models"☆261Jan 13, 2026Updated 6 months ago
- This is open-source implementation of MixedAE (https://arxiv.org/pdf/2303.17152.pdf)☆22Feb 14, 2025Updated last year
- Code for "Weakly-supervised Fingerspelling Recognition in British Sign Language Videos", BMVC 2022.☆12Jun 22, 2023Updated 3 years ago
- ☆17Feb 20, 2025Updated last year
- Open source AI/ML capabilities for the FiftyOne ecosystem☆160Updated this week
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Awesome Visual Tokenizers/Autoencoders☆20Nov 19, 2025Updated 8 months ago
- ☆11Sep 1, 2024Updated last year
- [ACM MM 2024] PyTorch Implementation of "ARTS: Semi-Analytical Regressor using Disentangled Skeletal Representations for Human Mesh Recov…☆16Feb 27, 2025Updated last year
- “YOLOLite — lightweight YOLO in PyTorch. ONNX export + CPU inference (Raspberry Pi friendly).”☆82May 25, 2026Updated 2 months ago
- Calibration RGB / Depth for the 7-Scenes dataset☆15Mar 18, 2024Updated 2 years ago
- ☆10Nov 27, 2024Updated last year
- Reasoning Guided Embeddings: Leveraging MLLM Reasoning for Improved Multimodal Retrieval☆15Nov 29, 2025Updated 7 months ago
- [ICML 2026] Text Before Vision: Staged Knowledge Injection Matters for Agentic RLVR in Ultra-High-Resolution Remote Sensing Understanding☆16Mar 13, 2026Updated 4 months ago
- Official Repository for "Communication Efficient Federated Learning with Generalized Heavy-Ball Momentum", accepted at TMLR 2025☆28Jul 14, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- This is a Streamlit application that allows two local Ollama models to chat with each other.☆57Oct 21, 2025Updated 9 months ago
- This repository contains a curated list of research papers and resources focusing on saliency and scanpath prediction, human attention, h…☆66May 9, 2025Updated last year
- [EMNLP 2024] Preserving Multi-Modal Capabilities of Pre-trained VLMs for Improving Vision-Linguistic Compositionality☆23Oct 8, 2024Updated last year
- Official code for Range-Agnostic Multi-View Depth Estimation with Keyframe Selection (3DV 2024)☆44Feb 19, 2024Updated 2 years ago
- [ICCV 2025] Official repository of the paper "Talking to DINO: Bridging Self-Supervised Vision Backbones with Language for Open-Vocabular…☆194Nov 10, 2025Updated 8 months ago
- A curated list of papers that focus on how to represent Earth data in embedding space — spatial, temporal, or semantic — and how those em…☆103Jul 17, 2026Updated last week
- [ICCV 2025] MissRAG: Addressing the Missing Modality Challenge in Multimodal Large Language Models☆26May 12, 2026Updated 2 months ago