AMoE: Agglomerative Mixture-of-Experts Vision Foundation Models
☆59Jun 11, 2026Updated 3 months ago
Alternatives and similar repositories for siglino
Users that are interested in siglino are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A set of tools and examples for converting and utilizing powerful vision models, DINOv3 and EdgeTAM (SAM2), within the ONNX ecosystem.☆15Nov 5, 2025Updated 10 months ago
- Efficient Universal Perception Encoder: a single on-device vision encoder with versatile representations that match or exceed specialized…☆718Apr 14, 2026Updated 5 months ago
- ☆10Dec 12, 2023Updated 2 years ago
- SPAR: Single-Pass Any-Resolution ViT for Open-vocabulary Segmentation☆36Aug 13, 2026Updated last month
- OpenEarthMap-SAR: A benchmark dataset for land cover mapping under all-weather conditions☆25Jun 26, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆16Dec 15, 2025Updated 9 months ago
- Video-R2: Reinforcing Consistent and Grounded Reasoning in Multimodal Language Models☆19Sep 5, 2026Updated 2 weeks ago
- ☆34Jun 10, 2026Updated 3 months ago
- Official code "Taming SAM3 under Distribution Shift: ConceptBank for Support-Assisted Open-Vocabulary Segmentation"☆119Aug 14, 2026Updated last month
- ☆20Aug 13, 2026Updated last month
- Code and updates for the ScoreRS project.☆44Sep 19, 2025Updated last year
- Code for "Scaling Language-Free Visual Representation Learning" paper (Web-SSL).☆216Mar 20, 2026Updated 5 months ago
- Efficient image to 3D geometry foundation models from Meta Reality Labs for monocular depth, point maps, and surface normals. Featuring H…☆69May 20, 2026Updated 3 months ago
- SteerViT is a framework that equips any ViT with the ability to steer both its global and local visual representations with natural langu…☆148Jun 13, 2026Updated 3 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official code of Franca: Nested Matryoshka Clustering for Scalable Visual Representation Learning☆282Aug 26, 2026Updated 3 weeks ago
- ☆24Nov 22, 2024Updated last year
- [ICCV 2025 Oral] CorrCLIP: Reconstructing Patch Correlations in CLIP for Open-Vocabulary Semantic Segmentation☆73Updated this week
- [CVPR2025] Project for "HyperSeg: Towards Universal Visual Segmentation with Large Language Model".☆184Dec 13, 2024Updated last year
- [ICLR '26 Oral] Official repository of the paper "AnyUp: Universal Feature Upsampling".☆588Apr 17, 2026Updated 5 months ago
- Official repository for "AM-RADIO: Reduce All Domains Into One"☆1,954May 29, 2026Updated 3 months ago
- Public release of the code for "Accelerating Vision Transformers with Adaptive Patches"☆119May 6, 2026Updated 4 months ago
- ☆48Apr 16, 2026Updated 5 months ago
- [ICCV2025] Harnessing CLIP, DINO and SAM for Open Vocabulary Segmentation☆129Nov 22, 2025Updated 9 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- [ECCV'24] Official PyTorch implementation of In Defense of Lazy Visual Grounding for Open-Vocabulary Semantic Segmentation☆51Sep 24, 2024Updated last year
- Depth-Guided Scale-Aware Global Structure-from-Motion☆28Jul 10, 2026Updated 2 months ago
- Clean Micropython implementation of the Huskylens protocol for UART and I2C☆34Jul 22, 2026Updated last month
- FlowWM stochastic world modeling via flow matching in DINOv3 feature space, with the FuturePerception (Waymo) benchmark.☆83Sep 9, 2026Updated last week
- Apple's Cut Cross Entropy☆35Jan 19, 2025Updated last year
- [IEEE TGRS 2025] Be the Change You Want to See: Revisiting Remote Sensing Change Detection Practices☆41Dec 1, 2025Updated 9 months ago
- A real-time inferencing of multistreaming YOWOv3(Spatio Temporal Action Detection task) using (UCF101-24) dataset. The repo is extension …☆26May 15, 2026Updated 4 months ago
- ☆21Jan 17, 2025Updated last year
- [ACM MM24 Poster] Official implementation of paper "MVPbev: Multi-view Perspective Image Generation from BEV with Test-time Controllabili…☆20Sep 6, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Project COLON-X (Shaping the neXt frontier in intelligent COLONoscopy)☆23May 2, 2026Updated 4 months ago
- [NeurIPS'25] A work to improve CLIP's visual detail capturing ability by inverting the unCLIP generative model.☆27Mar 19, 2026Updated 6 months ago
- The official implementation of the paper "Understanding and Harnessing Sparsity in Unified Multimodal Models" (TMLR).☆24Sep 2, 2026Updated 2 weeks ago
- ☆16Sep 25, 2025Updated 11 months ago
- DOFA-CLIP: Multimodal Vision–Language Foundation Models for Earth Observation☆43Jul 30, 2025Updated last year
- Vision-Language Dataset for Remote Sensing☆42May 27, 2025Updated last year
- [CVPR 2026 Oral] "INSID3: Training-Free In-Context Segmentation with DINOv3"☆754Jun 26, 2026Updated 2 months ago