Official implementation of "Surgical-VQLA: Transformer with Gated Vision-Language Embedding for Visual Question Localized-Answering in Robotic Surgery", ICRA 2023.
☆27Jul 7, 2024Updated 2 years ago
Alternatives and similar repositories for Surgical-VQLA
Users that are interested in Surgical-VQLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆17Nov 19, 2020Updated 5 years ago
- ☆21Dec 19, 2025Updated 7 months ago
- Official implementation of "EndoUIC: Promptable Diffusion Transformer for Unified Illumination Correction in Capsule Endoscopy", MICCAI 2…☆12Jan 29, 2026Updated 6 months ago
- Official Implementation of "Surgical-VQLA++: Adversarial Contrastive Learning for Calibrated Robust Visual Question Localized-Answering i…☆15May 6, 2025Updated last year
- ☆17Jul 5, 2021Updated 5 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Implementation of the paper LIMITR: Leveraging Local Information for Medical Image-Text Representation☆17Jul 21, 2026Updated 3 weeks ago
- ☆16May 31, 2024Updated 2 years ago
- [MICCAI'22] AutoLaparo: A New Dataset of Integrated Multi-tasks for Image-guided Surgical Automation in Laparoscopic Hysterectomy☆21Mar 21, 2025Updated last year
- S2ME: Spatial-Spectral Mutual Teaching and Ensemble Learning for Scribble-supervised Polyp Segmentation (MICCAI 2023)☆21Dec 1, 2023Updated 2 years ago
- Implementation of ''VPUFormer: Visual Prompt Unified Transformer for Interactive Image Segmentation''☆15Sep 16, 2025Updated 11 months ago
- [NeurIPS 2024 Workshop AIM-FM] Official code implementation for paper: Surgical SAM 2☆78Apr 15, 2025Updated last year
- [IEEE RA-L&ICRA2025] Think Step by Step: Chain-of-Gesture Prompting for Error Detection in Robotic Surgical Videos☆15Aug 26, 2025Updated 11 months ago
- Code repository for paper: "General surgery vision transformer: A video pre-trained foundation model for general surgery"☆53Apr 19, 2024Updated 2 years ago
- SASVi - Segment Any Surgical Video (IPCAI 2025)☆16Jul 3, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- VQA-Med 2021☆24May 13, 2026Updated 3 months ago
- A repository for surgical action triplet dataset. Data are videos of laparoscopic cholecystectomy that have been annotated with <instrume…☆86Sep 17, 2025Updated 11 months ago
- ☆13Jul 20, 2023Updated 3 years ago
- This repository is made for the paper: Self-supervised vision-language pretraining for Medical visual question answering☆44Apr 8, 2023Updated 3 years ago
- [MICCAI 2025] Official code implementation for paper: ReSurgSAM2: Referring Segment Anything in Surgical Video via Credible Long-term Tra…☆43Nov 4, 2025Updated 9 months ago
- ☆16Feb 5, 2024Updated 2 years ago
- ☆15Mar 11, 2023Updated 3 years ago
- PyTorch implements `Image Super-Resolution Using Very Deep Residual Channel Attention Networks` paper.☆15Dec 6, 2022Updated 3 years ago
- ☆14Jun 26, 2022Updated 4 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆19May 27, 2025Updated last year
- Official Repository for the Endoscapes Dataset for Surgical Scene Segmentation, Object Detection, and Critical View of Safety Assessment☆64Sep 17, 2025Updated 11 months ago
- Pytorch implementation of the MICCAI 2020 paper ISINet: An Instance-Based Approach for Surgical Instrument Segmentation.☆27Oct 12, 2021Updated 4 years ago
- Official Repository for the ICML 2023 paper "BiRT: Bio-inspired Replay in Vision Transformers for Continual Learning"☆16Oct 11, 2023Updated 2 years ago
- ☆38Apr 5, 2025Updated last year
- 丁立中的大模型算法工程作品集。聚焦 LLM / VLM 全链路的复现与优化,涵盖:① 预训练与微调(Pretrain / SFT / MoE / 多模态对齐);② 强化学习对齐(PPO / GRPO / DAPO,含多奖励函数与 GAE);③ 知识蒸馏(离线 KL 蒸馏、在…☆17May 26, 2026Updated 2 months ago
- Official code of the paper "EgoExOR: EgoExOR: An Ego-Exo-Centric Operating Room Dataset for Surgical Activity Understanding" accepted at …☆29May 6, 2026Updated 3 months ago
- Fine-Grained Knowledge Fusion for Retrieval-Augmented Medical Visual Question☆11Jul 18, 2024Updated 2 years ago
- A simulation project on Dynamic Movement Primitive (DMP) , containing 1D numerical simulation, 2D learning from demonstration and 3D simu…☆12Feb 27, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- The official repository for "SurgNet: Self-supervised Pretraining with Semantic Consistency for Vessel and Instrument Segmentation in Sur…☆15Dec 30, 2024Updated last year
- Fleming-VL: Towards Universal Medical Visual Understanding with Multimodal LLMs☆15Nov 6, 2025Updated 9 months ago
- Open-H-Embodiment is a community‑driven dataset initiative building the open, shared foundation needed to train and evaluate a generalist…☆134Updated this week
- Repository for the paper: Open-Ended Medical Visual Question Answering Through Prefix Tuning of Language Models (https://arxiv.org/abs/23…☆19Sep 2, 2023Updated 2 years ago
- ☆39Updated this week
- MICCAI 2022: Free Lunch for Surgical Video Understanding by Distilling Self-Supervisions☆13Sep 17, 2022Updated 3 years ago
- ☆66Apr 21, 2026Updated 3 months ago