Official implementation of "Surgical-VQLA: Transformer with Gated Vision-Language Embedding for Visual Question Localized-Answering in Robotic Surgery", ICRA 2023.
☆27Jul 7, 2024Updated 2 years ago
Alternatives and similar repositories for Surgical-VQLA
Users that are interested in Surgical-VQLA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official implementation of “CAT-ViL: Co-Attention Gated Vision-Language Embedding for Visual Question Localized-Answering in Robotic Surg…☆18Jul 7, 2024Updated 2 years ago
- Globally reasoned multi-task model for surgical scene understanding. A multi-task model for segmentation and scene graph. Offical Impleme…☆15May 5, 2022Updated 4 years ago
- ☆17Nov 19, 2020Updated 5 years ago
- ☆15Jul 4, 2023Updated 3 years ago
- Official implementation of “LLCaps: Learning to Illuminate Low-Light Capsule Endoscopy with Curved Wavelet Attention and Reverse Diffusio…☆21Jul 7, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Official Implementation of "Surgical-VQLA++: Adversarial Contrastive Learning for Calibrated Robust Visual Question Localized-Answering i…☆15May 6, 2025Updated last year
- ☆16May 31, 2024Updated 2 years ago
- [MICCAI'22] AutoLaparo: A New Dataset of Integrated Multi-tasks for Image-guided Surgical Automation in Laparoscopic Hysterectomy☆21Mar 21, 2025Updated last year
- ☆18Jan 20, 2026Updated 7 months ago
- S2ME: Spatial-Spectral Mutual Teaching and Ensemble Learning for Scribble-supervised Polyp Segmentation (MICCAI 2023)☆21Dec 1, 2023Updated 2 years ago
- [NeurIPS 2024 Workshop AIM-FM] Official code implementation for paper: Surgical SAM 2☆78Aug 19, 2026Updated 2 weeks ago
- [IEEE RA-L&ICRA2025] Think Step by Step: Chain-of-Gesture Prompting for Error Detection in Robotic Surgical Videos☆15Aug 26, 2025Updated last year
- Code repository for paper: "General surgery vision transformer: A video pre-trained foundation model for general surgery"☆54Apr 19, 2024Updated 2 years ago
- SASVi - Segment Any Surgical Video (IPCAI 2025)☆17Jul 3, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- ☆13Jul 20, 2023Updated 3 years ago
- This repository is made for the paper: Self-supervised vision-language pretraining for Medical visual question answering☆44Apr 8, 2023Updated 3 years ago
- [MICCAI 2025] Official code implementation for paper: ReSurgSAM2: Referring Segment Anything in Surgical Video via Credible Long-term Tra…☆43Nov 4, 2025Updated 10 months ago
- ☆16Feb 5, 2024Updated 2 years ago
- code for Expert Knowledge-Aware Image Difference Graph Representation Learning for Difference-Aware Medical Visual Question Answering☆29May 30, 2025Updated last year
- PyTorch implements `Image Super-Resolution Using Very Deep Residual Channel Attention Networks` paper.☆15Dec 6, 2022Updated 3 years ago
- Pytorch implementation of the MICCAI 2020 paper ISINet: An Instance-Based Approach for Surgical Instrument Segmentation.☆27Oct 12, 2021Updated 4 years ago
- Official Repository for the ICML 2023 paper "BiRT: Bio-inspired Replay in Vision Transformers for Continual Learning"☆16Oct 11, 2023Updated 2 years ago
- ☆16Dec 14, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- RUArt: A Novel Text-Centered Solution for Text-Based Visual Question Answering☆10Nov 27, 2022Updated 3 years ago
- ☆28Feb 7, 2024Updated 2 years ago
- ☆39Apr 5, 2025Updated last year
- Pytorch Implementation for paper "Adversarial Graph Disentanglement"☆13Jul 18, 2023Updated 3 years ago
- Our first tutorial, make your own Augmented Reality app.☆12Jul 25, 2024Updated 2 years ago
- Fine-Grained Knowledge Fusion for Retrieval-Augmented Medical Visual Question☆11Jul 18, 2024Updated 2 years ago
- A simulation project on Dynamic Movement Primitive (DMP) , containing 1D numerical simulation, 2D learning from demonstration and 3D simu…☆12Feb 27, 2025Updated last year
- [IPCAI'24 Best Paper] Advancing Surgical VQA with Scene Graph Knowledge☆52May 23, 2025Updated last year
- Fleming-VL: Towards Universal Medical Visual Understanding with Multimodal LLMs☆16Nov 6, 2025Updated 10 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ECCV 2020] Temporal Aggregate Representations for Long-Range Video Understanding☆11Sep 13, 2021Updated 4 years ago
- ☆10Oct 20, 2022Updated 3 years ago
- ☆47Updated this week
- Repository for the paper: Open-Ended Medical Visual Question Answering Through Prefix Tuning of Language Models (https://arxiv.org/abs/23…☆19Sep 2, 2023Updated 3 years ago
- Motion-aware 3D Gaussian Splatting for Efficient Dynamic Scene Reconstruction☆25Jul 29, 2024Updated 2 years ago
- Official repository for 'Algorithmic encoding of protected characteristics in chest X-ray disease detection models'☆24Feb 15, 2023Updated 3 years ago
- [ICCV 2023] Rethinking Point Cloud Registration as Masking and Reconstruction☆10Aug 14, 2023Updated 3 years ago