☆24Apr 29, 2025Updated last year
Alternatives and similar repositories for visfocus
Users that are interested in visfocus are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- STVQA and TextVQA OCR results from Amazon Text in Image pipeline☆12Jul 18, 2022Updated 4 years ago
- TextAdaIN: Paying Attention to Shortcut Learning in Text Recognizers☆21Jul 26, 2022Updated 4 years ago
- An implementation of the Holistic Pursuit for the Multi-Layer Sparse Coding model. Contains a comparison to the projection pursuit algori…☆19Dec 19, 2018Updated 7 years ago
- ☆73Jul 17, 2024Updated 2 years ago
- Unofficial implementation of CVPR 2020 paper "SCATTER: Selective Context Attentional Scene Text Recognizer"☆66Mar 3, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Official implementation for "GLASS: Global to Local Attention for Scene-Text Spotting" (ECCV'22)☆102Jun 28, 2024Updated 2 years ago
- Ada-LISTA: Learned Solvers Adaptive to Varying Models☆11Feb 18, 2020Updated 6 years ago
- [AAAI 2025] DocKylin: A Large Multimodal Model for Visual Document Understanding with Efficient Visual Slimming☆36Jun 1, 2025Updated last year
- OCR Annotations from Amazon Textract for Industry Documents Library☆104Aug 20, 2022Updated 4 years ago
- Implementation of the paper "It's All in the Head: Representation Knowledge Distillation through Classifier Sharing"☆37Aug 23, 2022Updated 4 years ago
- ☆11Jun 3, 2023Updated 3 years ago
- The code and data for "Summary-Oriented Vision Modeling for Multimodal Abstractive Summarization"☆11May 16, 2023Updated 3 years ago
- ☆25Nov 22, 2024Updated last year
- ☆12Jul 21, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Demo for Multi-Layer ISTA and Multi-Layer FISTA algorithms for convolutional neural networks, as described in J. Sulam, A. Aberdam, A. Be…☆29Nov 20, 2018Updated 7 years ago
- RareAct: A video dataset of unusual interactions☆35Aug 4, 2020Updated 6 years ago
- Official Code of ICCV 2021 Paper: Learning to Cut by Watching Movies☆51Nov 9, 2022Updated 3 years ago
- Evaluation of concept erasing diffusion models should include latent likelihood☆22Nov 3, 2025Updated 9 months ago
- Best Performing System of SemEval2023 Task 4 - ValueEval: Identification of Human Values behind Arguments☆14Mar 26, 2024Updated 2 years ago
- ☆11Oct 16, 2023Updated 2 years ago
- The public reproducible analysis code used for the gaze project☆11May 16, 2026Updated 3 months ago
- [ECCV 2024] Official PyTorch implementation of LUT "Learning with Unmasked Tokens Drives Stronger Vision Learners"☆14Dec 1, 2024Updated last year
- Uncertainty-Guided Pseudo-Labelling with Model Averaging☆11Mar 17, 2026Updated 5 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Pytorch Implementation of PermutedAdaIN☆36Mar 29, 2021Updated 5 years ago
- Code for the "Long Context Needs Some R&R" paper.☆12Mar 11, 2024Updated 2 years ago
- Concise Reasoning via Reinforcement Learning☆13Apr 16, 2025Updated last year
- CAD - Memory Efficient Convolutional Adapter for Segment Anything☆12Oct 4, 2024Updated last year
- The implementation of our NeurIPS 2024 paper "DarkSAM: Fooling Segment Anything Model to Segment Nothing".☆14Nov 4, 2024Updated last year
- ☆16Sep 20, 2022Updated 3 years ago
- Code for "Saliency Prediction of Sports Videos: A Large-Scale Database and a Self-Adaptive Approach", ICASSP 2024☆14May 28, 2024Updated 2 years ago
- Reading list for research topics in intent analysis.☆16Oct 23, 2023Updated 2 years ago
- [ICLR 2025] DGQ: Distribution-Aware Group Quantization for Text-to-Image Diffusion Models☆20Mar 25, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Q-resafe:Assessing Safety Risks and Quantization-aware Safety Patching for Quantized Large Language Models (ICML'2025)☆16Jun 28, 2025Updated last year
- OneVOS: Unifying Video Object Segmentation with All-in-One Transformer Framework☆13Feb 27, 2025Updated last year
- This is the GitHub repository for Data Augmentation for Saliency Prediction via Latent Diffusion paper in ECCV 2024, Milano, Italy☆15Nov 7, 2024Updated last year
- Water Network-Augmented Two-State model for Protein−Ligand Binding Affinity Prediction☆12Jun 10, 2023Updated 3 years ago
- SST-Sal: A spherical spatio-temporal approach for saliency prediction in 360º videos☆15Aug 31, 2023Updated 3 years ago
- Source code for the paper "A Medical Semantic-Assisted Transformer for Radiographic Report Generation"☆25Jun 23, 2023Updated 3 years ago
- Implementation of LaTr: Layout-aware transformer for scene-text VQA,a novel multimodal architecture for Scene Text Visual Question Answer…☆56Jul 22, 2026Updated last month