[CVPR 2025] Adaptive Markup Language Generation for Contextually-Grounded Visual Document Understanding
☆16Jun 16, 2025Updated last year
Alternatives and similar repositories for DocMark
Users that are interested in DocMark are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [NeurIPS 2025] UI-Genie: A Self-Improving Approach for Iteratively Boosting MLLM-based Mobile GUI Agents☆60Nov 27, 2025Updated 7 months ago
- PyTorch implementation of "UNIT: Unifying Image and Text Recognition in One Vision Encoder", NeurlPS 2024.☆34Sep 26, 2024Updated last year
- Official repository of "SeGA: Preference-Aware Self-Contrastive Learning with Prompts for Anomalous User Detection on Twitter" @ AAAI 202…☆10Nov 30, 2024Updated last year
- 苏州大学每日健康情况自动化打卡脚本☆13Mar 30, 2022Updated 4 years ago
- 综合项目实践项目学习记录+代码☆11Jun 18, 2022Updated 4 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- 一个桌面宠物程序,现在似乎发展成为桌面便签了。桌面便签程序见develop-todolist分支。☆11Nov 17, 2024Updated last year
- [3DV2026] Official repository for "CamC2V: Context-aware Controllable Video Generation"☆14Nov 11, 2025Updated 8 months ago
- 为visinger SVS系统写的展示系统~本质仍然是个音乐播放器☆11Apr 18, 2023Updated 3 years ago
- 测试 https://huggingface.co/OFA-Sys/gsm8k-rft-llama7b-u13b 的 GSM8K 分数☆15Aug 10, 2023Updated 2 years ago
- This project uses Centernet and Conditional Convolutions for Instance Segmentation☆13Sep 17, 2020Updated 5 years ago
- The Soft Cosine Measure system developed for the ARQMath-3 shared task evaluation of math information retrieval systems☆13Sep 8, 2022Updated 3 years ago
- [Paper] Code for the EMNLP2023 (Findings) paper "Global Structure Knowledge-Guided Relation Extraction Method for Visually-Rich Document"☆17Dec 1, 2023Updated 2 years ago
- Create PDF animations from graphics files and inline graphics using LaTeX☆12Jun 8, 2018Updated 8 years ago
- Official PyTorch implementation of “MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation”☆18Dec 5, 2024Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- It's the code for the paper Pushing the Performance Limit of Scene Text Recognizer without Human Annotation, CVPR 2022.☆28Jul 6, 2022Updated 4 years ago
- ☆42Sep 2, 2023Updated 2 years ago
- Official implementation of "LOCATEdit: Graph Laplacian Optimized Cross Attention for Localized Guided Image Editing☆16May 27, 2025Updated last year
- PFCC 社区博客☆14Updated this week
- AI-powered slide workspace for creating, editing, versioning, and presenting beautiful reveal.js decks from prompts and source files.☆15Apr 14, 2026Updated 3 months ago
- ☆11Jan 27, 2020Updated 6 years ago
- pytorch crnn with centerloss to solve the near word problem☆16Jan 27, 2022Updated 4 years ago
- Open source implementation of the Mamba architecture in TensorFlow☆20Jul 15, 2024Updated 2 years ago
- 《카카오 아레나 데이터 경진대회 1등 노하우》 예제 코드☆17Jan 15, 2021Updated 5 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Repository for ACL2020 paper "Refer360° A Referring Expression Recognition Dataset in 360°Images"☆14Jun 26, 2021Updated 5 years ago
- Official repository of "CoMP: Continual Multimodal Pre-training for Vision Foundation Models"☆48Apr 3, 2025Updated last year
- [TPAMI] Locating and Counting Heads in Crowds With a Depth Prior☆10Jan 7, 2022Updated 4 years ago
- Code release for "Weakly Supervised Open-Vocabulary Object Detection", AAAI2024☆36Sep 9, 2024Updated last year
- 音乐可视化前端框架☆18Feb 27, 2015Updated 11 years ago
- The official repository for paper Evaluating Financial Relational Graphs: Interpretation Before Prediction☆20Jan 2, 2026Updated 6 months ago
- [CVPR 2025] Docopilot: Improving Multimodal Models for Document-Level Understanding☆37Jul 22, 2025Updated 11 months ago
- [NeurIPS 2024] Official Code for the Paper "Multimodal Task Vectors Enable Many-Shot Multimodal In-Context Learning"☆27Apr 8, 2025Updated last year
- [ICLR 2026] Code for Evolutionary Caching to Accelerate Your Off-the-Shelf Diffusion Model☆30Mar 1, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆16May 20, 2026Updated 2 months ago
- ☆16Nov 12, 2025Updated 8 months ago
- 音乐可视化 canvas☆18Jan 4, 2023Updated 3 years ago
- Official implementation for [ICML 2026] Scalable GANs with Transformers☆17Jul 1, 2026Updated 2 weeks ago
- [ICRA 2026] Official implementation of the paper: “EgoTraj-bench: Towards robust trajectory prediction under ego-view noisy observations”☆20Jul 6, 2026Updated 2 weeks ago
- [ICCV 2025] CHORDS: Diffusion Sampling Accelerator with Multi-core Hierarchical ODE Solvers☆17Mar 3, 2026Updated 4 months ago
- ☆28Mar 17, 2025Updated last year