[CVPR 2025] Adaptive Markup Language Generation for Contextually-Grounded Visual Document Understanding
☆17Jun 16, 2025Updated last year
Alternatives and similar repositories for DocMark
Users that are interested in DocMark are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- PyTorch implementation of "UNIT: Unifying Image and Text Recognition in One Vision Encoder", NeurlPS 2024.☆34Sep 26, 2024Updated last year
- Benchmarking End-to-End Photographed Document Parsing and Translation☆18Dec 4, 2025Updated 9 months ago
- Official repository of "SeGA: Preference-Aware Self-Contrastive Learning with Prompts for Anomalous User Detection on Twitter" @ AAAI 202…☆10Nov 30, 2024Updated last year
- weixin125个人健康数据管理系统的设计与实现微信小程序+ssm后端毕业源码案例设计☆10Feb 28, 2024Updated 2 years ago
- SceneCompleter: Dense 3D Scene Completion for Generative Novel View Synthesis☆37Jun 13, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- This project uses Centernet and Conditional Convolutions for Instance Segmentation☆13Sep 17, 2020Updated 6 years ago
- 测试 https://huggingface.co/OFA-Sys/gsm8k-rft-llama7b-u13b 的 GSM8K 分数☆15Aug 10, 2023Updated 3 years ago
- The Soft Cosine Measure system developed for the ARQMath-3 shared task evaluation of math information retrieval systems☆13Sep 8, 2022Updated 4 years ago
- [Paper] Code for the EMNLP2023 (Findings) paper "Global Structure Knowledge-Guided Relation Extraction Method for Visually-Rich Document"☆17Dec 1, 2023Updated 2 years ago
- [ECML-PKDD 2025] Official Implementation of "Trajectory Imputation in Multi-Agent Sports with Derivative-Accumulating Self-Ensemble".☆15Jun 20, 2025Updated last year
- It's the code for the paper Pushing the Performance Limit of Scene Text Recognizer without Human Annotation, CVPR 2022.☆28Jul 6, 2022Updated 4 years ago
- ☆12Mar 8, 2022Updated 4 years ago
- 清华大学校园网客户端与联网库,适用于命令行环境,Windows、Linux、Mac OS X桌面平台与UWP、iOS、Android移动平台☆12Mar 3, 2020Updated 6 years ago
- PFCC 社区博客☆14Updated this week
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- ☆11Jan 27, 2020Updated 6 years ago
- [ICCV 2025] Official implementation of "Anchor Token Matching: Implicit Structure Locking for Training-free AR Image Editing"☆28Apr 15, 2025Updated last year
- deepseek and langchain are used together to realize RAG☆15Jul 20, 2026Updated 2 months ago
- ☆18Mar 27, 2020Updated 6 years ago
- TL;DR: We propose a large-scale cross-domain persuasion dataset covers 13,000 scenarios in 35 domains, with the developed PersuGPT model …☆17Feb 12, 2025Updated last year
- ☆19Jul 7, 2025Updated last year
- Repository for ACL2020 paper "Refer360° A Referring Expression Recognition Dataset in 360°Images"☆15Jun 26, 2021Updated 5 years ago
- Official PyTorch implementation of “MaskRIS: Semantic Distortion-aware Data Augmentation for Referring Image Segmentation”☆18Dec 5, 2024Updated last year
- DCEN☆13Aug 12, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [WAICA-26 Best Student Paper] Official repository of "Enhancing Vision Foundation Models via Multimodal Continual Pre-Training"☆50Jul 21, 2026Updated last month
- Code release for "Weakly Supervised Open-Vocabulary Object Detection", AAAI2024☆36Sep 9, 2024Updated 2 years ago
- Code for paper "Automatic Neural Network Compression by Sparsity-Quantization Joint Learning: A Constrained Optimization-based Approach"☆21Jul 9, 2020Updated 6 years ago
- [ICML2025] A Non-isotropic Time Series Diffusion Model with Moving Average Transitions☆17Jun 23, 2025Updated last year
- 记录一些学习生活中的收集☆18Sep 7, 2026Updated last week
- The official repository for paper Evaluating Financial Relational Graphs: Interpretation Before Prediction☆20Jan 2, 2026Updated 8 months ago
- [CVPR 2025] Docopilot: Improving Multimodal Models for Document-Level Understanding☆37Jul 22, 2025Updated last year
- [ICLR 2026] Code for Evolutionary Caching to Accelerate Your Off-the-Shelf Diffusion Model☆30Mar 1, 2026Updated 6 months ago
- The official code for the CVPR 2024 paper: Multi-modal In-Context Learning Makes an Ego-evolving Scene Text Recognizer☆55Jun 14, 2024Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Pedestrian Trajectory Prediction with Missing Data: Datasets, Imputation, and Benchmarking (NeurIPS 2024)☆18Feb 16, 2025Updated last year
- ☆16Nov 12, 2025Updated 10 months ago
- This project aims to collect and collate various datasets for multimodal large model training, including but not limited to pre-training …☆78May 7, 2025Updated last year
- ☆23Jun 9, 2019Updated 7 years ago
- Official implementation for our paper: Rethinking Video Tokenization: A Conditioned Diffusion-based Approach☆17Apr 2, 2025Updated last year
- A professional list on deep learning models for Trajectory Similarity Measurement and Trajectory Distance Computation☆28Apr 9, 2024Updated 2 years ago
- A PyTorch implementation of "VectorSynth: Fine-Grained Satellite Image Synthesis with Structured Semantics"☆20Aug 18, 2026Updated last month