houking-can / PDFConverter
Best PDF Converter! PDF to any format, pdf2word/excel/xml/html/txt...
☆152Updated 4 years ago
Alternatives and similar repositories for PDFConverter:
Users that are interested in PDFConverter are comparing it to the libraries listed below
- 该项目主要是抽取病历文件中的一些关键信息。并将抽取的内容进行streamlit前端的展示。目前支持的文件类型:图片,pdf文件,word文件☆23Updated 2 years ago
- FinanceEventGraph,金融领域事件图谱开放数据集,可用于事件图谱搭建于实验,包括3865个acquire并购事件、9093个invest投资事件,总计12960的事件☆19Updated last year
- 基于simhash的文本去重算法☆20Updated 3 years ago
- 中英文敏感词、语言检测、中外手机/电话归属地/运营商查询、名字推断性别、手机号抽取、身份证抽取、 邮箱抽取、中日文人名库、中文缩写库、拆字词典、词汇情感值、停用词、反动词表、暴恐词表、繁简体转换、英文模拟中文发音、汪峰歌词生成器、职业名称词库、同义词库、反义词库、否定词库、汽…☆32Updated 6 years ago
- It's for a research for AI and law☆43Updated 4 years ago
- ☆44Updated 5 years ago
- CCKS2019评测任务五-公众公司公告信息抽取,第3名☆122Updated 5 years ago
- ☆22Updated 4 years ago
- 通用版面分 析 | 中文文档解析 |Document Layout Analysis | layout paser☆46Updated 10 months ago
- 一个简单易用的 Python 模块,用于通过字符串来操作日期/时间。正则时间提取,字符串时间解析,字符串时间提取。中文时间提取,一句话里面提取时间☆75Updated 9 months ago
- 简单的表格图片内容ocr☆38Updated 5 years ago
- It's a python script that convert PDF to txt using PDFMiner☆46Updated 3 years ago
- RelExt: A Tool for Relation Extraction from Text. 文本实体关系抽取工具。☆50Updated 2 years ago
- Recognize tables and text from scanned images that contain tables. 从包含表格的扫描图片中识别表格和文字☆254Updated last year
- A Multi-Modal Dataset of Chinese Governmental Docunments☆32Updated 4 years ago
- 金庸小说人物关系图谱构建☆61Updated 5 years ago
- 🌳CED: Catalog Extraction from Documents☆16Updated last year
- 【间隙·树·排序算法】 对OCR结果或PDF提取的文本进行版面分析,按人类阅读顺序进行排序。☆130Updated last year
- CLUEWSC2020: WSC Winograd模式挑战中文版,中文指代消解任务☆75Updated 4 years ago
- WordForm,针对中文词语的笔画拆解,偏旁查询,拼音转换接口☆65Updated 6 years ago
- PaddleOCR 输出结果的行对齐,表格制式图像OCR行对齐☆44Updated 3 years ago
- CausalKnowledgeBase, causal knowledge base including causal pairs extracted from web text using the methods like PMI, Collocation。基于网络文本的…☆48Updated 6 years ago
- chinese anti semantic word search interface based on dict crawled from online resources, ChineseAntiword,针对中文词语的反义词查询接口☆59Updated 6 years ago
- EventKGNELL, event knowlege graph never end learning system, a event-centric knowledge base search system,实时事理逻辑知识库终身学习系统项目和事件为核心的知识库搜索系统…☆71Updated 5 years ago
- 中文PDF转TXT的实用工具☆30Updated 3 years ago
- 时间关键词正则提取以及标准化☆21Updated 3 years ago
- 错别字纠正算法。调用pycorrector接口,使用规则。☆68Updated 5 years ago
- 好未来第二周-自动评分☆30Updated 7 years ago
- 知识图谱的小demo☆17Updated 6 years ago
- Let ChatGPT (Large Language Models) Serve As Data Annotator and Zero-shot/few-shot Information Extractor.☆31Updated 2 years ago