Evaluating the performance and accuracy of ABBYY FineReader's OCR on Senate Financial Disclosure scanned forms
☆135Mar 22, 2016Updated 10 years ago
Alternatives and similar repositories for abbyy-finereader-ocr-senate
Users that are interested in abbyy-finereader-ocr-senate are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- NICAR 2016 talk about PDFs!☆63Mar 12, 2016Updated 10 years ago
- Investigative tool for extracting relevant areas from many documents☆14Nov 17, 2015Updated 10 years ago
- ☆23Mar 7, 2015Updated 11 years ago
- A collection of lists of forms maintained by local, state and federal policing organizations. If you have a form name, you have a FOIA re…☆22Updated this week
- Handouts/Tipsheets for the 2015 Global Investigative Journalism Conference☆11Oct 9, 2015Updated 11 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- A re-useable, stand-alone version of LittleSis network storytelling tool☆12Jan 30, 2016Updated 10 years ago
- Mapping the growth of Wal-Mart in urban areas.☆14Apr 1, 2015Updated 11 years ago
- A financial disclosure data extraction tool.☆23Aug 2, 2023Updated 3 years ago
- Tools for working with Optical Character Recognition output☆16Mar 7, 2014Updated 12 years ago
- Python 3.x notebooks about real-world data cleaning and visualization☆71May 4, 2016Updated 10 years ago
- NICAR 2019 workshop on using Python and PDFplumber to extract text from PDFs☆12Mar 9, 2019Updated 7 years ago
- Example data project for NICAR15 presentation☆13Mar 3, 2015Updated 11 years ago
- Scraper for financial disclosure reports from the US Senate☆23May 7, 2024Updated 2 years ago
- Lightweight data store library featuring two way binding & 1-n views☆15Jul 9, 2015Updated 11 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- POLITICO's system for managing civic data☆20Dec 7, 2022Updated 3 years ago
- Obtained in December 2014 through a Freedom of Information request☆15Jan 29, 2016Updated 10 years ago
- A responsive iframe built via custom elements☆28Mar 4, 2023Updated 3 years ago
- Code for the Deep Learning HackerEarth Challenge #1☆12Nov 1, 2017Updated 8 years ago
- assorted text data☆34Jan 10, 2019Updated 7 years ago
- Various NLP-related stuff☆10Apr 13, 2017Updated 9 years ago
- ☆14Feb 8, 2024Updated 2 years ago
- Workbook to teach the concept of risk ratios for data journalism applications☆33Apr 15, 2022Updated 4 years ago
- Machine Learning Hackathon organized by Hackerearth☆13Feb 2, 2016Updated 10 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- Site for Melody Kramer's Visiting Nieman Fellowship☆13May 15, 2015Updated 11 years ago
- Tooling to extract data from scanned paper forms OCR-ed by Tesseract using the HOCR standard.☆83Mar 1, 2016Updated 10 years ago
- A cookbook for installing and configuring Apache Spark☆11Sep 6, 2018Updated 8 years ago
- Create and manage Shlink short links from WordPress☆17Jun 30, 2026Updated 3 months ago
- 🗄 Bot powering the @LinkArchiver Twitter tool to send tweeted URLs to the Wayback Machine☆44Sep 21, 2017Updated 9 years ago
- 3rd place solution to the Mars Express Power Challenge hosted by the European Space Agency☆13Sep 13, 2016Updated 10 years ago
- The Berkeley Document Summarizer is a learning-based, single-document summarization system that extracts source document content, exploit…☆746Feb 25, 2019Updated 7 years ago
- A Rails application to lookup congressional and state legislators by latitude and longitude☆30Jun 17, 2022Updated 4 years ago
- A quick and dirty command line tool for bulk uploading documents to DocumentCloud.☆22Mar 17, 2011Updated 15 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Code for extracting data from a large number of PDFs, particularly FCC political ad documents☆15Oct 26, 2017Updated 8 years ago
- [DEPRECATED] Unofficial Python Pandas DataReader objects with requests and requests_cache☆16Mar 5, 2018Updated 8 years ago
- A set of tools for extracting tables from PDF files helping to do data mining on (OCR-processed) scanned documents.☆2,255Jun 24, 2022Updated 4 years ago
- A catalogue of public national and supranational open data portals.☆12May 19, 2017Updated 9 years ago
- Slides and supporting material for Chase's data journalism ignite talk from NewsFoo 2012.☆29Nov 30, 2012Updated 13 years ago
- Simple library for storing Scrapy Items in sqlite database☆12Jan 28, 2016Updated 10 years ago
- A set of Python modules for downloading, parsing, and outputting data related to the Supreme Court.☆40Jun 20, 2019Updated 7 years ago