Library to populate items using XPath and CSS with a convenient API
☆49Sep 16, 2026Updated last week
Alternatives and similar repositories for itemloaders
Users that are interested in itemloaders are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- https://mimesniff.spec.whatwg.org/ implementation for Python☆13Jul 9, 2026Updated 2 months ago
- Common interface for data container classes☆69Sep 8, 2026Updated 2 weeks ago
- A browser extension to monitor your spiders deployed on Scrapy Cloud.☆16Mar 8, 2025Updated last year
- A library to make it easier to load input URLs to start scrapy processes☆14Feb 21, 2021Updated 5 years ago
- Analyze scraped data☆47Dec 9, 2019Updated 6 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Python library of web-related functions☆417Updated this week
- A pure-Python robots.txt parser with support for modern conventions.☆94Updated this week
- Convert Javascript code to an XML document☆186Mar 14, 2022Updated 4 years ago
- Web scraping Page Objects core library☆106Sep 7, 2026Updated 2 weeks ago
- Page Object pattern for Scrapy☆129Sep 16, 2026Updated last week
- 🕶 Awesome list of Scrapy tools and libraries☆60Jul 6, 2020Updated 6 years ago
- Library for annotation-based dependency injection☆24Updated this week
- Parse numbers written in natural language☆130Oct 23, 2024Updated last year
- Automatic unit test generation for Scrapy.☆58Jul 12, 2021Updated 5 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Python clients for Zyte AutoExtract API☆41Jan 17, 2022Updated 4 years ago
- Collection of persistent (disk-based) and non-persistent (memory-based) queues for Python☆300Updated this week
- Python client for Zyte API☆30Aug 11, 2026Updated last month
- A scrapy extension to sync `.scrapy` folder to an S3 bucket☆18Mar 28, 2022Updated 4 years ago
- Scrapinghub Command Line Client☆129Jul 22, 2026Updated 2 months ago
- Web application to help categorize and aggregate subscriptions of media channels for easy access. (working only with Youtube channels at …☆16Aug 27, 2020Updated 6 years ago
- Extract text from HTML☆135Apr 8, 2026Updated 5 months ago
- Parsel lets you extract data from XML/HTML documents using XPath or CSS selectors☆1,356Updated this week
- A linter for Scrapy projects.☆22Updated this week
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Remove DIVs, style stuff and normalize HTML preserving structure information☆14Aug 13, 2026Updated last month
- Guidelines for Software Development Projects☆22Apr 1, 2026Updated 5 months ago
- ☆23Sep 3, 2026Updated 2 weeks ago
- Scrapy Extension for monitoring spiders execution.☆562Sep 9, 2026Updated 2 weeks ago
- Extract price amount and currency symbol from a raw text string☆348Aug 6, 2026Updated last month
- Scrapy spider middleware to split an item into multiple items using a multi-valued key☆21Feb 8, 2017Updated 9 years ago
- ☆16Apr 10, 2026Updated 5 months ago
- Performance-focused replacement for Python urllib☆21Apr 13, 2026Updated 5 months ago
- Parsing JavaScript objects into Python data structures☆220May 17, 2026Updated 4 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Spider templates for automatic crawlers.☆34Mar 26, 2026Updated 5 months ago
- Example frontera project☆12Aug 10, 2017Updated 9 years ago
- CSS Selectors for Python☆310Sep 16, 2026Updated last week
- A simple, Qt-Webengine powered web browser with built in functionality for basic scrapy webscraping support.☆110May 21, 2024Updated 2 years ago
- Scrapy downloader middleware that stores response HTMLs to disk.☆18Apr 14, 2026Updated 5 months ago
- Google Tink's critical Ed25519 bug related to Java "final" keyword☆11Apr 5, 2020Updated 6 years ago
- A complimentary proxy to help to use SPM with headless browsers☆109May 20, 2026Updated 4 months ago