The dataset contains 3 million attribute-value annotations across 1257 unique categories on 2.2 million cleaned Amazon product profiles. It is a large, multi-sourced, diverse dataset for product attribute extraction study.
☆157Dec 16, 2022Updated 3 years ago
Alternatives and similar repositories for MAVE
Users that are interested in MAVE are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ACL19-Scaling Up Open Tagging from Tens to Thousands☆17Aug 23, 2019Updated 7 years ago
- ☆89Sep 15, 2020Updated 6 years ago
- Unofficial implementation of the paper "OpenTag: Open Attribute Value Extraction from Product Profiles"☆33Aug 22, 2018Updated 8 years ago
- Attribute Value Extraction using Large Language Models☆30May 24, 2024Updated 2 years ago
- Code for paper OA-Mine: Open-World Attribute Mining for E-Commerce Products with Weak Supervision☆30May 9, 2022Updated 4 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- [ACL 2023] Codes and Datasets for Paper: FolkScope: Intention Knowledge Graph Construction for Discovering E-commerce Commonsense☆42Mar 3, 2025Updated last year
- ☆11Sep 7, 2021Updated 5 years ago
- ☆14Apr 18, 2020Updated 6 years ago
- Code for paper "Conversational Product Search Based on Negative Feedback"☆12Jun 26, 2020Updated 6 years ago
- K-PLUG: Knowledge-injected Pre-trained Language Model for Natural Language Understanding and Generation in E-Commerce (Findings of EMNLP …☆30Jan 6, 2023Updated 3 years ago
- Code for our EMNLP 2020 Paper "AIN: Fast and Accurate Sequence Labeling with Approximate Inference Network"☆19Nov 14, 2022Updated 3 years ago
- The implementation our EMNLP 2021 paper "Enhanced Language Representation with Label Knowledge for Span Extraction".☆115May 22, 2023Updated 3 years ago
- Large online shopping companies need to automatically populate their product descriptions supplied by the sellers. Many a times the text …☆11Jul 4, 2018Updated 8 years ago
- 基于预训练模型的中文关键词抽取方法(论文SIFRank: A New Baseline for Unsupervised Keyphrase Extraction Based on Pre-trained Language Model 的中文版代码)☆12May 17, 2020Updated 6 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- This is the official repository of the revised datasets FUNSD-r and CORD-r, introduced in EMNLP 2023 paper Reading Order Matters: Informa…☆16Mar 20, 2024Updated 2 years ago
- ☆13Sep 5, 2021Updated 5 years ago
- TER-plus Machine Translation metric.☆31May 23, 2022Updated 4 years ago
- This repository holds the annotated spreadsheet files, comprising the DECO dataset.☆13Mar 21, 2019Updated 7 years ago
- code for our WWW 2019 paper: "Open-world Learning and Application to Product Classification"☆37Apr 19, 2019Updated 7 years ago
- The 1st place solution for SIGIR 2020 E-Commerce Workshop Multimodal Product Classification Challenge☆21Aug 3, 2020Updated 6 years ago
- Learning Cross-modal Retrieval with Noisy Labels (CVPR 2021, PyTorch Code)☆13Apr 7, 2021Updated 5 years ago
- clip retrieval benchmark☆17May 4, 2022Updated 4 years ago
- Code for ACL 2019 paper "Multi-hop Reading Comprehension across Multiple Documents by Reasoning over Heterogeneous Graphs"☆18Feb 9, 2020Updated 6 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- ☆20Mar 25, 2023Updated 3 years ago
- ☆23May 25, 2022Updated 4 years ago
- High-level Semantic Feature Detection: A New Perspective for Pedestrian Detection, CVPR, 2019☆13Aug 20, 2019Updated 7 years ago
- ICLR 2019 paper: "textTOvec: DEEP CONTEXTUALIZED NEURAL AUTOREGRESSIVE TOPIC MODELS OF LANGUAGE WITH DISTRIBUTED COMPOSITIONAL PRIOR"☆25Dec 30, 2018Updated 7 years ago
- CrossWeigh: Training Named Entity Tagger from Imperfect Annotations☆177Jul 25, 2024Updated 2 years ago
- This is the repository of code and dataset for paper "The Rise of Guardians: Fact-checking URL Recommendation to Combat Fake News", SIGIR…☆18Feb 19, 2022Updated 4 years ago
- WebRED is a large and diverse manually annotated dataset for extracting relationships from a variety of text found on the World Wide Web.☆22Mar 11, 2021Updated 5 years ago
- ☆37Sep 22, 2021Updated 5 years ago
- Iterative Rank-Aware Open IE☆30Jun 24, 2019Updated 7 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- OPUS (opus.nlpl.eu) Python3 API☆18Nov 23, 2024Updated last year
- ☆31Dec 24, 2021Updated 4 years ago
- Dataset and scripts for HRDoc☆42Jun 21, 2023Updated 3 years ago
- Refined Commonsense Knowledge from Large-Scale Web Contents (TKDE 2022)☆14Oct 19, 2022Updated 3 years ago
- AI-based web extractor☆13Feb 25, 2023Updated 3 years ago
- [ACL-IJCNLP 2021] Improving Named Entity Recognition by External Context Retrieving and Cooperative Learning☆91Nov 20, 2022Updated 3 years ago
- Winner system (DAMO-NLP) of SemEval 2022 MultiCoNER shared task over 10 out of 13 tracks.☆185Jan 10, 2023Updated 3 years ago