Use multi-threaded crawler to crawl the idiom data
☆14Dec 11, 2020Updated 5 years ago
Alternatives and similar repositories for Crawl
Users that are interested in Crawl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- python class for elasticsearch , including add, batch add, update, delete, query, and scan query. also with a demo that put Wikipedia in…☆17Sep 3, 2022Updated 3 years ago
- A based-bert baseline for Chinese idiom cloze test with pytorch.☆18Dec 24, 2020Updated 5 years ago
- tf-idf 模型封装类,包含计算所有文档的tf-idf值,实现了基于tf-idf搜索引擎功能。根据query,计算与每个文档的相似度,返回与query相似度最高的topk文档☆15Nov 20, 2020Updated 5 years ago
- semantic similarity, word2vec + wmd, bert+wmd, pytorch☆31Jan 29, 2024Updated 2 years ago
- ☆12Mar 26, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Datafountain-Epidemic government affairs quiz assistant competition. We divided this task into two parts: document retrieval and answer e…☆14Aug 21, 2022Updated 3 years ago
- DataFountain 疫情政务问答助手解决方案分享☆16May 2, 2020Updated 6 years ago
- 文档记录☆15Mar 16, 2021Updated 5 years ago
- Implementation of AAAI2021 paper "Writing Polishment with Simile: Task, Dataset and A Neural Approach"☆21Dec 25, 2020Updated 5 years ago
- ☆21Oct 15, 2022Updated 3 years ago
- A chinese simile recognition dataset of "Xiang".☆24Oct 5, 2022Updated 3 years ago
- Moss Vortex is a lightweight and high-performance deployment and inference backend engineered specifically for MOSS 003, providing a weal…☆37Apr 25, 2023Updated 3 years ago
- BBPE 底层实现☆38Apr 29, 2024Updated 2 years ago
- ☆129Nov 22, 2022Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ChineseBert用于中文拼写纠错☆43Mar 14, 2023Updated 3 years ago
- Reference Implementation for WSDM 2018 Paper "Hyperbolic Representation Learning for Fast and Efficient Neural Question Answering"☆68Nov 16, 2018Updated 7 years ago
- 教务管理系统javaweb项目 运行环境:window系统,Apache Tomcat v7.0.84、JDK1.8 开发环境:J2EE eclipse、navicat for mysql 运用的技术:MVC设计模式、DAO模式、Servlet、JSP、Filter、MyS…☆136Jul 12, 2023Updated 3 years ago
- ☆166May 26, 2020Updated 6 years ago
- MEASURING MASSIVE MULTITASK CHINESE UNDERSTANDING☆90Mar 24, 2024Updated 2 years ago
- 基于capsule的观点型阅读理解模型☆88Aug 8, 2019Updated 6 years ago
- SG-Net: Syntax-guided machine reading comprehension (AAAI 2020)☆83Dec 16, 2022Updated 3 years ago
- A Massive Multi-Level Multi-Subject Knowledge Evaluation benchmark☆106Jul 20, 2023Updated 3 years ago
- ChID: A Large-scale Chinese IDiom Dataset for Cloze Test☆150May 8, 2023Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Neural word segmentation with rich pretraining, code for ACL 2017 paper☆165Jan 10, 2019Updated 7 years ago
- TensorFlow code and pre-trained models for BERT and ERNIE☆147Jun 5, 2019Updated 7 years ago
- Bicleaner is a parallel corpus classifier/cleaner that aims at detecting noisy sentence pairs in a parallel corpus.☆160Jun 18, 2024Updated 2 years ago
- 法研杯2019 阅读理解赛道 top3☆151Nov 13, 2023Updated 2 years ago
- ☆344Dec 11, 2018Updated 7 years ago
- Naive Bayes-based Context Extension☆328Dec 9, 2024Updated last year
- A pytorch implementation of the ACL2019 paper "Simple and Effective Text Matching with Richer Alignment Features".☆305Aug 24, 2022Updated 3 years ago
- keras example of seq2seq, auto title☆331Dec 9, 2019Updated 6 years ago
- Python wrapper for Stanford CoreNLP.☆916Dec 7, 2021Updated 4 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆368Jul 19, 2023Updated 3 years ago
- MuCGEC中文纠错数据集及文本纠错SOTA模型开源;Code & Data for our NAACL 2022 Paper "MuCGEC: a Multi-Reference Multi-Source Evaluation Dataset for Chinese Gr…☆570Jun 9, 2023Updated 3 years ago
- XVERSE-13B: A multilingual large language model developed by XVERSE Technology Inc.☆641Apr 9, 2024Updated 2 years ago
- An implementation of TransE and its extended models for Knowledge Representation Learning on TensorFlow☆512Nov 3, 2022Updated 3 years ago
- 📦 快速转化「中文数字」和「阿拉伯数字」~ (最新特性:分数,日期、温度等转化)☆765Apr 23, 2026Updated 3 months ago
- A full Python Implementation of the ROUGE Metric (not a wrapper)☆720Nov 19, 2024Updated last year
- Code for the ICML 2023 paper "SparseGPT: Massive Language Models Can Be Accurately Pruned in One-Shot".☆891Aug 20, 2024Updated last year