Use multi-threaded crawler to crawl the idiom data
☆14Dec 11, 2020Updated 5 years ago
Alternatives and similar repositories for Crawl
Users that are interested in Crawl are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- python class for elasticsearch , including add, batch add, update, delete, query, and scan query. also with a demo that put Wikipedia in…☆17Sep 3, 2022Updated 3 years ago
- A based-bert baseline for Chinese idiom cloze test with pytorch.☆18Dec 24, 2020Updated 5 years ago
- tf-idf 模型封装类,包含计算所有文档的tf-idf值,实现了基于tf-idf搜索引擎功能。根据query,计算与每个文档的相似度,返回与query相似度最高的topk文档☆15Nov 20, 2020Updated 5 years ago
- semantic similarity, word2vec + wmd, bert+wmd, pytorch☆31Jan 29, 2024Updated 2 years ago
- ☆12Mar 26, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Datafountain-Epidemic government affairs quiz assistant competition. We divided this task into two parts: document retrieval and answer e…☆14Aug 21, 2022Updated 3 years ago
- DataFountain 疫情政务问答助手解决方案分享☆16May 2, 2020Updated 6 years ago
- Implementation of AAAI2021 paper "Writing Polishment with Simile: Task, Dataset and A Neural Approach"☆21Dec 25, 2020Updated 5 years ago
- ☆22Oct 15, 2022Updated 3 years ago
- run chatglm3-6b in BM1684X☆38Mar 1, 2024Updated 2 years ago
- BBPE 底层实现☆38Apr 29, 2024Updated 2 years ago
- ChineseBert用于中文拼写纠错☆43Mar 14, 2023Updated 3 years ago
- [AAAI 2024] LLMEval Phase II dataset — professional domain evaluation across 12 academic disciplines☆71May 21, 2026Updated 2 months ago
- MEASURING MASSIVE MULTITASK CHINESE UNDERSTANDING☆90Mar 24, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 基于capsule的观点型阅读理解模型☆88Aug 8, 2019Updated 7 years ago
- SG-Net: Syntax-guided machine reading comprehension (AAAI 2020)☆82Dec 16, 2022Updated 3 years ago
- ☆98Dec 5, 2023Updated 2 years ago
- Blog of programming☆351Oct 30, 2022Updated 3 years ago
- 科赛网-莱斯杯:全国第二届“军事智能机器阅读”挑战赛 前十团队PPT文档代码总结☆131Feb 5, 2020Updated 6 years ago
- Neural word segmentation with rich pretraining, code for ACL 2017 paper☆165Jan 10, 2019Updated 7 years ago
- TensorFlow code and pre-trained models for BERT and ERNIE☆147Jun 5, 2019Updated 7 years ago
- Bicleaner is a parallel corpus classifier/cleaner that aims at detecting noisy sentence pairs in a parallel corpus.☆159Jun 18, 2024Updated 2 years ago
- 法研杯2019 阅读理解赛道 top3☆151Nov 13, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Naive Bayes-based Context Extension☆328Dec 9, 2024Updated last year
- A pytorch implementation of the ACL2019 paper "Simple and Effective Text Matching with Richer Alignment Features".☆305Aug 24, 2022Updated 3 years ago
- Python wrapper for Stanford CoreNLP.☆916Dec 7, 2021Updated 4 years ago
- ☆368Jul 19, 2023Updated 3 years ago
- KgCLUE: 大规模中文开源知识图谱问答☆457Jul 5, 2022Updated 4 years ago
- 以词为基本单位的中文BERT☆475Nov 18, 2021Updated 4 years ago
- Reject complicated operations for incorporating lexicon for Chinese NER.☆437Jan 22, 2022Updated 4 years ago
- C++ implementation of Qwen-LM☆628Dec 6, 2024Updated last year
- A prize for finding tasks that cause large language models to show inverse scaling☆622Oct 11, 2023Updated 2 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- [ICML 2024] SqueezeLLM: Dense-and-Sparse Quantization☆724Aug 13, 2024Updated 2 years ago
- XVERSE-13B: A multilingual large language model developed by XVERSE Technology Inc.☆641Apr 9, 2024Updated 2 years ago
- 🎯🗯 Dataset generation for AI chatbots, NLP tasks, named entity recognition or text classification models using a simple DSL!☆889Sep 3, 2023Updated 2 years ago
- FP16xINT4 LLM inference kernel that can achieve near-ideal ~4x speedups up to medium batchsizes of 16-32 tokens.☆1,130Sep 4, 2024Updated last year
- LongBench v2 and LongBench (ACL 25'&24')☆1,225Jan 15, 2025Updated last year
- Must-read papers on Machine Reading Comprehension☆888Jul 9, 2020Updated 6 years ago
- A Tensorflow implementation of QANet for machine reading comprehension☆985May 30, 2018Updated 8 years ago