中文预训练ModernBert
☆100Apr 11, 2025Updated last year
Alternatives and similar repositories for ChineseModernBert
Users that are interested in ChineseModernBert are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆53Feb 10, 2025Updated last year
- A hybrid retrieval-based question answering system built on BERT, Faiss, and ElasticSearch.☆25May 29, 2025Updated last year
- Code for KaLM-Embedding models☆118Jun 30, 2025Updated last year
- ModernVBERT is a 250M-parameter vision–language encoder that aligns a text-encoder (Ettin-150M) with a vision-encoder (SigLIP2-B) through…☆16Oct 16, 2025Updated 9 months ago
- A massively multilingual modern encoder language model☆145Jan 20, 2026Updated 6 months ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Papers about event extraction and event relation extraction☆13May 17, 2023Updated 3 years ago
- 中华经典文献数据集☆22Jun 29, 2023Updated 3 years ago
- ☆18Aug 9, 2024Updated last year
- It includes various question-answering technology sub-projects☆25Aug 23, 2025Updated 11 months ago
- 基于树形条件随机场的高阶句法分析☆16Apr 28, 2022Updated 4 years ago
- ☆30Aug 19, 2024Updated last year
- 基于Pytorch的知识蒸馏(中文文本分类)☆23Jan 12, 2023Updated 3 years ago
- ☆63Jul 21, 2024Updated 2 years ago
- The official repo of "WebExplorer: Explore and Evolve for Training Long-Horizon Web Agents"☆120Sep 29, 2025Updated 10 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Source codes for paper "BlockFFN: Towards End-Side Acceleration-Friendly Mixture-of-Experts with Chunk-Level Activation Sparsity".☆19Jan 10, 2026Updated 6 months ago
- GlimpRouter: Efficient Collaborative Inference by Glimpsing One Token of Thoughts☆16Apr 24, 2026Updated 3 months ago
- Vocabulary Trimming (VT) is a model compression technique, which reduces a multilingual LM vocabulary to a target language by deleting ir…☆67Oct 25, 2024Updated last year
- Use the tokenizer in parallel to achieve superior acceleration☆20Mar 21, 2024Updated 2 years ago
- ☆10Apr 17, 2023Updated 3 years ago
- [NeurIPS 2024] The official implementation of "Image Copy Detection for Diffusion Models"☆18Oct 1, 2024Updated last year
- Code utilized in the paper "A phenome-wide association and Mendelian randomization study for Alzheimer's disease: a prospective cohort st…☆14Jun 10, 2022Updated 4 years ago
- A Python Package for PheWAS Execution, Visualization, and Analysis☆14Mar 13, 2026Updated 4 months ago
- ☆109Jun 2, 2025Updated last year
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Unraveling the metabolic underpinnings of frailty using multicohort observational and Mendelian randomization analyses☆13May 17, 2023Updated 3 years ago
- Zeta implementation of a reusable and plug in and play feedforward from the paper "Exponentially Faster Language Modeling"☆16Nov 11, 2024Updated last year
- Official Implementation of "Probing Language Models for Pre-training Data Detection"☆20Dec 4, 2024Updated last year
- My NER Experiments with ModernBERT and Ettin☆29Jul 17, 2025Updated last year
- ☆26Jul 2, 2026Updated 3 weeks ago
- 受到self-instruct启发,除了通用LLM还能做垂直领域的小LLM实现定制效果,通过GPT获得question和answer来作为训练数据☆18May 12, 2023Updated 3 years ago
- ☆13Mar 26, 2026Updated 4 months ago
- 一种面向中文复杂问句的查询图生成方法,以及一份含有多种复杂句的中文知识图谱问答数据集☆18Mar 16, 2023Updated 3 years ago
- Python package for serving a local search engine. One command to download and serve a datastore---that's it 😎.☆26Jun 6, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- A RAG that can scale 🧑🏻💻☆11May 28, 2024Updated 2 years ago
- Reproduced the DFT method without using Verl. https://arxiv.org/abs/2508.05629☆24Oct 14, 2025Updated 9 months ago
- Generative Pretraining from Transcriptomes☆17Feb 6, 2023Updated 3 years ago
- GMEG☆33Nov 21, 2024Updated last year
- Top Picks for Data Science Self-Study: From Newbies to Pros!☆11Apr 2, 2024Updated 2 years ago
- Diffusion-generated Facial Forgery Dataset☆58Apr 7, 2026Updated 3 months ago
- GISTEmbed: Guided In-sample Selection of Training Negatives for Text Embeddings☆45Mar 6, 2024Updated 2 years ago