baichuan and baichuan2 finetuning and alpaca finetuning
☆32Mar 10, 2025Updated last year
Alternatives and similar repositories for baichuan_finetuning
Users that are interested in baichuan_finetuning are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- qwen models finetuning☆106Mar 9, 2025Updated last year
- ChatGLM2-6B微调, SFT/LoRA, instruction finetune☆106Jul 19, 2023Updated 3 years ago
- 基于serverless实现的《图片艺术化应用》☆10Sep 8, 2020Updated 6 years ago
- 基于大模型ChatGLM,微调方式为LORA,集SFT、RM、PPO算法为一体项目☆14Jun 20, 2023Updated 3 years ago
- Implementation of StyleTTS for Mandarin☆11Jun 22, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Codes and data for CIKM 2022 paper "RuDi: Explaining Behavior Sequence Models by Automatic Statistics Generation and Rule Distillation"☆12Aug 16, 2022Updated 4 years ago
- Xcbwin - a simple C++ class for graphical outputs using XCB☆12May 12, 2015Updated 11 years ago
- Enhancing LangChain prompts to work better with RWKV models☆32May 30, 2023Updated 3 years ago
- ☆13Jun 17, 2024Updated 2 years ago
- ☆15Jun 6, 2023Updated 3 years ago
- This repository open-sources our GEC system submitted by THU KELab (sz) in the CCL2023-CLTC Track 1: Multidimensional Chinese Learner Tex…☆15Nov 25, 2023Updated 2 years ago
- Code & Data for our Paper "NaSGEC: Multi-Domain Chinese Grammatical Error Correction for Native Speaker Texts" (ACL 2023 Findings)☆97Feb 18, 2025Updated last year
- deep version SentiBank☆12Dec 16, 2014Updated 11 years ago
- 本项目利用深度学习技术,实时检测人体3D姿态,并基于此预测未来人体动作。采用mmpose框架与多进程技术实现后端快速预测,利用混合现实Hololens2头戴显示器显示人物动作,做到实时抓取,实时预测,实时显示。☆12Oct 30, 2023Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- [EMNLP 2025 Findings] Retrieval-Augmented Machine Translation with Unstructured Knowledge☆15Sep 4, 2025Updated last year
- ☆12Dec 22, 2024Updated last year
- Source code for "BLOOM-Net: Blockwise Optimization for Masking Networks Toward Scalable and Efficient Speech Enhancement"☆14Feb 13, 2022Updated 4 years ago
- AMI and ICSI Corpora in JSON format.☆37Sep 29, 2023Updated 2 years ago
- pytorch版损失函数,改写自科学空间文章,【通过互信息思想来缓解类别不平衡问题】、【将“softmax+交叉熵”推广到多标签分类问题】☆12Aug 22, 2021Updated 5 years ago
- A Language-consistent Open Relation Extraction Model.☆16Mar 24, 2023Updated 3 years ago
- Yet another Bloomfilter implementation in Python, compatible with Java's Guava library☆12Aug 10, 2024Updated 2 years ago
- GAIIC2022商品标题实体识别Baseline,使用GlobalPointer实现,线上0.80349☆52Apr 9, 2022Updated 4 years ago
- 基于c++ muduo网络库的集群聊天服务器,使用nginx实现负载均衡,使用reids消息队列实现跨服务器通信☆13Feb 23, 2024Updated 2 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- the classic abstract strategy game☆13Jan 9, 2011Updated 15 years ago
- Rescoring methods for end-to-end Automatic Speech Recognition☆27Sep 23, 2020Updated 6 years ago
- The repository to keep supporting files for my blog posts.☆16Sep 20, 2025Updated last year
- [ICASSP 2022] AISHELL-NER: Named Entity Recognition from Chinese Speech☆27Apr 20, 2022Updated 4 years ago
- The code and data for the paper "Lost-in-the-Middle in Long-Text Generation: Synthetic Dataset, Evaluation Framework, and Mitigation"☆15Oct 8, 2025Updated 11 months ago
- Baichuan-13B 指令微调☆87Jul 14, 2023Updated 3 years ago
- ☆17Nov 3, 2024Updated last year
- 简单问答系统优化版本☆12Dec 10, 2019Updated 6 years ago
- 一个基于trie树的具有联想功能的文本编辑器。采用python和pyqt☆10Sep 7, 2016Updated 10 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 自然语言处理之中文文本分类(以垃圾短信识别为例)☆24Jun 4, 2020Updated 6 years ago
- lightgbm☆14Jun 21, 2022Updated 4 years ago
- Dynamic vision-guided speaker embedding for audio-visual speaker diarization☆12Jul 5, 2022Updated 4 years ago
- Code and data for the paper "Steering Conversational Large Language Models for Long Emotional Support Conversations" along with a UI to v…☆15Apr 14, 2025Updated last year
- MSU: CSE 251 Programming in C☆13Oct 1, 2022Updated 3 years ago
- [ICLR26] An open-source project dedicated to training LRMs to achieve "just-enough" thinking☆18Mar 25, 2026Updated 6 months ago
- NS3 simulator for RDMA load balancing☆12Jan 31, 2025Updated last year