chatglm_rlhf_finetuning
☆30Oct 10, 2023Updated 2 years ago
Alternatives and similar repositories for chatglm_rlhf
Users that are interested in chatglm_rlhf are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A full pipeline to finetune ChatGLM LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Huma…☆138Apr 28, 2023Updated 3 years ago
- realize the reinforcement learning training for gpt2 llama bloom and so on llm model☆27Sep 19, 2023Updated 2 years ago
- 对ChatGLM直接使用RLHF提升或降低目标输出概率|Modify ChatGLM output with only RLHF☆195May 23, 2023Updated 3 years ago
- aigc_serving lightweight and efficient Language service model reasoning☆24Jun 12, 2024Updated 2 years ago
- share data, prompt data , pretraining data☆36Nov 30, 2023Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Solution of team funny in WSDM2020☆13Jan 17, 2020Updated 6 years ago
- deep learning☆150May 6, 2025Updated last year
- moss chat finetuning☆51Apr 23, 2024Updated 2 years ago
- Evaluation for AI apps and agent☆46Jan 18, 2024Updated 2 years ago
- 🛰️ 基于真实医疗对话数据在ChatGLM上进行LoRA、P-Tuning V2、Freeze、RLHF等微调,我们的眼光不止于医疗问答☆340Sep 2, 2023Updated 2 years ago
- 基于bert的中文实体链接☆30Nov 24, 2021Updated 4 years ago
- Config files for my GitHub profile.☆11Updated this week
- ☆11Nov 16, 2019Updated 6 years ago
- clue chatyuan finetuning☆17Mar 10, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ChatGLM2-6B微调, SFT/LoRA, instruction finetune☆107Jul 19, 2023Updated 3 years ago
- Combine Tecent's bert as service model and rasa_nlu for text classification☆20Oct 29, 2022Updated 3 years ago
- ☆11Aug 10, 2022Updated 4 years ago
- chatglm 6b finetuning and alpaca finetuning☆1,526Mar 9, 2025Updated last year
- 在kaggle部署ChatGLM API,和ChatGPT api使用相同的调用方式☆14Jun 30, 2023Updated 3 years ago
- Code for the paper "Closing the Curious Case of Neural Text Degeneration"☆12Apr 9, 2025Updated last year
- Latin texts annotated for named entities and NER tagger used for the Herodotos Project (Ohio State University / Ghent University)☆12Sep 26, 2022Updated 3 years ago
- A Toolkit for Fine-Tuning Large Language Models with LoRA and DeepSpeed☆11Apr 14, 2023Updated 3 years ago
- 目前只有阅读理解赛道的☆13Mar 31, 2021Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- A Complete Self-Assessment Mobile App for Cricket Players☆27Sep 25, 2015Updated 10 years ago
- ChatGLM2-6B 全参数微调,支持多轮对话的高效微调。☆400Aug 17, 2023Updated 2 years ago
- Pytorch implementation of vision models.☆12Dec 8, 2022Updated 3 years ago
- 用于大模型 RLHF 进行人工数据标注排序的工具。A tool for manual response data annotation sorting in RLHF stage.☆254Aug 1, 2023Updated 3 years ago
- use chatGLM to perform text embedding☆45Apr 9, 2023Updated 3 years ago
- ☆14Dec 26, 2022Updated 3 years ago
- 人工精调的中文对话数据集和一段chatglm的微调代码☆1,189May 3, 2025Updated last year
- ☆15Aug 4, 2024Updated 2 years ago
- A simple generate script utils using fastchat conv template for generation of Large Language Models☆21Jun 21, 2023Updated 3 years ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- 支持rasa-nlu 的bert finetune☆46Jul 9, 2024Updated 2 years ago
- Implementation of Chinese ChatGPT☆285Nov 20, 2023Updated 2 years ago
- Files from the published Alpha Star paper by DeepMind☆18Nov 14, 2019Updated 6 years ago
- An implementation of Factoid Question Answering presented in Large-scale Simple Question Answering with Memory Networks☆15Oct 15, 2019Updated 6 years ago
- Finetuning Stable Diffusion from Diffusers☆11Mar 11, 2024Updated 2 years ago
- We introduce OpenStory++, a large-scale open-domain dataset focusing on enabling MLLMs to perform storytelling generation tasks.☆18Aug 30, 2024Updated last year
- The SMAPH system for query entity linking.☆20Jul 29, 2018Updated 8 years ago