☆17Jul 10, 2023Updated 3 years ago
Alternatives and similar repositories for LLM-Performance-Improvement-Paper
Users that are interested in LLM-Performance-Improvement-Paper are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆34Sep 14, 2024Updated last year
- TOD-Flow: Modeling the Structure of Task-Oriented Dialogues☆13Feb 7, 2024Updated 2 years ago
- ☆10May 28, 2023Updated 3 years ago
- ☆13Sep 21, 2019Updated 6 years ago
- ☆14May 23, 2022Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- MXNet finetune baseline (res152) for challenger.ai/competition/scene☆11Sep 24, 2017Updated 8 years ago
- llms related stuff , including code, docs☆13Feb 25, 2025Updated last year
- (NBCE)Naive Bayes-based Context Extension on ChatGLM-6b☆15Jun 7, 2023Updated 3 years ago
- Improving Pseudo Labels with Global-Local Denoising Framework for Cross-lingual Named Entity Recognition (IJCAI 2024)☆11Aug 18, 2024Updated last year
- ✨✨ Official repo for "Comparative Analysis of Demonstration Selection Algorithms for LLM In-Context Learning"☆16Nov 8, 2024Updated last year
- (TG'2023) Official code for the paper "Revisiting of AlphaStar" (previously called "Rethinking of AlphaStar"). It compares the raw interf…☆10Sep 6, 2021Updated 4 years ago
- (1)弹性区间标准化的旋转位置词嵌入编码器+peft LORA量化训练,提高万级tokens性能支持。(2)证据理论解释学习,提升模型的复杂逻辑推理能力(3)兼容alpaca数据格式。☆43Jul 19, 2023Updated 3 years ago
- 中文原生多层次文生视频测评基准☆18Jul 8, 2024Updated 2 years ago
- C# Windows Form Control--Support HikVision,DAHDENG IMAGING,and Basler cameras.☆14Dec 21, 2018Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆12Apr 24, 2025Updated last year
- Make wonderful CVs with web technologies.☆14Nov 5, 2021Updated 4 years ago
- Making large AI models cheaper, faster and more accessible☆15Apr 20, 2023Updated 3 years ago
- Dataset for Pinyin Regularization in Error Correction for Chinese Speech Recognition with Large Language Models in Interspeech 2024.☆16Jul 4, 2024Updated 2 years ago
- The source code and models for our paper PNP: Robust Learning from Noisy Labels by Probabilistic Noise Prediction☆14Jan 30, 2023Updated 3 years ago
- [CVPR 2024] Targeted Representation Alignment for Open-World Semi-Supervised Learning☆14Sep 23, 2024Updated last year
- ☆12Apr 23, 2018Updated 8 years ago
- ☆14Mar 1, 2023Updated 3 years ago
- [NeurIPS 2022] Latency-aware Spatial-wise Dynamic Networks☆25Aug 21, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ EMNLP 2025 Main ] Enhancing Efficiency and Exploration in Reinforcement Learning for LLMs☆18Nov 7, 2025Updated 9 months ago
- BMInf demos.☆14Oct 14, 2021Updated 4 years ago
- Look, Compare, Decide: Alleviating Hallucination in Large Vision-Language Models via Multi-View Multi-Path Reasoning☆24Sep 9, 2024Updated last year
- Source code for the NeurIPS 2023 paper: "CSOT: Curriculum and Structure-Aware Optimal Transport for Learning with Noisy Labels"☆19Dec 11, 2023Updated 2 years ago
- ☆18Dec 22, 2020Updated 5 years ago
- 《寒蝉鸣泣之时》系列简体中文汉化补丁网站☆17Jan 6, 2026Updated 7 months ago
- Udacity Self-Driving Car Engineer Nanodegree. Project: Unscented Kalman Filters☆11Apr 25, 2017Updated 9 years ago
- ☆22Apr 22, 2025Updated last year
- A comparison of pretraining framework for LLM☆22Feb 6, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A large-scale, fine-grained, diverse preference dataset (and models).☆369Dec 29, 2023Updated 2 years ago
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆19Jul 20, 2023Updated 3 years ago
- This repository collects various works that reproduce DeepSeek R1, as well as works related to DeepSeek R1 and the DeepSeek series.☆19Apr 27, 2025Updated last year
- ☆15Jun 20, 2024Updated 2 years ago
- Dialogue Action Tokens: Steering Language Models in Goal-Directed Dialogue with a Multi-Turn Planner☆31Jun 27, 2024Updated 2 years ago
- ☆15Apr 7, 2024Updated 2 years ago
- Bidirectional Autoregressive Talker from Generative Pre-trained Transformer☆39Jul 27, 2023Updated 3 years ago