Another ChatGLM2 implementation for GPTQ quantization
☆55Oct 15, 2023Updated 2 years ago
Alternatives and similar repositories for chatglm-q
Users that are interested in chatglm-q are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆90Jun 30, 2023Updated 3 years ago
- run ChatGLM2-6B in BM1684X☆49Mar 1, 2024Updated 2 years ago
- 基于自由度(熵)、凝固度 新词发现算法实现☆12Oct 7, 2018Updated 7 years ago
- fastertransformer for codegeex model☆64Jun 6, 2023Updated 3 years ago
- A simple cycle-accurate DaDianNao simulator☆13Mar 27, 2019Updated 7 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- (1)弹性区间标准化的旋转位置词嵌入编码器+peft LORA量化训练,提高万级tokens性 能支持。(2)证据理论解释学习,提升模型的复杂逻辑推理能力(3)兼容alpaca数据格式。☆43Jul 19, 2023Updated 3 years ago
- LLaMa/RWKV onnx models, quantization and testcase☆367Jul 6, 2023Updated 3 years ago
- This sample shows how to use the oneAPI Video Processing Library (oneVPL) to perform a single and multi-source video decode and preproces…☆15Jun 15, 2023Updated 3 years ago
- Simple Structured Perceptron tagger in Python☆10May 30, 2017Updated 9 years ago
- Demo on iGPU for FFmpeg decode and scale, OpenVINO inference. this is zero-copy solution, which means No frame data copy from CPU to iGPU…☆17Jan 25, 2023Updated 3 years ago
- 本项目采用PyTorch和transformers模块实现英语序列标注,其中对BERT进行微调。☆17Feb 1, 2021Updated 5 years ago
- Python library for adding visual effects to video streams☆12Dec 20, 2019Updated 6 years ago
- chatglm-6b微调/LORA/PPO/推理, 样本为自动生成的整数/小数加减乘除运算, 可gpu/cpu☆165Aug 24, 2023Updated 3 years ago
- ☆23Aug 7, 2023Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Simulation codes for over-the-air federated learning via second-order optimization☆14Jan 27, 2022Updated 4 years ago
- A high-throughput and memory-efficient inference and serving engine for LLMs☆12Nov 14, 2025Updated 9 months ago
- Tools for easier OpenVINO development/debugging☆10Jul 16, 2025Updated last year
- A package for filtering sensitive data (parameters, keys) from a variety of JS objects☆10Feb 17, 2026Updated 6 months ago
- Asynchronous event I/O driven quantitative trading framework.☆12Aug 29, 2020Updated 6 years ago
- 纯c++的全平台llm加速库,支持python调用,支持chatglm-6B, llama, baichuan, moss基座,x86 / ARM☆13Aug 21, 2026Updated 2 weeks ago
- Echelon Blockchain Node - Cosmos SDK, IBC, and EVM compatible☆17Jan 19, 2026Updated 7 months ago
- PyTorch Implementation of FILM: Frame Interpolation for Large Motion☆25Jan 27, 2023Updated 3 years ago
- This project is targeted to detect which parking lot (actually any user defined polygon) are occupied by any object.☆12Aug 30, 2018Updated 8 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 使用peft库,对chatGLM-6B/chatGLM2-6B实现4bit的QLoRA高效微调,并做lora model和base model的merge及4bit的量化(quantize)。☆356Aug 22, 2023Updated 3 years ago
- Mujoco Playground experiment for Open Duck Mini V2☆36Feb 24, 2025Updated last year
- Firefly中文LLaMA-2大模型,支持增量预训练Baichuan2、Llama2、Llama、Falcon、Qwen、Baichuan、InternLM、Bloom等大模型☆415Oct 21, 2023Updated 2 years ago
- This the code of paper "Generative Adversarial Network Based Abnormal Behavior Detection in Massive Crowd Videos: A Hajj Case Study"☆11Jun 8, 2021Updated 5 years ago
- Adaptive floating-point based numerical format for resilient deep learning☆14Apr 11, 2022Updated 4 years ago
- C++ implementation of ChatGLM-6B & ChatGLM2-6B & ChatGLM3 & GLM4(V)☆2,962Jul 31, 2024Updated 2 years ago
- ☆17Jun 1, 2022Updated 4 years ago
- use chatGLM to perform text embedding☆45Apr 9, 2023Updated 3 years ago
- Asynchronous event I/O driven quantitative trading framework.☆20Mar 15, 2021Updated 5 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆15Aug 21, 2023Updated 3 years ago
- A corpus of short answers written by learners of English and graded with CEFR levels☆13Dec 17, 2021Updated 4 years ago
- ☆23Apr 21, 2023Updated 3 years ago
- fastllm是后端无依赖的高性能大模型推理库。同时支持张量并行推理稠密模型和混合模式推理MOE模型,任意10G以上显卡即可推理满血DeepSeek。双路9004/9005服务器+单显卡部署DeepSeek满血满精度原版模型,单并发20tps;INT4量化模型单并发30tp…☆4,954Updated this week
- Fine-tuning ChatGLM-6B with PEFT | 基于 PEFT 的高效 ChatGLM 微调☆3,713Oct 12, 2023Updated 2 years ago
- ☆11Sep 9, 2024Updated last year
- [ICML 2023] Optimizing the Collaboration Structure in Cross-Silo Federated Learning. Wenxuan Bao, Haohan Wang, Jun Wu, Jingrui He.☆21Jul 25, 2023Updated 3 years ago