A more efficient GLM implementation!
☆54Feb 18, 2023Updated 3 years ago
Alternatives and similar repositories for one-glm
Users that are interested in one-glm are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- TensorRT☆11Sep 22, 2020Updated 5 years ago
- OPD: Chinese Open-Domain Pre-trained Dialogue Model☆73Jun 5, 2023Updated 3 years ago
- Transformer related optimization, including BERT, GPT☆39Feb 10, 2023Updated 3 years ago
- A toolkit for developers to simplify the transformation of nn.Module instances. It's now corresponding to Pytorch.fx.☆13Apr 7, 2023Updated 3 years ago
- a single-header math library☆17Nov 7, 2025Updated 8 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A multi-task learning approach for conditioned response generation (NAACL 2021)☆12Nov 18, 2022Updated 3 years ago
- CUDA 12.2 HMM demos☆21Jul 26, 2024Updated 2 years ago
- [CVPR-2023] Towards Any Structural Pruning☆18Apr 27, 2023Updated 3 years ago
- [ICML 2023] "Data Efficient Neural Scaling Law via Model Reusing" by Peihao Wang, Rameswar Panda, Zhangyang Wang☆14Jan 4, 2024Updated 2 years ago
- A Structured Span Selector (NAACL 2022). A structured span selector with a WCFG for span selection tasks (coreference resolution, semanti…☆21Jul 11, 2022Updated 4 years ago
- ☆13Apr 15, 2024Updated 2 years ago
- A Framework for Multimodal Parsing, Contextual Narration, and Hierarchical Labeling of ESG Reports☆16Nov 14, 2025Updated 8 months ago
- Chrome extension for OA sites like arxiv, openreivew: 1. PDF back to abstract page, 2. Rename PDF page with paper title.☆18Oct 12, 2023Updated 2 years ago
- ⚡ boost inference speed of GPT models in transformers by onnxruntime☆51Aug 20, 2023Updated 2 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- The code implementation of HyGRAG, accepted by WWW'26.☆15May 31, 2026Updated last month
- [Findings of ACL 2022] Meta-Path Guided Contrastive Learning for Logical Reasoning of Text☆28Apr 8, 2026Updated 3 months ago
- Code for paper "Dependency-based Mixture Language Models" by Zhixian Yang, and Xiaojun Wan. This paper is accepted by ACL 2022 Main Confe…☆26May 27, 2022Updated 4 years ago
- Code for the paper: https://arxiv.org/pdf/2309.06979.pdf☆21Jul 29, 2024Updated 2 years ago
- ☆36Nov 22, 2024Updated last year
- LiBai(李白): A Toolbox for Large-Scale Distributed Parallel Training☆403Jul 31, 2025Updated 11 months ago
- LORA微调BLOOMZ,参考BELLE☆25Mar 24, 2023Updated 3 years ago
- Template Filling with Generative Transformers☆22Jun 8, 2021Updated 5 years ago
- ☆20Jul 8, 2026Updated 3 weeks ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- 使用bert训练MRPC数据集,写成API接口模式以及简易的html界面☆21Jun 30, 2019Updated 7 years ago
- Pretrain CPM-1☆53Apr 20, 2021Updated 5 years ago
- End-to-End Speech Processing Toolkit☆16Jan 20, 2025Updated last year
- Code and data for COLING 2022 paper titled "Structural Bias For Aspect Sentiment Triplet Extraction"☆26May 28, 2023Updated 3 years ago
- ☆31May 13, 2024Updated 2 years ago
- Resources for our IJCAI 2020 paper, TopicKA: Generating Commonsense Knowledge-Aware Dialogue Responses Towards the Recommended Topic Fact☆12Nov 30, 2020Updated 5 years ago
- Source code for "Domain-Aware Dialogue State Tracker for Multi-Domain Dialogue Systems"☆10Oct 5, 2020Updated 5 years ago
- Ongoing research training transformer language models at scale, including: BERT & GPT-2☆69Jul 20, 2023Updated 3 years ago
- Performed document clustering using the DBSCAN clustering algorithm☆14Oct 21, 2020Updated 5 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- transformers implement (architecture, task example, serving and more)☆96Mar 23, 2022Updated 4 years ago
- alpaca中文指令微调数据集☆395Mar 26, 2023Updated 3 years ago
- GLM (General Language Model)☆3,622Nov 3, 2023Updated 2 years ago
- ☆14Aug 28, 2018Updated 7 years ago
- A Slot-filling based Dialog Manager for Task-oriented Bot☆13Dec 29, 2016Updated 9 years ago
- ☆44Mar 29, 2023Updated 3 years ago
- The official codes for "Aurora: Activating chinese chat capability for Mixtral-8x7B sparse Mixture-of-Experts through Instruction-Tuning"☆261May 9, 2024Updated 2 years ago