Longitudinal Evaluation of LLMs via Data Compression
☆32May 29, 2024Updated 2 years ago
Alternatives and similar repositories for llm-compressive
Users that are interested in llm-compressive are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Latest Evaluation Toolkit (LatestEval). Assessing the language models with latest, uncontaminated materials.☆29Feb 17, 2025Updated last year
- Implementation for IceFormer: Accelerated Inference with Long-Sequence Transformers on CPUs (ICLR 2024).☆25Jun 9, 2026Updated last month
- A high-throughput and memory-efficient inference and serving engine for LLMs☆17Jun 3, 2024Updated 2 years ago
- Download full or partial git-lfs repos without temporarily using 2x disk space☆32Oct 13, 2023Updated 2 years ago
- An approximate implementation of the OpenAI paper - An Empirical Model of Large-Batch Training for MNIST☆11Nov 19, 2022Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A chinese simile recognition dataset of "Xiang".☆24Oct 5, 2022Updated 3 years ago
- Code for paper 'Are We Falling in a Middle-Intelligence Trap? An Analysis and Mitigation of the Reversal Curse'☆14Aug 2, 2024Updated last year
- Quantized Attention on GPU☆45Nov 22, 2024Updated last year
- QAQ: Quality Adaptive Quantization for LLM KV Cache☆55Mar 27, 2024Updated 2 years ago
- ☆77Feb 22, 2024Updated 2 years ago
- Research without Re-search: Maximal Update Parametrization Yields Accurate Loss Prediction across Scales☆32Jul 17, 2023Updated 3 years ago
- Code for EMNLP'24 paper - On Diversified Preferences of Large Language Model Alignment☆16Aug 6, 2024Updated last year
- ☆32May 26, 2024Updated 2 years ago
- my poc☆16Oct 28, 2020Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Implementation of OpenAI paper with Simple Noise Scale on Fastai V2☆19Apr 16, 2021Updated 5 years ago
- VibeRL is a Reinforcement Learning framework built essentially through vibe coding with Kimi K2.☆17Jul 20, 2026Updated last week
- Mini Model Daemon☆13Nov 9, 2024Updated last year
- 使用OpenCV部署CoupledTPS,包含了肖像矫正,不规则边界的图像矩形化,旋转图像矫正,三个模型。依然是包含C++和Python两个版本的程序☆21Jul 4, 2024Updated 2 years ago
- WIKIGENBENCH: Exploring Full-length Wikipedia Generation under Real-World Scenario (COLING 2025)☆13Jan 5, 2025Updated last year
- ☆11Apr 3, 2023Updated 3 years ago
- Training code repo of the paper "DeepDance: Music-to-Dance Motion Choreography with Adversarial Learning"☆11May 18, 2021Updated 5 years ago
- MMM 2021: Crossed-Time Delay Neural Network for Speaker Recognition☆11Dec 4, 2021Updated 4 years ago
- MNBVC项目-ShareGPT语料清洗☆16Oct 4, 2023Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Training a reward model for RLHF using RWKV.☆15Jun 5, 2023Updated 3 years ago
- test images with not appropriate labels in MNIST dataset☆10Mar 3, 2018Updated 8 years ago
- aigc evals☆10Dec 2, 2023Updated 2 years ago
- 英文文献的《中国图书馆分类法》自动标注小程序☆13Oct 29, 2024Updated last year
- RWKV-7 mini☆12Mar 29, 2025Updated last year
- Korean Abstract Meaning Representation (AMR) Corpus☆10Feb 27, 2022Updated 4 years ago
- Code for ACL22 short Paper "Hierarchical Curriculum Learning for AMR Parsing"☆13Jun 1, 2022Updated 4 years ago
- Direct Preference Optimization for RWKV, aiming for RWKV-5 and 6.☆11Mar 1, 2024Updated 2 years ago
- Source code of our paper "Focus on the Target’s Vocabulary: Masked Label Smoothing for Machine Translation" @ ACL 2022☆13Apr 13, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Decoding Attention is specially optimized for MHA, MQA, GQA and MLA using CUDA core for the decoding stage of LLM inference.☆48Jun 11, 2025Updated last year
- GoldFinch and other hybrid transformer components☆16Dec 9, 2025Updated 7 months ago
- Code for MERMAID : Metaphor Generation with Symbolism and Discriminative Decoding☆11May 2, 2022Updated 4 years ago
- Source code for paper "On the Pareto Front of Multilingual Neural Machine Translation" @ NeurIPS 2023☆17Sep 27, 2023Updated 2 years ago
- Utilize BERT model for multi task including ABSA (aspect based sentiment analysis) task and AE (Aspect Extraction) task☆10May 31, 2019Updated 7 years ago
- ☆14Jul 13, 2025Updated last year
- 清华大学学生健康和出行情况报告每日自动提交☆14Jan 30, 2021Updated 5 years ago