Lab 5 project of MIT-6.5940, deploying LLaMA2-7B-chat on one's laptop with TinyChatEngine.
☆18Dec 1, 2023Updated 2 years ago
Alternatives and similar repositories for LLaMA2-7B-on-laptop
Users that are interested in LLaMA2-7B-on-laptop are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Structured Pruning Adapters in PyTorch☆19Aug 30, 2023Updated 2 years ago
- Code release for AdapMoE accepted by ICCAD 2024☆38Apr 28, 2025Updated 11 months ago
- channel pruning for accelerating very deep neural networks☆13Mar 8, 2021Updated 5 years ago
- ☆180Aug 9, 2023Updated 2 years ago
- Code for Adaptive Deep Neural Network Inference Optimization with EENet☆12Mar 28, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- ☆10Feb 7, 2022Updated 4 years ago
- In this project, we provide a strong template for a PyTorch project. The purpose of this repository is to provide an example (and strong …☆15Apr 21, 2021Updated 4 years ago
- [ICLR 2023] PyTorch code for DFPC: Data flow driven pruning of coupled channels without data.☆15Aug 25, 2023Updated 2 years ago
- [OSDI 2025] DecDEC: A Systems Approach to Advancing Low‑Bit LLM Quantization☆23Jan 29, 2026Updated 2 months ago
- ☆77Nov 5, 2024Updated last year
- An implementation of Distortion-Free Wide-Angle Portraits on Camera Phones☆10Dec 24, 2019Updated 6 years ago
- Rust bindings for SPDK☆12Mar 5, 2020Updated 6 years ago
- Source code of our TNNLS paper "Boosting Convolutional Neural Networks with Middle Spectrum Grouped Convolution"☆12Apr 14, 2023Updated 3 years ago
- LeetGPU Solutions☆114Oct 9, 2025Updated 6 months ago
- Simple, predictable pricing with DigitalOcean hosting • AdAlways know what you'll pay with monthly caps and flat pricing. Enterprise-grade infrastructure trusted by 600k+ customers.
- [NeurIPS 2024] Search for Efficient LLMs☆16Jan 16, 2025Updated last year
- A basic deep learning library, comparable to a very minimal version of PyTorch.☆19Mar 1, 2023Updated 3 years ago
- This is a repository of coursework project for the Stanford Compilers MOOC course. The result is a fully-working compiler for the COOL Pr…☆18Sep 11, 2023Updated 2 years ago
- A cpp threadpool for c++11 c++14 c++17 c++20☆15Jun 30, 2023Updated 2 years ago
- ☆12Jun 12, 2025Updated 10 months ago
- a port forwarding tool similar to lcx☆10Mar 14, 2019Updated 7 years ago
- TinyML and Efficient Deep Learning Computing☆20Apr 26, 2024Updated last year
- ☆15Feb 1, 2016Updated 10 years ago
- code for the paper "A Statistical Framework for Low-bitwidth Training of Deep Neural Networks"☆29Oct 31, 2020Updated 5 years ago
- Deploy open-source AI quickly and easily - Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- 🦙🦙.🦀☆28Sep 24, 2023Updated 2 years ago
- Implementation of the paper : Not all attention is needed - Gated Attention Network for Sequence Data (GA-Net) [https://arxiv.org/abs/191…☆13Aug 20, 2020Updated 5 years ago
- A docker image for One Student One Chip's debug exam☆10Sep 22, 2023Updated 2 years ago
- 南京大学软件分析作业☆14Jul 30, 2022Updated 3 years ago
- A fast implementation of Leiserchess AI for MIT 6.172`16 http://scrimmage.csail.mit.edu/☆12Dec 22, 2016Updated 9 years ago
- Leaderboard implementations for datasets produced by the Mosaic Team.☆20Jul 6, 2023Updated 2 years ago
- Sirius, an efficient correction mechanism, which significantly boosts Contextual Sparsity models on reasoning tasks while maintaining its…☆21Sep 10, 2024Updated last year
- 哈尔滨工业大学(深圳)2021年球季学期深度学习体系结构实验☆17Oct 1, 2022Updated 3 years ago
- ☆10Nov 14, 2023Updated 2 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Asynchronous Rust bindings for SPDK.☆18Nov 1, 2022Updated 3 years ago
- 2023秋PKU编译原理lab,以及Koopa IR C++接口的文档☆16Feb 12, 2024Updated 2 years ago
- 基于Tensorflow2卷积神经网络即插即用模块实现☆11Dec 13, 2022Updated 3 years ago
- CUDA_C编程权威指南示例代码☆13Mar 22, 2023Updated 3 years ago
- ☆58May 4, 2024Updated last year
- MLIR dialect for libgccjit☆23Dec 3, 2024Updated last year
- sample grpc client in java for use in jmeter for performance testing of grpc API☆25Jan 30, 2018Updated 8 years ago