run ChatGLM2-6B in BM1684X
☆49Mar 1, 2024Updated 2 years ago
Alternatives and similar repositories for ChatGLM2-TPU
Users that are interested in ChatGLM2-TPU are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A whisper repo for TPU☆11Jun 4, 2024Updated 2 years ago
- Sophgo AI chips driver and runtime library.☆26Updated this week
- run chatglm3-6b in BM1684X☆38Mar 1, 2024Updated 2 years ago
- Text2speech & tone color conversion demo running on SG2300x 结合openvoice和emotivoice的TTS+即时克隆☆22Oct 30, 2024Updated last year
- Run generative AI models in sophgo BM1684X/BM1688☆293Updated this week
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- DETR tensor去除推理过程无用辅助头+fp16部署再次加速+解决转tensorrt 输出全为0问题的新方法。☆12Jan 9, 2024Updated 2 years ago
- ☆12Dec 16, 2021Updated 4 years ago
- Machine learning compiler based on MLIR for Sophgo TPU.☆952Updated this week
- simplify >2GB large onnx model☆72Nov 30, 2024Updated last year
- For 2022 Nvidia Hackathon☆22Jun 28, 2022Updated 4 years ago
- 使用onnxruntime部署夜间雾霾图像的可见度增强,包含C++和Python两个版本的程序☆13Feb 17, 2024Updated 2 years ago
- PyTorch in Go, using LibTorch.☆15May 21, 2019Updated 7 years ago
- ☆31Jun 2, 2022Updated 4 years ago
- NVIDIA® TensorRT™, an SDK for high-performance deep learning inference, includes a deep learning inference optimizer and runtime that del…☆26Jul 21, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- OpenVINO™ optimization for PointPillars*☆32May 5, 2025Updated last year
- cvitek ai compiler base on MLIR☆23Mar 14, 2022Updated 4 years ago
- Python scripts performing Open Vocabulary Object Detection using the YOLO-World model in ONNX. And Export the ONNX model for AXera's NPU☆12Aug 11, 2025Updated 11 months ago
- Bert TensorRT模型加速部署☆10Apr 1, 2022Updated 4 years ago
- ☆53Mar 27, 2023Updated 3 years ago
- Examples for SophonSDK☆107Aug 11, 2022Updated 3 years ago
- CASTER: Predicting Drug Interactions with Chemical Substructure Representation (AAAI 2020)☆25Oct 28, 2020Updated 5 years ago
- ffmpeg+cuvid+tensorrt+multicamera☆12Dec 31, 2024Updated last year
- h264的软解和硬解,基于FFmpeg和MPP☆11Mar 23, 2022Updated 4 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- TensorRT-FastSAM(https://github.com/CASIA-IVA-Lab/FastSAM)☆23Feb 29, 2024Updated 2 years ago
- SAM and lama inpaint,包含QT的GUI交互界面,实现了交互式可实时显示结果的画点、画框进行SAM,然后通过进行Inpaint,具体操作看readme里的视频。☆54Jan 30, 2024Updated 2 years ago
- Multiple Lidar preprocessor for BEVfusion☆11Aug 25, 2023Updated 2 years ago
- PointPillars TensorRT version pretrained on MMDetection3d with WaymoOpenDataset☆23Aug 11, 2022Updated 3 years ago
- ☆26Feb 2, 2024Updated 2 years ago
- 🎉My Collections of CUDA Kernels~☆11Jun 25, 2024Updated 2 years ago
- ChatTTS is a generative speech model for daily dialogue.☆14Oct 21, 2024Updated last year
- ☆36Mar 29, 2023Updated 3 years ago
- 基于 CUDA Driver API 的 cuda 运行时环境☆16Jul 30, 2025Updated 11 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- pose estimation code with deepstream and yolo-pose☆13Oct 14, 2022Updated 3 years ago
- Model Quantization Benchmark☆20Apr 17, 2026Updated 3 months ago
- stable diffusion using mnn☆68Sep 28, 2023Updated 2 years ago
- This is a TensorRT based deepsort project☆79Sep 24, 2021Updated 4 years ago
- ☆61Nov 21, 2024Updated last year
- ☆47Jun 30, 2026Updated 3 weeks ago
- ☆15Apr 18, 2023Updated 3 years ago