GLM Series Edge Models
☆163Jun 12, 2025Updated last year
Alternatives and similar repositories for GLM-Edge
Users that are interested in GLM-Edge are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- fast randomized PCA for sparse data☆14Mar 14, 2023Updated 3 years ago
- Our 2nd-gen LMM☆34May 22, 2024Updated 2 years ago
- Valley is a cutting-edge multimodal large model designed to handle a variety of tasks involving text, images, video, and audio data.☆295May 8, 2026Updated 3 months ago
- ☆190Mar 13, 2026Updated 4 months ago
- Fast instruction tuning with Llama2☆10Apr 8, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- GLM-4-Voice | 端到端中英语音对话模型☆3,212Dec 5, 2024Updated last year
- [EMNLP 2024] Official repository for paper "From the Least to the Most: Building a Plug-and-Play Visual Reasoner via Data Synthesis"☆22Oct 15, 2024Updated last year
- Stable Diffusion in TensorRT 8.5+☆15Mar 19, 2023Updated 3 years ago
- Strong and Open Vision Language Assistant for Mobile Devices☆1,368Apr 15, 2024Updated 2 years ago
- ✨✨[NeurIPS 2025] VITA-1.5: Towards GPT-4o Level Real-Time Vision and Speech Interaction☆2,526Mar 28, 2025Updated last year
- GLM-4 series: Open Multilingual Multimodal Chat LMs | 开源多语言多模态对话模型☆7,073Updated this week
- An open-sourced end-to-end VLM-based GUI Agent☆1,192Apr 4, 2025Updated last year
- ☆323Sep 18, 2024Updated last year
- ☆51Oct 29, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- YOLOv5在高通AI Engine Direct环境下进行QNN量化,CPU推理的项目☆17Sep 10, 2024Updated last year
- a benckmark for evaluating logical reasoning of LLMs☆23Jan 25, 2024Updated 2 years ago
- Seed1.5-VL, a vision-language foundation model designed to advance general-purpose multimodal understanding and reasoning, achieving stat…☆1,583Jun 14, 2025Updated last year
- An acceleration library that supports arbitrary bit-width combinatorial quantization operations☆247Sep 30, 2024Updated last year
- Less is More: High-value Data Selection for Visual Instruction Tuning☆20Jan 18, 2025Updated last year
- GPT4V-level open-source multi-modal model based on Llama3-8B☆2,434Mar 3, 2025Updated last year
- 【TMM 2025🔥】 Mixture-of-Experts for Large Vision-Language Models☆2,322Jul 15, 2025Updated last year
- ☆18Dec 7, 2023Updated 2 years ago
- A tool convert TensorRT engine/plan to a fake onnx☆41Nov 22, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- MiniCPM5-1B: A SOTA 1B on-device LLM, small yet powerful.☆10,141Jul 27, 2026Updated 2 weeks ago
- Fast LLM Training CodeBase With dynamic strategy choosing [Deepspeed+Megatron+FlashAttention+CudaFusionKernel+Compiler];☆41Jan 4, 2024Updated 2 years ago
- Repository for Skill Set Optimization☆14Jul 26, 2024Updated 2 years ago
- [ICML 2023] SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models☆11Dec 13, 2023Updated 2 years ago
- Demo for Qwen2.5-VL-3B-Instruct on Axera device.☆16Sep 3, 2025Updated 11 months ago
- ☆71Jul 8, 2025Updated last year
- Web app for makeup transfer using Stable Diffusion☆10Sep 11, 2023Updated 2 years ago
- Official repository for ACL 2025 paper "Model Extrapolation Expedites Alignment"☆75May 20, 2025Updated last year
- GLM-4.6V/4.5V/4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning☆2,363Jul 21, 2026Updated 2 weeks ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [EMNLP Findings 2024] MobileQuant: Mobile-friendly Quantization for On-device Language Models☆69Sep 22, 2024Updated last year
- C++ implementation of Qwen-LM☆627Dec 6, 2024Updated last year
- ☆20Jan 7, 2024Updated 2 years ago
- Run Chinese MobileBert model on SNPE.☆15May 19, 2023Updated 3 years ago
- Gaussian Embedding of Large-scale Attributed Graphs☆10Mar 13, 2020Updated 6 years ago
- CogView4, CogView3-Plus and CogView3(ECCV 2024)☆1,100Mar 29, 2025Updated last year
- Open deep learning compiler stack for cpu, gpu and specialized accelerators☆20Updated this week