蚂蚁开源技术沙龙演讲材料归档
☆29Jun 16, 2026Updated 3 months ago
Alternatives and similar repositories for antoss-meetup
Users that are interested in antoss-meetup are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SGLang is a fast serving framework for large language models and vision language models.☆35Updated this week
- A high-throughput and memory-efficient inference and serving engine for LLMs☆11Sep 4, 2025Updated last year
- The NVIDIA Driver Manager is a Kubernetes component which assist in seamless upgrades of NVIDIA Driver on each node of the cluster.☆55Updated this week
- Tile-Based Runtime for Ultra-Low-Latency LLM Inference☆1,794Aug 13, 2026Updated last month
- Write events for TensorBoard☆13Apr 27, 2026Updated 4 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Tired of creating Helm Charts from scratch? Try Kustohelmize today!☆20Jan 20, 2025Updated last year
- [ICML 2023] SmoothQuant: Accurate and Efficient Post-Training Quantization for Large Language Models☆11Dec 13, 2023Updated 2 years ago
- Agentman: A tool for building and managing AI agents☆20Jul 10, 2025Updated last year
- Build a feature-less eBPF vm on eBPF, just for fun.☆16Mar 10, 2024Updated 2 years ago
- [VLDB 2024] Source code for FusionQuery: On-demand Fusion Queries over Multi-source Heterogeneous Data☆15Mar 11, 2025Updated last year
- Kubernetes CSI Driver for serving OCI model artifacts☆30May 25, 2026Updated 3 months ago
- accelerate generating vector by using onnx model☆18Jan 23, 2024Updated 2 years ago
- NVLink microbenchmark with IBM Power8 and NVIDIA P100 GPU - Master Thesis☆11Aug 23, 2017Updated 9 years ago
- ☆15Mar 22, 2026Updated 6 months ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Hooked CUDA-related dynamic libraries by using automated code generation tools.☆173Dec 12, 2023Updated 2 years ago
- [ICML 2025 Spotlight] ShadowKV: KV Cache in Shadows for High-Throughput Long-Context LLM Inference☆314May 1, 2025Updated last year
- A repo to learn HPC(high performance computing)☆21Sep 5, 2026Updated 2 weeks ago
- ☆13Sep 3, 2026Updated 2 weeks ago
- lessfs - An inline data deduplicating filesystem☆22May 31, 2011Updated 15 years ago
- Codebase for RecSys 2024 paper, The Elephant in the Room: Rethinking the Usage of Pre-trained Language Model in Sequential Recommendation☆19Aug 7, 2024Updated 2 years ago
- ☆19Apr 3, 2024Updated 2 years ago
- ☆15Oct 2, 2025Updated 11 months ago
- ☆23May 6, 2022Updated 4 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- 电子版书籍☆15Dec 23, 2019Updated 6 years ago
- ☆20May 8, 2024Updated 2 years ago
- The Best Open Source LLM Code Interpreter☆17Sep 2, 2023Updated 3 years ago
- I read papers, and here are my highlights.☆16Jun 7, 2020Updated 6 years ago
- ☆13Feb 16, 2021Updated 5 years ago
- MacOS Finder Sync Extension to Allow Adding Custom Actions☆14Feb 15, 2022Updated 4 years ago
- 【KDD2021】"HALO: Hierarchy-aware Fault Localization for Cloud Systems" code reproduction☆10Aug 23, 2021Updated 5 years ago
- Probabilistic Forecasting with Temporal Convolutional Neural Network☆11Nov 7, 2019Updated 6 years ago
- MacOS, Linux and Windows Clipboard Management App☆13Sep 1, 2026Updated 3 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- An adapter layer that ensures torch_musa🔦 delivers a CUDA-compatible PyTorch experience.☆42Updated this week
- Xcode MCP Server xcf is a 100% Swift based allowing you to integrate Xcode with your favorite AI IDE or MCP Client☆12Mar 30, 2026Updated 5 months ago
- Deep learning algorithms: A sparse autoencoder (and someday more algorithms), implemented in Common Lisp.☆27Jun 10, 2010Updated 16 years ago
- code for DOMI☆13Mar 24, 2023Updated 3 years ago
- Blog for dataclouds@thoughtworks.☆10Jun 19, 2016Updated 10 years ago
- Lightning-fast LLM inference engine - Built with Rust (inspiration from https://github.com/GeeeekExplorer/nano-vllm)☆38Jun 24, 2025Updated last year
- An Effective and Scalable Framework for Multimodal Search of Target Modality☆26Nov 7, 2023Updated 2 years ago