Dynolog is a telemetry daemon for performance monitoring and tracing. It exports metrics from different components in the system like the linux kernel, CPU, disks, Intel PT, GPUs etc. Dynolog also integrates with pytorch and can trigger traces for distributed training applications.
☆385Oct 6, 2026Updated this week
Alternatives and similar repositories for dynolog
Users that are interested in dynolog are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A library to analyze PyTorch traces.☆561Updated this week
- A CPU+GPU Profiling library that provides access to timeline traces and hardware performance counters.☆997Sep 28, 2026Updated last week
- PArametrized Recommendation and Ai Model benchmark is a repository for development of numerous uBenchmarks as well as end to end nets for…☆156Sep 24, 2026Updated 2 weeks ago
- Meta's fleetwide profiler framework☆352Jul 7, 2026Updated 3 months ago
- NCCL Profiling Kit☆156Jul 1, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- GPUd automates monitoring, diagnostics, and issue identification for GPUs☆496Oct 2, 2026Updated last week
- CUDA checkpoint and restore utility☆498Jul 6, 2026Updated 3 months ago
- Collection of scripts to build PyTorch and the domain libraries from source.☆14Sep 18, 2026Updated 3 weeks ago
- ☆27Jun 29, 2026Updated 3 months ago
- BDC is the eBPF powered DNS caching mechanism in kernel inspired by BMC☆10May 13, 2022Updated 4 years ago
- CUPTI based GPU profiling library exposing usdt hooks☆40Updated this week
- Fault tolerance for PyTorch (HSDP, LocalSGD, DiLoCo, Streaming DiLoCo)☆542Updated this week
- [OSDI'24] Serving LLM-based Applications Efficiently with Semantic Variable☆227Updated this week
- NVIDIA Inference Xfer Library (NIXL)☆1,293Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- MSLK (Meta Superintelligence Labs Kernels) is a collection of PyTorch GPU operator libraries that are designed and optimized for GenAI tr…☆155Updated this week
- Lightweight daemon for monitoring CUDA runtime API calls with eBPF uprobes☆157Mar 29, 2025Updated last year
- ☆18May 16, 2022Updated 4 years ago
- ☆24Sep 29, 2026Updated last week
- A tool for bandwidth measurements on NVIDIA GPUs.☆785Jul 28, 2026Updated 2 months ago
- CUDA Kernel Benchmarking Library☆935Updated this week
- A Datacenter Scale Distributed Inference Serving Framework☆8,251Updated this week
- NVIDIA Data Center GPU Manager (DCGM) is a project for gathering telemetry and measuring the health of NVIDIA GPUs☆799Aug 19, 2026Updated last month
- GPU-CR: GPU Checkpoint & Restore☆34Jun 4, 2026Updated 4 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Microsoft Collective Communication Library☆400Aug 25, 2026Updated last month
- TransferBench is a utility capable of benchmarking simultaneous copies between user-specified devices (CPUs/GPUs)☆80Sep 28, 2026Updated last week
- MSCCL++: A GPU-driven communication stack for scalable AI applications☆562Updated this week
- [DEPRECATED] Moved to ROCm/rocm-systems repo☆153May 28, 2026Updated 4 months ago
- Collective communications library with various primitives for multi-machine training.☆1,458Oct 1, 2026Updated last week
- The NVIDIA® Tools Extension SDK (NVTX) is a C-based Application Programming Interface (API) for annotating events, code ranges, and resou…☆562Updated this week
- ☆14Sep 13, 2026Updated 3 weeks ago
- A low-latency & high-throughput serving engine for LLMs☆529Jan 8, 2026Updated 9 months ago
- TritonParse: A Compiler Tracer, Visualizer, and Reproducer for Triton Kernels☆229Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.☆6,746Updated this week
- Fast OS-level support for GPU checkpoint and restore☆296Sep 28, 2025Updated last year
- Userspace eBPF runtime for Observability, Network, GPU & General Extensions Framework☆1,584Updated this week
- FlashInfer: Kernel Library for LLM Serving☆6,578Updated this week
- LLTFI is a tool, which is an extension of LLFI, allowing users to run fault injection experiments on C/C++, TensorFlow and PyTorch applic…☆46Jul 9, 2026Updated 3 months ago
- Byted PyTorch Distributed for Hyperscale Training of LLMs and RLs☆1,041Mar 3, 2026Updated 7 months ago
- Optimized primitives for collective multi-GPU communication☆5,147Updated this week