Advancing the frontier of efficient AI
☆69Sep 1, 2026Updated this week
Alternatives and similar repositories for skylight-research
Users that are interested in skylight-research are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Sep 23, 2025Updated 11 months ago
- ☆24Dec 6, 2025Updated 8 months ago
- ☆66May 19, 2025Updated last year
- A benchmark for evaluating LLMs on open-ended CS problems. Exploring the Next Frontier of Computer Science.☆312Aug 21, 2026Updated 2 weeks ago
- Code for Fast-weight Product Key Memory (FwPKM)☆22Mar 18, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [AAAI26]: DS SERVE: The Largest Open Vector Store over Pretain Data; A Framework for Efficient and Scalable Neural Retrieval☆56Jan 28, 2026Updated 7 months ago
- PoC for "SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning" [NeurIPS '25]☆75Oct 2, 2025Updated 11 months ago
- [NeurIPS 2025] Multipole Attention for Efficient Long Context Reasoning☆25Dec 5, 2025Updated 8 months ago
- The official implementation of paper: SimLayerKV: A Simple Framework for Layer-Level KV Cache Reduction.☆55Oct 18, 2024Updated last year
- ⛔ DEPRECATED -- use flash-head instead (pip install flash-head)☆29Apr 10, 2026Updated 4 months ago
- MinT-2M: Long-context training system for resident-prefix GRPO☆45Jul 24, 2026Updated last month
- Accelerating Large-Scale Reasoning Model Inference with Sparse Self-Speculative Decoding☆119Dec 2, 2025Updated 9 months ago
- ☆66May 7, 2026Updated 3 months ago
- ☆28May 21, 2026Updated 3 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Systematic evaluation framework that automatically rates overthinking behavior in large language models.☆103May 16, 2025Updated last year
- ☆18Jul 1, 2025Updated last year
- Gecko Architecture☆18Jan 13, 2026Updated 7 months ago
- build and benchmark deep research☆245Mar 28, 2026Updated 5 months ago
- Official Repository of VisGym: Diverse, Customizable, Scalable Environments for Multimodal Agents☆115May 3, 2026Updated 4 months ago
- LLM KV cache compression made easy☆1,202Updated this week
- An LLM inference engine, written in C++☆20Mar 30, 2026Updated 5 months ago
- Repo for "AlphaResearch: Accelerating New Algorithm Discovery with Language Models"☆61Nov 12, 2025Updated 9 months ago
- Code repository for the SOSP'25 paper DCP: Addressing Input Dynamism In Long-Context Training via Dynamic Context Parallelism.☆22Nov 28, 2025Updated 9 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- AI-Driven Scientific, Algorithmic, and Systems Discovery☆631Updated this week
- ☆253Nov 19, 2025Updated 9 months ago
- ☆14Aug 3, 2026Updated last month
- Structured Primitives for Efficient Architecture Research☆21Dec 22, 2025Updated 8 months ago
- ☆15Apr 26, 2022Updated 4 years ago
- ☆333Jul 10, 2025Updated last year
- [ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.☆39Mar 5, 2026Updated 5 months ago
- ☆35Dec 10, 2025Updated 8 months ago
- Preview Code for Continuum Paper☆102Aug 13, 2026Updated 3 weeks ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- FeRTOS is a simple "operating system" that currently supports ARM Cortex-M CPUs☆12Jul 9, 2022Updated 4 years ago
- Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers☆44Jul 1, 2026Updated 2 months ago
- LM engine is a library for pretraining/finetuning LLMs☆190Aug 10, 2026Updated 3 weeks ago
- Fast and memory-efficient classical machine learning operators☆584Updated this week
- Compiler-R1: Towards Agentic Compiler Auto-tuning with Reinforcement Learning☆37Jul 14, 2025Updated last year
- Building the Virtuous Cycle for AI-driven LLM Systems☆281May 1, 2026Updated 4 months ago
- Fast and memory-efficient exact kmeans☆726Aug 28, 2026Updated last week