Advancing the frontier of efficient AI
☆69Sep 16, 2026Updated last week
Alternatives and similar repositories for skylight-research
Users that are interested in skylight-research are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆18Sep 23, 2025Updated last year
- ☆24Dec 6, 2025Updated 9 months ago
- ☆68May 19, 2025Updated last year
- A benchmark for evaluating LLMs on open-ended CS problems. Exploring the Next Frontier of Computer Science.☆318Sep 13, 2026Updated last week
- Code for Fast-weight Product Key Memory (FwPKM)☆25Mar 18, 2026Updated 6 months ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [AAAI26]: DS SERVE: The Largest Open Vector Store over Pretain Data; A Framework for Efficient and Scalable Neural Retrieval☆57Jan 28, 2026Updated 7 months ago
- PoC for "SpecReason: Fast and Accurate Inference-Time Compute via Speculative Reasoning" [NeurIPS '25]☆75Oct 2, 2025Updated 11 months ago
- [NeurIPS 2025] Multipole Attention for Efficient Long Context Reasoning☆26Dec 5, 2025Updated 9 months ago
- The official implementation of paper: SimLayerKV: A Simple Framework for Layer-Level KV Cache Reduction.☆55Oct 18, 2024Updated last year
- How to Train Your Advisor: Steering Black-Box LLMs with Advisor Models☆87Feb 5, 2026Updated 7 months ago
- ⛔ DEPRECATED -- use flash-head instead (pip install flash-head)☆29Apr 10, 2026Updated 5 months ago
- Accelerating Large-Scale Reasoning Model Inference with Sparse Self-Speculative Decoding☆121Dec 2, 2025Updated 9 months ago
- ☆66May 7, 2026Updated 4 months ago
- MinT-2M: Long-context training system for resident-prefix GRPO☆46Jul 24, 2026Updated 2 months ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆30May 21, 2026Updated 4 months ago
- Experiments Notebook of "Understanding the Skill Gap in Recurrent Language Models: The Role of the Gather-and-Aggregate Mechanism"☆17Apr 30, 2025Updated last year
- Systematic evaluation framework that automatically rates overthinking behavior in large language models.☆103May 16, 2025Updated last year
- Simulator for comparing memory allocation policies for caches.☆20May 15, 2019Updated 7 years ago
- ☆18Jul 1, 2025Updated last year
- build and benchmark deep research☆247Mar 28, 2026Updated 5 months ago
- Official Repository of VisGym: Diverse, Customizable, Scalable Environments for Multimodal Agents☆115May 3, 2026Updated 4 months ago
- LLM KV cache compression made easy☆1,214Updated this week
- An LLM inference engine, written in C++☆20Mar 30, 2026Updated 5 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Repo for "AlphaResearch: Accelerating New Algorithm Discovery with Language Models"☆61Nov 12, 2025Updated 10 months ago
- ☆255Nov 19, 2025Updated 10 months ago
- ☆14Updated this week
- AI-Driven Scientific, Algorithmic, and Systems Discovery☆687Sep 14, 2026Updated last week
- Structured Primitives for Efficient Architecture Research☆21Dec 22, 2025Updated 9 months ago
- ☆16Sep 25, 2025Updated 11 months ago
- Code release for "MDQE: Mining Discriminative Query Embeddings to Segment Occluded Instances on Challenging Videos"(CVPR2023)☆14Dec 14, 2023Updated 2 years ago
- Code for "SePPO: Semi-Policy Preference Optimization for Diffusion Alignment."☆18Oct 7, 2024Updated last year
- ☆334Jul 10, 2025Updated last year
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- ☆35Dec 10, 2025Updated 9 months ago
- [ICML 2026] Code for V1: Unifying Generation and Self-Verification for Parallel Reasoners.☆39Mar 5, 2026Updated 6 months ago
- Preview Code for Continuum Paper☆107Aug 13, 2026Updated last month
- Piecewise Sparse Attention Is Wiser for Efficient Diffusion Transformers☆44Jul 1, 2026Updated 2 months ago
- LM engine is a library for pretraining/finetuning LLMs☆193Updated this week
- Compiler-R1: Towards Agentic Compiler Auto-tuning with Reinforcement Learning☆37Jul 14, 2025Updated last year
- Fast and memory-efficient classical machine learning operators☆595Aug 31, 2026Updated 3 weeks ago