☆72Aug 6, 2025Updated last year
Alternatives and similar repositories for gpt-oss-reverse-engineering
Users that are interested in gpt-oss-reverse-engineering are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution☆105Sep 24, 2025Updated 10 months ago
- [NeurIPS 2022] Your Transformer May Not be as Powerful as You Expect (official implementation)☆35Aug 6, 2023Updated 3 years ago
- Aims for memory-efficient training (24GB VRAM) on consumer GPUs. Optimizing language models through guidance tokens in reasoning chains, …☆28Feb 23, 2025Updated last year
- ☆37Aug 7, 2025Updated last year
- ☆32Jul 2, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆40Dec 14, 2025Updated 7 months ago
- Efficient Scaling laws and collaborative pretraining.☆23Jul 19, 2026Updated 3 weeks ago
- Implementation of the paper Knowledge-Enhanced Dual-stream Zero-shot Composed Image Retrieval (CVPR 2024)☆21Nov 4, 2024Updated last year
- Simple (fast) transformer inference in PyTorch with torch.compile + lit-llama code☆11Aug 29, 2023Updated 2 years ago
- Accelerate LLM preference tuning via prefix sharing with a single line of code☆52Jul 4, 2025Updated last year
- Code for "Towards Revealing the Mystery behind Chain of Thought: a Theoretical Perspective"☆21Jul 16, 2023Updated 3 years ago
- ☆17Jun 24, 2024Updated 2 years ago
- Understanding the correlation between different LLM benchmarks☆30Jan 11, 2024Updated 2 years ago
- Homepage for ProLong (Princeton long-context language models) and paper "How to Train Long-Context Language Models (Effectively)"☆263Sep 12, 2025Updated 11 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆17Aug 23, 2025Updated 11 months ago
- Benchmark tests supporting the TiledCUDA library.☆19Nov 19, 2024Updated last year
- AI conference deadline countdowns + Calendar overview with deadlines and conference dates.☆16Aug 15, 2024Updated last year
- Adversarially Robust Generalization Just Requires More Unlabeled Data☆11Aug 8, 2019Updated 7 years ago
- Efficient Finetuning for OpenAI GPT-OSS☆24Oct 2, 2025Updated 10 months ago
- NVIDIA cuTile learn☆169Dec 9, 2025Updated 8 months ago
- [COLM 2024] TriForce: Lossless Acceleration of Long Sequence Generation with Hierarchical Speculative Decoding☆281Aug 31, 2024Updated last year
- ☆15Jan 27, 2025Updated last year
- Two Stones Hit One Bird: Bilevel Positional Encoding for Better Length Extrapolation, ICML 2024☆25Jun 26, 2024Updated 2 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- Quantized Attention on GPU☆45Nov 22, 2024Updated last year
- Quartet II Official Code☆80May 1, 2026Updated 3 months ago
- ☆20Sep 28, 2024Updated last year
- Low overhead tracing library and trace visualizer for pipelined CUDA kernels☆135Jul 29, 2026Updated 2 weeks ago
- A survey of manufacturer-provided DRAM operating parameters and timings as specified by DRAM chip datasheets from between 1970 and 2021. …☆11May 4, 2022Updated 4 years ago
- ☆157Jun 22, 2023Updated 3 years ago
- Measuring Thinking Efficiency in Reasoning Models - Research Repository☆40Dec 2, 2025Updated 8 months ago
- Code of ICML paper arxiv.org/abs/2302.08105☆14May 4, 2023Updated 3 years ago
- ☆110May 29, 2026Updated 2 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for the paper "The Journey, Not the Destination: How Data Guides Diffusion Models"☆26Dec 12, 2023Updated 2 years ago
- ☆41May 26, 2026Updated 2 months ago
- ☆25May 20, 2025Updated last year
- Generating artificial disfluencies from fluent text easily and promptly☆16Sep 28, 2022Updated 3 years ago
- ☆890Sep 15, 2025Updated 10 months ago
- [ASPLOS'26] Taming the Long-Tail: Efficient Reasoning RL Training with Adaptive Drafter☆175Feb 27, 2026Updated 5 months ago
- [NeurIPS 2024] Fast Best-of-N Decoding via Speculative Rejection☆56Oct 29, 2024Updated last year