☆54Mar 14, 2025Updated last year
Alternatives and similar repositories for cse234-w25-PA
Users that are interested in cse234-w25-PA are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Minimal (truly) muP implementation, consistent with TP4 and TP5 papers notation☆14Jan 2, 2026Updated 7 months ago
- ☆19Mar 29, 2026Updated 4 months ago
- ECE408 (Applied Parallel Programming) Fall 2022 MP☆21Mar 24, 2023Updated 3 years ago
- Official repository of "Distort, Distract, Decode: Instruction-Tuned Model Can Refine its Response from Noisy Instructions", ICLR 2024 Sp…☆21Mar 7, 2024Updated 2 years ago
- ☆13Jan 7, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- [ICML 2026 Spotlight] Official implementation of TetraJet-v2: Accurate NVFP4 Training for LLMs, with fully-NVFP4 linear layer with unbias…☆16Jul 3, 2026Updated last month
- Problems from IOITC'16 (India)☆10Jan 12, 2022Updated 4 years ago
- This repositories contains the reference implementation for the Sparse Delta Memory paper.More precisely, it contains the model definitio…☆34Jul 9, 2026Updated last month
- ☆10Mar 25, 2024Updated 2 years ago
- Contains all relevant scripts & tools developed to assist in conducting the first iteration of the Indian Inter-college Competitive Progr…☆10Sep 14, 2024Updated last year
- Examples and problems accompanying Daniel Kirschen's Power Systems Textbook☆12Nov 8, 2023Updated 2 years ago
- The official implementation for the intra-stage fusion technique introduced in https://arxiv.org/abs/2409.13221☆31Apr 22, 2025Updated last year
- Surgical GPU kernel benchmark: 7 hard problems, frontier coding agents, roofline-graded against hardware peak.☆19Jun 12, 2026Updated last month
- C++ library for finding Strongly Connected Components in parallel, based on paper: https://dl.acm.org/citation.cfm?id=2851161☆12May 22, 2018Updated 8 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- 快来生成你的浏览记录年度总结!☆18Dec 12, 2024Updated last year
- cc98爬虫☆15Sep 1, 2013Updated 12 years ago
- ☆60Updated this week
- ☆28Oct 2, 2025Updated 10 months ago
- ☆20Dec 24, 2024Updated last year
- ☆46Oct 15, 2025Updated 9 months ago
- High performance RMSNorm Implement by using SM Core Storage(Registers and Shared Memory)☆30Jan 22, 2026Updated 6 months ago
- [NeurIPS 23] Characterizing OOD Error via Optimal Transport☆13Nov 19, 2023Updated 2 years ago
- Copy as an OS Service☆27Nov 20, 2025Updated 8 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation of Minimum Spanning Trees on Apache Spark.☆10May 25, 2015Updated 11 years ago
- An efficient implementation of the NSA (Native Sparse Attention) kernel☆134Jun 24, 2025Updated last year
- The Newton-Muon optimizer☆29Jun 5, 2026Updated 2 months ago
- TopoTrans: Optimal Transport meets Topological Data Analysis☆14Apr 20, 2023Updated 3 years ago
- Code for Tangent Model Composition for Ensembling and Continual Fine-tuning (ICCV 2023) and Tangent Transformers for Composition, Privacy…☆14May 14, 2024Updated 2 years ago
- ☆19May 9, 2025Updated last year
- Code and results accompanying our paper titled Leveraging Unlabeled Data to Predict Out-of-Distribution Performance at ICLR 2022☆11Dec 8, 2022Updated 3 years ago
- Implementation of Fast Weight Attention☆33Jun 3, 2026Updated 2 months ago
- ☆21Mar 17, 2026Updated 4 months ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- A Triton-only attention backend for vLLM☆28Jul 14, 2026Updated 3 weeks ago
- Stick-breaking attention☆63Jul 1, 2025Updated last year
- Tigon: A Distributed Database for a CXL Pod [OSDI '25]☆51Jul 13, 2026Updated 3 weeks ago
- My notebook using mkdocs☆19Mar 12, 2025Updated last year
- Nex Venus Communication Library☆75Nov 17, 2025Updated 8 months ago
- ☆16Jan 5, 2024Updated 2 years ago
- High-performance GPU kernels for LLM inference in OpenAI Triton. Fused RMSNorm, SwiGLU, INT8 GEMM with benchmarks and roofline analysis.☆37Jul 22, 2026Updated 2 weeks ago