A Learnable LSH Framework for Efficient NN Training
☆35Jul 22, 2021Updated 5 years ago
Alternatives and similar repositories for mongoose
Users that are interested in mongoose are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆15Jan 7, 2022Updated 4 years ago
- A Sparse-tensor Communication Framework for Distributed Deep Learning☆13Nov 1, 2021Updated 4 years ago
- A compressed adaptive optimizer for training large-scale deep learning models using PyTorch☆25Nov 26, 2019Updated 6 years ago
- [NSDI'26] PolyRL is a reinforcement learning framework for LLM that harvest spot instances on the cloud to reduce cost.☆19Mar 30, 2026Updated 3 months ago
- Run ONNX RWKV-v4 models with GPU acceleration using DirectML [Windows], or just on CPU [Windows AND Linux]; Limited to 430M model at this…☆21Mar 16, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆11Apr 3, 2023Updated 3 years ago
- ☆21Mar 7, 2024Updated 2 years ago
- ☆44Mar 29, 2023Updated 3 years ago
- FPGA-based HyperLogLog Accelerator☆12Jul 13, 2020Updated 6 years ago
- Proximal Asynchronous SAGA☆13Nov 30, 2017Updated 8 years ago
- A2C training of Relational Deep Reinforcement Learning Architecture☆13Jun 22, 2022Updated 4 years ago
- Code repository for "Spatiotemporal Traffic Matrix Synthesis", Paul Tune and Matthew Roughan, ACM SIGCOMM 2015, London, UK, August 2015.☆15Jan 13, 2016Updated 10 years ago
- ☆11Jun 29, 2021Updated 5 years ago
- ☆13Mar 27, 2019Updated 7 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- TiledLower is a Dataflow Analysis and Codegen Framework written in Rust.☆13Nov 23, 2024Updated last year
- A source-to-source compiler for optimizing CUDA dynamic parallelism by aggregating launches☆15Jun 21, 2019Updated 7 years ago
- ☆19Sep 10, 2019Updated 6 years ago
- FA4-based Relative Attention Kernel developed by TML and Colfax☆17Jul 17, 2026Updated last week
- PyTorch compilation tutorial covering TorchScript, torch.fx, and Slapo☆17Mar 13, 2023Updated 3 years ago
- bigcomputing☆33Nov 3, 2020Updated 5 years ago
- Multi-index hashing for the resolution of ANN search problem on large datasets☆15Oct 16, 2018Updated 7 years ago
- The NYU Systems Seminar☆24Feb 26, 2024Updated 2 years ago
- Better profiling reports for Julia☆14Feb 8, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- ☆45Apr 30, 2018Updated 8 years ago
- Advanced Retrieval Algorithms for Decomposing Large-Scale Candidate Set into Pieces.☆73Apr 13, 2025Updated last year
- This is an official GitHub repository for the paper, "Towards timeout-less transport in commodity datacenter networks.".☆17Oct 12, 2021Updated 4 years ago
- A GPU-accelerated DNN inference serving system that supports instant kernel preemption and biased concurrent execution in GPU scheduling.☆43May 29, 2022Updated 4 years ago
- Printable (A4) overview of tactics in Lean 4 for beginners☆31Sep 19, 2024Updated last year
- Foundry materializes CUDA graphs along with its execution context to disk to support fast cold start of serving engines.☆46Jul 8, 2026Updated 2 weeks ago
- Jax like function transformation engine but micro, microjax☆34Oct 25, 2024Updated last year
- Code for our ACL '20 paper "Representation Engineering with Natural Language Explanations"☆30Jun 15, 2020Updated 6 years ago
- Code for paper 'Minimizing FLOPs to Learn Efficient Sparse Representations' published at ICLR 2020☆20Feb 14, 2020Updated 6 years ago
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- tinygrad port of the RWKV large language model.☆44Mar 9, 2025Updated last year
- ☆226Feb 21, 2023Updated 3 years ago
- Implementation of Hyena Hierarchy in JAX☆10Apr 30, 2023Updated 3 years ago
- An FPGA integration and acceleration of the popular FAISS framework for approximate similarity search☆25Jul 20, 2019Updated 7 years ago
- [WIP] Better (FP8) attention for Hopper☆33Feb 24, 2025Updated last year
- A Julia wrapper for the OpenSpiel project☆12Sep 5, 2024Updated last year
- Measuring the Signal to Noise Ratio in Language Model Evaluation☆31Aug 19, 2025Updated 11 months ago