ML/DL Math and Method notes
☆67Dec 2, 2023Updated 2 years ago
Alternatives and similar repositories for ml-ways
Users that are interested in ml-ways are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Python tools☆14Oct 22, 2023Updated 2 years ago
- A minimal PyTorch re-implementation of GPT (Generative Pretrained Transformer) language model training☆19Sep 15, 2023Updated 2 years ago
- Helper scripts and notes that were used while porting various nlp models☆51Mar 22, 2022Updated 4 years ago
- Demo of fine-tuning QA models for answering FAQ of cloud providers documentation☆11Jun 20, 2026Updated last month
- My Gen AI research☆11Jun 3, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- The Art of Debugging Open Book☆1,686Updated this week
- This repository shows how to implement a basic model for multimodal entailment.☆10Aug 17, 2021Updated 4 years ago
- Automatic GPU+CPU memory profiling, re-use and memory leaks detection using jupyter/ipython experiment containers☆236Dec 15, 2023Updated 2 years ago
- a version of baby agi using dspy and typed predictors☆15Mar 9, 2024Updated 2 years ago
- This repository hosts the code to port NumPy model weights of BiT-ResNets to TensorFlow SavedModel format.☆14Dec 21, 2021Updated 4 years ago
- ☆16Apr 28, 2023Updated 3 years ago
- ☆13Apr 16, 2021Updated 5 years ago
- ☆30Feb 11, 2022Updated 4 years ago
- The backend behind the LLM-Perf Leaderboard☆11May 5, 2024Updated 2 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Implementation of CaiT models in TensorFlow and ImageNet-1k checkpoints. Includes code for inference and fine-tuning.☆12Jun 9, 2023Updated 3 years ago
- This repository contains code for the paper "Uncertainty Estimation and Calibration with Finite-State Probabilistic RNNs" (Wang, Lawrence…☆17Mar 8, 2021Updated 5 years ago
- Helper scripts I use to run many experiments in the morning to check at night☆20Jun 14, 2021Updated 5 years ago
- Experiments on GPT-3's ability to fit numerical models in-context.☆14Aug 11, 2022Updated 4 years ago
- EMNLP 2021: Single-dataset Experts for Multi-dataset Question-Answering☆67Nov 26, 2021Updated 4 years ago
- Transfer Learning in Dialogue Benchmarking Toolkit☆14Mar 31, 2023Updated 3 years ago
- Lego for GRPO☆30May 27, 2025Updated last year
- A puzzle to learn about prompting☆141May 12, 2023Updated 3 years ago
- Evaluate Transformers from the Hub 🔥☆14May 26, 2026Updated 2 months ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆46Apr 13, 2022Updated 4 years ago
- polyglot task runner☆21Jan 13, 2026Updated 7 months ago
- Code for a workshop hosted at the MLOps World Summit '22☆18Jun 14, 2022Updated 4 years ago
- YouTube Assistant☆12May 15, 2023Updated 3 years ago
- ☆19Apr 5, 2022Updated 4 years ago
- Accepted by AAAI2022☆21Apr 10, 2022Updated 4 years ago
- Alpaca-lora for huggingface implementation using Deepspeed and FullyShardedDataParallel☆24Apr 3, 2023Updated 3 years ago
- A chatbot using the Vaswani transformer as it's sequence-to-sequence module☆22Jul 27, 2023Updated 3 years ago
- Multilingual Compositional Wikidata Questions (MCWQ)☆20Jun 12, 2023Updated 3 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆14Aug 3, 2021Updated 5 years ago
- An ultra-lightweight JAX implementation of sparse Gaussian processes via pathwise sampling.☆22Mar 31, 2021Updated 5 years ago
- ☆13Nov 21, 2021Updated 4 years ago
- [NAACL 2021] This is the code for our paper `Fine-Tuning Pre-trained Language Model with Weak Supervision: A Contrastive-Regularized Self…☆205Aug 17, 2022Updated 3 years ago
- The TaskBench500 dataset and code for generating tasks.☆16Jul 16, 2022Updated 4 years ago
- An implementation of (Induced) Set Attention Block, from the Set Transformers paper☆71Jun 8, 2026Updated 2 months ago
- A Likelihood framework brought to you with from the Weizmann stat. team☆14Feb 8, 2020Updated 6 years ago