☆63Oct 1, 2021Updated 4 years ago
Alternatives and similar repositories for gpt-j-6b
Users that are interested in gpt-j-6b are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Repo for fine-tuning Casual LLMs☆467Mar 27, 2024Updated 2 years ago
- KoGPT-2 finetuning Based Kiosk chatbot☆12Dec 12, 2023Updated 2 years ago
- Repo for paper: Examining LLMs' Uncertainty Expression Towards Questions Outside Parametric Knowledge☆14Feb 20, 2024Updated 2 years ago
- ☆25Aug 10, 2021Updated 5 years ago
- ☆16Nov 16, 2022Updated 3 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Eval LLMs☆11May 12, 2024Updated 2 years ago
- ☆10Jul 6, 2023Updated 3 years ago
- ☆14Dec 12, 2019Updated 6 years ago
- [IROS 2024] "ComTraQ-MPC: Meta-Trained DQN-MPC Integration for Trajectory Tracking with Limited Active Localization Updates" by Gokul Put…☆13Apr 10, 2025Updated last year
- ☆26Sep 13, 2022Updated 3 years ago
- A Deep Reinforcement Learning model for high volume and frequency Forex Portfolio Management☆13Jan 11, 2023Updated 3 years ago
- ☆11Jul 15, 2020Updated 6 years ago
- GreenLIT: Using GPT-J with Multi-Task Learning to Create New Screenplays☆16Nov 27, 2022Updated 3 years ago
- Lucene open-domain QA retrieval in python☆11Feb 18, 2021Updated 5 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆11Jun 1, 2021Updated 5 years ago
- Fine-tuning GPT-J-6B on colab or equivalent PC GPU with your custom datasets: 8-bit weights with low-rank adaptors (LoRA)☆73Jun 18, 2022Updated 4 years ago
- ☆13May 27, 2021Updated 5 years ago
- ☆50Jan 4, 2023Updated 3 years ago
- Exploring actor critic deep reinforcement learning methods for maximizing profits by learning stock trading strategies☆11Mar 24, 2023Updated 3 years ago
- A PS4 Controlled Holonomic Drive train equipped with Mecanum wheels.☆21Dec 4, 2022Updated 3 years ago
- A profitable cryptocurrency trading environment using deep reinforcement learning and OpenAI's gym☆11May 3, 2019Updated 7 years ago
- Extends inputfield dependencies so that inputfield visibility or required status may be determined at runtime by selector or custom PHP c…☆12Dec 3, 2025Updated 8 months ago
- 🧙 Playground☆12Apr 13, 2021Updated 5 years ago
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- Code, data, and pretrained models for the paper "Generating Wikipedia Article Sections from Diverse Data Sources"☆21Feb 5, 2021Updated 5 years ago
- Fine-tuning 6-Billion GPT-J (& other models) with LoRA and 8-bit compression☆70Oct 5, 2022Updated 3 years ago
- DGL implementation of GRAND(Graph Random Neural Network, NeurIPS 2020)☆18Mar 19, 2021Updated 5 years ago
- Automatic Cardiac MRI Segmentation via Context Aware Recurrent Generative Adversarial Neural Network☆12Feb 6, 2018Updated 8 years ago
- ☆12Jun 19, 2022Updated 4 years ago
- ☆16May 25, 2021Updated 5 years ago
- Codebase for Math Neurosurgery: Isolating LLMs' Math Reasoning Abilities Using Only Forward Passes☆24Jun 15, 2025Updated last year
- Machine learning is changing the world and if you want to be a part of the ML revolution, this is a great place to start! This repository…☆14Aug 25, 2019Updated 6 years ago
- Synthesizer Self-Attention is a very recent alternative to causal self-attention that has potential benefits by removing this dot product…☆14Dec 29, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Model parallel transformers in JAX and Haiku☆6,378Jan 21, 2023Updated 3 years ago
- Official code for the paper: DRA-GRPO: Exploring Diversity-Aware Reward Adjustment for R1-Zero-Like Training of Large Language Models☆24Jan 6, 2026Updated 7 months ago
- This demo scene uses Evergine with .NET 6 support. The new Post-processing graph is used with several effects.☆16Oct 23, 2024Updated last year
- A repository with various methods for clustering mixed datasets in python☆22Nov 7, 2021Updated 4 years ago
- ☆34Aug 10, 2021Updated 5 years ago
- 🫠 check your data, before you wreck your model☆16Aug 11, 2022Updated 4 years ago
- ☆55Jan 18, 2023Updated 3 years ago