Data preparation code for Amber 7B LLM
☆96May 10, 2024Updated 2 years ago
Alternatives and similar repositories for amber-data-prep
Users that are interested in amber-data-prep are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Data preparation code for CrystalCoder 7B LLM☆45May 10, 2024Updated 2 years ago
- Pre-training code for Amber 7B LLM☆176May 10, 2024Updated 2 years ago
- Pre-training code for CrystalCoder 7B LLM☆59May 10, 2024Updated 2 years ago
- Open Implementations of LLM Analyses☆111Oct 8, 2024Updated last year
- QuoteSum is a textual QA dataset containing Semi-Extractive Multi-source Question Answering (SEMQA) examples written by humans, based on …☆13Mar 25, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- MobileLLM Optimizing Sub-billion Parameter Language Models for On-Device Use Cases. In ICML 2024.☆1,456Apr 30, 2026Updated 3 months ago
- ☆26May 30, 2023Updated 3 years ago
- QAmeleon introduces synthetic multilingual QA data using PaLM, a 540B large language model. This dataset was generated by prompt tuning P…☆34Aug 15, 2023Updated 2 years ago
- ☆23Jun 16, 2026Updated last month
- LMTuner: Make the LLM Better for Everyone☆38Sep 21, 2023Updated 2 years ago
- ☆209Apr 19, 2025Updated last year
- instruction-following benchmark for large reasoning models☆49Apr 19, 2026Updated 3 months ago
- ☆15Feb 21, 2024Updated 2 years ago
- ☆56Jun 26, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆60Jun 6, 2024Updated 2 years ago
- Code repository for the public reproduction of the language modelling experiments on "MatFormer: Nested Transformer for Elastic Inference…☆31Nov 14, 2023Updated 2 years ago
- Code of our paper "Method-Level Bug Severity Prediction using Source Code Metrics and LLMs" which is accepted to ISSRE 2023.☆10Nov 12, 2023Updated 2 years ago
- Reaching LLaMA2 Performance with 0.1M Dollars☆986Jul 23, 2024Updated 2 years ago
- ☆93Oct 5, 2023Updated 2 years ago
- Data and tools for generating and inspecting OLMo pre-training data.☆1,534Nov 5, 2025Updated 9 months ago
- ☆41May 2, 2024Updated 2 years ago
- codebase release for EMNLP2023 paper publication☆19Sep 18, 2025Updated 10 months ago
- ☆13Oct 20, 2022Updated 3 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Documenting large text datasets 🖼️ 📚☆14Dec 17, 2024Updated last year
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 3 months ago
- Language models scale reliably with over-training and on downstream tasks☆103Apr 2, 2024Updated 2 years ago
- Hugging Face and Pyserini interoperability☆20May 18, 2023Updated 3 years ago
- This repository contains code and tooling for the Abacus.AI LLM Context Expansion project. Also included are evaluation scripts and bench…☆605Nov 17, 2023Updated 2 years ago
- A family of open-sourced Mixture-of-Experts (MoE) Large Language Models☆1,694Mar 8, 2024Updated 2 years ago
- Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.☆3,258Updated this week
- Official PyTorch Implementation of EMoE: Unlocking Emergent Modularity in Large Language Models [main conference @ NAACL2024]☆39May 28, 2024Updated 2 years ago
- Package and scripts used to build a dataset of Wikipedia articles in Markdown.☆20Sep 11, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [NeurIPS 2024] Goldfish Loss: Mitigating Memorization in Generative LLMs☆98Nov 17, 2024Updated last year
- PyTorch code for System-1.x: Learning to Balance Fast and Slow Planning with Language Models☆25Jul 22, 2024Updated 2 years ago
- ☆415Nov 2, 2023Updated 2 years ago
- Klear-Reasoner: Advancing Reasoning Capability via Gradient-Preserving Clipping Policy Optimization☆82Dec 25, 2025Updated 7 months ago
- A toolkit for inference and evaluation of 'mixtral-8x7b-32kseqlen' from Mistral AI☆770Dec 15, 2023Updated 2 years ago
- Reimplementation of the task generation part from the Alpaca paper☆118Apr 4, 2023Updated 3 years ago
- Datasets and code from our paper, where we use machine learning to predict if ChatGPT will refuse a given prompt.☆38Sep 23, 2023Updated 2 years ago