π’ Data Toolkit for Sailor Language Models
β94Feb 24, 2025Updated last year
Alternatives and similar repositories for sailcraft
Users that are interested in sailcraft are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [EMNLP-2024] βοΈ Sailor: Open Language Models for South-East Asiaβ139Dec 21, 2024Updated last year
- β21Apr 16, 2025Updated last year
- π± Sailor2: Sailing in South-East Asia with Inclusive Multilingual LLMsβ74Mar 21, 2025Updated last year
- Implementation of the paper: "Turning Tables: Generating Examples from Semi-structured Tables for Endowing Language Models with Reasoningβ¦β22Nov 2, 2021Updated 4 years ago
- Dataset for TACL 2022 paper: "FeTaQA: Free-form Table Question Answering"β90May 11, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- This repository contains source code for the PASTA model, a pre-trained language model for table-based fact verification.β18Dec 27, 2022Updated 3 years ago
- QAmeleon introduces synthetic multilingual QA data using PaLM, a 540B large language model. This dataset was generated by prompt tuning Pβ¦β34Aug 15, 2023Updated 3 years ago
- Can LLMs generate code-mixed sentences through zero-shot prompting?β11Apr 18, 2023Updated 3 years ago
- β13Sep 6, 2022Updated 4 years ago
- β15Mar 12, 2024Updated 2 years ago
- The official repository of 'Unnatural Language Are Not Bugs but Features for LLMs'β25May 20, 2025Updated last year
- Improved Few-Shot Jailbreaking Can Circumvent Aligned Language Models and Their Defenses (NeurIPS 2024)β66Jan 11, 2025Updated last year
- An Empirical Study of Memorization in NLP (ACL 2022)β13Jun 22, 2022Updated 4 years ago
- [ICLR 2025] A Closer Look at Machine Unlearning for Large Language Modelsβ49Dec 4, 2024Updated last year
- Deploy on Railway without the complexity - Free Credits Offer β’ AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- β15Oct 13, 2025Updated 11 months ago
- [ACL 2024 Demo] SeaLLMs - Large Language Models for Southeast Asiaβ176Jul 30, 2024Updated 2 years ago
- Official code for "MAmmoTH2: Scaling Instructions from the Web" [NeurIPS 2024]β146Oct 27, 2024Updated last year
- Code for "Mixed Cross Entropy Loss for Neural Machine Translation"β20Jul 23, 2021Updated 5 years ago
- [ICLR 2025] Cheating Automatic LLM Benchmarks: Null Models Achieve High Win Rates (Oral)β86Oct 23, 2024Updated last year
- Difference-based Contrastive Learning for Korean Sentence Embeddingsβ23Mar 11, 2026Updated 6 months ago
- Official code repo for paper "Great Memory, Shallow Reasoning: Limits of kNN-LMs"β24Apr 30, 2025Updated last year
- β14Sep 30, 2021Updated 4 years ago
- NusaWrites is an in-depth analysis of corpora collection strategy and a comprehensive language modeling benchmark for underrepresented anβ¦β30Sep 27, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI β’ AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [ICLR 2025] 𧬠RegMix: Data Mixture as Regression for Language Model Pre-training (Spotlight)β209Feb 17, 2025Updated last year
- Tiny evaluation of leading LLMs on competitive programming problemsβ14Apr 10, 2026Updated 5 months ago
- The source code of our work "Prepacking: A Simple Method for Fast Prefilling and Increased Throughput in Large Language Models" [AISTATS β¦β62Oct 11, 2024Updated last year
- Implementations of online merging optimizers proposed by Online Merging Optimizers for Boosting Rewards and Mitigating Tax in Alignmentβ82Jun 19, 2024Updated 2 years ago
- Optimizing Anytime Reasoning via Budget Relative Policy Optimizationβ54Jul 15, 2025Updated last year
- NAACL 2024: SeaEval for Multilingual Foundation Models: From Cross-Lingual Alignment to Cultural Reasoningβ26Mar 3, 2025Updated last year
- Code base for the EMNLP 2021 Findings paper: Cartography Active Learningβ14Jun 3, 2025Updated last year
- A Python implementation of an agent swarm system that works with local LLM servers. The system allows you to create multiple agents that β¦β14Nov 20, 2024Updated last year
- Learning to Rewrite for Non-Autoregressive Neural Machine Translationβ21Dec 23, 2021Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI β’ AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ICLR 2021: Pre-Training for Context Representation in Conversational Semantic Parsingβ31Aug 30, 2021Updated 5 years ago
- Minimal (truly) muP implementation, consistent with TP4 and TP5 papers notationβ14Jan 2, 2026Updated 8 months ago
- This is the code for our paper: PLACES: Prompting Language Models for Social Conversation Synthesisβ11Feb 17, 2023Updated 3 years ago
- The official repository for the paper "From Zero to Hero: Examining the Power of Symbolic Tasks in Instruction Tuning".β64Apr 18, 2023Updated 3 years ago
- Freeing data processing from scripting madness by providing a set of platform-agnostic customizable pipeline processing blocks.β3,341Updated this week
- β26Nov 20, 2021Updated 4 years ago
- The dataset and source code for our paper: "Did You Ask a Good Question? A Cross-Domain Question IntentionClassification Benchmark for Teβ¦β32Jul 5, 2021Updated 5 years ago