A LLaMA1/LLaMA12 Megatron implement.
☆28Dec 13, 2023Updated 2 years ago
Alternatives and similar repositories for LLaMA-Megatron
Users that are interested in LLaMA-Megatron are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆84Sep 9, 2023Updated 3 years ago
- ☆52Mar 5, 2025Updated last year
- ☆21Sep 5, 2023Updated 3 years ago
- Love 2d based Tetris game☆13May 21, 2013Updated 13 years ago
- Demonstration of a factory pattern where the types automatically register themselves☆13Mar 13, 2019Updated 7 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- distributed trainer for LLMs☆590May 20, 2024Updated 2 years ago
- Nested Named Entity Recognition for Chinese Biomedical Text☆12Jan 25, 2024Updated 2 years ago
- A reimplementation of KOSMOS-1 from "Language Is Not All You Need: Aligning Perception with Language Models"☆27Mar 3, 2023Updated 3 years ago
- ☆11Jun 5, 2025Updated last year
- a within-document event coreference resolution system, trained and evaluated on the KBP corpus.☆10May 15, 2023Updated 3 years ago
- Example of binding a TF32 CUTLASS GEMM kernel to PyTorch☆12Jun 7, 2024Updated 2 years ago
- Code Roberta version of RetroMAE: Pre-Training Retrieval-oriented Language Models Via Masked Auto-Encoder☆10Mar 16, 2023Updated 3 years ago
- [XLLM@ACL2025] Official Code for "Less is More: Enhancing Structured Multi-Agent Reasoning via Quality-Guided Distillation"☆22Jul 29, 2025Updated last year
- 本项目包含几种常用 NLP算法的实现:关键词(keyword)、命名实体(named entity)、自动摘要(abstract)、文本相似度比较(text similarity)等☆16Jan 16, 2022Updated 4 years ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- PyTorch implementation of the Reinforced Mnemonic Reader + Answer Verifier model (https://arxiv.org/abs/1808.05759)☆10Nov 23, 2018Updated 7 years ago
- [Findings of ACL'2023] Improving Contrastive Learning of Sentence Embeddings from AI Feedback☆40Aug 14, 2023Updated 3 years ago
- Repository for "Scaling Evaluation-time Compute with Reasoning Models as Process Evaluators"☆12Mar 25, 2025Updated last year
- Resources for our ACL 2023 paper: Distilling Script Knowledge from Large Language Models for Constrained Language Planning☆36Aug 19, 2023Updated 3 years ago
- ☆15Mar 5, 2023Updated 3 years ago
- GoldFinch and other hybrid transformer components☆46Jul 20, 2024Updated 2 years ago
- ☆12Jun 7, 2019Updated 7 years ago
- 2021sodic企业隐患排查赛道——top6水煮毛血旺方案分享☆11Jul 17, 2021Updated 5 years ago
- Co-training for Policy Learning☆13Aug 8, 2019Updated 7 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- 🤔 When in Doubt: Improving Classification Performance with Alternating Normalization [Findings of EMNLP2021]☆15Oct 29, 2021Updated 4 years ago
- AutodiffEngine☆13Apr 1, 2019Updated 7 years ago
- ☆11Nov 21, 2024Updated last year
- ☆11Apr 23, 2023Updated 3 years ago
- Nsight Compute In Docker☆13Dec 21, 2023Updated 2 years ago
- ☆11Sep 18, 2020Updated 6 years ago
- ☆41Sep 2, 2021Updated 5 years ago
- [AAAI 2024] LLMEval Phase I dataset — 17 categories, 453 questions, 2186 annotators for Chinese LLM evaluation☆114May 21, 2026Updated 3 months ago
- ☆10Dec 8, 2022Updated 3 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- Recursive Abstractive Processing for Tree-Organized Retrieval☆10May 30, 2024Updated 2 years ago
- This repository contains source code for the PASTA model, a pre-trained language model for table-based fact verification.☆18Dec 27, 2022Updated 3 years ago
- Langchain Agent finetuning using 7B - LLAMA 2 , on hotpotQA (Retroformer framework)☆16Sep 5, 2023Updated 3 years ago
- 回声Echo:AI文案助手☆10May 6, 2023Updated 3 years ago
- ☆11Jul 3, 2023Updated 3 years ago
- UniDoc-RL: Unified Document Understanding with Reinforcement Learning☆18May 21, 2026Updated 3 months ago
- keras encoder-decoder☆17Apr 3, 2018Updated 8 years ago