This repository provides a comprehensive library for parallel training and LoRA algorithm implementations, supporting multiple parallel strategies and a rich collection of LoRA variants. It serves as a flexible and efficient model fine-tuning toolkit for researchers and developers. Please contact hehn@mail.ustc.edu.cn for detailed information.
☆67Mar 25, 2026Updated 5 months ago
Alternatives and similar repositories for MyTransformers
Users that are interested in MyTransformers are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆12Oct 24, 2024Updated last year
- [ACL 2025] SeedBench: A Multi-task Benchmark for Evaluating Large Language Models in Seed Science🌾☆24Dec 19, 2025Updated 8 months ago
- One Initialization to Rule them All: Fine-tuning via Explained Variance Adaptation☆52Oct 20, 2025Updated 10 months ago
- molly, an LLM designed to understand multi-omics data.☆26Dec 23, 2025Updated 8 months ago
- code for ACL24 "MELoRA: Mini-Ensemble Low-Rank Adapter for Parameter-Efficient Fine-Tuning"☆34Feb 19, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [NeurIPS 2024 Spotlight] EMR-Merging: Tuning-Free High-Performance Model Merging☆82Mar 1, 2025Updated last year
- The first high school physics Olympiad benchmark for evaluating (M)LLMs with step-level grading and human-level comparison.☆26Dec 19, 2025Updated 8 months ago
- ☆17Apr 29, 2025Updated last year
- ☆24Sep 10, 2024Updated last year
- This is a summary of research on general-purpose Medical Image Restoration. Please raise an issue if you suggest new qualified project.☆15Dec 16, 2025Updated 8 months ago
- Model Merging with SVD to Tie the KnOTS [ICLR 2025]☆94Apr 3, 2025Updated last year
- An Android Application for GLCC☆11Sep 30, 2022Updated 3 years ago
- (ICCV-2025 Official Code)) Improving Generalist Model with Domain-Specific Experts☆87Oct 29, 2025Updated 10 months ago
- Not All Patches Are Equal: Hierarchical Dataset Condensation for Single Image Super-Resolution☆10May 7, 2024Updated 2 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ICLR 2025☆30May 21, 2025Updated last year
- A Maximal Mutual Information Criterion for Manipulation Concept Discovery☆14Sep 26, 2024Updated last year
- Official codes for COLING 2024 paper "Robust and Scalable Model Editing for Large Language Models": https://arxiv.org/abs/2403.17431v1☆14Mar 27, 2024Updated 2 years ago
- A repository of Python & PyTorch scripts which (currently) converts .safetensors models into scaled FP8 variants, utilizing gradient desc…☆26Aug 8, 2025Updated last year
- Bayesian Low-Rank Adaptation for Large Language Models☆43Jun 22, 2024Updated 2 years ago
- HydraRNA is a full-length RNA language model.☆18Dec 15, 2025Updated 8 months ago
- (ACL-2025 main conference) Dolphin: Moving Towards Closed-loop Auto-research through Thinking, Practice, and Feedback☆44Jun 24, 2025Updated last year
- (IROS 2025) RRT*former: Environment-Aware Sampling-Based Motion Planning using Transformer☆16Feb 28, 2025Updated last year
- Implementation for the paper "StraTA: Incentivizing Agentic Reinforcement Learning with Strategic Trajectory Abstraction".☆47May 8, 2026Updated 3 months ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [ICLR 2025 Oral🔥] SD-LoRA: Scalable Decoupled Low-Rank Adaptation for Class Incremental Learning☆95Jun 27, 2025Updated last year
- P1-VL: Bridging Visual Perception and Scientific Reasoning in Physics Olympiads☆15Feb 11, 2026Updated 6 months ago
- [ICLR 2026] Code, benchmark and environment for "ScienceBoard: Evaluating Multimodal Autonomous Agents in Realistic Scientific Workflows"☆137Feb 2, 2026Updated 7 months ago
- PhyGaP: Physically-Grounded Gaussians with Polarization Cues(CVPR 2026 Oral)☆17May 21, 2026Updated 3 months ago
- Implementation of E2-TTS, "Embarrassingly Easy Fully Non-Autoregressive Zero-Shot TTS", in MLX☆21Oct 8, 2024Updated last year
- [ICLR 26] The official code repository for the paper "Mirage or Method? How Model–Task Alignment Induces Divergent RL Conclusions".☆19Feb 9, 2026Updated 6 months ago
- Localize-and-Stitch: Efficient Model Merging via Sparse Task Arithmetic☆33Feb 18, 2026Updated 6 months ago
- ☆13Apr 23, 2025Updated last year
- Adaptive Non-Uniform Timestep Sampling for Diffusion Model Training (CVPR 2025)☆17May 22, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [CVPR 2025 Highlight] SoMA: Singular Value Decomposed Minor Components Adaptation for Domain Generalizable Representation Learning☆71Jun 26, 2025Updated last year
- This repo contains the source code for VB-LoRA: Extreme Parameter Efficient Fine-Tuning with Vector Banks (NeurIPS 2024).☆43Oct 15, 2024Updated last year
- Enhance robot task understanding ability through visual semantic graph☆10May 20, 2021Updated 5 years ago
- [DAI 2025] Beyond GPT-5: Making LLMs Cheaper and Better via Performance–Efficiency Optimized Routing☆223Dec 11, 2025Updated 8 months ago
- ☆23Jan 23, 2024Updated 2 years ago
- The Official PyTorch implementation of CorrFill: Enhancing Faithfulness in Reference-based Inpainting with Correspondence Guidance in Dif…☆16Jan 14, 2025Updated last year
- [Findings@ACL'26] LLMRouterBench: A Massive Benchmark and Unified Framework for LLM Routing☆101Apr 6, 2026Updated 5 months ago