☆16Jun 4, 2025Updated last year
Alternatives and similar repositories for coto
Users that are interested in coto are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICML 2025] Open Your Eyes: Vision Enhances Message Passing Neural Networks in Link Prediction☆16May 28, 2025Updated last year
- Official implementation of Nemesis: Normalizing the Soft-prompt Vectors of Vision-Language Models (ICLR 2024 Spotlight)☆16Jul 19, 2026Updated last month
- The code of RouterDC☆76Apr 14, 2025Updated last year
- A tool for traditional performing arts in MEDIAN Lab☆11Feb 22, 2025Updated last year
- DiTASK: Multi-Task Fine-Tuning with Diffeomorphic Transformations (CVPR 2025)☆14Jun 1, 2025Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- [ICML 2025🔥] ParallelComp: Parallel Long-Context Compressor for Length Extrapolation☆30Jun 16, 2025Updated last year
- The repo for HiRA paper☆38Jan 9, 2026Updated 7 months ago
- [NeurIPS 2024] GITA: Graph to Image-Text Integration for Vision-Language Graph Reasoning☆55Nov 25, 2025Updated 9 months ago
- Activation-Steered Compression☆18Jan 30, 2026Updated 7 months ago
- A reinforcement learning framework with verifiable aesthetic rewards for improving aesthetic slide generation capabilities in LLM agents.…☆30May 19, 2026Updated 3 months ago
- MOTIF: Modular Thinking via Reinforcement Fine-tuning in LLMs☆17Jul 6, 2025Updated last year
- Official Code Repository for [AutoScale📈: Scale-Aware Data Mixing for Pre-Training LLMs] Published as a conference paper at **COLM 2025*…☆14Aug 8, 2025Updated last year
- [ACL 2024] Official PyTorch implementation of "IntactKV: Improving Large Language Model Quantization by Keeping Pivot Tokens Intact"☆46May 24, 2024Updated 2 years ago
- ☆20Nov 3, 2024Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆30May 24, 2025Updated last year
- [NAACL 2025] Representing Rule-based Chatbots with Transformers☆23Feb 9, 2025Updated last year
- [ICLR 2026] M2-Miner: Multi-Agent Enhanced MCTS for Mobile GUI Agent Data Mining☆55Apr 22, 2026Updated 4 months ago
- Official PyTorch implementation of Domain-Guided Conditional Diffusion Model for Unsupervised Domain Adaptation☆22Dec 27, 2024Updated last year
- User-friendly implementation of the Mixture-of-Sparse-Attention (MoSA). MoSA selects distinct tokens for each head with expert choice rou…☆30May 3, 2025Updated last year
- ☆18Jun 10, 2022Updated 4 years ago
- ☆43Nov 1, 2024Updated last year
- Implementation of the paper "Neural Wasserstein Gradient Flows for Discrepancies with Riesz Kernels"☆11Mar 18, 2024Updated 2 years ago
- PyTorch Reimplementation of LoRA (featuring with supporting nn.MultiheadAttention in OpenCLIP)☆79Jun 10, 2025Updated last year
- Deploy open-source AI quickly and easily - Special Bonus Offer • AdRunpod Hub is built for open source. One-click deployment and autoscaling endpoints without provisioning your own infrastructure.
- ☆22Jan 2, 2026Updated 7 months ago
- ☆47Jun 11, 2025Updated last year
- PLM: Efficient Peripheral Language Models Hardware-Co-Designed for Ubiquitous Computing☆21Mar 18, 2025Updated last year
- We introduce DreamPRM-1.5, an instance-reweighted framework that adaptively adjusts the importance of each training example via bi-level …☆16Nov 13, 2025Updated 9 months ago
- Source code for SWIFT, an efficient reward model.☆21Jan 13, 2026Updated 7 months ago
- [ACL 2026 Main] Official Repo for Paper "Which Reasoning Trajectories Teach Students to Reason Better? A Simple Metric of Informative Ali…☆16Jul 1, 2026Updated last month
- Code for the preprint "Cache Me If You Can: How Many KVs Do You Need for Effective Long-Context LMs?"☆48Jul 29, 2025Updated last year
- [SIGIR'24] The official implementation code of MOELoRA.☆37Aug 3, 2024Updated 2 years ago
- Qwen-WisdomVast is a large model trained on 1 million high-quality Chinese multi-turn SFT data, 200,000 English multi-turn SFT data, and …☆17Apr 12, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ACL 2024] Code for the paper "ALaRM: Align Language Models via Hierarchical Rewards Modeling"☆25Mar 28, 2024Updated 2 years ago
- The official implementation for MTLoRA: A Low-Rank Adaptation Approach for Efficient Multi-Task Learning (CVPR '24)☆75Jul 3, 2025Updated last year
- A family of efficient edge language models in 100M~1B sizes.☆19Feb 14, 2025Updated last year
- The code for LaRA Benchmark☆52May 28, 2025Updated last year
- GluonTS - Probabilistic Time Series Modeling in Python☆27Sep 19, 2022Updated 3 years ago
- This repository contains the code for the paper: SirLLM: Streaming Infinite Retentive LLM☆60May 28, 2024Updated 2 years ago
- Official repo of Promoting Efficient Reasoning with Verifiable Stepwise Reward☆16Sep 9, 2025Updated 11 months ago