Aims for memory-efficient training (24GB VRAM) on consumer GPUs. Optimizing language models through guidance tokens in reasoning chains, based on DeepSeekRL-Extended.
☆28Feb 23, 2025Updated last year
Alternatives and similar repositories for Guide-GRPO
Users that are interested in Guide-GRPO are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Implementation for "DeltaPhi: Learning Physical Trajectory Residual for PDE Solving"☆13Jun 17, 2024Updated 2 years ago
- (VillagerAgent ACL 2024) A Graph based Minecraft multi agents framework☆95Jun 5, 2026Updated last month
- [ACL2023] WhitenedCSE: Whitening-based Contrastive Learning of Sentence Embeddings☆18Sep 12, 2023Updated 2 years ago
- Code for "Holistic Physics Solver: Learning PDEs in a Unified Spectral-Physical Space"☆25Mar 25, 2026Updated 4 months ago
- [ICLR 2025] VideoGrain: This repo is the official implementation of "VideoGrain: Modulating Space-Time Attention for Multi-Grained Video …☆159Mar 24, 2025Updated last year
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [AAAI 2024] DGL: Dynamic Global-Local Prompt Tuning for Text-Video Retrieval.☆49Oct 14, 2024Updated last year
- [TIP 2023] Co-Learning Meets Stitch-Up for Noisy Multi-label Visual Recognition.☆13Aug 19, 2023Updated 2 years ago
- The official repository for paper "FlexSelect: Flexible Token Selection for Efficient Long Video Understanding".☆31Sep 19, 2025Updated 10 months ago
- [ICLR 2024] Test-Time RL with CLIP Feedback for Vision-Language Models.☆102Oct 20, 2025Updated 9 months ago
- Official repository of DoraemonGPT: Toward Understanding Dynamic Scenes with Large Language Models☆91Jun 19, 2026Updated last month
- Self-supervised Point Cloud Representation Learning via Separating Mixed Shapes☆21May 23, 2023Updated 3 years ago
- The official code for [ACM MM 2022] 'In-N-Out Generative Learning for Dense Unsupervised Video Segmentation'.☆20Feb 22, 2023Updated 3 years ago
- Global-to-Local Modeling for Video-based 3D Human Pose and Shape Estimation☆59Jun 21, 2023Updated 3 years ago
- Code for SEEG: Semantic Energized Co-speech Gesture Generation☆33Dec 3, 2022Updated 3 years ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- ☆27Apr 5, 2024Updated 2 years ago
- ☆31Mar 1, 2024Updated 2 years ago
- MediaPipeを用いたハンドジェスチャーによる簡単なマウス操作を行うプログラムです。☆12Mar 17, 2021Updated 5 years ago
- ☆14May 5, 2019Updated 7 years ago
- The project page of paper: Aha! Adaptive History-driven Attack for Decision-based Black-box Models [ICCV 2021]☆10Feb 23, 2022Updated 4 years ago
- (ICML 2024) Improve Context Understanding in Multimodal Large Language Models via Multimodal Composition Learning☆28Sep 27, 2024Updated last year
- [CVPR2024] CapHuman: Capture Your Moments in Parallel Universes☆99Nov 20, 2024Updated last year
- A tool to download and format PASCAL VOC 2007 dataset for multilabel classification☆10Jul 17, 2017Updated 9 years ago
- Official code for Attention-driven GUI Grounding, AAAI2025☆16Dec 17, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Smooth Variational Graph Embeddings for Efficient Neural Architecture Search☆14Feb 2, 2023Updated 3 years ago
- RETROAGENT: From Solving to Evolving via Retrospective Dual Intrinsic Feedback☆26Mar 30, 2026Updated 3 months ago
- ☆17Jul 10, 2022Updated 4 years ago
- (AAAI2024) Controllable 3D Face Generation with Conditional Style Code Diffusion☆40Apr 17, 2024Updated 2 years ago
- [ICCV 23]This is a Pytorch implementation of our paper "SMMix: Self-Motivated Image Mixing for Vision Transformers"☆16Jul 14, 2023Updated 3 years ago
- Fast LLM Training CodeBase With dynamic strategy choosing [Deepspeed+Megatron+FlashAttention+CudaFusionKernel+Compiler];☆41Jan 4, 2024Updated 2 years ago
- High-performance ASR tool using Faster Whisper, supporting custom models, multi-language transcription, and real-time processing feedback…☆10Sep 17, 2025Updated 10 months ago
- Official repository for ALT (ALignment with Textual feedback).☆10Jul 25, 2024Updated 2 years ago
- Code for our ACL'23 paper on how to identify metaphor mappings with the help of GPT-3☆12May 21, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- implementation of " Discovering Causal Signals in Images "☆13Oct 7, 2021Updated 4 years ago
- Multi-Domain Multi-Scale Diffusion Model for Low-Light Image Enhancement (AAAI'24)☆45Mar 1, 2025Updated last year
- UPDATE: All future changes will be pushed to https://github.com/HICAI-ZJU/PromptProtein☆15Apr 23, 2023Updated 3 years ago
- ☆20Jun 16, 2026Updated last month
- 本项目设计了一个基于UDP的网络拍卖行程序,包含客户端和服务端。使用语言:python3;UI设计:pyqt5;采用多线程。☆11Mar 27, 2020Updated 6 years ago
- Resources for our paper: "Agent-R: Training Language Model Agents to Reflect via Iterative Self-Training"☆174Oct 20, 2025Updated 9 months ago
- TRAIL: Simulating the Impact of Human Locomotion on Natural Landscapes - 2024 - Computer Graphics International (CGI)☆13Apr 7, 2025Updated last year