Agent-Omit: Training Efficient LLM Agents for Adaptive Thought and Observation Omission via Reinforcement Learning
☆33May 11, 2026Updated 2 months ago
Alternatives and similar repositories for Agent-Omit
Users that are interested in Agent-Omit are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The code for USTBench.☆24Mar 29, 2026Updated 4 months ago
- [arXiv] "Linear Dynamics in the RLVR Training of Large Language Models"☆17May 25, 2026Updated 2 months ago
- Official code for article "LLMLight: Large Language Models as Traffic Signal Control Agents".☆287Aug 12, 2025Updated 11 months ago
- Codes for replicating the dataset described in "A Satellite Imagery Dataset for Long-Term Sustainable Development in United States Cities…☆12Jun 23, 2023Updated 3 years ago
- JAILJUDGE: A comprehensive evaluation benchmark which includes a wide range of risk scenarios with complex malicious prompts (e.g., synth…☆65Dec 13, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆20Apr 25, 2023Updated 3 years ago
- UUKG: Unified Urban Knowledge Graph Dataset for Knowledge-Enhanced Urban Spatiotemporal Prediction☆120Apr 29, 2025Updated last year
- CoLLMLight: Cooperative Large Language Model Agents for Network-Wide Traffic Signal Control☆28Apr 23, 2026Updated 3 months ago
- [NeurIPS2025] Foundation Models for Scientific Discovery: From Paradigm Enhancement to Paradigm Transition☆36Oct 21, 2025Updated 9 months ago
- Bag of Tricks: Benchmarking of Jailbreak Attacks on LLMs. Empirical tricks for LLM Jailbreaking. (NeurIPS 2024)☆167Nov 30, 2024Updated last year
- Co-evolving policy actors and experience extractors for efficient experience-driven agent RL☆51May 12, 2026Updated 2 months ago
- 🔥🔥🔥 DSLighting is an LLM-driven autonomous data science execution engine that turns task descriptions and datasets into iterative cod…☆52Jun 14, 2026Updated last month
- Official implementation for ICML24 paper "Irregular Multivariate Time Series Forecasting: A Transformable Patching Graph Neural Networks …☆138Nov 28, 2025Updated 8 months ago
- ☆34Aug 24, 2023Updated 2 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- ☆30Mar 17, 2026Updated 4 months ago
- An Awesome Collection of Urban Foundation Models (UFMs).☆224Apr 7, 2026Updated 3 months ago
- OR-R1: Automating Modeling and Solving of Operations Research Optimization Problem via Test-Time Reinforcement Learning☆17Nov 12, 2025Updated 8 months ago
- ☆105Jul 18, 2026Updated 2 weeks ago
- ☆12Apr 13, 2017Updated 9 years ago
- ☆23Jul 10, 2025Updated last year
- code for "Data Might be Enough: Bridge Real-World Traffic Signal Control Using Offline Reinforcement Learning"☆11May 2, 2024Updated 2 years ago
- Identifying Nuances in Fake News vs. Satire: Using Semantic and Linguistic Cues (NLP4IF, EMNLP-IJCNLP 2019)☆11Dec 21, 2020Updated 5 years ago
- Official codes of KDD'24 paper "HiFGL: A Hierarchical Framework for Cross-silo Cross-device Federated Graph Learning"☆10Sep 4, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Official implementation of BandPO: Bridging Trust Regions and Ratio Clipping via Probability-Aware Bounds for LLM Reinforcement Learning.…☆49Apr 8, 2026Updated 3 months ago
- ☆12Feb 19, 2024Updated 2 years ago
- [ICML 2026] Milestone-Guided Policy Learning for Long-Horizon Language Agents☆41May 29, 2026Updated 2 months ago
- [ICLR 2026] Meta-RL Induces Exploration in Language Agents☆45Feb 1, 2026Updated 6 months ago
- Paper: “MEMRL: SELF-EVOLVING AGENTS VIA RUNTIME REINFORCEMENT LEARNING ON EPISODIC MEMORY” Open-Source Code☆162Jul 18, 2026Updated 2 weeks ago
- Official implementation for KDD25 paper "GraphLoRA: Structure-Aware Contrastive Low-Rank Adaptation for Cross-Graph Transfer Learning"☆22Jul 10, 2025Updated last year
- [ACL 2026 Oral] From Word to World: Can Large Language Models be Implicit Text-based World Models?☆68Apr 13, 2026Updated 3 months ago
- MiroEval: A benchmark and evaluation framework for deep research agents — 100 tasks (70 text, 30 multimodal) assessed across synthesis qu…☆46Jul 6, 2026Updated 3 weeks ago
- Official repo: “Oh LLM, I’m Asking Thee, Please Give Me a Decision Tree”: Zero-Shot Decision Tree Induction and Embedding with Large Lang…☆16Jul 17, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- [ICLR 2026] LongRLVR: Long-Context Reinforcement Learning Requires Verifiable Context Rewards.☆19Mar 16, 2026Updated 4 months ago
- TwiMed: Twitter and PubMed Comparable Corpus of Drugs, Diseases, Symptoms and their Relations☆11May 24, 2017Updated 9 years ago
- Official implementation of "G-LNS: Generative Large Neighborhood Search for LLM-Based Automatic Heuristic Design"(arXiv:2602.08253).☆29Feb 10, 2026Updated 5 months ago
- LatentMem: Customizing Latent Memory for Multi-Agent Systems☆49Feb 9, 2026Updated 5 months ago
- Investigating Rumor News using Agreement-Aware Search☆12Feb 25, 2018Updated 8 years ago
- Code for accepted paper at ICLR 2026☆15May 19, 2026Updated 2 months ago
- Official code for paper "GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable R…☆66Mar 29, 2026Updated 4 months ago