☆15Oct 28, 2024Updated last year
Alternatives and similar repositories for play-urts
Users that are interested in play-urts are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR☆21Apr 7, 2026Updated 4 months ago
- [AAAI 2023 Oral] Official code for "PiCor: Multi-Task Deep Reinforcement Learning with Policy Correction".☆21Jul 26, 2025Updated last year
- 集成SQL语句测评和学生管理的数据库系统实验平台,可用于毕设和大作业,欢迎提issue和pr。☆19Oct 17, 2022Updated 3 years ago
- Official repository for ODQA experiments from Decomposed Prompting: A Modular Approach for Solving Complex Tasks, ICLR23☆15Jul 28, 2023Updated 3 years ago
- UCAS 每日自动打卡。好用可以 star 😉☆13Dec 15, 2022Updated 3 years ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- [REALM25 @ ACL25] - "StateAct" Official Paper Repo (SOTA LLM Agent)☆19Updated this week
- Fast and controllable text-to-image model.☆41Jun 16, 2023Updated 3 years ago
- ☆17May 25, 2023Updated 3 years ago
- The Eureka Lab Series is designed for learners at all levels of experience and interest in security concepts and technologies.☆10Nov 30, 2025Updated 8 months ago
- A python client library for microRTS.☆20Feb 5, 2020Updated 6 years ago
- A record of coursework in AI Computing Systems, mainly focusing on high performance computing development for MLU.☆14Jul 14, 2022Updated 4 years ago
- This repository contains statistics about the AI Infrastructure products.☆16Feb 27, 2025Updated last year
- train AI agents to master Free-style Gomoku(五子棋)☆25Mar 2, 2024Updated 2 years ago
- A LLM-powered agent for NetHack☆28Nov 4, 2024Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- [SIGIR 2024] TRAD: Enhancing LLM Agents with Step-Wise Thought Retrieval and Aligned Decision☆20Mar 28, 2024Updated 2 years ago
- Official code for the paper "ADaPT: As-Needed Decomposition and Planning with Language Models"☆93Jan 3, 2024Updated 2 years ago
- Evaluating Large Language Models with Grid-Based Game Competitions: An Extensible LLM Benchmark and Leaderboard☆25Dec 14, 2024Updated last year
- (ACL2025 Findings) Official code for the paper "STeCa: Step-level Trajectory Calibration for LLM Agent Learning"☆29Mar 2, 2026Updated 5 months ago
- ☆15Nov 28, 2024Updated last year
- ☆11Jul 14, 2021Updated 5 years ago
- 这是 对基于大模型的多智能体系统论文的总结☆10Jun 23, 2024Updated 2 years ago
- This is for EMNLP 2024 Paper: AppBench: Planning of Multiple APIs from Various APPs for Complex User Instruction☆16Nov 4, 2024Updated last year
- Neural Network Custom Module for Godot Engine☆20Aug 11, 2020Updated 6 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Implementation for our TOIS paper --- Attentive Long Short-Term Preference Modeling for Personalized Product Search.☆19Feb 14, 2020Updated 6 years ago
- ☆15Jul 9, 2026Updated last month
- Code for Fast Ternary Large Language Model Inference with Addition-Based Sparse GEMM on Edge Devices☆17Apr 17, 2026Updated 3 months ago
- [RAL 2024 & IROS 2024] Enhancing Visual Place Recognition in Spatial Domain on Aerial Vehicle Platforms.☆16Oct 21, 2024Updated last year
- LLM finetuning☆41Aug 9, 2023Updated 3 years ago
- ☆21Aug 13, 2021Updated 5 years ago
- A flexible Multi-Agent Reinforcement Learning (MARL) environment for Collective Robotic Construction (CRC) systems☆13Mar 22, 2023Updated 3 years ago
- Code and data for the paper: Competing Large Language Models in Multi-Agent Gaming Environments☆98Jan 26, 2026Updated 6 months ago
- ☆234Dec 20, 2024Updated last year
- End-to-end encrypted email - Proton Mail • AdSpecial offer: 40% Off Yearly / 80% Off First Month. All Proton services are open source and independently audited for security.
- use python to manipuate PSASP☆25Jan 2, 2020Updated 6 years ago
- common used Recommend Baseline Model, including the traditional statistical model, and the Nerual Network Model. Focus on the SRS (Sequen…☆17Dec 5, 2021Updated 4 years ago
- A collection of product search embedding models☆19Jan 17, 2020Updated 6 years ago
- ☆42Jun 15, 2026Updated last month
- [ICLR'26] R-HORIZON: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?☆26May 9, 2026Updated 3 months ago
- 南京信息工程大学硕士毕业论文LaTeX模板,由单滢滢、耿祥编写,由耿祥整理发布,请大家给该项目打星。☆33Apr 10, 2021Updated 5 years ago
- ☆34Sep 19, 2025Updated 10 months ago