Code for the paper Xiangqi-R1: Enhancing Spatial Strategic Reasoning in LLMs for Chinese Chess via Reinforcement Learning
☆15Jul 23, 2025Updated last year
Alternatives and similar repositories for Xiangqi-R1
Users that are interested in Xiangqi-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A toolkit for automated alignment research.☆16Jul 3, 2026Updated 2 months ago
- The dataset and baseline code for ASC23 LLM inference optimization challenge.☆35Dec 20, 2023Updated 2 years ago
- ☆14Oct 17, 2024Updated last year
- (CVPR 26 Findings) Official implementation of the paper "Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-…☆34Apr 7, 2026Updated 5 months ago
- ☆11May 17, 2024Updated 2 years ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- A Workbench for Autograding Retrieve/Generate Systems☆15Jun 30, 2025Updated last year
- ☆11Jun 7, 2023Updated 3 years ago
- A Jieqi AI #揭棋AI☆27Aug 29, 2022Updated 4 years ago
- 个人项目,中国象棋Qt界面与AI象棋引擎☆49May 11, 2023Updated 3 years ago
- Flutter 中国象棋,从0到上架☆30Dec 5, 2020Updated 5 years ago
- 欢迎参加中文讽刺计算评测任务!☆15Nov 4, 2024Updated last year
- ☆23Dec 17, 2024Updated last year
- 中文文本近似计算☆12Jan 22, 2019Updated 7 years ago
- The VirtuaNES EXtended version☆38Mar 13, 2020Updated 6 years ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- LLM Safeguarding with Internal Representations☆21Apr 27, 2026Updated 4 months ago
- ☆11Nov 23, 2024Updated last year
- [ACL26-Findings]💁📲 Self-evolving customer service framework, SEAD, operates without any human-labeled data. It can be quickly launched…☆26Aug 21, 2026Updated last month
- ☆11Updated this week
- Code to reproduce the experiments in the paper: Does CLIP Bind Concepts? Probing Compositionality in Large Image Models.☆16Oct 14, 2023Updated 2 years ago
- A webrtc de-noising module encapsulation.☆12Dec 14, 2016Updated 9 years ago
- A parallel coordinates plot using matplotlib☆14Aug 13, 2021Updated 5 years ago
- ☆15Dec 10, 2021Updated 4 years ago
- Source code for "TeaRAG: A Token-Efficient Agentic Retrieval-Augmented Generation Framework"☆45Jul 2, 2026Updated 2 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- 率土之滨辅助☆31Oct 15, 2025Updated 11 months ago
- Use seaborn to draw RL picture☆25Jun 20, 2023Updated 3 years ago
- ☆24May 14, 2025Updated last year
- [ACL 2023] PyTorch Implementation of Zero-and Few-Shot Event Detection via Prompt-Based Meta Learning☆17Jun 6, 2023Updated 3 years ago
- The source code for paper: HEProto: A Hierarchical Enhancing ProtoNet based on Multi-Task Learning for Few-shot Named Entity Recognition☆11Jan 11, 2024Updated 2 years ago
- 中国象棋开源AI整理, 包括ElephantEye, BitStronger, 撤蛋 夢入神蛋 MRSD2, mars☆38Sep 8, 2017Updated 9 years ago
- Fast Memorization of Prompt Improves Context Awareness of Large Language Models (Findings of EMNLP 2024)☆22Oct 22, 2024Updated last year
- This is a cool hover interactive view, Integrated the commonly used waithud and refresh manager. Do not need to rely on third-party libra…☆16Jul 11, 2017Updated 9 years ago
- Analyzing partial dimensional collapse in non-contrastive self-supervised learning. "Understanding Collapse in Non-Contrastive Siamese Re…☆16Nov 12, 2023Updated 2 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Human-Aligned Chess With a Bit of Search☆23Nov 15, 2025Updated 10 months ago
- [WSDM 2024 Best Paper Honorable Mention] Debiasing Sequential Recommenders through Distributionally Robust Optimization over System Expos…☆16Jun 20, 2024Updated 2 years ago
- [Communications Medicine' 25 (Nature Portfolio) ] Tuning Vision Foundation Models for Rectal Cancer Segmentation from CT Scans☆14Jul 11, 2025Updated last year
- The dataset for paper "Why Do We Click: Visual Impression-aware News Recommendation", ACM MM 2021☆16Feb 24, 2022Updated 4 years ago
- Turn ideas into autonomous research with Git-like evidence histories—inspectable, reproducible, and reversible.☆128Sep 1, 2026Updated 2 weeks ago
- My full implementation the NachOS code base for CS162 Fall 2010. Please see wiki for more details.☆12Mar 2, 2018Updated 8 years ago
- ☆14Sep 22, 2022Updated 3 years ago