Code for the paper Xiangqi-R1: Enhancing Spatial Strategic Reasoning in LLMs for Chinese Chess via Reinforcement Learning
☆15Jul 23, 2025Updated 11 months ago
Alternatives and similar repositories for Xiangqi-R1
Users that are interested in Xiangqi-R1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The dataset and baseline code for ASC23 LLM inference optimization challenge.☆34Dec 20, 2023Updated 2 years ago
- ☆14Oct 17, 2024Updated last year
- ChatCFD☆56Apr 9, 2026Updated 3 months ago
- (CVPR 26 Findings) Official implementation of the paper "Bind-Your-Avatar: Multi-Talking-Character Video Generation with Dynamic 3D-mask-…☆34Apr 7, 2026Updated 3 months ago
- A Workbench for Autograding Retrieve/Generate Systems☆15Jun 30, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆11Jun 7, 2023Updated 3 years ago
- ☆13Feb 8, 2025Updated last year
- This repository is the official implementation of Bidirectional Learning for Offline Infinite-width Model-based Optimization (NeurIPS 202…☆14Jan 19, 2023Updated 3 years ago
- Robust active flow control over a range of Reynolds numbers using artificial neural network trained through deep reinforcement learning☆34Dec 28, 2020Updated 5 years ago
- PyTorch Implementation: Code for the paper "Generalizing to Unseen Domains via Adversarial Data Augmentation", NeurIPS 2018. Origin Tenso…☆14Sep 17, 2020Updated 5 years ago
- ☆12Aug 28, 2025Updated 10 months ago
- [ICCV 2021] Multimodal Knowledge Expansion☆10Aug 28, 2021Updated 4 years ago
- ☆23Dec 17, 2024Updated last year
- LLM Safeguarding with Internal Representations☆19Apr 27, 2026Updated 2 months ago
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- ☆11Nov 23, 2024Updated last year
- ☆24Feb 16, 2022Updated 4 years ago
- GINOT is a deep learning model that combines transformers with neural operators for accurate forward predictions on arbitrary 2D and 3D g…☆37Jan 20, 2026Updated 6 months ago
- This is the official code repository of ICLR 2023 Tiny Paper, "Hierarchical Dialogue Understanding with Special Tokens and Turn-level Att…☆19Jun 20, 2023Updated 3 years ago
- ☆11May 1, 2022Updated 4 years ago
- This repository is the official implementation of Generalized Data Weighting via Class-level Gradient Manipulation (NeurIPS 2021)(http://…☆23Oct 8, 2022Updated 3 years ago
- Code to reproduce the experiments in the paper: Does CLIP Bind Concepts? Probing Compositionality in Large Image Models.☆16Oct 14, 2023Updated 2 years ago
- Data and Code for EMNLP 2023 paper "QTSumm: Query-Focused Summarization over Tabular Data"☆23Mar 29, 2024Updated 2 years ago
- A parallel coordinates plot using matplotlib☆14Aug 13, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Benchmark for Biophysical Sequence Optimization Algorithms☆24Apr 15, 2026Updated 3 months ago
- ☆15Dec 10, 2021Updated 4 years ago
- Offcial Repo of Paper "Eliminating Position Bias of Language Models: A Mechanistic Approach""☆23Jun 13, 2025Updated last year
- ☆23Oct 10, 2025Updated 9 months ago
- ☆21May 14, 2025Updated last year
- Combining CFDQuery, CFDCodeBench and FoamBench☆42Jun 10, 2026Updated last month
- 率土之滨辅助☆30Oct 15, 2025Updated 9 months ago
- The dataset for paper "Why Do We Click: Visual Impression-aware News Recommendation", ACM MM 2021☆15Feb 24, 2022Updated 4 years ago
- [ACL 2023] PyTorch Implementation of Zero-and Few-Shot Event Detection via Prompt-Based Meta Learning☆17Jun 6, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- The source code for paper: HEProto: A Hierarchical Enhancing ProtoNet based on Multi-Task Learning for Few-shot Named Entity Recognition☆11Jan 11, 2024Updated 2 years ago
- The source code and dataset of paper "Time-sensitive Retrieval-Augmented Generation for Question Answering"☆15Jan 3, 2025Updated last year
- Fast Memorization of Prompt Improves Context Awareness of Large Language Models (Findings of EMNLP 2024)☆22Oct 22, 2024Updated last year
- Analyzing partial dimensional collapse in non-contrastive self-supervised learning. "Understanding Collapse in Non-Contrastive Siamese Re…☆16Nov 12, 2023Updated 2 years ago
- [WSDM 2024 Best Paper Honorable Mention] Debiasing Sequential Recommenders through Distributionally Robust Optimization over System Expos…☆16Jun 20, 2024Updated 2 years ago
- [Communications Medicine' 25 (Nature Portfolio) ] Tuning Vision Foundation Models for Rectal Cancer Segmentation from CT Scans☆14Jul 11, 2025Updated last year
- COS-PLAY: Co-Evolving LLM Decision and Skill Bank Agents for Long-Horizon Game Play☆29Jul 11, 2026Updated last week