A collection on the recent reproduction papers and projects on DeepSeek-R1
☆31Feb 27, 2025Updated last year
Alternatives and similar repositories for awesome-deepseek-r1
Users that are interested in awesome-deepseek-r1 are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- A novel template-free retrosynthesizer that can generate diverse sets of reactants for a desired product via discrete conditional variati…☆15Aug 7, 2022Updated 4 years ago
- The code of paper Sample-Efficient Reinforcement Learning via Conservative Model-Based Actor-Critic. Zhihai Wang, Jie Wang*, Qi Zhou, Bin…☆21May 26, 2022Updated 4 years ago
- This is the source code of our ICML25 paper, titled "Accelerating Large Language Model Reasoning via Speculative Search".☆24Jun 1, 2025Updated last year
- The code of paper "Learning Rule-Induced Subgraph Representations for Inductive Relation Prediction" in NeurIPS 2023.☆14Nov 25, 2023Updated 2 years ago
- The code of paper LMC: Fast Training of GNNs via Subgraph Sampling with Provable Convergence. Zhihao Shi, Xize Liang, Jie Wang. ICLR 2023…☆48Feb 15, 2023Updated 3 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- NeurIPS-2025☆21Nov 4, 2025Updated 9 months ago
- Must-read papers on Knowledge Graph Embedding☆29Oct 15, 2020Updated 5 years ago
- This repository collects various works that reproduce DeepSeek R1, as well as works related to DeepSeek R1 and the DeepSeek series.☆19Apr 27, 2025Updated last year
- The training implementation of the paper "FitDiT: Advancing the Authentic Garment Details for High-fidelity Virtual Try-on."☆12Apr 8, 2025Updated last year
- Testing Theory of Mind (ToM) in language models with epistemic logic☆21Jul 3, 2026Updated last month
- The official code repo and data hub of top_nsigma sampling strategy for LLMs.☆28Feb 11, 2025Updated last year
- ☆10Jul 13, 2024Updated 2 years ago
- ☆20Dec 24, 2024Updated last year
- Code and data repository for two papers (ACL & EMNLP 2024) on the topic of collapse in model editing.☆10Dec 20, 2024Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- [ICME 2019] Source code and datasets for "Semi-supervised Compatibility Learning Across Categories for Clothing Matching"☆11Apr 26, 2024Updated 2 years ago
- ☆15Jun 3, 2019Updated 7 years ago
- ☆17Jul 30, 2024Updated 2 years ago
- ☆25Aug 3, 2026Updated 3 weeks ago
- On the Complementarity between Pre-Training and Back-Translation for Neural Machine Translation (Findings of EMNLP 2021))☆13Nov 21, 2021Updated 4 years ago
- This repository mirrors the principal Gitlab repository of the Chebyshev Accelerated Subspace iteration Eigensolver. If you want to contr…☆21Jul 8, 2026Updated last month
- ☆19Nov 10, 2024Updated last year
- DuoDecoding: Hardware-aware Heterogeneous Speculative Decoding with Dynamic Multi-Sequence Drafting☆19Mar 4, 2025Updated last year
- Use contrastive learning to train a large language model (LLM) as a retriever☆12Jul 19, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Must-read papers on Knowledge Graph Reasoning (KGR)☆21Mar 16, 2020Updated 6 years ago
- Coco is a proactive co-assistant that connects user workspace with a broader ecosystem of AI agents.☆33Updated this week
- Independent robustness evaluation of Improving Alignment and Robustness with Short Circuiting☆18Apr 15, 2025Updated last year
- Scaling Agentic Environments Automatically.☆68Mar 26, 2026Updated 5 months ago
- [ICML 2024] Junk DNA Hypothesis: A Task-Centric Angle of LLM Pre-trained Weights through Sparsity; Lu Yin*, Ajay Jaiswal*, Shiwei Liu, So…☆16Apr 21, 2025Updated last year
- Code for "A Multi-Task BERT Model for Schema-Guided Dialogue State Tracking"☆14May 26, 2023Updated 3 years ago
- Synthetic data library used in operator learning for PDE problems that overcomes dependence on classical solvers such as finite differenc…☆18Aug 8, 2024Updated 2 years ago
- Source code of ACL 2023 accepted paper "AD-KD: Attribution-Driven Knowledge Distillation for Language Model Compression"☆13Jun 14, 2023Updated 3 years ago
- ☆50Mar 6, 2026Updated 5 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- The official implementation of Preference Data Reward-Augmentation.☆18May 1, 2025Updated last year
- ☆16May 22, 2025Updated last year
- ☆14Jan 6, 2025Updated last year
- This is the repository for EMNLP 2022 paper "Efficient Zero-shot Event Extraction with Context-Definition Alignment"☆11Dec 27, 2024Updated last year
- SDD 规格驱动开发 + Harness 多 Agent 编排;支持在 OpenCode、Claude Code、Codex CLI 间切换执行引擎,完成从规格到落地的自动化开发任务。☆15Apr 8, 2026Updated 4 months ago
- Implementation for "DeltaPhi: Learning Physical Trajectory Residual for PDE Solving"☆13Jun 17, 2024Updated 2 years ago
- ☆38Dec 26, 2022Updated 3 years ago