Kimi K2 Thinking Agentic Search Unofficial Implementation
☆15Nov 9, 2025Updated 8 months ago
Alternatives and similar repositories for kimi-deepresearch
Users that are interested in kimi-deepresearch are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Pushing Test-Time Scaling Limits of Deep Search with Asymmetric Verification☆21Oct 8, 2025Updated 9 months ago
- RedSearcher's framework for deep search agent trajectory synthesis, QA filtering, and model evaluation, supporting ReACT and DeepSeek-sty…☆23Feb 26, 2026Updated 4 months ago
- LongSeeker: Elastic Context Orchestration for Long-Horizon Search Agents☆24May 29, 2026Updated last month
- Fully open reproduction of DeepSeek-R1☆11Mar 24, 2025Updated last year
- 集中管理所有的prompt。☆14Nov 27, 2024Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of Browse-Master, a tool-augmented web-search agent.☆35Aug 22, 2025Updated 10 months ago
- Official Codebase for "Aligning Diffusion Behaviors with Q-functions for Efficient Continuous Control" (NeurIPS 2024)☆15Oct 29, 2024Updated last year
- Implementation for the paper "Fictitious Synthetic Data Can Improve LLM Factuality via Prerequisite Learning"☆11Jan 10, 2025Updated last year
- The official repo of "WebExplorer: Explore and Evolve for Training Long-Horizon Web Agents"☆120Sep 29, 2025Updated 9 months ago
- ☆35Jun 3, 2025Updated last year
- RewardAnything: Generalizable Principle-Following Reward Models☆44Jun 11, 2025Updated last year
- REDSearch: A scalable, cost-efficient framework for long-horizon search agents. Features complex task synthesis, optimized mid-training, …☆128Feb 26, 2026Updated 4 months ago
- VitaBench: Benchmarking LLM Agents with Versatile Interactive Tasks in Real-world Applications☆23Oct 17, 2025Updated 9 months ago
- Adaptable Agent Populations via a Generative Model of Policies☆12Oct 14, 2021Updated 4 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official eval scripts for JobBench☆26Updated this week
- A simple implementation of ReasonGenRM.☆19Apr 21, 2025Updated last year
- [ACL 2024] Making Long-Context Language Models Better Multi-Hop Reasoners☆20May 28, 2024Updated 2 years ago
- A cog model for the all-mpnet-base-v2 sentence-transformers embedding model.☆15Jan 3, 2024Updated 2 years ago
- ☆16Jul 29, 2025Updated 11 months ago
- This is a detailed code demo on how to conduct Full-Param Supervised Fine-tuning (SFT) and DPO (Direct Preference Optimization)☆20Jan 9, 2025Updated last year
- HierSearch: A Hierarchical Enterprise Deep Search Framework Integrating Local and Web Searches☆40Oct 9, 2025Updated 9 months ago
- 酒馆一键docker启动命令☆17Jun 10, 2026Updated last month
- A preliminary platform for up to 1 million reinforcement learning agents☆11Aug 27, 2017Updated 8 years ago
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- BrowseComp-Plus: A More Fair and Transparent Evaluation Benchmark of Deep-Research Agent (ACL 2026 Main)☆316May 28, 2026Updated last month
- 🔍 Awesome Agentic Search is a curated list of papers, tools, and resources on agentic search—where AI agents plan, search, and reason to…☆60Aug 28, 2025Updated 10 months ago
- gpt-o1 like chain of thoughts with local LLMs in R☆31Oct 15, 2024Updated last year
- a benckmark for evaluating logical reasoning of LLMs☆23Jan 25, 2024Updated 2 years ago
- 🔥🔥🔥Latest Papers, Codes on Uncertainty-based RL☆58Aug 24, 2025Updated 10 months ago
- (ICLR 2025) AgentRefine: Enhancing Agent Generalization through Refinement Tuning☆20Nov 22, 2025Updated 7 months ago
- [ICCAD 2025] Squant☆15Jul 3, 2025Updated last year
- [ACL 2026 Main Conference] WebAggregator☆69Oct 18, 2025Updated 9 months ago
- Run multiple commands in a docker container.☆37Oct 5, 2014Updated 11 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Neural Fictitious Self-Play in Leduc Holdem☆11Jul 4, 2018Updated 8 years ago
- [ICASSP 2025] "FLowHigh: Towards efficient and high-quality audio super-resolution with single-step flow matching"☆33May 12, 2025Updated last year
- Unsupervised Differentiable Multi-aspect Network Embedding (asp2vec) (KDD2020)☆19Aug 19, 2021Updated 4 years ago
- Reinforcement learning - Batched Impala - PyTorch - Mario Kart☆13Jul 21, 2020Updated 5 years ago
- Start a docker image and attach to it☆58Nov 10, 2019Updated 6 years ago
- [ICML'26] Scaling Long-Horizon LLM Agent via Context-Folding☆173May 18, 2026Updated 2 months ago
- Implementation of Direct Preference Optimization☆17Jul 17, 2023Updated 3 years ago