agent skill as memory layer
☆21Jan 30, 2026Updated 8 months ago
Alternatives and similar repositories for skill-memory
Users that are interested in skill-memory are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆11Oct 22, 2023Updated 2 years ago
- To mitigate position bias in LLMs, especially in long-context scenarios, we scale only one dimension of LLMs, reducing position bias and …☆12Jun 18, 2024Updated 2 years ago
- Code for the paper "Faster Neural Network Training with Approximate Tensor Operations"☆10Oct 23, 2021Updated 4 years ago
- Build event-driven workflows with python async functions☆39Sep 18, 2024Updated 2 years ago
- Contrastive self-supervised learning using Rényi divergence☆14Oct 21, 2022Updated 3 years ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism☆18Nov 18, 2025Updated 10 months ago
- ☆13Aug 25, 2021Updated 5 years ago
- The official implementation of "LightTransfer: Your Long-Context LLM is Secretly a Hybrid Model with Effortless Adaptation"☆22Apr 22, 2025Updated last year
- Python Implementation for Rethinking RAG based Decoding☆17Sep 10, 2025Updated last year
- Direct preference optimization with f-divergences.☆17Nov 3, 2024Updated last year
- [npj Digital Medicine] Official repository for paper "Automating Expert-Level Medical Reasoning Evaluation of Large Language Models"☆15Jan 2, 2026Updated 9 months ago
- 本项目展示了2022年部分信息检索/数据挖掘顶会论文分类。☆17Jun 13, 2022Updated 4 years ago
- [SIGKDD 2024] Rethinking Fair Graph Neural Networks from Re-balancing☆10Jul 15, 2024Updated 2 years ago
- 知予人工智能:从学习者到研究者☆14Jan 20, 2025Updated last year
- Open source password manager - Proton Pass • AdSecurely store, share, and autofill your credentials with Proton Pass, the end-to-end encrypted password manager trusted by millions.
- 国家统计局中国省市县乡村5级地址抓取,http://www.stats.gov.cn/tjsj/tjbz/tjyqhdmhcxhfdm/2018/index.html☆12Jan 8, 2020Updated 6 years ago
- [ICML 2025] Official code of "DAMA: Data- and Model-aware Alignment of Multi-modal LLMs"☆16May 24, 2025Updated last year
- ☆34Jul 15, 2025Updated last year
- Temperature Schedules for self-supervised contrastive methods on long-tail data (ICLR'23)☆18Apr 25, 2023Updated 3 years ago
- ☆18Dec 24, 2024Updated last year
- Empowering RAG with a versatile model-driven data interface for all-purpose applications!☆17Sep 10, 2024Updated 2 years ago
- Codebase for Temporal SAEs paper☆27Nov 14, 2025Updated 10 months ago
- The contrastive token loss function for reducing generative repetition of autoregressive neural language models.☆13May 11, 2022Updated 4 years ago
- An RL Recipe for Building Agentic LLMs via Self-Imitation on Long-Horizon Agentic Tasks☆40Jan 30, 2026Updated 8 months ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- "Found in the Middle: How Language Models Use Long Contexts Better via Plug-and-Play Positional Encoding" Zhenyu Zhang, Runjin Chen, Shiw…☆35May 7, 2024Updated 2 years ago
- ☆11Oct 14, 2021Updated 4 years ago
- Executive Memory for Coherent Long-Horizon Reasoning!☆87Jan 14, 2026Updated 8 months ago
- ☆21Mar 23, 2022Updated 4 years ago
- [NeurIPS 2024] Official code of $\beta$-DPO: Direct Preference Optimization with Dynamic $\beta$☆51Oct 23, 2024Updated last year
- ☆10Oct 23, 2021Updated 4 years ago
- Reflect-RL: Two-Player Online RL Fine-Tuning for LMs☆18Jul 19, 2025Updated last year
- Probabilistic Alignment of Relations, Instances, and Schema☆21Nov 17, 2025Updated 10 months ago
- A Reading List of logic rules-based reasoning☆17Jan 17, 2020Updated 6 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- [ACL 2025 Findings] Understanding the Repeat Curse in Large Language Models from a Feature Perspective☆22Jun 13, 2025Updated last year
- ☆24Dec 21, 2025Updated 9 months ago
- [CVPR 2024] KEPP: Why Not Use Your Textbook? Knowledge-Enhanced Procedure Planning of Instructional Videos☆12Sep 24, 2024Updated 2 years ago
- ☆13Jul 25, 2024Updated 2 years ago
- [NeurIPS 2026 Pre-to-Post ] From P(y|x) to P(y): Investigating Reinforcement Learning in Pre-train Space☆18Apr 16, 2026Updated 5 months ago
- ☆16Jul 15, 2025Updated last year
- [EMNLP 2024] The official GitHub repo for the paper "Course-Correction: Safety Alignment Using Synthetic Preferences"☆20Oct 2, 2024Updated 2 years ago