☆65Dec 6, 2024Updated last year
Alternatives and similar repositories for llms_can_learn_rules
Users that are interested in llms_can_learn_rules are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆22Nov 26, 2024Updated last year
- ☆20Nov 20, 2024Updated last year
- ☆34Nov 21, 2023Updated 2 years ago
- ☆18Oct 16, 2024Updated last year
- 同济大学信息安全原理课程作业(Project 1-2 l2tp&ipsec 协议分析,2023). Project 1-2 l2tp and ipsec of Information Security Principles Course of Tongji Universi…☆12Jun 30, 2023Updated 3 years ago
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- ☆20Nov 3, 2024Updated last year
- Steganographic text via language models☆27Nov 21, 2022Updated 3 years ago
- Official Code Repo for the paper "Learning to Play Atari in a World of Tokens" accepted at ICML, 2024☆11Jun 6, 2024Updated 2 years ago
- ☆11Mar 31, 2024Updated 2 years ago
- ☆10Dec 10, 2024Updated last year
- ☆17Feb 12, 2025Updated last year
- NAACL '24 (Best Demo Paper RunnerUp) / MlSys @ NeurIPS '23 - RedCoast: A Lightweight Tool to Automate Distributed Training and Inference☆69Dec 9, 2024Updated last year
- Code for the paper: Sparsely Changing Latent States for Prediction and Planning in Partially Observable Domains☆12Nov 12, 2021Updated 4 years ago
- ☆70Mar 30, 2025Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- Implementation of the Quiet-STAR paper (https://arxiv.org/pdf/2403.09629.pdf)☆57Aug 8, 2024Updated 2 years ago
- Code for Paper 'DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards'☆18May 21, 2026Updated 2 months ago
- a benchmark to evaluate the situated inductive reasoning☆19Jan 7, 2025Updated last year
- High-performance C++ inference engine for Diffusion Language Models (LLaDA, SEDD, MDLM)☆17Apr 5, 2026Updated 4 months ago
- Code for "An Empirical Study of Retrieval Augmented Generation with Chain-of-Thought"☆18Jul 27, 2024Updated 2 years ago
- cdn 解决最优网络流传输分配的服务器选点问题☆11Nov 11, 2018Updated 7 years ago
- LLM training in simple, raw C/CUDA☆15Dec 5, 2024Updated last year
- ☆105Dec 6, 2024Updated last year
- ☆53May 11, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Source code for GreaTer ICLR 2025 - Gradient Over Reasoning makes Smaller Language Models Strong Prompt Optimizers☆36Apr 18, 2025Updated last year
- ☆33Sep 27, 2024Updated last year
- V-MPO torch version with DMLab30 and GTrXL☆13Mar 1, 2021Updated 5 years ago
- Codev-Bench (Code Development Benchmark), a fine-grained, real-world, repository-level, and developer-centric evaluation framework. Codev…☆49Nov 6, 2024Updated last year
- [NeurIPS'24] Weak-to-Strong Search: Align Large Language Models via Searching over Small Language Models☆67Dec 10, 2024Updated last year
- In the consideration that the college student behavior dataset is scarce, we set up a novel college students’ action dataset in the class…☆22Aug 1, 2022Updated 4 years ago
- KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality☆48May 19, 2026Updated 2 months ago
- ☆20Jul 2, 2026Updated last month
- Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision☆18Apr 1, 2025Updated last year
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- Official implementation of “Code Recommendation for Open Source Software Developers" at The Web Conference 2023 (WWW 2023).☆18Apr 29, 2023Updated 3 years ago
- Camera streaming on Android using ffmpeg, x264, live555, forked from https://github.com/parizene/android-streamer ,but some function re…☆11Aug 26, 2018Updated 7 years ago
- ☆26Jan 27, 2026Updated 6 months ago
- Commit0: Library Generation from Scratch☆191Feb 24, 2026Updated 5 months ago
- Code for the paper LeanReasoner: Boosting Complex Logical Reasoning with Lean: https://arxiv.org/pdf/2403.13312.pdf☆27May 25, 2024Updated 2 years ago
- Source code of "Reasons to Reject? Aligning Language Models with Judgments"☆58Feb 29, 2024Updated 2 years ago
- ☆19Feb 18, 2026Updated 5 months ago