☆65Dec 6, 2024Updated last year
Alternatives and similar repositories for llms_can_learn_rules
Users that are interested in llms_can_learn_rules are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- This the implementation of LeCo☆33Jan 20, 2025Updated last year
- ☆20Nov 20, 2024Updated last year
- ☆18Oct 16, 2024Updated last year
- 同济大学信息安全原理课程作业(Project 1-2 l2tp&ipsec 协议分析,2023). Project 1-2 l2tp and ipsec of Information Security Principles Course of Tongji Universi…☆12Jun 30, 2023Updated 3 years ago
- ☆20Nov 3, 2024Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆64Apr 9, 2024Updated 2 years ago
- Let's Sample Step by Step: Adaptive-Consistency for Efficient Reasoning with LLMs☆41Jan 30, 2024Updated 2 years ago
- Official Code Repo for the paper "Learning to Play Atari in a World of Tokens" accepted at ICML, 2024☆11Jun 6, 2024Updated 2 years ago
- ☆10Dec 10, 2024Updated last year
- Neuro-Symbolic Hierarchical Rule Induction☆15Dec 31, 2022Updated 3 years ago
- Generate the WizardCoder Instruct from the CodeAlpaca☆21Jun 27, 2023Updated 3 years ago
- Code for the paper: Sparsely Changing Latent States for Prediction and Planning in Partially Observable Domains☆11Nov 12, 2021Updated 4 years ago
- ☆70Mar 30, 2025Updated last year
- Implementation of the Quiet-STAR paper (https://arxiv.org/pdf/2403.09629.pdf)☆57Aug 8, 2024Updated last year
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Code for Paper 'DelTA: Discriminative Token Credit Assignment for Reinforcement Learning from Verifiable Rewards'☆17May 21, 2026Updated 2 months ago
- a benchmark to evaluate the situated inductive reasoning☆16Jan 7, 2025Updated last year
- ☆12Jul 30, 2025Updated 11 months ago
- Testing paligemma2 finetuning on reasoning dataset☆18Dec 28, 2024Updated last year
- Pile Deduplication Code☆18May 15, 2023Updated 3 years ago
- ☆16Jul 23, 2024Updated 2 years ago
- LLM training in simple, raw C/CUDA☆15Dec 5, 2024Updated last year
- [ICML 2025] Official implementation of the paper "SkipGPT: Dynamic Layer Pruning Reinvented with Token Awareness and Module Decoupling". …☆21Nov 17, 2025Updated 8 months ago
- ☆105Dec 6, 2024Updated last year
- Virtual machines for every use case on DigitalOcean • AdGet dependable uptime with 99.99% SLA, simple security tools, and predictable monthly pricing with DigitalOcean's virtual machines, called Droplets.
- CVPR 2019 paper "Disentangling Adversarial Robustness and Generalization".☆14Oct 28, 2019Updated 6 years ago
- ☆53May 11, 2025Updated last year
- Source code for GreaTer ICLR 2025 - Gradient Over Reasoning makes Smaller Language Models Strong Prompt Optimizers☆36Apr 18, 2025Updated last year
- ☆33Sep 27, 2024Updated last year
- V-MPO torch version with DMLab30 and GTrXL☆13Mar 1, 2021Updated 5 years ago
- Learn online intrinsic rewards from LLM feedback☆45Dec 17, 2024Updated last year
- [NeurIPS'24] Weak-to-Strong Search: Align Large Language Models via Searching over Small Language Models☆67Dec 10, 2024Updated last year
- KnowRL: Exploring Knowledgeable Reinforcement Learning for Factuality☆48May 19, 2026Updated 2 months ago
- Evaluation on Logical Reasoning and Abstract Reasoning Challenges☆30Apr 21, 2025Updated last year
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- Chain-of-Thought Matters: Improving Long-Context Language Models with Reasoning Path Supervision☆19Apr 1, 2025Updated last year
- Commit0: Library Generation from Scratch☆189Feb 24, 2026Updated 5 months ago
- Source code of "Reasons to Reject? Aligning Language Models with Judgments"☆58Feb 29, 2024Updated 2 years ago
- A benchmark for evaluating the efficiency of LLM-generated code☆17Apr 17, 2025Updated last year
- ☆21Oct 31, 2024Updated last year
- Automatic prompt optimization framework for multi-step agent tasks.☆37Nov 12, 2024Updated last year
- TARGET is a benchmark for evaluating Table Retrieval for Generative Tasks such as Fact Verification and Text-to-SQL☆29Jul 14, 2025Updated last year