Llemma formal2formal (tactic prediction) theorem proving experiments
☆20Oct 17, 2023Updated 2 years ago
Alternatives and similar repositories for llemma_formal2formal
Users that are interested in llemma_formal2formal are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- https://albertqjiang.github.io/Portal-to-ISAbelle/☆57Sep 6, 2023Updated 3 years ago
- Code for the paper LeanReasoner: Boosting Complex Logical Reasoning with Lean: https://arxiv.org/pdf/2403.13312.pdf☆28May 25, 2024Updated 2 years ago
- ☆18Oct 12, 2022Updated 3 years ago
- ☆17May 31, 2023Updated 3 years ago
- ☆45Sep 19, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The is the official implementation of "Lyra: Orchestrating Dual Correction in Automated Theorem Proving"☆15Jul 2, 2024Updated 2 years ago
- ☆73Sep 30, 2023Updated 2 years ago
- ☆30Dec 27, 2024Updated last year
- Retrieval-Augmented Theorem Provers for Lean☆338Jan 30, 2025Updated last year
- The official repository for the paper Multilingual Mathematical Autoformalization☆39May 20, 2024Updated 2 years ago
- Git for "Stepwise Self-Consistent Mathematical Reasoning with Large Language Models"☆12Nov 26, 2024Updated last year
- The code of CIKM 2023 (Oral Presentation) : A Multi-Task Semantic Decomposition Framework with Task-specific Pre-training for Few-Shot NE…☆14Jul 19, 2024Updated 2 years ago
- ☆18Jul 12, 2025Updated last year
- Code & data for ICLR 2024 spotlight paper: 🍯MUSTARD: Mastering Uniform Synthesis of Theorem and Proof Data☆43May 29, 2024Updated 2 years ago
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Model-Agnostic Adaptive Testing☆10Dec 16, 2020Updated 5 years ago
- 🤖ConvRe🤯: An Investigation of LLMs’ Inefficacy in Understanding Converse Relations (EMNLP 2023)☆24Oct 10, 2023Updated 2 years ago
- From Accuracy to Robustness: A Study of Rule- and Model-based Verifiers in Mathematical Reasoning.☆25Oct 7, 2025Updated 11 months ago
- llmstep: [L]LM proofstep suggestions in Lean 4.☆155Nov 11, 2023Updated 2 years ago
- [EMNLP 2022] Official Pytorch implementation for "Tiny-NewsRec: Efficient and Effective PLM-based News Recommendation"☆19Sep 18, 2023Updated 3 years ago
- [ACL 2024 Findings] The official repo for "ConceptMath: A Bilingual Concept-wise Benchmark for Measuring Mathematical Reasoning of Large …☆26May 29, 2024Updated 2 years ago
- ☆10Apr 16, 2024Updated 2 years ago
- ☆26Aug 23, 2024Updated 2 years ago
- ☆201Jan 23, 2023Updated 3 years ago
- End-to-end encrypted cloud storage - Proton Drive • AdSpecial offer: 40% Off Yearly / 80% Off First Month. Protect your most important files, photos, and documents from prying eyes.
- Official Code Repository for paper "HYDRA: Model Factorization Framework for Black-Box LLM Personalization"☆16Oct 7, 2024Updated last year
- PreAct: Prediction Enhances Agent's Planning Ability (Coling2025)☆31Dec 12, 2024Updated last year
- [ACL 2024, Main Conference] CoGenesis: A Framework Collaborating Large and Small Language Models for Secure Context-Aware Instruction Fol…☆15Aug 7, 2024Updated 2 years ago
- https://scale.com/research/mrt☆20Mar 16, 2026Updated 6 months ago
- PyTorch implementation of experiments in the paper Aligning Language Models with Human Preferences via a Bayesian Approach☆32Nov 6, 2023Updated 2 years ago
- ☆73Apr 2, 2024Updated 2 years ago
- ☆25Feb 3, 2026Updated 7 months ago
- ☆14Apr 27, 2022Updated 4 years ago
- ☆28May 8, 2024Updated 2 years ago
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆14Jul 17, 2025Updated last year
- EQUATE (Evaluating Quantitative Understanding Aptitude in Textual Entailment), framework for evaluating quantitative reasoning ability in…☆14Feb 13, 2022Updated 4 years ago
- ☆83Apr 18, 2024Updated 2 years ago
- ☆12Jun 19, 2025Updated last year
- [NLPCC 2024] Shared Task 10: Regulating Large Language Models☆14Jun 12, 2024Updated 2 years ago
- Must-read papers on network representation learning (NRL)/network embedding (NE)☆12Nov 9, 2017Updated 8 years ago
- [ACL 2024 Findings] CriticBench: Benchmarking LLMs for Critique-Correct Reasoning☆31Mar 5, 2024Updated 2 years ago