☆19Nov 24, 2025Updated 8 months ago
Alternatives and similar repositories for how-do-llms-use-their-depth
Users that are interested in how-do-llms-use-their-depth are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- [ICLR 2025] Monet: Mixture of Monosemantic Experts for Transformers☆79Jun 23, 2025Updated last year
- ☆25Oct 22, 2025Updated 9 months ago
- Official codebase for our paper "Do Language Models Use Their Depth Efficiently?"☆29Jun 25, 2025Updated last year
- Paper Implementation of Self-Rewarding Language Models☆13Feb 1, 2024Updated 2 years ago
- SNS Hashtag Offilne Event Managing Platform☆12Feb 11, 2022Updated 4 years ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- A Jensen-Shannon Divergence Driven Mechanistic Study of Context Attribution in Retrieval-Augmented Generation☆15Aug 28, 2025Updated 11 months ago
- a Jax/Flax inference code of StarCoder☆12Jun 12, 2023Updated 3 years ago
- Files used for the evaluation of uiCA☆19Dec 14, 2022Updated 3 years ago
- docTR by Mindee (Document Text Recognition) - a seamless, high-performing & accessible library for OCR-related tasks powered by Deep Lear…☆11May 19, 2026Updated 2 months ago
- Code for Negation Neglect☆16May 22, 2026Updated 2 months ago
- Official implementation of ICLR 2026 paper "LUMINA: Detecting Hallucinations in RAG System with Context–Knowledge Signals"☆18Jan 31, 2026Updated 6 months ago
- Train a SmolLM-style llm on fineweb-edu in JAX/Flax with an assortment of optimizers.☆19Jul 24, 2025Updated last year
- ☆22Dec 4, 2025Updated 8 months ago
- AgentOpt automatically finds the best LLM model combination for each step of your agent — optimizing for accuracy, cost, and latency.☆86Jul 18, 2026Updated 3 weeks ago
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- AAAI'23 Workshop, A-ColViT☆11Jun 16, 2023Updated 3 years ago
- Learning to Generate STRUCTURED Output with Schema Reinforcement Learning☆26Mar 2, 2025Updated last year
- TCube generates rich and fluent narratives that describes the characteristics, trends, and anomalies of any time-series data (domain-agno…☆15Sep 8, 2021Updated 4 years ago
- ☆14Jun 11, 2021Updated 5 years ago
- Mixture of Cognitive Reasoners: Modular Reasoning with Brain-Like Specialization☆46Feb 7, 2026Updated 6 months ago
- 🧠Plan-and-Budget: Training-free test-time reasoning framework for adaptive token allocation in large language models (ICLR 2026).☆15Mar 2, 2026Updated 5 months ago
- Design for Error Detection in Deep-Research Agents Trajectories.☆22Jun 4, 2026Updated 2 months ago
- ☆31Sep 17, 2024Updated last year
- Welcome to the LLM Tutorials and RAG Implementations repository! This repository provides tutorials, guides, and implementations for work…☆13Jul 1, 2025Updated last year
- Proton VPN Special Offer - Get 70% off • AdSpecial partner offer. Trusted by over 100 million users worldwide. Tested, Approved and Recommended by Experts.
- 53 implementations of synthetic learning problems from Geoffrey Hinton's experimental papers (1981-2022). Pure numpy, laptop-runnable, pa…☆33May 11, 2026Updated 2 months ago
- An interactive playbook for learning Claude Code through source analysis, architecture breakdowns, and guided code-reading paths.☆18Apr 1, 2026Updated 4 months ago
- The PsiloQA pipeline automates the construction of a multilingual, span-level hallucination detection dataset with contexts.☆16Apr 24, 2026Updated 3 months ago
- ☆11Jan 25, 2021Updated 5 years ago
- [CVPR'26] SafeGRPO: Self-Rewarded Multimodal Safety Alignment via Rule-Governed Policy Optimization☆22Feb 19, 2026Updated 5 months ago
- 광운대학교 컴퓨터 비전 AI 경진대회 1등 솔루션입니다.☆15Oct 5, 2022Updated 3 years ago
- We conduct a preregistered experiment to investigate whether fact checks provided by a large language model can serve as an effective mis…☆13Dec 14, 2024Updated last year
- Submission to the inverse scaling prize☆23Jul 23, 2023Updated 3 years ago
- ☆19May 19, 2024Updated 2 years ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- [ACL '26] source code for the paper: "Long-Chain Reasoning Distillation via Adaptive Prefix Alignment"☆16Jan 21, 2026Updated 6 months ago
- A collection of lightweight interpretability scripts to understand how LLMs think☆91Mar 18, 2026Updated 4 months ago
- [COLM 2024] Early Weight Averaging meets High Learning Rates for LLM Pre-training☆19Oct 12, 2024Updated last year
- 🥉171st place in Google brain solution🥉☆10Jul 25, 2022Updated 4 years ago
- experimental port of nervana neon kernels in OpenCL☆11Jul 24, 2016Updated 10 years ago
- An Adaptive Multi-Agent Framework for Dynamic Fact-Checking Evaluation of Large Language Models☆18Feb 27, 2025Updated last year
- [ICLR 2026] RPG: KL-Regularized Policy Gradient (https://arxiv.org/abs/2505.17508)☆76Jun 29, 2026Updated last month