☆62Mar 3, 2025Updated last year
Alternatives and similar repositories for LLM-Microscope
Users that are interested in LLM-Microscope are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆72Aug 27, 2024Updated last year
- The code for "VISTA: Enhancing Long-Duration and High-Resolution Video Understanding by VIdeo SpatioTemporal Augmentation" [CVPR2025]☆20Feb 27, 2025Updated last year
- Official PyTorch Implementation for Vision-Language Models Create Cross-Modal Task Representations, ICML 2025☆34May 1, 2025Updated last year
- Convert MUSE from TensorFlow to PyTorch and ONNX☆11May 22, 2024Updated 2 years ago
- The official code repo and data hub of top_nsigma sampling strategy for LLMs.☆26Feb 11, 2025Updated last year
- GPUs on demand by Runpod - Special Offer Available • AdRun AI, ML, and HPC workloads on powerful cloud GPUs—without limits or wasted spend. Deploy GPUs in under a minute and pay by the second.
- Cramming 1568 Tokens into a Single Vector and Back Again: Exploring the Limits of Embedding Space Capacity (ACL 2025, oral)☆35Jun 14, 2025Updated last year
- ☆14Apr 10, 2025Updated last year
- Official implementation of RMoE (Layerwise Recurrent Router for Mixture-of-Experts)☆33Aug 4, 2024Updated last year
- Skoltech NLA 2024 course.☆37Dec 10, 2024Updated last year
- ☆25Dec 13, 2024Updated last year
- [ICLR 2026] Rectifying LLM Thought From Lens of Optimization☆15Dec 5, 2025Updated 7 months ago
- ☆14Jun 10, 2023Updated 3 years ago
- Seminars from 2024 Machine Learning course☆14Mar 9, 2024Updated 2 years ago
- OpenSportsLib is a library designed for advanced video understanding in soccer. It provides state-of-the-art tools for action recognition…☆18Updated this week
- Deploy to Railway using AI coding agents - Free Credits Offer • AdUse Claude Code, Codex, OpenCode, and more. Autonomous software development now has the infrastructure to match with Railway.
- ☆151Sep 12, 2025Updated 10 months ago
- Intentional is an open-source framework to build reliable LLM chatbots that actually talk and behave as you expect.☆12Dec 31, 2024Updated last year
- ☆34Apr 14, 2025Updated last year
- Fast and customizable framework for automatic and quick Causal Inference in Python☆129Updated this week
- Repository containing lectures from 2024 Machine Learning course☆18Feb 29, 2024Updated 2 years ago
- [EMNLP 2025 Main] Official implementation of VRoPE: Rotary Position Embedding for Video Large Language Models.☆28Nov 18, 2025Updated 8 months ago
- Bunch of notebooks for pre-training custom Saiga-like LLM☆12Feb 9, 2024Updated 2 years ago
- Official Repo for DAC-RL: Training LLMs for Divide-and-Conquer Reasoning Elevates Test-Time Scalability☆16Feb 26, 2026Updated 4 months ago
- [COLM'25] Missing Premise exacerbates Overthinking: Are Reasoning Models losing Critical Thinking Skill?☆39Jun 5, 2025Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- Code repository for the paper "The Inherent Limits of Pretrained LLMs: The Unexpected Convergence of Instruction Tuning and In-Context Le…☆14Jan 16, 2025Updated last year
- ☆19Mar 5, 2024Updated 2 years ago
- Top 3 solution for CVPR24 SEGMENT ANYTHING IN MEDICAL IMAGES ON LAPTOP Challenge☆11Apr 8, 2025Updated last year
- ☆15Apr 14, 2025Updated last year
- Official repository for the paper Number Cookbook: Number Understanding of Language Models and How to Improve It.☆21Mar 31, 2025Updated last year
- RAG benchmark☆31Feb 6, 2026Updated 5 months ago
- [NeurIPS 2024] How do Large Language Models Handle Multilingualism?☆52Nov 8, 2024Updated last year
- Top ML papers of the week.☆46Updated this week
- This repository contains a deep learning-based approach for improving A* search efficiency on grid graphs. By learning instance-dependent…☆17Oct 2, 2024Updated last year
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆99Mar 28, 2025Updated last year
- ☆31Sep 23, 2024Updated last year
- [NeurIPS 2024] Goldfish Loss: Mitigating Memorization in Generative LLMs☆98Nov 17, 2024Updated last year
- AI-generated text boundary detection with RoFT☆26Sep 9, 2024Updated last year
- Paper dataset for "Factored Verification: Detecting and Reducing Hallucination in Summaries of Academic Papers"☆13Oct 20, 2024Updated last year
- Train punctuation and capitalization models for different languages☆26Apr 2, 2022Updated 4 years ago
- ☆13Jan 22, 2025Updated last year