☆44Jul 16, 2025Updated last year
Alternatives and similar repositories for information_flow
Users that are interested in information_flow are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆19Mar 10, 2026Updated 5 months ago
- [NeurIPS'25] Towards Interpretability Without Sacrifice: Faithful Dense Layer Decomposition with Mixture of Decoders☆16May 28, 2025Updated last year
- ☆17Apr 7, 2025Updated last year
- [ICLR 2025] Code and Data Repo for Paper "Latent Space Chain-of-Embedding Enables Output-free LLM Self-Evaluation"☆101Dec 19, 2024Updated last year
- [ICLR 2025] RaSA: Rank-Sharing Low-Rank Adaptation☆10May 19, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Implementation and evaluation of Scaling Embedding Layers in Language Models research paper☆15Feb 2, 2026Updated 6 months ago
- ☆22Jun 2, 2026Updated 2 months ago
- SPAA'21: Efficient Stepping Algorithms and Implementations for Parallel Shortest Paths☆21Aug 10, 2024Updated 2 years ago
- Scalable Computation of Hessian Diagonals☆14Jun 2, 2024Updated 2 years ago
- Official implementation for Text Generation Beyond Discrete Token Sampling☆26Aug 11, 2025Updated last year
- ☆21Jan 21, 2026Updated 6 months ago
- Towards Meta-Pruning via Optimal Transport, ICLR 2024 (Spotlight)☆18Dec 5, 2024Updated last year
- Official implementation of the paper "Pretraining Language Models to Ponder in Continuous Space"☆27Jul 21, 2025Updated last year
- Codebase for EMNLP 2025 Findings paper "Text or Pixels? Evaluating Efficiency and Understanding of LLMs with Visual Text Inputs"☆19Nov 14, 2025Updated 9 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- Multi-head Recurrent Layer Attention for Vision Network☆23Mar 2, 2023Updated 3 years ago
- ☆13Apr 15, 2024Updated 2 years ago
- Official implementation of Latent-SFT: teaching LLMs to reason with vocabulary-space latent chains.☆58May 18, 2026Updated 2 months ago
- MonarchAttention: Zero-Shot Conversion to Fast, Hardware-Aware Structured Attention (NeurIPS'25 Spotlight)☆26Feb 22, 2026Updated 5 months ago
- Generic library for neural collapse and several derivative works on the phenomenon.☆17Apr 14, 2025Updated last year
- RocketSimu's repository☆10Feb 18, 2020Updated 6 years ago
- AdaSkip: Adaptive Sublayer Skipping for Accelerating Long-Context LLM Inference☆21Jan 24, 2025Updated last year
- [PACT'24] GraNNDis. A fast and unified distributed graph neural network (GNN) training framework for both full-batch (full-graph) and min…☆10Aug 13, 2024Updated 2 years ago
- [ICLR'26] "Nabla-Reasoner: LLM Reasoning via Test-Time Gradient Descent in Latent Space" by Peihao Wang*, Ruisi Cai*, Zhen Wang, Hongyuan…☆35Mar 10, 2026Updated 5 months ago
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- [ACL'25] We propose a novel fine-tuning method, Separate Memory and Reasoning, which combines prompt tuning with LoRA.☆88Nov 2, 2025Updated 9 months ago
- Official implementation of Gumbel Distillation for Parallel Text Generation☆20Mar 24, 2026Updated 4 months ago
- WavSpA: Wavelet Space Attention for Enhancing Transformer's Long Sequence Learning☆13Feb 24, 2024Updated 2 years ago
- Large language models, physics-based modeling, experimental measurements: the trinity of data-scarce learning of polymer properties☆14Sep 4, 2025Updated 11 months ago
- ☆15Apr 20, 2025Updated last year
- ☆14Oct 23, 2025Updated 9 months ago
- ☆63Jul 21, 2024Updated 2 years ago
- ☆14Dec 13, 2024Updated last year
- ☆11May 26, 2023Updated 3 years ago
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- Official code for PLoP☆20Mar 6, 2026Updated 5 months ago
- ☆22Jun 25, 2023Updated 3 years ago
- codes and plots for "Active-Dormant Attention Heads: Mechanistically Demystifying Extreme-Token Phenomena in LLMs"☆11Dec 30, 2024Updated last year
- [NeurIPS 2024] A Novel Rank-Based Metric for Evaluating Large Language Models☆59May 28, 2025Updated last year
- Learning to Count without Annotations☆24May 24, 2024Updated 2 years ago
- Dataset of Organic Photovoltaics☆17Updated this week
- [NeurIPS 2025] Think Silently, Think Fast: Dynamic Latent Compression of LLM Reasoning Chains☆98Jun 29, 2026Updated last month