[NeurIPS 2025] L-MTP: Leap Multi-Token Prediction Beyond Adjacent Context for Large Language Models
☆33May 8, 2026Updated 3 months ago
Alternatives and similar repositories for L-MTP
Users that are interested in L-MTP are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- The implementation of paper "Strategy-aware Bundle Recommender System", SIGIR'23.☆17Sep 4, 2023Updated 2 years ago
- [TPAMI 2026] Principled Multimodal Representation Learning☆39Jul 26, 2026Updated 3 weeks ago
- [KDD 2025] Fine-tuning Multimodal Large Language Models for Product Bundling☆16Sep 20, 2025Updated 10 months ago
- The implementation of paper "EliMRec: Eliminating single-modal bias in multimedia recommendation", MM'22.☆24Dec 7, 2023Updated 2 years ago
- [MM 2025] Towards Modality Generalization: A Benchmark and Prospective Analysis☆31May 22, 2025Updated last year
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- A curated list of papers, tools, and resources on Multi-Token Prediction (MTP) and related techniques in Large Language Models (LLMs), Sp…☆196Aug 9, 2026Updated last week
- The implementation of paper "Self-supervised learning for multimedia recommendation", TMM'22.☆11Jul 4, 2022Updated 4 years ago
- Official implementation for "RSafe: Incentivizing proactive reasoning to build robust and adaptive LLM safeguards"☆17Jan 31, 2026Updated 6 months ago
- Regularly Truncated M-estimators for Learning with Noisy Labels☆11Apr 24, 2024Updated 2 years ago
- NeurIPS'2022: Pluralistic Image Completion with Gaussian Mixture Models☆14Jan 28, 2023Updated 3 years ago
- A terminal agent powered by OpenAI/Claude/Gemini API☆18Sep 20, 2024Updated last year
- A curated list of Vision (video/image) to Audio Generation☆107Feb 10, 2026Updated 6 months ago
- [NeurIPS 2025] The implementation of paper "The Emergence of Abstract Thought in Large Language Models Beyond Any Language"☆20Jun 9, 2025Updated last year
- Independent robustness evaluation of Improving Alignment and Robustness with Short Circuiting☆18Apr 15, 2025Updated last year
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- ☆18Oct 12, 2022Updated 3 years ago
- Official implementation for "ALI-Agent: Assessing LLMs'Alignment with Human Values via Agent-based Evaluation"☆21Jan 31, 2026Updated 6 months ago
- ☆16Sep 10, 2024Updated last year
- QuantClaw is a plug-and-play task-type routing quantization plugin for OpenClaw.☆116Apr 27, 2026Updated 3 months ago
- ☆15Apr 13, 2023Updated 3 years ago
- ☆29Jun 2, 2026Updated 2 months ago
- [CVPR 2026 Oral] A training-free, mask-free framework for 3D shape editing.☆52May 9, 2026Updated 3 months ago
- (ICLR 2025 Spotlight) DEEM: Official implementation of Diffusion models serve as the eyes of large language models for image perception.☆51Jul 1, 2025Updated last year
- About face technology☆20Feb 9, 2023Updated 3 years ago
- Managed Database hosting by DigitalOcean • AdPostgreSQL, MySQL, MongoDB, Kafka, Valkey, and OpenSearch available. Automatically scale up storage and focus on building your apps.
- Code for ICLR 2023 Harnessing Out-Of-Distribution Examples via Augmenting Content and Style☆13Jul 3, 2023Updated 3 years ago
- Improving large language models with concept-aware fine-tuning (CAFT)☆29Jan 31, 2026Updated 6 months ago
- Reading comprehension based question-answering model for news articles.☆11Jun 22, 2022Updated 4 years ago
- TPAMI: Classification with noisy labels by importance reweighting.☆39Oct 4, 2019Updated 6 years ago
- [AAAI26] Trade-offs in Large Reasoning Models: An Empirical Analysis of Deliberative and Adaptive Reasoning over Foundational Capabilitie…☆11Feb 7, 2026Updated 6 months ago
- ☆14Oct 7, 2023Updated 2 years ago
- ☆17May 2, 2024Updated 2 years ago
- ☆22Mar 11, 2026Updated 5 months ago
- [ICCV2025 Highlight] Where, What, Why: Towards Explainable Driver Attention Prediction☆54Oct 31, 2025Updated 9 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The code of the paper of "A Differentiable Semantic Metric Approximation in Probabilistic Embedding for Cross-Modal Retrieval" accepted b…☆19Jan 16, 2024Updated 2 years ago
- Improving Adversarial Robustness via Mutual Information Estimation☆11Apr 2, 2024Updated 2 years ago
- Towards Defending against Adversarial Examples via Attack-Invariant Features☆13Oct 12, 2023Updated 2 years ago
- A Chinese-focused PyTorch framework for exploring Attention Residuals in Qwen3-style causal LMs, with baseline, Block AttnRes, Full AttnR…☆21May 3, 2026Updated 3 months ago
- Marathon: A Multiple-choice Long Context Evaluation Benchmark for Large Language Models.☆10May 16, 2024Updated 2 years ago
- ☆14Aug 28, 2019Updated 6 years ago
- [ICML 2025] DreamDPO: Aligning Text-to-3D Generation with Human Preferences via Direct Preference Optimization☆22May 24, 2025Updated last year