顾名思义,用手搓的方式学LLM
☆21Apr 14, 2026Updated 3 months ago
Alternatives and similar repositories for learn-LLM-manually
Users that are interested in learn-LLM-manually are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- ☆21Apr 30, 2026Updated 2 months ago
- Official Implementation of Trajectory-Refined Distillation☆26Jun 9, 2026Updated last month
- TVRBench: Target Viewpoint Reproduction Benchmark for Active Spatial Intelligence☆25Jun 2, 2026Updated last month
- SPUQ: Perturbation-Based Uncertainty Quantification for Large Language Models☆17Jun 24, 2024Updated 2 years ago
- Unsupervised Learning of Generalizable Robot Motion from Compact State Representation☆40Jun 10, 2026Updated last month
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- [ICML 2025] Official Implementation of Hierarchical Masked Autoregressive Models with Low-Resolution Token Pivots☆30May 28, 2025Updated last year
- [ICLR'26] Official PyTorch implementation of "Time Is a Feature: Exploiting Temporal Dynamics in Diffusion Language Models".☆66Mar 5, 2026Updated 4 months ago
- Author's implementation of DemoDiffusion.☆70Jan 14, 2026Updated 6 months ago
- Official Implementation of "Geometrically-Constrained Agent for Spatial Reasoning"☆89Apr 7, 2026Updated 3 months ago
- ☆53May 19, 2022Updated 4 years ago
- [arXiv: 2502.05178] QLIP: Text-Aligned Visual Tokenization Unifies Auto-Regressive Multimodal Understanding and Generation☆97Mar 1, 2025Updated last year
- 使用tensorflow2.1实现联合学习推荐模型,并加入差分隐私噪声进行隐私保护。☆54Mar 25, 2023Updated 3 years ago
- ☆98May 6, 2025Updated last year
- [ICCV2025]LeanVAE: An Ultra-Efficient Reconstruction VAE for Video Diffusion Models☆111Updated this week
- Serverless GPU API endpoints on Runpod - Get Bonus Credits • AdSkip the infrastructure headaches. Auto-scaling, pay-as-you-go, no-ops approach lets you focus on innovating your application.
- [NeurIPS 2025 Spotlight] A Generalist Diffusion Model for Vision Perception☆318Sep 21, 2025Updated 10 months ago
- Official Repository for ACL 2024 Paper SafeDecoding: Defending against Jailbreak Attacks via Safety-Aware Decoding☆154Jul 19, 2024Updated 2 years ago
- [NeurIPS 2025] Vision as a Dialect: Unifying Visual Understanding and Generation via Text-Aligned Representations☆202Sep 18, 2025Updated 10 months ago
- code for GaussianShader: 3D Gaussian Splatting with Shading Functions for Reflective Surfaces☆405May 20, 2024Updated 2 years ago
- FedMD: Heterogenous Federated Learning via Model Distillation☆166Jun 3, 2021Updated 5 years ago
- Bilingual (中文+EN) ML / LLM / diffusion / agent interview cheat sheets for AI 秋招 — generated by ARIS /interview-cheatsheet, rendered by /r…☆313Jul 14, 2026Updated last week
- Video Generation, Physical Commonsense, Semantic Adherence, VideoCon-Physics☆205Jan 30, 2026Updated 5 months ago
- Official code for "SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization"☆353May 20, 2026Updated 2 months ago
- World Simulator Assistant for Physics-Aware Text-to-Video Generation☆276Sep 22, 2025Updated 10 months ago
- Managed Kubernetes at scale on DigitalOcean • AdDigitalOcean Kubernetes includes the control plane, bandwidth allowance, container registry, automatic updates, and more for free.
- QiZhenGPT: An Open Source Chinese Medical Large Language Model|一个开源的中文医疗大语言模型☆776Aug 9, 2024Updated last year
- [ECCV 2024]"FSGS: Real-Time Few-Shot View Synthesis using Gaussian Splatting", Zehao Zhu*, Zhiwen Fan*, Yifan Jiang, Zhangyang Wang☆544Feb 27, 2024Updated 2 years ago
- TensorFlow Implementation of Attentional Factorization Machine☆405Jul 16, 2018Updated 8 years ago
- [CVPR 2024 Highlight] OPERA: Alleviating Hallucination in Multi-Modal Large Language Models via Over-Trust Penalty and Retrospection-Allo…☆411Aug 24, 2024Updated last year
- [ECCV2024] Relightable 3D Gaussian: Real-time Point Cloud Relighting with BRDF Decomposition and Ray Tracing☆678Aug 14, 2024Updated last year
- Code for the IJCAI'19 paper "Deep Session Interest Network for Click-Through Rate Prediction"☆450May 23, 2023Updated 3 years ago
- TriAttention — Efficient long reasoning with trigonometric KV cache compression. Enables OpenClaw local deployment on memory-constrained …☆828Jul 14, 2026Updated last week
- Awesome List for On-Policy Distillation☆760Jun 23, 2026Updated 3 weeks ago
- Rethinking On-Policy Distillation of Large Language Models: Phenomenology, Mechanism, and Recipe☆835Jun 29, 2026Updated 3 weeks ago
- Bare Metal GPUs on DigitalOcean Gradient AI • AdPurpose-built for serious AI teams training foundational models, running large-scale inference, and pushing the boundaries of what's possible.
- GenEval: An object-focused framework for evaluating text-to-image alignment☆470Mar 3, 2025Updated last year
- 🔥 大模型 & Agent 面试八股文完全指南 | LLM & Agent Interview Preparation Guide☆576Feb 28, 2026Updated 4 months ago
- [ICML2024] Official code for GaussianPro: 3D Gaussian Splatting with Progressive Propagation☆841Aug 24, 2025Updated 10 months ago
- Inference-Time Intervention: Eliciting Truthful Answers from a Language Model☆581Jan 28, 2025Updated last year
- SimCSE在中文任务上的简单实验☆605Aug 7, 2023Updated 2 years ago
- [ICLR 2024] Real-time Photorealistic Dynamic Scene Representation and Rendering with 4D Gaussian Splatting☆1,017Jan 31, 2026Updated 5 months ago
- 🎓 系统性大语言模型构建课程|🛠️ 覆盖预训练数据工程、Tokenizer、Transformer、MoE、GPU 编程 (CUDA/Triton)、分布式训练、Scaling Laws、推理优化及对齐 (SFT/RLHF/GRPO)|🚀 6 个渐进式作业 + 代码驱…☆1,061Jun 26, 2026Updated 3 weeks ago