☆16Apr 15, 2026Updated 5 months ago
Alternatives and similar repositories for expert-upcycling
Users that are interested in expert-upcycling are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Official Code for What Makes and Breaks Safety Fine-tuning? A Mechanistic Study (NeurIPS 2024)☆11Oct 31, 2024Updated last year
- [CVPR'25] Attention IoU: Examining Biases in CelebA using Attention Maps☆13Mar 26, 2025Updated last year
- A framework to meta-train transformers for causal ICL☆11Jul 15, 2026Updated 2 months ago
- Official Repository of Paper "Watch the Weights: Unsupervised monitoring and control of fine-tuned LLMs"☆15Sep 25, 2025Updated last year
- ☆23Jun 2, 2026Updated 4 months ago
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- Allows two LLMs to communicate and run code in the terminal☆28Dec 8, 2024Updated last year
- ☆20Sep 6, 2025Updated last year
- ☆492Updated this week
- Find similarity between two strings, based on Dice Similarity Coefficient DSC☆13Jan 23, 2023Updated 3 years ago
- The Full Spectrum of Deepnet Hessians at Scale: Dynamics with SGD Training and Sample Size☆19May 19, 2019Updated 7 years ago
- ☆28Feb 20, 2026Updated 7 months ago
- A Unified Framework for Detecting Point and Collective Anomalies in Operating System Logs via Collaborative Transformers☆23Jan 28, 2026Updated 8 months ago
- Code for "What really matters in matrix-whitening optimizers?"☆25Oct 31, 2025Updated 11 months ago
- [ACL 2025] "CoT-UQ: Improving Response-wise Uncertainty Quantification in LLMs with Chain-of-Thought"☆17Apr 3, 2025Updated last year
- AI Agents on DigitalOcean Gradient AI Platform • AdBuild production-ready AI agents using customizable tools or access multiple LLMs through a single endpoint. Create custom knowledge bases or connect external data.
- Postgres Book, one topic at a time.☆109Sep 7, 2026Updated last month
- Design and analyze optimal deep learning models.☆30Aug 2, 2025Updated last year
- FastCuRL: Curriculum Reinforcement Learning with Stage-wise Context Scaling for Efficient LLM Reasoning (EMNLP 2025)☆62Oct 10, 2025Updated 11 months ago
- This repository contains example code to build models on TPUs☆30Feb 17, 2023Updated 3 years ago
- The official baseline implementations for Chronocept. Published at EACL 2026.☆10Aug 4, 2026Updated 2 months ago
- Unauthenticated enumeration of AWS IAM Roles.☆28Apr 18, 2026Updated 5 months ago
- RoPE attention is an exact forward-pass gradient step with softmax intact☆57Sep 6, 2026Updated last month
- PAL: Predictive Analysis & Laws of Large Language Models☆40Jun 29, 2026Updated 3 months ago
- ☆33Jan 7, 2025Updated last year
- 1-Click AI Models by DigitalOcean Gradient • AdDeploy popular AI models on DigitalOcean Gradient GPU virtual machines with just a single click. Zero configuration with optimized deployments.
- ☆29Nov 6, 2025Updated 11 months ago
- Code for "Echo Chamber: RL Post-training Amplifies Behaviors Learned in Pretraining"☆30Oct 14, 2025Updated 11 months ago
- Stable-DiffCoder is a family of lightweight open-source code DLLMs(diffusion large language models) comprising base and instruct models, …☆85Mar 9, 2026Updated 7 months ago
- A tool to audit Hex dependencies, to make sure your ✨ gleam projects really sparkle!☆26Jun 21, 2026Updated 3 months ago
- Code for "Evidence of Learned Look-Ahead in a Chess-Playing Neural Network"☆31Jun 4, 2024Updated 2 years ago
- resources, links for OCR & greek☆11Mar 8, 2021Updated 5 years ago
- Cyber Threat Defense World Modeling☆26May 13, 2026Updated 4 months ago
- ☆35Jul 5, 2023Updated 3 years ago
- FlashMemory DS-V4 Retriever: a lightweight retriever that sparsifies DeepSeek-V4 CSA KV-cache. Weights available on Hugging Face.☆112Jul 18, 2026Updated 2 months ago
- GPU virtual machines on DigitalOcean Gradient AI • AdGet to production fast with high-performance AMD and NVIDIA GPUs you can spin up in seconds. The definition of operational simplicity.
- ☆28Apr 15, 2025Updated last year
- Low-rank Matrix Completion using Alternating Minimization☆23May 1, 2018Updated 8 years ago
- Codebase exploration with AI research agents☆21Feb 25, 2025Updated last year
- Minimum viable code for the Decodable Information Bottleneck paper. Pytorch Implementation.☆12Oct 20, 2020Updated 5 years ago
- An awesome collection of articles, papers, conferences, guides, and tools relating to deception in cybersecurity.☆131Sep 18, 2026Updated 2 weeks ago
- Improving Your Model Ranking on Chatbot Arena by Vote Rigging (ICML 2025)☆27Feb 25, 2025Updated last year
- AstroAgents: Multi-Agent AI for Hypothesis Generation from Mass Spectrometry Data☆15Apr 1, 2025Updated last year