Uncensored/abliterated Ornith-1.0-35B (AEON Ultimate): 0% refusal, 0 coding-capability loss. BF16 + FP8 for vLLM.
☆83Jun 29, 2026Updated 3 months ago
Alternatives and similar repositories for Ornith-1.0-35B-AEON-Ultimate-Uncensored
Users that are interested in Ornith-1.0-35B-AEON-Ultimate-Uncensored are comparing it to the libraries listed below. We may earn a commission when you buy through links labeled 'Ad' on this page.
Sorting:
- Profile repo — categorized index of NVFP4 model releases, DGX Spark inference stacks, Apple Silicon MLX builds, the voice-AI stack, and t…☆40Sep 21, 2026Updated last week
- ☆19May 10, 2026Updated 4 months ago
- DeepSeek-V4-Flash-DSpark abliterated (uncensored) · ~100% refusal bypass · C1 ~57 tok/s · 1M ctx · 2× DGX Spark · HF weights☆51Aug 17, 2026Updated last month
- Deploy DeepSeek V4 Flash (MoE reasoning model) on dual DGX Spark nodes with 1M token context, InfiniBand, and FP8 KV-cache☆107Aug 20, 2026Updated last month
- Fully uncensored, capability-enhanced abliteration of Qwen3.6-27B. NVFP4 + z-lab DFlash speculative decoding (n=12) on the unified ghcr.i…☆469Sep 15, 2026Updated 2 weeks ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- DGX Spark / GB10 vLLM image for Gemma 4 31B Deckard Heretic Uncensored NVFP4 with z-lab DFlash speculative decoding.☆34May 15, 2026Updated 4 months ago
- DeepSeek-v4-Flash 0731 recipe for 2x DGX Sparks☆1,439Sep 26, 2026Updated last week
- Deploy concurrent Hermes Agent workers on unified-memory GPUs (GB10, DGX Spark) for maximum total tok/s. Profile-isolated, kanban-coordin…☆81Jul 18, 2026Updated 2 months ago
- NVFP4 Gemma-4 26B-A4B MoE for DGX Spark — optimal recipe: DFlash n=10 (flex) on AEON vLLM Ultimate. 144 tok/s single / 1,724 peak (Coding…☆44Jun 28, 2026Updated 3 months ago
- MiMo-V2.5 NVFP4 4-bit weights + NVFP4 4-bit KV cache + DFlash speculative decoding on 2× NVIDIA DGX Spark — 1M context, 3.4M-token KV poo…☆18Jul 13, 2026Updated 2 months ago
- Run the AEON Bench suite on your own hardware: verified HuggingFace pull → serve → benchmark (text · agentic ×3 harnesses · vision · audi…☆27Sep 11, 2026Updated 3 weeks ago
- sparkDash ⚡ — Multi-DGX Spark Monitoring Dashboard☆515Updated this week
- DeepSeek V4 Flash DSpark 1M NVFP4 KV recipe for 2x DGX Spark☆490Sep 10, 2026Updated 3 weeks ago
- MiniMax-M3 (428B, no pruning) at 36 tok/s on 2× NVIDIA DGX Spark — W4A16 GPTQ + NVFP4 KV + EAGLE-3 speculative decoding on vLLM. Three se…☆45Jul 13, 2026Updated 2 months ago
- Wordpress hosting with auto-scaling - Free Trial Offer • AdFully Managed hosting for WordPress and WooCommerce businesses that need reliable, auto-scalable performance. Cloudways SafeUpdates now available.
- Recipe: GLM-5.2 (unpruned QuantTrio Int4-Int8Mix) at 200K ctx with MTP spec decode on a 4x NVIDIA DGX Spark (GB10) cluster☆76Jul 25, 2026Updated 2 months ago
- Fast, interactive markdown viewer with powerful search, full customization, and clickable links for terminal-centric workflows.☆19Sep 26, 2026Updated last week
- Tool-calling quality benchmark for LLM serving stacks. 80+ deterministic scenarios testing multi-turn orchestration, safety boundaries, a…☆354Updated this week
- Autonomous self-improving 4x DGX Spark (GB10) MoA stack + LoRA loop (DSV4F router, Qwen3.6/Omni/TwoTower/Gemma). Hermes MoA routing, ~90%…☆24Aug 17, 2026Updated last month
- Setup wizard for Hermes Gateway Desktop☆32Jul 2, 2026Updated 3 months ago
- Self-contained DeepSeek V4 Flash DSpark TP=2 recipe for 2x DGX Spark with 62 tok/s benchmark☆24Jul 13, 2026Updated 2 months ago
- Qwen3.5-122B-A10B on DGX Spark: 28.3 → 51 tok/s (+80%)☆315Aug 23, 2026Updated last month
- Pure Rust Inference Engine☆700Updated this week
- sparkrun - launch, manage, and stop LLM inference workloads on NVIDIA DGX Spark systems. Live support: https://discord.com/invite/GH5kRg…☆526Updated this week
- Deploy on Railway without the complexity - Free Credits Offer • AdConnect your repo and Railway handles the rest with instant previews. Quickly provision container image services, databases, and storage volumes.
- The best benchmark for LLMs on Apple's MLX framework knowledge and coding tasks.☆38Jun 12, 2026Updated 3 months ago
- Dual-engine (llama.cpp + vLLM) LLM benchmarking pipeline for GGUF & safetensors on NVIDIA GPUs — speed, quality, live dashboard, publisha…☆40Updated this week
- Native MLX quants of Qwen3.6-27B-AEON-Ultimate-Uncensored for Apple Silicon (Metal): 8-bit, FP4, and MTP self-speculation. Preserves the …☆42Sep 15, 2026Updated 2 weeks ago
- 4-5x faster Qwen3.5 on ASUS GX10 / DGX Spark — Hybrid INT4+FP8 + MTP via one shell script☆32Apr 16, 2026Updated 5 months ago
- Code Execution plugin for TypingMind.☆12Aug 21, 2023Updated 3 years ago
- import documents for LLMs☆49Updated this week
- Application that can be used to demo networking SDN load balancing and security policy solutions (ACI,NSX)☆10Aug 9, 2017Updated 9 years ago
- The agent that grows with you☆180May 8, 2026Updated 4 months ago
- Docker configuration for running VLLM on dual DGX Sparks☆2,339Updated this week
- Managed hosting for WordPress and PHP on Cloudways • AdManaged hosting for WordPress, Magento, Laravel, or PHP apps, on multiple cloud providers. Deploy in minutes on Cloudways by DigitalOcean.
- ☆17Dec 1, 2022Updated 3 years ago
- BH hackathon☆14Apr 4, 2024Updated 2 years ago
- ☆17Aug 13, 2026Updated last month
- ☆16Jun 27, 2026Updated 3 months ago
- Encrypted vault, ZK settlement archiving, x402 payment receipts, Agent Passport binding, and Solana anchoring for OpenClaw and agent runt…☆22Updated this week
- Offline macOS meeting transcription that runs in your menu bar, uses Whisper locally, and makes transcripts searchable in Claude Desktop …☆18Feb 19, 2026Updated 7 months ago
- Static Encrypt is an crate that encrypts string literals at compile time and only decrypted at runtime when needed.☆60Jan 17, 2026Updated 8 months ago