0xSero

Orchestrator

GitHub 125 repos 2.1k followers

Explained projects

Yet to be explained

DeepSeek-V4.1-Flash-Two-Sparks
DeepSeek-V4.1-Flash on two DGX Sparks: EXL3 routed experts, 262k context, 2M-token KV, vision, tools, MTP, bundled with Pi
Shell90 stars
Explain
harness-bridge
Point any coding harness (Claude Code, Codex, OpenCode, Pi, OMP, Crush, Copilot, Grok…) at any OpenAI/Anthropic/Responses-compatible inference endpoint. Core library + CLI + macOS tray + web UI.
TypeScript90 stars
Explain
glm-5.3-low-bit-tr3-wiki
Wiki for the GLM-5.3 Flash low-bitrate TR3 quantization campaign: architecture, EXL3/TR3 encoding, K2/K3/K4 tiers, validation gates, and quality metrics.
8 stars
Explain
glm-5.3-flash-spark-mosaic
GLM-5.3 Flash on one DGX Spark: M288 mixed-precision mosaic (mul1) + native MTP (mcg) serving recipe with measured evidence
Python4 stars
Explain
glm-5.3-flash-4x-rtx-pro-6000
Pinned Local Inference Lab R35 GLM-5.3-Flash Docker deployment for 4x RTX PRO 6000: vision, video, 400k context, image history pruning and measured evidence
Python4 stars
Explain
model-toolkit
Practical model evaluation and compression tools: Terminal-Bench, DeepSWE, GPQA, KL divergence, REAP, EXL3, and reproducible evidence.
Python4 stars
Explain
local-ai-data
Local AI data: every model, build, hardware record, price, benchmark and community measurement behind the Local AI registry (moved from local-ai-registry, full history)
HTML2 stars
Explain
dgx-spark-fleet
Agent skill: set up and run 1 to 4 NVIDIA DGX Sparks. Linking, best models per fleet size, recipes, fixes.
2 stars
Explain
omarchy-hf-agent
Open any coding agent on a Hugging Face Inference Providers model, from one button on the Omarchy bar
Shell2 stars
Explain
scripture-art-galleries
Six galleries of public-domain paintings with a short line of scripture set into each plate
HTML1 stars
Explain
augustine-wiki
Augustine of Hippo: his writing broken into six arcs, with every passage quoted verbatim from public-domain sources
JavaScript1 stars
Explain
Step-5-Preview-Four-Sparks
Serve StepFun Step-5-Preview (EXL3) on four NVIDIA DGX Spark with vLLM TP4, plus Pi integration
Shell0 stars
Explain
local-ai-hub
JavaScript0 stars
Explain
dynaprune
Observation-driven dynamic pruning: configure what a model loses from external observations. Verification-first system design, schemas, and minimal reference code.
0 stars
Explain
ai-hardware-dates
Source-checked dates for the 'A Decade of AI Hardware Innovation' chart
HTML0 stars
Explain