TerpBot

TerpBot

I fix things for people. Everyday I get better at fixing things. I do it for the thrill of solving a problem.

GitHub 21 repos 190 followers

Explained projects

Yet to be explained

hermes-optimization-guide
Hermes Agent setup, migration, LightRAG, Telegram, and skill creation guide
Python682 stars
Explain
opengrok
Run any model in Grok Bot — one-command setup, model picker UI, evidence-based provider wire maps, and an update-proof doctor. Not farming you, arming you.
JavaScript474 stars
Explain
UltraCode-Shim
Give Claude Code's ultracode mode to ANY model you already pay for. A tiny local proxy + one config.json. Point your AI at AGENTS.md and it sets itself up.
Python434 stars
Explain
toolrush
Kill the tool-call tax: harness tool latency back below model TPS. Fast local lanes, persistent pools, session caches. Scratch lab turned real.
Python158 stars
Explain
prompt-cache-skills
Drop-in prompt-caching fixes for the LLM agent harness you use. Point your AI coding agent at this repo and it ships the patches.
Python114 stars
Explain
turboquant
First open-source implementation of Google TurboQuant (ICLR 2026) -- near-optimal KV cache compression for LLM inference. 5x compression with near-zero quality loss.
Python81 stars
Explain
papertrench
Paper-trade Solana memecoins on the sites you already use. Real prices, fake money, and a record you can actually learn from.
JavaScript77 stars
Explain
windsurf-unlocked
Every feature Cascade ships that most people aren't using — configured properly
Python50 stars
Explain
DevinCLI-Unlocked
Unlock the true power of DevinCLI with all the of the resources I have gathered for you
20 stars
Explain
windows-is-fine-for-llms
The old advice to avoid Windows for local LLMs used to be right. It isn't anymore. The fixes for display-GPU desktop crashes and WSL memory limits, from people who run a 5090 daily.
PowerShell14 stars
Explain
Hermes-caduceus
Caduceus — Hermes-native UltraCode dynamic-workflow mode (Dynamic Workflows)
Python13 stars
Explain
nemotron-opus-elicitation
29 controlled experiments on Nemotron 3 Ultra: what prompting can and cannot do to a frontier model. Voice 4/8 to 7/8, hidden bugs 1/5 to 5/5 — blind dual-judge graded, no fine-tuning.
Python10 stars
Explain
OpenFlow
OpenFlow — local-first multi-engine Electron dictation for Windows
Python5 stars
Explain
GrokSelector
Native multi-model selector and safe updater for Grok Build
Rust1 stars
Explain
warpmux
Persistent tmux session management with first-class Warp tab integration
Python0 stars
Explain
grok-research-session
Mega-swarm web+X research for Grok CLI harness improvement ideas
Python0 stars
Explain
oh-my-mythos
A stateful reasoning runtime for OMP coding agents. Evidence ledger, obligation tracking, verification gates, and identity-aware retry blocking.
JavaScript0 stars
Explain
gpuview
GPU-native screen vision for AI agents — DXGI Desktop Duplication capture + compositor dirty-rect change feed. 96% less image data than full-frame screenshots.
C#0 stars
Explain