MoralityLabAI
GitHub
24 repos
0 followers
Explained projects
Pixieology: Steering the Fae in the Machine
How MoralityLabAI uses mechanistic interpretability to surgically toggle between cold logic and lyrical whimsy.
7 min read ยท Mar 25, 2026
Yet to be explained
Hermes-Skills
Python
1 stars
Explain
trm_observability_harness
Reasoning Trace Extractor for TRM Training Data across Envs
Python
1 stars
Explain
jev-qwen
Experiments with small JEV-likes
Python
0 stars
Explain
ThoughtLeader
Python
0 stars
Explain
jinn-or-beast-paper
Python
0 stars
Explain
AICourt
Unified strategic-disposition corpus tooling, expanded Court intrigue environment, and multi-agent replay desk.
JavaScript
0 stars
Explain
ALife
Experimental artificial life sandbox with multi-plane cellular dynamics and evidence-backed experiment tooling.
Python
0 stars
Explain
FurlingExperiments
Long-context storyworld, Verifiers, SAE, and VPD evaluation research harnesses.
Python
0 stars
Explain
RSITopology
Identity-attestation, spectral lineage, holonomy, and patchwise edit-control tools for mechanistic interpretability research
Python
0 stars
Explain
TheySing
They Sing: an AI-native strategy game and ASI diplomacy research harness.
TypeScript
0 stars
Explain
Control-Harness
Small Scale Prototype of LLMs subordinated to LDT/TRM
Python
0 stars
Explain
param-decomp
Python
0 stars
Explain
HybridTRMLDT
Python
0 stars
Explain
metta-storyworld
Adapting Storyworld building to the Metta reasoning/world modeling language
Python
0 stars
Explain
BlueBeam
SAE Control Protocol, Antivirus for Small Models
Python
0 stars
Explain
Adict
Alternative to SAE Mechinterp for ANE
Python
0 stars
Explain
morality-lab-site
Website for MoralityLab
TypeScript
0 stars
Explain
ConstitutionalAlignment
Harness to teach some moral framework to LLMs via storyworld role-play RL
Python
0 stars
Explain
TRMStoryworld
Training Tiny Recursive Reasoning Models to Play Storyworlds
Python
0 stars
Explain