Agentic, long-horizon visual generation: a fuzzy story → a cross-model-audited image-based movie. Brings ARIS's research-wiki + multi-agent debate to multimodal generation (intelligence lives in the agent; the diffusion model just renders). Image-based today, video next.
66 stars
Python
Your first custom repo explanation is free. Reading existing public explainers always stays free.
This will take 10-20 minutes. You can close the tab and come back later.