Test and eval harness for LLM-generated UI. Samples N generations, renders in headless Chromium, checks fidelity + a11y + layout, LLM-as-judge scoring, regression gates for CI.
By chatting or signing in you agree to the Terms and chat-message logging (revocable in History).