Your agent test passed. Would it pass again? Checks whether AI agent test results are repeatable and varied enough to trust as a regression baseline. Reads Promptfoo and DeepEval runs you already have.
AI vision GUI automation for browsers, mobile apps, desktops, and games — no DOM, no selectors. Standalone or on top of Playwright / Selenium / Appium / pytest.
Model-agnostic multi-modality generation server with plugin discovery: OpenAI-compatible TTS, embeddings, and image generation. Drop a Python file in ~/.muse/models/ to add a backend.