An agnostic AI-driven exploratory test framework that intelligently explores, tests, and validates any application
-
Updated
Sep 20, 2026 - Python
An agnostic AI-driven exploratory test framework that intelligently explores, tests, and validates any application
Evidence Layer For AI Devs
AI Testing Agent Framework · 16 experts + 32 skills + 49 utils · Multi-LLM (Claude/OpenAI/Qwen/etc) · MCP-native · Open-source · Learn-while-using
The first MCP server for autonomous Flutter testing on real iPhones and Android devices. 110 tools across Android (uiautomator2+adb), iOS (WebDriverAgent+pymobiledevice3), Flutter (Patrol + flutter run --machine). Works with Claude Desktop, Claude Code, Cursor.
Use LLM to convert requirement in plain english to Playwright script
This project implements an AI agent that verifies if automated Hercules test runs were executed as intended by comparing planning logs, video evidence, and final outputs. It uses open-source LLMs and computer vision tools to flag deviations, providing detailed reports with technical insights.
Autonomous QA MCP that tests web and macOS apps like a real testing engineer—and verifies every bug it reports.
AI-driven exploratory browser testing — LLM observes live DOM and decides next action; no fixed scripts, self-healing selectors, Allure + Jenkins CI
Add a login page - done in 5 minutes. Autonomous AI agent for webapp feature development: planning, implementation, testing, debugging, delivery. Built on OpenCode superpowers workflow.
Repository for Github action running the Klarent flows.
Native Rust and Mojo recursive self-improvement engine with paired bootstrap evaluation and cryptographic ledger
Autonomous QA loop for Claude Code & Codex — an AI agent uses your finished app like a real user in a sandbox, finds bugs, safely fixes them with regression tests, and loops until it converges. Hands-off, self-healing, honest.
Autonomous browser testing that turns discovered bugs into deterministic regression tests. A model may choose what to try; only browser evidence decides pass or fail. CLI + MCP server.
AI agent that validates REST/SOAP APIs autonomously — reads OpenAPI specs, generates test cases, executes, and reports failures
AI Agent for Mobile QA. Write plain language scenarios, run them on iOS and Android, get structured step-by-step report with screenshots as evidence.
Sentinel QA Phase 3 — full cognitive QA ecosystem: AI generates test plans, Streamlit HITL approval, parallel Playwright execution via pytest-xdist, autonomous Jira bug triage with SQLite memory
🧪 You don't know what to test. This skill does. Test judgment for vibe coders — decides what needs testing, writes the tests, finds bugs.
Open-source MCP server + Claude Code skill that gives Claude the ability to autonomously explore any macOS application, analyze its user experience, map its full feature scope, and produce a rich human-readable audit report.
Autonomous browser testing framework with Playwright
To associate your repository with the autonomous-testing topic, visit your repo's landing page and select "manage topics."