AI Awesome
Home
Discover the Best AI Tools
Your ultimate directory for finding the right artificial intelligence solutions for any task.
Search
AI Testing
(71)
S
open source
SnapDiff MCP
Calls SnapDiff’s hosted API for screenshots, visual comparisons, change checks and HTML/CSS image rendering. Intent-aware UI verification needs a project and baseline and can return a human-review verdict. Requires an API key; visual differences alone do not verify application behavior.
AI Testing
A
open source
Awarse Healwright MCP
Proposes replacement Playwright locators from ARIA snapshots using Gemini, checks uniqueness/visibility and can patch test files. Requires model credentials and browser dependencies; mock mode is only for offline tests. Snapshot data goes to the model, and a unique locator does not prove the repaired test still checks the intended behavior.
AI Testing
M
open source
MotionLint
Audits live-page animation timing, easing and reduced-motion behavior using deterministic browser measurements, with optional vision-model design review. Requires Chromium; model review needs a configured provider or local model. Its rule scores and model findings are review aids, not proof of complete accessibility or universal design quality.
AI Testing
C
open source
CIAgent
A tool for evaluating AI agents with features like stability reports, flip attribution, LLM judge audits, and deterministic checks derived from a knowledge base.
AI Testing
stability reports
flip attribution
LLM judge audits
+1
C
open source
ClawBench
ClawBench is a benchmark for evaluating browser AI agents on everyday tasks.
AI Testing
evaluating browser AI agents
everyday tasks
M
open source
Multi-SWE-bench
Multilingual benchmark for evaluating software agents on issue resolution.
AI Testing
A
open source
AgentLeak
AgentLeak is a Python SDK, CLI, and MCP tools for testing AI agent data leaks across tool calls, memory, messages, and logs, with redacted reports and CI gates.
AI Testing
Python SDK for testing AI agent data leaks
CLI for testing AI agent data leaks
MCP tools for testing AI agent data leaks
+2
D
open source
DOS (dos-kernel)
Tool for checking AI agents' completion claims against Git evidence.
AI Testing
A
open source
Agent-Wiz
Agent-Wiz is a command-line interface (CLI) for threat modeling and visualizing AI agents built with frameworks including LangGraph, AutoGen, and CrewAI.
AI Testing
threat modeling
visualizing AI agents
A
open source
AgentSkeptic
AgentSkeptic is a tool for detecting silent failures in AI agent workflows by verifying actual system state.
AI Testing
detects silent failures
verifies system state
B
open source
BrowserTrace
Local replay debugger for Browser Use failures, with screenshots, model input/output, failed-step timelines, and HTML exports.
AI Testing
screenshots
model input/output
failed-step timelines
+1
L
open source
Lians
A tool for checking AI coding-agent work and binding check results to the current Git state for human review.
AI Testing
checks AI coding-agent work
binds check results to Git state
facilitates human review
A
unknown
agent-qa
Open-source self-improving QA agent for software teams. A test harness with memory. Write tests in natural language for web and mobile. agent-qa learns from every run, adapts to UI changes, and catches regressions before you ship.
AI Testing
self-improving QA agent
test harness with memory
write tests in natural language
+3
R
open source
RAMPART
RAMPART is a pytest-native safety and security testing framework for agentic AI applications.
AI Testing
pytest-native safety and security testing framework
designed for agentic AI applications
MIT licensed
H
open source
hidai25/eval-view
A regression testing framework for AI agents that saves golden baselines and detects behavioral drift.
AI Testing
Save golden baselines
Detect behavioral drift
Block regressions in CI
+1
K
open source
KryptosAI/mcp-observatory
Regression testing tool for MCP servers that auto-discovers servers, checks capabilities, and detects schema drift.
AI Testing
Auto-discovers MCP servers from Claude configs
Checks server capabilities
Invokes tools
+3
T
open source
tsilverberg/webapp-uat
A full browser UAT skill for Playwright testing with various features.
AI Testing
Playwright testing with console/network error capture
WCAG 2.2 AA accessibility checks
i18n validation
+2
C
open source
claude-bug-bounty
Claude Code skill for AI-assisted bug bounty hunting that automates security testing.
AI Testing
automates reconnaissance, IDOR, XSS, SSRF, OAuth, GraphQL, and LLM injection testing
4-gate validation checklist
report generation
S
open source
Souzix76/n8n-workflow-tester-safe
A tool for testing, scoring, and inspecting n8n workflows via MCP.
AI Testing
Config-driven test suites
Two-tier scoring
Lightweight execution traces
+1
R
unknown
Rhesis
Testing infrastructure for LLM and agentic applications with collaborative evaluation.
AI Testing
Testing infrastructure for LLMs
Agentic applications support
Collaborative evaluation
T
unknown
test-driven-development
A tool for automated testing using AI to improve software quality.
AI Testing
automated testing
AI-driven quality assurance
software development
A
unknown
Ai-Api-Testing
Ai-Api-Testing is a tool for testing AI APIs.
AI Testing
api testing
ai testing
W
unknown
webapp-testing
A GitHub repository for webapp-testing skills.
AI Testing
web application testing
AI skills
Previous
Page 3 of 3
Next