Back to Categories

Coding & Assistance shortlist

AI Testing & Debugging

Automated testing, debugging, and quality assurance

Tool List

46 tools
Diffsmith

Diffsmith

Local code review studio for AI agent-generated changes with line comments and MCP handoff

Manta AI

Manta AI

Autonomous web app testing agent that explores flows, finds bugs, and generates self-healing tests from a URL

Keywords AI / Respan

Keywords AI / Respan

LLM engineering platform for gateway routing, observability, evaluations, and prompt optimization

SlimSnap

SlimSnap

Mac tool that turns screenshots into JSON readable by CLI agents

Bugpilot

Bugpilot

Automated bug tracking, error monitoring, and contextual reporting platform for modern web applications

Expect

Expect

Expect lets agents test code in a real browser by scanning diffs, generating plans, and running execution workflows from one command

TestSprite

TestSprite

TestSprite is an AI testing agent that can plan, write, execute, debug, and report software tests end to end

Ogoron

Ogoron

Ogoron is an automated testing platform where autonomous agents plan, generate, and maintain tests directly from your codebase

QA.tech

QA.tech

QA.tech uses AI agents to explore web apps, generate complete test suites, and run QA checks on release, schedule, or manual triggers

Cekura

Cekura

End-to-end testing and observability for conversational AI. Run pre-production simulations and monitor production conversations for voice and chat agents

Glassbrain

Glassbrain

Glassbrain helps teams visually debug LLM apps by capturing every OpenAI, Anthropic, and LangChain call and replaying failed runs

Future AGI

Future AGI

Future AGI helps teams build, evaluate, optimize, and monitor LLM and AI agent applications with multimodal quality testing and observability

LangWatch

LangWatch

Coding-agent observability for session costs, tokens, cache usage, tool calls, file edits, and terminal replays across Claude Code, Codex, and more

Waydev

Waydev

Waydev is an engineering analytics platform for measuring delivery performance, DORA metrics, and productivity trends

Breadcrumb

Breadcrumb

Simple, open-source LLM tracing for AI agents. Track prompts, completions, latency, token usage, and cost

Euphony

Euphony

Render AI chat data and Codex logs into filterable browser timelines

Regent

Regent

Regression testing and semantic diff platform for agent apps to catch behavior changes before release

AgentPeek

AgentPeek

Open-source debugging and run-inspection tool for AI agent workflows, prompts, tool calls, and runtime state

PandaProbe Cloud

PandaProbe Cloud

Open-source agent engineering platform with traces, evals, and metrics for debugging AI agents

ReleaseDock

ReleaseDock

AI-assisted release management for release notes, launch checklists, and rollout communication

Osloq

Osloq

AI agent that reproduces GitHub issues and returns evidence-backed reports

Retrace

Retrace

Debug AI agents by recording, replaying, forking, and sharing runs

Latitude

Latitude

Open-source AI agent monitoring and quality platform that detects failure modes and helps developers fix agent issues before production

SwiftScale Software

SwiftScale Software

Dedicated QraftAI agent for generating and running cross-browser QA tests

Prelint

Prelint

Product-drift review for AI-written pull requests

Prefactor

Prefactor

Real-time production evaluation for AI agent runs

TraceLLM

TraceLLM

OpenTelemetry-based tracing and observability for prompts, model calls, tool spans, tokens, latency, and errors in production AI applications

AMP by CanyonTechs AI

AMP by CanyonTechs AI

AI agent that monitors production logs, detects incidents, and opens reviewable fix PRs

Gitar

Gitar

AI tool that reviews pull requests, analyzes CI failures, and submits fixes

Ito

Ito

Runs pull requests in ephemeral environments and reviews them with runtime evidence

Media Sharing by Argos

Media Sharing by Argos

Lets coding agents and CI publish screenshots, recordings, and ready-to-paste Markdown to pull requests

CodeBurn

CodeBurn

Local cost observability for 40 AI coding tools, broken down by task, model, project, and pull request

Inferock Bench

Inferock Bench

Local LLM proxy that records usage, failures, retries, and billing-integrity receipts per call

Superflow AI

Superflow AI

Turns website QA checklists from spreadsheets, CSVs, or PDFs into parallel testing agents

FetchSandbox MCP

FetchSandbox MCP

Validate integration code inside AI coding assistants with runnable sandbox APIs, webhooks, and failure scenarios

Port Radar for macOS

Port Radar for macOS

Inspect, explain, stop local port processes, and share localhost from the macOS menu bar

Agnost AI

Agnost AI

Find and cluster AI agent failures from real conversations and traces, then turn them into actionable evals

ArcReel

ArcReel

AI Agent 驱动的开源视频生成工作台 — 小说→角色/场景/道具设计→剧本→分镜图→视频,跨镜头角色与场景一致 | Open-source AI video workspace powered by AI Agents, Nano Banana 2 & Veo 3.1 / Grok / Seedance / OpenAI

CorridorKey-Runtime

CorridorKey-Runtime

Native AI keying runtime and OFX plugin for DaVinci Resolve, built in collaboration with Corridor Digital

mcporter

mcporter

Call MCPs via TypeScript, masquerading as simple TypeScript API. Or package them as cli

OpenCLI

OpenCLI

Make Any Website & Tool Your CLI. A universal CLI Hub and AI-native runtime. Transform any website, Electron app, or local binary into a standardized command-line interface. Built for AI Agents to discover, learn, and execute tools seamlessly via a unified AGENT.md integration

mcp-memory-service

mcp-memory-service

Open-source persistent memory for AI agent pipelines (LangGraph, CrewAI, AutoGen) and Claude. REST API + knowledge graph + autonomous consolidation

ai

ai

🤖 Type-safe, provider-agnostic TypeScript AI SDK for streaming chat, tool calling, agents, and multimodal apps across OpenAI, Anthropic, Gemini, React, Vue, Svelte, and Solid

Reference

Reference

Local semantic search for AI agents

Coldtea.ai

Coldtea.ai

Automate software delivery with coding, visual QA, and production monitoring agents

Progress AI Observability

Progress AI Observability

Trace, evaluate, and improve AI agents in production

Related categories