Canary is an AI QA engineer that reads your source code to understand developer intent, then automatically generates and runs end-to-end tests on every pull request. Instead of flaky DOM scraping or screenshot analysis, it reads the diff, works out what a change is meant to do, and tests the real user flows it touches - catching broken checkout, auth, and billing before code reaches production. Founded in 2026 by ex-Windsurf, Cognition, and Google engineers, the two-person San Francisco team is part of Y Combinator's Winter 2026 batch.
Confident AI is a San Francisco startup (YC W25) that builds the evaluation and testing layer for LLM applications. Its open-source frameworks DeepEval and DeepTeam - often described as 'Pytest for LLMs' - let engineers unit-test, benchmark, red team, and monitor chatbots, agents, and RAG pipelines with research-backed metrics. The cloud platform adds dataset management, observability, adversarial testing, and governance so teams can prove AI quality before and after shipping.
Docket is a San Francisco AI QA testing platform that lets teams write end-to-end tests in plain English and keeps them running as the app changes. Instead of brittle CSS selectors, its multimodal agents look at the screen and click where a human would, recording pixel coordinates and self-healing when the UI shifts. Founded in 2025 by ex-Stripe and ex-Citadel engineers Nishant Hooda and Boris Skurikhin, Docket went through Y Combinator's Spring 2025 batch and covers web, mobile, and desktop apps.
TesterArmy is a Y Combinator-backed (P26) startup building AI agents that test web and mobile apps the way a human would. Teams describe tests in plain English; the agent launches a real browser or device, clicks and types through user flows, handles logins, OAuth and OTP, then returns screenshots, recordings and actionable bug reports. It runs on every GitHub pull request and on a schedule against production, aiming to catch broken flows before customers do.
Opkey is an AI-powered test automation and ERP lifecycle optimization company that helps large enterprises test, configure, and maintain packaged cloud applications like Oracle Cloud, Workday, SAP, and Salesforce. Its no-code platform, powered by an enterprise-specific AI model called Argus and a set of agentic 'virtual agents,' automates test discovery, self-healing test scripts, impact analysis, configuration migration, and end-user training. Founded in 2015 and backed by a $47M Series B led by PeakSpan Capital, Opkey serves more than 250 customers, a majority of them Fortune 1000 firms, and positions itself as the platform that pulls enterprises out of what its CEO calls 'testing hell.'
Katalon is an Atlanta-based software company that builds an AI-augmented, all-in-one quality management platform for software testing. Its flagship products - Katalon Studio, TestOps, TestCloud, and the newer TrueTest - let QA and DevOps teams create, run, and analyze automated tests across web, mobile, API, and desktop applications with both low-code and full-code approaches. Spun out of KMS Technology, Katalon is used by more than 30,000 teams across 80+ countries and has repeatedly been named a G2 Leader in software testing.
Provar is a UK-founded software company that builds test automation and quality management tools purpose-built for the Salesforce ecosystem. Its low-code platform - anchored by Provar Automation and Provar Manager - lets both manual testers and developers create resilient, metadata-aware tests across UI, APIs and data layers that survive Salesforce's three-times-a-year release cadence. Founded in 2014 by veterans of highly regulated banking projects, Provar has grown into the category leader for Salesforce QA, and was first to market with automated end-to-end testing for Salesforce's Agentforce AI agents.
Checksum.ai is an AI software-testing company that automatically writes and maintains end-to-end tests for web applications. Its AI agents learn from real user sessions to generate production-ready Playwright and Cypress tests, then auto-heal them when the app changes - so engineering teams get broad test coverage in days instead of months and stop babysitting flaky test suites. Founded in 2022 and built inside the super{set} startup studio, the San Francisco company serves developers and QA teams who want to ship fast without breaking things.
Ottometric is a Waltham, Massachusetts software company that uses AI to automate the validation and training of Advanced Driver Assistance Systems (ADAS) and autonomous-vehicle software. Its platform distills petabyte-scale, multimodal sensor data into decision-ready KPIs, cutting validation cost and time by more than half for the Tier-1 suppliers and OEMs that build the cars' eyes and reflexes.
Functionize is a San Francisco-based AI test automation company building an agentic platform that lets enterprise QA teams design, run, and maintain software tests in natural language. Founded by Tamas Cser in 2014, the company uses machine learning and computer vision to slash test maintenance and accelerate releases for Fortune 500 customers.
Rainforest QA is an AI-powered, no-code software testing platform that helps product teams run regression and functional tests without building or maintaining brittle test frameworks. Founded in 2012 in San Francisco and backed by Y Combinator, the company has executed more than 42 million tests for over 10,000 startups and product teams.