My mission is to apply AI to make the world better, higher quality, and personalized.
All three are powered by the same underlying technology: the CARBON test framework.
Jank.ai →
For developer scenarios: QA that keeps up with developers and their AI coding agents.
Watches coding, finds changes, automatically runs tests to verify them, and auto-fixes bugs.
IcebergQA →
For engineering managers who want full-service AI QA—the fastest way to add it to your processes.
Testers.ai →
For software testers: run entire teams of sub-agents inside a coding agent.
Accelerate testing to keep up with developers.
Books on software testing and quality — from how Google does it to shipping high-quality apps, and a new one on testing AI.
A field guide to using AI to inspect, explore, test, challenge, and qualify the software that people and AI agents build.
A new book on how to test AI systems — strategies, techniques, and hard-earned lessons for evaluating quality in the age of AI. Currently in draft.
Co-authored with James Whittaker and Jeff Carollo. The definitive inside look at how Google approaches testing at massive scale — best practices you can apply at any size.
Best practices and hard-earned lessons from analyzing hundreds of millions of app store reviews and thousands of testers — for developers, testers, and product managers.
Chrome extensions and iOS apps that bring AI capabilities directly into your browsing workflow.
Improves AI responses and tunes them to your company's policies. Get better, more relevant outputs aligned with your organization's standards.
Optimizes your life based on an optical view of your world. Snap a photo and get personally suggested next actions powered by AI.
Interactive worlds and experiments built with AI.
An open-flight arcade adventure over real Maple Valley terrain. Fly planes and helicopters, explore the landscape, and take on targets in the air and on the ground.
Decks and talks on building with AI — from shipping a business to making AI feel personal.
How AI can adapt to who you are — tailoring responses, content, and experiences to each individual.
I'm Jason Arbon — entrepreneur, technical cofounder of Testers.ai, co-author of How Google Tests Software, and obsessed with AI and Quality. I've spent my career building and scaling test infrastructure at the world's largest software companies, was Head of Product & Engineering at Applause, built a team of some of the best testers in the world, and was funded by Google's Gradient Ventures for test.ai. Now I'm building AI that tests software autonomously.
These are personal projects and experiments — Chrome extensions and iOS apps that explore how large language models can reshape the way we consume and evaluate content online. Each app leverages multiple AI providers (Claude, OpenAI, Gemini) so you can choose the model that works best for you.

Founded before the LLM era and funded by Google's Gradient Ventures. Built AI that learned to navigate and test apps autonomously — pioneering AI-driven testing years ahead of the curve. Same mission: to test the world's software.

The post-LLM evolution — AI agents that test software the way a human tester would, powered by modern large language models. Same mission, new capabilities: to test the world's software.

A co-venture with Phil Lew of XBOSOFT, tackling the hardest testing problems at scale by blending humans and machines — pairing AI with expert human judgment to go deeper than either can alone.

Built for GenAI developers to find the jank that AI creates — and fix it. GenAI has turned every developer into a tester, and jank.ai gives them the tools to keep up.

Managed testing bots across 100K+ machines nightly on the latest Chrome builds. Quantified quality of AI training infrastructure at scale on the relevance and personalization team. Tested hardware-software interfaces at the OS layer for ChromeOS.

Tested mission-critical enterprise products: BizTalk Server, SQL Server, Integration Server. Worked on Bing's relevance team and Windows CE hardware-software interfaces. Some of the largest-scale testing challenges in the industry.

Led product and engineering for the platform powering crowd testing at scale. Built software to orchestrate thousands of testers worldwide across devices, locales, and test types for the world's top brands.

Tested thousands of apps in the Google Play store using AI-driven execution. Managed a fully autonomous lights-out testing lab — AI launched, navigated, and evaluated apps at scale with rich analytics reporting.

Built AI that autonomously launched and played Xbox Cloud Gaming titles to validate the streaming platform. End-to-end quality coverage for one of the most complex real-time distributed systems in gaming.

Tested Fortnite's monetization and character customization paths — high-revenue, high-complexity flows where bugs directly impact the bottom line. Custom AI test automation for one of the world's most popular games.