autony

Real-phone testing for apps your agent builds.

Describe the core flow in one sentence. A tester uses your release APK on real Samsung phones, runs the same path again with the network off, fonts at 130% and the phone in Japanese, and returns a report your coding agent can act on. No scripts, no SDK, no device slots.

See a real report Continue with Google
claude mcp add autony https://api.autony.dev/mcp \
  --header "Authorization: Bearer autony_sk_…"
> "upload app-release.apk to autony and test the core flow"

No coding agent? Upload the APK in your browser — same report. · 2 free runs on your own app · prices below

A goal, not a script

“Create a note and save it.” The tester reads the screen and finds the way, step by step, until it's done. One named flow, one phone per run, assigned automatically.

Run again under real conditions

The same path, again, with the network off, fonts at 130% and the phone in Japanese. On our Galaxy phones (Android 13–15), not an emulator.

A report built for agents

Markdown with evidence links for Claude Code or Cursor; HTML with screenshots for you. Your app's text can't steer the tester.

What a report looks like

A verdict, defects ranked by how much they hurt the user, each with a screenshot, the log line and a suggested fix — and the exact path the tester took, so “no defects” still shows what was covered.

Open the real, unedited sample →

example defect card · from our test app with a planted bugP1 · blocks the flow
Screenshot: only a spinner while offline
Offline handlingOnly a progress indicator with no notice for 14.9 s while offline.Galaxy A15 · screenshot at step 5 · suggested fix included

How it compares

CriterionDevice cloudsAI test authoringautony
You writeAppium / Espresso scriptsTest steps in plain EnglishOne sentence
Who runs itYour scripts, their phonesTheir AI, often their phonesOur tester, our phones, offline / fonts / locale re-runs
OutputPass/fail, videoPass/fail, screenshotsDefects with severity, suggested fixes, Markdown for your agent
PricingPer device slot or minutePer seat or creditsPer run, from [PRICE/RUN]
Coming from Firebase Test Lab Robo?Robo shuts down on Sep 30, 2027. Robo crawled randomly across a device matrix; autony follows one flow you name, on one phone per run, and re-runs it under conditions Robo never covered. Upload the same APK, no test package needed.

Pay per run. No seats, no slots.

One run = one test of one flow on one phone, about 5 minutes once it starts. Runs that reach the phone are charged, even with no defects. If the install fails or our lab can't complete the run, the run is returned. Runs never expire.

Free$02 runs on your app
5 runs[PRICE][PRICE/RUN] per run
20 runs[PRICE][PRICE/RUN] per run
100 runs[PRICE][PRICE/RUN] per run
Start freePrices in USD, charged by [PARTNER]. Beta: new accounts are approved within 1 business day.