Mobile GUI Use
Optimized for everyday tasks on a real phone, CoreAutomata’s hosted agent navigates changing Android and iOS apps — search, compare, schedule, check out — and asks before anything irreversible.
Tell it to categorize the video content using AI
Write a test or describe a task. CoreAutomata’s hosted AI runs it on a real Android or iOS device you pair — autoverifies what the screen shows, self-heals when the UI moves, and adapts when a popup gets in the way.
200 free actions · hosted AI · real devices · no credit card
One paired device and hosted AI cover natural-language tests, goal-driven tasks, and the repairs that keep both running.
Optimized for everyday tasks on a real phone, CoreAutomata’s hosted agent navigates changing Android and iOS apps — search, compare, schedule, check out — and asks before anything irreversible.
The same real-device connection and hosted AI power both — author natural-language cases or describe a goal.
Stop maintaining selector maps. Author cases in natural language against a build you captured once, keep them in a shared library, and let our AI re-find elements when the UI shifts — on real devices you already own.
Automate workflows that only exist inside a mobile app — internal tools, vendor portals, apps with no public API — by describing the outcome. Our hosted agent drives a real phone; you never plug in or host a model.
One connection powers natural-language tests and goal-driven tasks — three steps, no scripting, AI included.
Install the CoreAutomata Portal on an Android phone, or run the iOS Bridge on a Mac, and scan the code in your browser. The device dials out — no cable, no adb, no Xcode on your machine.
Upload an APK, XAPK, or IPA and CoreAutomata maps screens and drafts cases. Or skip the build and drive any app already on the phone by describing the outcome.
Accept discovery’s drafts, write your own checks, or name a goal. Every run executes on a live device; our AI repairs moved targets, handles unexpected screens, and keeps screenshots plus per-step outcomes.
If it’s installed on the paired phone, CoreAutomata can drive it — messaging, social, commerce, work. No SDK in the app, no private API required.
Chat and inbox flows on a real device.
Publish, manage, and QA authorized accounts.
Source, monitor, and QA on real shopper devices.
Maps, browsers, and field apps your agent can drive.
Selector frameworks ask you to describe your app twice — once in code, once in the test. CoreAutomata works from the real-device screen, the same thing your users see, with AI that self-heals the suite and drives the next tap.
| CAPABILITY | CoreAutomata | Appium / Detox |
|---|---|---|
| Writing a step | Natural language | XPath / element IDs |
| Your first suite | Discovery drafts it | Hand-written, screen by screen |
| When the UI moves | AI re-finds the element | Locator breaks |
| Where it runs | Real device, QR-paired | adb / Xcode / driver farm |
| AI / model setup | Hosted — included | None (scripted) or DIY |
| Who can author | Anyone on the team | Whoever knows the framework |
One capture or one paired phone becomes the shared context for natural-language cases and goal-driven tasks — on real devices, with our AI.
Point CoreAutomata at a build and it explores on a real device — opening screens, filling forms, mapping flows — then drafts candidate test cases from paths it actually reached.
A step is a sentence: "tap Checkout", "confirm the shipping form is empty". No XPath, no accessibility IDs, no Appium driver. The step text is the test — anyone on the team can author.
A rerun first replays the recorded step. If the target moved, our AI re-finds it. If the screen changed more than that, it works from the step’s intent. Only then does it fail — and the report says which tier saved it.
CoreAutomata’s hosted vision stack reads each frame — controls, labels, state, what is tappable. No DOM access, no accessibility-tree assumptions, no changes to the app, and no model API for you to wire up.
For a goal, the agent breaks it into sub-goals, tracks which are done, and re-plans as the screen changes. Long flows stay on target instead of drifting after the third tap.
Taps, swipes, scrolls, long-presses, and text entry resolve to real coordinates on the live screen and dispatch over the device’s own connection — nothing to install on a host machine.
Consent dialogs, permission prompts, rating nags, upsell interstitials, an unexpected logout. The agent deals with whatever appears rather than needing a branch written for each one.
If it renders, it can be driven. That includes apps you did not build — vendor portals, internal tools, third-party apps with no API and no chance of getting one.
Each run keeps the screenshot, the action taken, and the outcome for every step. When something fails you get the frame it failed on — not a stack trace pointing at a locator.
Account logins go in the project Vault, encrypted at rest and injected at run time. They never appear in step text, in run logs, or in the screenshots you share.
Runs are scoped to the app you name, stopped before irreversible actions, and recorded step by step so you can see exactly what the agent did and replay it.
Assertions read the live UI the way a tester does — labels, totals, empty states, clipped controls — not brittle pixel baselines. Catch visual and copy issues that selector checks miss.
Natural-language cases for QA. Goal descriptions for automation. Both run on the same paired device with hosted AI.
This is the whole artifact — no page objects, no locator file, no driver setup. Steps that act and steps that check sit side by side. The same case reruns on a real device against every later build, and our AI repairs what moved.
You state the outcome. Our hosted agent decides what to tap on the real phone, deals with whatever the app puts in the way, and shows you what it did. Nothing here names a coordinate, an element ID, or a model you own.
Teams that test and automate on real phones — and industries where the work lives in mobile apps.
QA, ops, research, and product teams using the same Board.
Natural-language cases on real devices — self-healing when the UI moves.
Mobile-first process automation for apps with no API.
A real phone for your agent — hosted vision, no model key to bring.
Routine in-app tasks on the phone your team already uses.
Walk competitor apps and map screens, paywalls, and flows.
Replay onboarding and catch where the funnel stalls.
If the work only exists inside an app, CoreAutomata can drive it.
Vendor portals and logistics tools that never ship an API.
Field notes and mobile EHR flows on a paired device.
Checkout QA, promo paths, and catalog checks at release pace.
Expense approvals, payments exports, and reconciliation in-app.
Dispatch and work-order apps your crews live in every day.
Legacy vendor apps your ops team will never get an integration for.
Rerun the library against each new build. Steps whose targets shifted get repaired in place.
Expense approvals, HR portals, dispatch tools — drive the app when there is no API.
Prices, listings, statuses, inventory — structured from what the app displays.
Launch, sign in, primary purchase path — run against a release candidate before it ships.
Map onboarding, paywalls, and pricing into a graph without a person tapping through.
Empty states, error toasts, totals that match line items — grounded in the screen.
CoreAutomata is in beta. Start free on a device you already own — early adopters get free upgrades when paid plans launch.
200 agent actions a month and 2 projects on a real device you pair yourself. Hosted AI included. No card required. Early adopters get free upgrades as paid plans ship.
Start freeA dedicated device pool, SSO, custom limits, and support terms that fit your release cycle — available now while we’re in beta.
Contact usYes. Pair a physical Android phone or iPhone (via the iOS Bridge on a Mac), or use a simulator. Tests and agent tasks execute against the live screen — not against a mocked UI tree alone.
It covers the same job without the selector layer. Appium tests address elements by ID or XPath, so they break whenever the UI shifts. CoreAutomata works from what is on screen, so a moved or renamed element is something our AI can re-find rather than a broken locator.
Reruns go through a three-tier repair ladder. First the exact recorded step is replayed. If the target moved, the step is re-grounded against the current screen. If the screen changed more substantially, the agent works from the step's intent instead. The rerun only fails after all three tiers fail, and the report shows which tier recovered it.
Builds and run artifacts live in your project's storage. Test-account credentials go in the project Vault, encrypted at rest, and are injected into a run without appearing in step text, logs, or screenshots.
No — it is a commercial product. The Portal app and the iOS Bridge are free companion apps you use to connect your own devices.
No. Vision, planning, and grounding run on CoreAutomata’s own hosted AI backend. You describe test cases or tasks in natural language and pair a device — you do not plug in a model provider, host inference, or expose an LLM endpoint.
200 free agent actions a month. Real devices. No card. Pair a phone or create a device farm and run your first case or goal today.