THE MOBILE AGENT LANDSCAPE

Compare mobile automation tools.

See what’s supported—and the evidence behind it.

Source review14 Sept 2026
14 agents indexed20 capabilities tracked
/
14 agentsSelect up to 4 to compare
Mobile agents and documented capabilities
SelectAgentPlatformsMCPFlowsDebuggingProfilingSkillsBest suited forDetails
iOSAndroidDeep app diagnostics
iOSAndroidCoding-agent feedback loops for mobile, TV, and web app verification
iOSStreaming Apple simulators into an agent or browser
iOSAndroidDirect device control via MCP tools
iOSAndroidRepeatable end-to-end tests
iOSNative Apple development with integrated Xcode build, test, and simulator tooling
iOSLightweight iOS simulator UI automation for MCP coding agents
iOSAndroidAppium-based mobile, tvOS, and remote-grid automation
iOSAndroidVision-driven UI testing across web, mobile, and desktop
iOSAndroidNatural-language phone task automation
iOSAndroidMulti-step task execution
AndroidSelf-hosted vision-model phone agent for Android, HarmonyOS, iPhone
AndroidMulti-agent GUI automation research on Android, desktop, and web
AndroidAndroid app exploration and documentation-driven automation research

A device toolkit for interaction, repeatable QA, and native or React Native diagnostics.

SupportedMCPSupportedFlowsSupportedDebugSupportedProfile
iOSAndroid

A shared CLI, MCP server, and typed Node.js runtime that lets coding agents inspect, control, and verify iOS, Android, HarmonyOS, TV, web, macOS, and Linux apps, and preserves evidence for review.

SupportedMCPSupportedFlowsSupportedDebugSupportedProfile
iOSAndroid

Apple simulator streaming and control through a browser preview, CLI, and portable agent skill. Includes accessibility inspection, camera injection, and rendering diagnostics.

UnverifiedMCPPartialFlowsSupportedDebugPartialProfile
iOS

An MCP server for automating iOS and Android apps via accessibility trees or screenshot-based taps, covering simulators, emulators, real devices, and optional Mobile Next Cloud remote devices.

SupportedMCPUnverifiedFlowsSupportedDebugUnverifiedProfile
iOSAndroid

A UI testing framework with agent access through its bundled MCP server.

SupportedMCPSupportedFlowsPartialDebugUnverifiedProfile
iOSAndroid

An MCP server and CLI for building, running, testing, and debugging Xcode projects on iOS simulators, physical Apple devices, and macOS.

SupportedMCPPartialFlowsSupportedDebugUnverifiedProfile
iOS

A Model Context Protocol server that lets an MCP-integrated AI assistant inspect and control an iOS simulator: accessibility tree queries, tap/type/swipe, screenshots, video recording, app install/launch/terminate, and deep links.

SupportedMCPUnverifiedFlowsUnverifiedDebugUnverifiedProfile
iOS

An MCP interface to the Appium ecosystem, driving embedded local or remote Android, iOS, and tvOS sessions with AI-assisted locators and test generation.

SupportedMCPPartialFlowsPartialDebugUnverifiedProfile
iOSAndroid

An AI-powered GUI agent and testing kit for writing, running, and debugging visual UI workflows across web, Android, iOS, and desktop interfaces.

UnverifiedMCPSupportedFlowsPartialDebugUnverifiedProfile
iOSAndroid

An open-source framework for controlling Android and iOS devices with LLM agents, offering a CLI, Python API, and an optional Mobilerun Cloud service for hosted or managed devices.

UnverifiedMCPSupportedFlowsPartialDebugUnverifiedProfile
iOSAndroid

A phone agent that decomposes natural-language tasks and acts on UI state.

UnverifiedMCPUnverifiedFlowsUnverifiedDebugUnverifiedProfile
iOSAndroid

An open phone-agent vision-language model and framework that automates Android, HarmonyOS, and iPhone (via WebDriverAgent) apps from natural-language instructions, using a self-hosted or third-party OpenAI-compatible model endpoint.

UnverifiedMCPUnverifiedFlowsPartialDebugUnverifiedProfile
Android

A family of visual GUI agents and models (Mobile-Agent v1-v3.5, GUI-Owl, PC-Agent, UI-S1, GUI-Critic-R1) from Alibaba's Tongyi Lab. Mobile-Agent-v3 deploys a multi-agent planning framework on real Android/HarmonyOS phones via ADB/HDC, while the GUI-Owl-1.5 models are described as supporting desktop, mobile, and browser automation.

UnverifiedMCPUnverifiedFlowsUnverifiedDebugUnverifiedProfile
Android

A CHI 2025 research framework where a multimodal LLM (e.g. GPT-4V or Qwen-VL-max) learns to operate Android apps through autonomous exploration or human demonstration, building a per-app documentation knowledge base of UI elements that it then uses to complete new tasks via tap/swipe/type actions over ADB.

UnverifiedMCPUnverifiedFlowsUnverifiedDebugUnverifiedProfile
Android

Different tools, different jobs. These are documented capabilities, not benchmark scores. “Unverified” means we haven’t established support from the sources.

Built to stay current.

Eve checks upstream repositories and proposes data updates for review.