Tapkit lets AI agents see, tap, type, and work across apps on real iPhones.
Every app becomes a tool your agent can use. Your agent sees the same screen you do.

TapKit gives you programmatic control of real, physical iPhones through Apple's accessibility features. You send actions (tap, swipe, type) via our API, and they execute on your iPhone. For framework differences, compare TapKit with Appium.
No. Simulators run virtual iOS environments that can't install or use apps. TapKit runs on real iPhones. See the full TapKit vs iOS simulator comparison.
Anything on the App Store you can use.
Use our MCP server to connect directly to Claude Code or Cursor, our REST API with a Python SDK for custom integrations, or our Mac and Web apps for no-code control. The typical agent loop is: screenshot → send to your vision model → get action → execute on device. See how TapKit compares with iPhone Mirroring MCP.
Yes, TapKit is bring-your-own-device. You connect your own iPhone to your Mac and TapKit turns it into an API. We're exploring hosted phone options for teams that don't want to manage their own hardware. If you're evaluating hosted hardware, compare TapKit with device farms.
TapKit starts at $99/month for one phone, with a five-phone plan at $399/month for parallel workflows and custom pricing for larger fleets.
Every plan includes the complete TapKit experience.
Choose how many iPhones your agents can use.
For your first workflow.
For parallel workflows.
For production deployments.
Use your own iPhones. No special hardware. Cancel anytime.