Introducing AutoGLM-Phone-MultilingualZ.AIReleased December 11, 2025

AutoGLM-Phone-Multilingual

Multimodal mobile automation framework that understands smartphone screen content and executes real user actions through ADB across 50+ mainstream apps.

autonomous-agentProprietary APIMobile OS Automation AgentContext: 64K (64,000 tokens)

Technical Specifications

Architecture Type
Mobile OS Automation Multimodal Agent
Total Parameters
Specialized Mobile Vision Scale
Context Window
64K (64,000 tokens)
Max Output Tokens
8.192K (8,192 tokens)
Knowledge Cutoff
December 2025
Supported Modalities
vision, text, actions
License & Access
Z.AI Terms of Service

Benchmark Evaluations

Adb Execution Accuracy
94.2
Mobile Task Success Rate
88.6

Deep Architectural Overview

AutoGLM-Phone-Multilingual enables end-to-end smartphone control from natural language instructions. It parses Android UI hierarchies and screenshots, plans multi-step interaction workflows, and dispatches ADB tap, swipe, and input gestures reliably.

Strengths & Considerations

Core Strengths
  • True end-to-end smartphone screen understanding and control
  • Direct ADB gesture execution across 50+ apps
  • Multilingual support (English & Chinese)
Known Limitations
  • Requires ADB access and device permissions to dispatch actions

Token & API Pricing

Input Tokens (1M)$0.60
Output Tokens (1M)$2.20
Cached Input (1M)$0.1100
Pricing is verified directly against Z.AI's developer documentation and API rate sheets.

Similar & Alternative Models

Explore other frontier models from Z.AI and comparable reasoning engines.

Browse all models
Z.AI

Ultra-low latency version of GLM-5.3-Flash engineered for high-concurrency real-time IDE code completion and interactive agents.

200K ctx$0.37 / 1M tok (Ultra-Low Latency MoE)
Z.AI

The first native multimodal model in the GLM-5 series: 320B parameters (18B active) combining linear and sparse attention for low-cost visual coding.

200K ctx$0.15 / 1M tok (320B MoE Visual Coding)
Z.AI

Z.AI’s premier flagship model, delivering a 50% performance gain on Code Bench and matching Claude Mythos 5 in cybersecurity and vulnerability discovery.

1M ctx$1.40 / 1M tok (Flagship Coding & Cyber)
Z.AI

Flagship model built for project-scale context, supporting truly usable 1M-token context with 128k output.

1M ctx$1.40 / 1M tok (1M Lossless Context)