AI-driven cross-platform UI automation engine with test script generation
ScreenForge defines 5 tools with complete input schemas and descriptions. Naming follows verb_noun conventions (ui_agent_*), but descriptions vary in quality and completeness. All tools have proper JSON Schema definitions with types and enum constraints. However, parameter descriptions are sparse, output schemas are undocumented, and error handling guidance is missing. The server is at the median for community tools but has gaps preventing production-grade adoption.
Returns a capability snapshot of supported platforms, modes, control planes, and actions.
Execute or dry-run a ScreenForge request. Agent entry supports workflow, action, and doctor control planes. Execution modes: run, doctor, plan_only, dry_run. Platforms: web, android, ios. Control planes: workflow, action, doctor. Actions requiring extra_value: click, input, scroll_to, swipe, wait, select, press, assert_text, assert_url, goto.
Connect to the target platform and return a cleaned XML/DOM tree for the Agent to analyze and locate elements.
Load cross-run test memory for the Agent to reuse successful actions, locator strategies, and pytest assets.
Load a historical run by run_id, returning its summary, resume_context, and replay assets.
Output schemas undocumented. Tool descriptions state what is returned in prose but do not define the structured response schema. LLMs cannot plan downstream operations without knowing field names and types.
Error handling lacks recovery guidance. Tool descriptions do not explain what errors are retryable, what the LLM should do if a call fails, or how to self-correct invalid inputs.
Parameter relationships undocumented. ui_agent_execute accepts both 'workflow' and 'action' objects, but descriptions do not clarify that these are mutually exclusive or how the mode/platform interact with control_planes.
Inferred effective spec: 2025-06-18+.
| Scored | Grade | Overall | Spec posture | Rubric |
|---|---|---|---|---|
| 2026-09-23 | C | 66 | 2025-06-18+ | v2 |
No pagination support visible. ui_agent_load_case_memory accepts a 'limit' parameter (default 20) but does not document offset/cursor mechanism, total count return, or what to do when >limit results exist.
Descriptions contain Chinese text ("执行模式", "目标平台") mixed with English, reducing clarity for non-Chinese LLMs and creating inconsistency.
No tool annotations (readOnlyHint, destructiveHint, idempotentHint). ui_agent_execute is clearly a WRITE operation, but this is not declared in the schema or via annotations.