MCP server for image generation supporting multiple API providers (Tencent Hunyuan, OpenAI, Doubao) with both stdio and HTTP transport modes
Server provides 3 tools with reasonable input schemas and descriptions, but has significant gaps in parameter documentation, output schema clarity, and error handling guidance. Tool names follow verb_noun convention (generate_image, get_image_data, reload_config), but descriptions lack specificity about return structure and error recovery paths. Parameter descriptions exist but are inconsistent in detail, some provide format guidance ('Format: provider:style'), while others are sparse. No documented output schemas, no error recovery guidance for LLM clients. Server operates over STDIO+HTTP but schema definitions are embedded in code without formal registration visible. Overall falls into 'Fair' category with noticeable gaps.
Generate image based on prompt using multiple API providers (Hunyuan, OpenAI, Doubao)
Get base64 text data for a previously generated image by image_id (for programmable artifact use)
Reload runtime configuration from environment/.env without restarting the process. Only provider/model related settings and a small safe subset are hot-reloadable.
Missing output schema documentation. generate_image returns image data but no documented structure (image_id format, dimensions, URLs, metadata). Clients cannot reliably extract and chain IDs.
Incomplete parameter descriptions. 'provider' param says 'Leave empty to use default provider' but no default value documented in input schema. 'style', 'resolution', 'background' parameters lack format constraints or examples. LLMs cannot infer valid enums without explicit documentation.
No error recovery guidance. When image generation fails, no documented error message format or next-step guidance. LLM has no clue whether to retry, check input, or ask user. Missing pattern:recovery-guide.
Lack of idempotency clarity. generate_image does not document whether repeated calls with identical params produce same image_id or new one. Agents will not know if safe to retry on ambiguous failures.
Inferred effective spec: 2025-06-18+.
| Scored | Grade | Overall | Spec posture | Rubric |
|---|---|---|---|---|
| 2026-09-22 | D | 53 | 2025-06-18+ | v2 |
| 2026-03-09 | D | 50 | - | v1 |
Tool parameters accept free-form strings where enums are warranted. 'provider' should enumerate hunyuan|openai|doubao. 'background' hints at OpenAI-only enums but does not formally constrain them. 'output_compression' has min/max but 'output_format' is unconstrained.
get_image_data description is terse (48 chars). Does not explain when to use it vs embedded image in generate_image response. No documentation of what base64 data is returned or expected downstream consumption.
reload_config tool permits runtime configuration changes but no documented permission gate or audit trail requirement. Sensitive operation with no authorization guard specified.