独立的指标计算与缓存项目 (Independent metrics calculation and caching project)
This server exposes only 2 logging/instrumentation tools with significant quality gaps. Both tools are defined via fastmcp decorators in log_tools.py with input schemas and descriptions present, but critical issues reduce usability: (1) descriptions are in Chinese, making them inaccessible to most LLM deployments and evaluators; (2) parameter descriptions lack actionable constraints or examples; (3) no output schema documentation; (4) error handling returns unstructured strings rather than guiding recovery; (5) tools write directly to local files without validation, posing security risks in agent contexts. The tools themselves are functional but poorly optimized for LLM reasoning and safety.
记录指标设计数据 功能: 1. 接收指标设计数据 2. 保存到日志文件 3. 返回确认信息 使用场景: - 记录不同阶段的指标设计结果 - 保存指标树结构 - 追踪设计变更
记录 LLM 的思考内容到日志 功能: 1. 接收 LLM 的思考内容 2. 记录到日志文件 3. 返回确认信息 使用场景: - 在 LangGraph 节点中记录 LLM 的思考过程 - 保存中间结果和推理过程 - 便于调试和分析
查询API服务列表 功能: 1. 从配置文件加载API服务列表 2. 支持按API服务ID过滤 数据来源: - config/api_schemas.yaml - API服务列表
查询现有指标库 功能: 1. 从配置文件加载现有指标库 2. 支持按类型过滤 3. 支持关键词搜索 数据来源: - config/metrics.yaml - 现有指标库
查询函数库 功能: 1. 从配置文件加载函数库 2. 支持按函数类型过滤 数据来源: - config/metrics_meta.yaml - 函数库
Descriptions in Chinese without English translation, LLMs trained on English struggle with non-English docstrings, reducing tool discoverability and selection accuracy.
No output schema documented. Callers cannot predict response structure (success/error fields, error message format, log file path format). LLMs must infer structure from example responses.
Parameter descriptions lack constraint information (e.g., what constitutes valid node_name, thought_type, design_stage values). No enum constraints provided despite enumerable values (design_stage: 'analysis', 'conceptual', 'concrete', 'validation').
Error handling returns plain strings ('success': False, 'error': str(e)). Does not guide recovery or categorize errors as retryable vs fatal. LLM receives no actionable guidance after failure.
| Scored | Grade | Overall | Spec posture | Rubric |
|---|---|---|---|---|
| 2026-09-22 | F | 9 | 2026-07-28+ | v2 |
| 2026-03-09 | F | 29 | - | v1 |
File writes to local log/ directory without validation, access control, or disk space checks. In a multi-agent environment, concurrent writes risk file corruption or permission errors. No cleanup/rotation policy visible.
Metadata parameter (log_thought_content) is optional Dict but lacks schema validation. Agents may pass arbitrary nested objects, risking JSON serialization failures or log pollution.
Tool names 'log_*' do not clearly distinguish their purpose in a logging context. 'record_*' or 'store_*' would be clearer, but naming is acceptable (verb + noun convention followed).