Context integrity infrastructure for AI agents and retrieval systems. Score, explain, and wrap candidate context before it reaches the model.
freshcontext-mcp is a context-evaluation and data-extraction server with 18 tools focused on scraping, searching, and scoring information from diverse sources (GitHub, arXiv, Reddit, finance, etc.). Tool naming is generally clear and verb-based (extract_*, search_*), and descriptions are present and reasonably detailed. However, critical gaps emerge: (1) Input schemas are visible but parameter descriptions vary significantly in quality and completeness, many parameters lack inline descriptions; (2) Output schemas are not documented anywhere in the source code provided, forcing LLMs to guess at response structure; (3) Error handling and recovery guidance are absent; (4) Several tools conflate or blur responsibilities (e.g., extract_changelog could be search_changelog). The evaluate_context tool shows good parameter structure (profile, intent, signals as an object array with typed fields), but most extraction tools expose only 'url' and 'max_length' with minimal semantic guidance. The server is production-usable but falls short of A-grade standards due to underdocumented outputs and generic parameter descriptions.
Evaluate caller-provided candidate context and return decision-ready output. This is the primary FreshContext judgment path: it does not fetch, crawl, scrape, browse, read folders, or call adapters.
Extract papers from an arXiv search or category URL. Returns title, abstract, authors, publication date, and PDF link per paper.
Extract a software project's changelog — releases, dates, and key changes. Returns timestamped version history.
Extract financial data from Yahoo Finance, SEC Edgar filings, or macroeconomic dashboards. Returns price data, earnings, and filing links — all timestamped.
Extract real-time data from a GitHub repository — README, stars, forks, language, topics, last commit. Returns timestamped freshcontext.
Extract top stories or search results from Hacker News. Accepts an HN/Algolia URL or a plain search query while preserving the url field for compatibility.
Output schemas completely undocumented. No source code shows what fields or structure extract_github, search_repos, or other tools return. LLMs cannot plan downstream operations or validate responses without documented output schemas.
Many parameter descriptions are generic or missing. Tools like extract_github, extract_scholar, extract_reddit specify only 'url' and 'max_length' with minimal guidance on format, constraints, or expected values. LLMs cannot validate or self-correct invalid URLs without format hints (e.g., 'Full GitHub repo URL, not API endpoint').
| Scored | Grade | Overall | Spec posture | Rubric |
|---|---|---|---|---|
| 2026-09-22 | C | 61 | 2026-07-28+ | v2 |
| 2026-03-09 | D | 53 | - | v1 |
Extract a CNCF or custom ecosystem landscape. Returns tool categories, projects, and maturity levels from the SVG.
Extract Product Hunt launches and discussions. Returns product name, tagline, maker, and comment summaries — all timestamped.
Extract posts and comments from a Reddit URL (subreddit, search, or thread). Returns post titles, scores, and comment summaries — all timestamped.
Extract research results from a Google Scholar search URL. Returns titles, authors, publication years, and snippets — all timestamped.
Scrape YC company listings. Use https://www.ycombinator.com/companies?query=KEYWORD to find startups in a space. Returns name, batch, tags, description per company. Freshness is unknown — YC listings carry no reliable per-company update date.
Look up npm and PyPI package metadata — version history, release cadence, last updated. Use to gauge ecosystem activity around a tool or dependency. Supports comma-separated list of packages.
Search the GDELT database for geopolitical events, media articles, and mentions of organizations or people. Returns event summaries and source links — all timestamped.
Search Singapore government procurement (GeBIZ). Returns tender title, buyer agency, and closing date.
Search U.S. government contracts (SAM.gov, FedBizOpps) and proposals. Returns contract title, agency, and award amount.
Search job boards (HN Jobs, We Work Remotely, etc.) for job postings. Returns title, company, location, and posting date.
Search GitHub for repositories matching a keyword or topic. Returns top results by stars with activity signals. Use to find competitors, similar tools, or related projects.
Search SEC filings (10-K, 8-K, S-1, etc.) for a public company. Returns filing type, date, and link to the full document.
No error handling or recovery guidance. Tools do not document what happens when a URL is invalid, the service is down, or content cannot be scraped. LLMs cannot determine whether to retry, fall back, or inform the user.
Parameter schema documentation incomplete. While Zod schemas likely exist in the implementation, the actual parameter constraints (enums, min/max lengths, regex patterns) are not visible in the source excerpt. Tools like search_govcontracts and search_gebiz accept free-form query strings with no validation hints.
Inconsistent parameter naming. Some tools use 'url', others infer source type via URL content (e.g., extract_hackernews accepts 'HN URL... or search query'). Mixed handling of URLs vs. search queries as the same parameter increases ambiguity.