Free Official Data Samples, Provenance, Aggregations & Insights

Official data with free samples, provenance, aggregations, freshness and agent-ready insights.

Should I use this

Quality & Safety

A
Description quality
100%
Schema completeness
69%
Naming quality
91%
Poisoning risk
100%
Permission match
100%
Protocol compliance
100%

Findings (1)

  • LOWTool 'request_dataset_materialization' name length outside 3-30 rangein request_dataset_materialization

Based on automated analysis of tool definitions and protocol compliance.

Context Cost

~1,836Tokens (tool definitions)
~618 BTypical response size
Moderate attention impact (1.43% of 128k context)

This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.

Install

One-Click Install

Add this to your `claude_desktop_config.json` file:

{
  "mcpServers": {
    "official-provenance-aggregations-insights-free-samples-open-data": {
      "url": "https://agentnative.cazimedia.com/mcp/free-official-data-insights"
    }
  }
}

Remote endpoints

https://agentnative.cazimedia.com/mcp/free-official-data-insightsstreamable-http

What it can do

Tool inventory

Tools (14)

🟢 Read-only🟡 Write🔴 Delete⚪ Unknown
🟢search_public_datasets(query)

Use this free tool first when an agent needs official US federal data but does not yet know the dataset ID. Searches normalized catalog metadata and returns matching datasets with provenance; it does not query dataset rows.

Input Schema

{
  "type": "object",
  "properties": {
    "query": {
      "type": "string",
      "minLength": 1,
      "description": "Plain-language topic, agency, or dataset keywords, such as employment, schools, or air quality."
    }
  },
  "required": [
    "query"
  ],
  "additionalProperties": false
}
🟢request_dataset_materialization(discovered_dataset_id)

Queue a discovered official dataset for prioritized detached ingestion. This free idempotent operation returns a durable job ID; poll get_materialization_status until it supplies sample and query URLs.

Input Schema

{
  "type": "object",
  "properties": {
    "discovered_dataset_id": {
      "type": "string",
      "pattern": "^disc_[a-f0-9]{24}$",
      "description": "Exact catalog ID returned by search_discovered_datasets."
    }
  },
  "required": [
    "discovered_dataset_id"
  ],
  "additionalProperties": false
}
🟢get_materialization_status(materialization_id)

Poll a durable materialization job. A complete result includes the query-ready dataset ID plus free sample and paid query URLs.

Input Schema

{
  "type": "object",
  "properties": {
    "materialization_id": {
      "type": "string",
      "pattern": "^job_[a-f0-9]{24}$"
    }
  },
  "required": [
    "materialization_id"
  ],
  "additionalProperties": false
}
🟢list_official_sources

Use this free tool to inspect which federal, state, city, police, education, and other official source scopes AgentNative currently covers. Returns each source's importer, discovery state, and materialized coverage; it does not return dataset rows.

Input Schema

{
  "type": "object",
  "properties": {},
  "additionalProperties": false
}
🟢search_discovered_datasets(query, source_id, limit)

Use this free tool to find datasets across all discovered official sources, including candidates not yet queryable. Returns catalog matches and materialization state; use list_imported_datasets when only query-ready data is acceptable.

Input Schema

{
  "type": "object",
  "properties": {
    "query": {
      "type": "string",
      "maxLength": 100,
      "description": "Optional title or description keywords."
    },
    "source_id": {
      "type": "string",
      "maxLength": 100,
      "description": "Optional exact official source ID returned by list_official_sources."
    },
    "limit": {
      "type": "integer",
      "minimum": 1,
      "maximum": 100,
      "description": "Maximum matches to return; defaults to the service limit."
    }
  },
  "additionalProperties": false
}
🟢get_coverage_status

Use this free operational tool to decide whether available public-data coverage is sufficient or whether to request a missing capability. Returns discovery, materialization, queryable-row, queue, failure, and freshness counts.

Input Schema

{
  "type": "object",
  "properties": {},
  "additionalProperties": false
}
🟢list_imported_datasets

Use this free tool when an agent needs only datasets that can be sampled or queried now. Returns fully published warehouse snapshots with dataset IDs, row counts, freshness, and provenance.

Input Schema

{
  "type": "object",
  "properties": {},
  "additionalProperties": false
}
🟢sample_imported_dataset(dataset_id)

Use this free tool to evaluate a query-ready official dataset before paying. Returns three normalized rows, deterministic summaries, freshness, and provenance; use query_imported_dataset only after the sample proves useful.

Input Schema

{
  "type": "object",
  "properties": {
    "dataset_id": {
      "type": "string",
      "description": "Exact query-ready dataset ID returned by list_imported_datasets."
    }
  },
  "required": [
    "dataset_id"
  ],
  "additionalProperties": false
}
🟢sample_us_federal_register

Use this free no-auth tool for current US regulatory activity or to evaluate AgentNative before paying. Returns three Federal Register records plus deterministic aggregates over the latest 25 documents, direct record URLs, provenance, and rate-limit status.

Input Schema

{
  "type": "object",
  "properties": {},
  "additionalProperties": false
}
🟢query_imported_dataset(dataset_id, top, skip, filter_field, filter_value)

Use this paid read-only tool after sample_imported_dataset confirms the data is suitable. Returns up to 100 normalized rows from a published snapshot with bounded pagination, one exact-match filter, freshness, and provenance.

Input Schema

{
  "type": "object",
  "properties": {
    "dataset_id": {
      "type": "string"
    },
    "top": {
      "type": "integer",
      "minimum": 1,
      "maximum": 100
    },
    "skip": {
      "type": "integer",
      "minimum": 0,
      "maximum": 5000
    },
    "filter_field": {
      "type": "string"
    },
    "filter_value": {
      "type": "string"
    }
  },
  "required": [
    "dataset_id"
  ],
  "additionalProperties": false
}
🟢aggregate_imported_dataset(dataset_id, group_by, metric, field, top)

Use this paid read-only tool for deterministic grouped statistics instead of downloading rows and calculating locally. Returns bounded count, sum, average, minimum, or maximum groups with dataset provenance.

Input Schema

{
  "type": "object",
  "properties": {
    "dataset_id": {
      "type": "string"
    },
    "group_by": {
      "type": "string"
    },
    "metric": {
      "enum": [
        "count",
        "sum",
        "avg",
        "min",
        "max"
      ]
    },
    "field": {
      "type": "string"
    },
    "top": {
      "type": "integer",
      "minimum": 1,
      "maximum": 50
    }
  },
  "required": [
    "dataset_id",
    "group_by"
  ],
  "additionalProperties": false
}
⚪request_capability(dataset, capability, source_url, example_query)

Use this state-changing tool only when existing discovery and coverage tools cannot satisfy the task. Records demand for a missing dataset, aggregation, insight, filter, freshness level, or export; repeated requests increase autonomous build priority.

Input Schema

{
  "type": "object",
  "properties": {
    "dataset": {
      "type": "string",
      "minLength": 2,
      "maxLength": 100
    },
    "capability": {
      "type": "string",
      "enum": [
        "dataset",
        "filter",
        "aggregation",
        "insight",
        "freshness",
        "export"
      ]
    },
    "source_url": {
      "type": "string",
      "format": "uri"
    },
    "example_query": {
      "type": "string",
      "maxLength": 240
    }
  },
  "required": [
    "dataset",
    "capability"
  ],
  "additionalProperties": false
}
⚪request_paid_access(plan)

Use this state-changing tool only after free samples demonstrate value and the human owner approves payment. Creates a Stripe Checkout plus one-time claim URL; choose the $5 pass for research or $5/month subscription for recurring workflows.

Input Schema

{
  "type": "object",
  "properties": {
    "plan": {
      "type": "string",
      "enum": [
        "pass",
        "subscription"
      ],
      "default": "pass"
    }
  },
  "additionalProperties": false
}
🟢query_us_federal_register(start_date, end_date, type, top)

Use this paid read-only tool when the free Federal Register briefing is insufficient. Returns up to 100 normalized official records filtered by date or document type, with deterministic provenance and direct source URLs.

Input Schema

{
  "type": "object",
  "properties": {
    "start_date": {
      "type": "string",
      "pattern": "^\\d{4}-\\d{2}-\\d{2}$"
    },
    "end_date": {
      "type": "string",
      "pattern": "^\\d{4}-\\d{2}-\\d{2}$"
    },
    "type": {
      "type": "string",
      "enum": [
        "NOTICE",
        "PRORULE",
        "RULE",
        "PRESDOCU"
      ]
    },
    "top": {
      "type": "integer",
      "minimum": 1,
      "maximum": 100
    }
  },
  "additionalProperties": false
}

Community

Rate this Server

Evidence

Recent observations

verifiedversion not recorded14 tools
verifiedversion not recorded14 tools