Free Official Data Samples, Provenance, Aggregations & Insights
Official data with free samples, provenance, aggregations, freshness and agent-ready insights.
Should I use this
Quality & Safety
Findings (1)
- LOWin request_dataset_materialization
Based on automated analysis of tool definitions and protocol compliance.
Context Cost
This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.
Install
One-Click Install
Add this to your `claude_desktop_config.json` file:
{
"mcpServers": {
"official-provenance-aggregations-insights-free-samples-open-data": {
"url": "https://agentnative.cazimedia.com/mcp/free-official-data-insights"
}
}
}Remote endpoints
https://agentnative.cazimedia.com/mcp/free-official-data-insightsstreamable-httpWhat it can do
Tool inventory
Tools (14)
🟢search_public_datasets(query)
Use this free tool first when an agent needs official US federal data but does not yet know the dataset ID. Searches normalized catalog metadata and returns matching datasets with provenance; it does not query dataset rows.
Input Schema
{
"type": "object",
"properties": {
"query": {
"type": "string",
"minLength": 1,
"description": "Plain-language topic, agency, or dataset keywords, such as employment, schools, or air quality."
}
},
"required": [
"query"
],
"additionalProperties": false
}🟢request_dataset_materialization(discovered_dataset_id)
Queue a discovered official dataset for prioritized detached ingestion. This free idempotent operation returns a durable job ID; poll get_materialization_status until it supplies sample and query URLs.
Input Schema
{
"type": "object",
"properties": {
"discovered_dataset_id": {
"type": "string",
"pattern": "^disc_[a-f0-9]{24}$",
"description": "Exact catalog ID returned by search_discovered_datasets."
}
},
"required": [
"discovered_dataset_id"
],
"additionalProperties": false
}🟢get_materialization_status(materialization_id)
Poll a durable materialization job. A complete result includes the query-ready dataset ID plus free sample and paid query URLs.
Input Schema
{
"type": "object",
"properties": {
"materialization_id": {
"type": "string",
"pattern": "^job_[a-f0-9]{24}$"
}
},
"required": [
"materialization_id"
],
"additionalProperties": false
}🟢list_official_sources
Use this free tool to inspect which federal, state, city, police, education, and other official source scopes AgentNative currently covers. Returns each source's importer, discovery state, and materialized coverage; it does not return dataset rows.
Input Schema
{
"type": "object",
"properties": {},
"additionalProperties": false
}🟢search_discovered_datasets(query, source_id, limit)
Use this free tool to find datasets across all discovered official sources, including candidates not yet queryable. Returns catalog matches and materialization state; use list_imported_datasets when only query-ready data is acceptable.
Input Schema
{
"type": "object",
"properties": {
"query": {
"type": "string",
"maxLength": 100,
"description": "Optional title or description keywords."
},
"source_id": {
"type": "string",
"maxLength": 100,
"description": "Optional exact official source ID returned by list_official_sources."
},
"limit": {
"type": "integer",
"minimum": 1,
"maximum": 100,
"description": "Maximum matches to return; defaults to the service limit."
}
},
"additionalProperties": false
}🟢get_coverage_status
Use this free operational tool to decide whether available public-data coverage is sufficient or whether to request a missing capability. Returns discovery, materialization, queryable-row, queue, failure, and freshness counts.
Input Schema
{
"type": "object",
"properties": {},
"additionalProperties": false
}🟢list_imported_datasets
Use this free tool when an agent needs only datasets that can be sampled or queried now. Returns fully published warehouse snapshots with dataset IDs, row counts, freshness, and provenance.
Input Schema
{
"type": "object",
"properties": {},
"additionalProperties": false
}🟢sample_imported_dataset(dataset_id)
Use this free tool to evaluate a query-ready official dataset before paying. Returns three normalized rows, deterministic summaries, freshness, and provenance; use query_imported_dataset only after the sample proves useful.
Input Schema
{
"type": "object",
"properties": {
"dataset_id": {
"type": "string",
"description": "Exact query-ready dataset ID returned by list_imported_datasets."
}
},
"required": [
"dataset_id"
],
"additionalProperties": false
}🟢sample_us_federal_register
Use this free no-auth tool for current US regulatory activity or to evaluate AgentNative before paying. Returns three Federal Register records plus deterministic aggregates over the latest 25 documents, direct record URLs, provenance, and rate-limit status.
Input Schema
{
"type": "object",
"properties": {},
"additionalProperties": false
}🟢query_imported_dataset(dataset_id, top, skip, filter_field, filter_value)
Use this paid read-only tool after sample_imported_dataset confirms the data is suitable. Returns up to 100 normalized rows from a published snapshot with bounded pagination, one exact-match filter, freshness, and provenance.
Input Schema
{
"type": "object",
"properties": {
"dataset_id": {
"type": "string"
},
"top": {
"type": "integer",
"minimum": 1,
"maximum": 100
},
"skip": {
"type": "integer",
"minimum": 0,
"maximum": 5000
},
"filter_field": {
"type": "string"
},
"filter_value": {
"type": "string"
}
},
"required": [
"dataset_id"
],
"additionalProperties": false
}🟢aggregate_imported_dataset(dataset_id, group_by, metric, field, top)
Use this paid read-only tool for deterministic grouped statistics instead of downloading rows and calculating locally. Returns bounded count, sum, average, minimum, or maximum groups with dataset provenance.
Input Schema
{
"type": "object",
"properties": {
"dataset_id": {
"type": "string"
},
"group_by": {
"type": "string"
},
"metric": {
"enum": [
"count",
"sum",
"avg",
"min",
"max"
]
},
"field": {
"type": "string"
},
"top": {
"type": "integer",
"minimum": 1,
"maximum": 50
}
},
"required": [
"dataset_id",
"group_by"
],
"additionalProperties": false
}⚪request_capability(dataset, capability, source_url, example_query)
Use this state-changing tool only when existing discovery and coverage tools cannot satisfy the task. Records demand for a missing dataset, aggregation, insight, filter, freshness level, or export; repeated requests increase autonomous build priority.
Input Schema
{
"type": "object",
"properties": {
"dataset": {
"type": "string",
"minLength": 2,
"maxLength": 100
},
"capability": {
"type": "string",
"enum": [
"dataset",
"filter",
"aggregation",
"insight",
"freshness",
"export"
]
},
"source_url": {
"type": "string",
"format": "uri"
},
"example_query": {
"type": "string",
"maxLength": 240
}
},
"required": [
"dataset",
"capability"
],
"additionalProperties": false
}⚪request_paid_access(plan)
Use this state-changing tool only after free samples demonstrate value and the human owner approves payment. Creates a Stripe Checkout plus one-time claim URL; choose the $5 pass for research or $5/month subscription for recurring workflows.
Input Schema
{
"type": "object",
"properties": {
"plan": {
"type": "string",
"enum": [
"pass",
"subscription"
],
"default": "pass"
}
},
"additionalProperties": false
}🟢query_us_federal_register(start_date, end_date, type, top)
Use this paid read-only tool when the free Federal Register briefing is insufficient. Returns up to 100 normalized official records filtered by date or document type, with deterministic provenance and direct source URLs.
Input Schema
{
"type": "object",
"properties": {
"start_date": {
"type": "string",
"pattern": "^\\d{4}-\\d{2}-\\d{2}$"
},
"end_date": {
"type": "string",
"pattern": "^\\d{4}-\\d{2}-\\d{2}$"
},
"type": {
"type": "string",
"enum": [
"NOTICE",
"PRORULE",
"RULE",
"PRESDOCU"
]
},
"top": {
"type": "integer",
"minimum": 1,
"maximum": 100
}
},
"additionalProperties": false
}Community
Evidence