Dataset Filter & Transform
Clean, filter and reshape JSON rows in one call: 26 transforms, 25 filters, sort, dedupe, limit.
Should I use this
Quality & Safety
Based on automated analysis of tool definitions and protocol compliance.
Context Cost
This is the approximate number of tokens consumed each time the server's tools are loaded into a model's context. Higher counts reduce the attention available for other tasks.
Install
One-Click Install
Add this to your `claude_desktop_config.json` file:
{
"mcpServers": {
"dataset-filter-transform": {
"url": "https://dataset-filter-transform.nerolabs.workers.dev/mcp"
}
}
}Remote endpoints
https://dataset-filter-transform.nerolabs.workers.dev/mcpstreamable-httpWhat it can do
Tool inventory
Tools (2)
🟢list_capabilities
Returns every filter operator and transform operation this server supports, the order the pipeline runs in, and the maximum rows per call. Call this first if you are unsure what is available. Free, processes no data.
Input Schema
{
"type": "object",
"properties": {},
"additionalProperties": false
}🔴process_rows(rows, transforms, filters, filterCombineMode, caseSensitiveFilters, ...)
Cleans a list of JSON rows in one call: transforms each row, filters out the rows you do not want, then optionally sorts, de-duplicates and limits them. Returns the kept rows plus a summary of exactly what was removed and why. Use it to reshape scraped or API data before passing it on: rename and drop fields, cast strings to numbers, compute new fields with arithmetic, extract with regex, then keep only the rows that match your rules. Transforms run before filters, so you can filter on a field you just created.
Input Schema
{
"type": "object",
"properties": {
"rows": {
"type": "array",
"description": "The rows to process. Each row is a JSON object. Keys may differ between rows.",
"items": {
"type": "object"
}
},
"transforms": {
"type": "array",
"description": "Transform steps applied to every row, in order. Each step is an object with an \"op\" key, for example {\"op\":\"cast\",\"field\":\"price\",\"to\":\"number\"} or {\"op\":\"compute\",\"field\":\"total\",\"expression\":\"price * qty\"} or {\"op\":\"rename\",\"from\":\"e_mail\",\"to\":\"email\"}. Call list_capabilities for all 26 operations.",
"items": {
"type": "object"
}
},
"filters": {
"type": "array",
"description": "Conditions a row must satisfy to be kept, for example {\"field\":\"country\",\"operator\":\"equals\",\"value\":\"UK\"} or {\"field\":\"price\",\"operator\":\"greaterThan\",\"value\":100}. Dot paths like \"address.city\" work. Call list_capabilities for all 25 operators.",
"items": {
"type": "object"
}
},
"filterCombineMode": {
"type": "string",
"enum": [
"AND",
"OR"
],
"description": "Whether a row must match every filter (AND, the default) or any one of them (OR)."
},
"caseSensitiveFilters": {
"type": "boolean",
"description": "Text comparisons are case-insensitive by default. Set true to make them exact."
},
"lenientNumbers": {
"type": "boolean",
"description": "On by default: values like \"$1,234.50\", \"49 USD\" and \"12%\" are read as numbers. Set false to require real numeric values."
},
"sortBy": {
"type": "array",
"description": "Sort the kept rows, for example [{\"field\":\"price\",\"direction\":\"desc\"}].",
"items": {
"type": "object"
}
},
"distinctBy": {
"type": "array",
"description": "Keep only the first row for each combination of these fields. Applied after sorting, so sorting first lets you keep the newest row per key.",
"items": {
"type": "string"
}
},
"offset": {
"type": "integer",
"description": "Skip this many rows after filtering and sorting."
},
"limit": {
"type": "integer",
"description": "Return at most this many rows."
}
},
"required": [
"rows"
],
"additionalProperties": false
}Community
Evidence