Dataset Filter & Transform
Clean, filter and reshape JSON rows in one call: 26 transforms, 25 filters, sort, dedupe, limit.
我该使用它吗
质量与安全性
基于对工具定义和协议合规性的自动分析。
上下文开销
这是每次将服务器的工具加载到模型上下文窗口时所消耗的大致 token 数。数值越高,可用于其他任务的注意力就越少。
安装
一键安装
将以下内容添加到你的 `claude_desktop_config.json` 文件中:
{
"mcpServers": {
"dataset-filter-transform": {
"url": "https://dataset-filter-transform.nerolabs.workers.dev/mcp"
}
}
}远程端点
https://dataset-filter-transform.nerolabs.workers.dev/mcpstreamable-http它能做什么
工具清单
工具(2)
🟢list_capabilities
Returns every filter operator and transform operation this server supports, the order the pipeline runs in, and the maximum rows per call. Call this first if you are unsure what is available. Free, processes no data.
输入模式
{
"type": "object",
"properties": {},
"additionalProperties": false
}🔴process_rows(rows, transforms, filters, filterCombineMode, caseSensitiveFilters, ...)
Cleans a list of JSON rows in one call: transforms each row, filters out the rows you do not want, then optionally sorts, de-duplicates and limits them. Returns the kept rows plus a summary of exactly what was removed and why. Use it to reshape scraped or API data before passing it on: rename and drop fields, cast strings to numbers, compute new fields with arithmetic, extract with regex, then keep only the rows that match your rules. Transforms run before filters, so you can filter on a field you just created.
输入模式
{
"type": "object",
"properties": {
"rows": {
"type": "array",
"description": "The rows to process. Each row is a JSON object. Keys may differ between rows.",
"items": {
"type": "object"
}
},
"transforms": {
"type": "array",
"description": "Transform steps applied to every row, in order. Each step is an object with an \"op\" key, for example {\"op\":\"cast\",\"field\":\"price\",\"to\":\"number\"} or {\"op\":\"compute\",\"field\":\"total\",\"expression\":\"price * qty\"} or {\"op\":\"rename\",\"from\":\"e_mail\",\"to\":\"email\"}. Call list_capabilities for all 26 operations.",
"items": {
"type": "object"
}
},
"filters": {
"type": "array",
"description": "Conditions a row must satisfy to be kept, for example {\"field\":\"country\",\"operator\":\"equals\",\"value\":\"UK\"} or {\"field\":\"price\",\"operator\":\"greaterThan\",\"value\":100}. Dot paths like \"address.city\" work. Call list_capabilities for all 25 operators.",
"items": {
"type": "object"
}
},
"filterCombineMode": {
"type": "string",
"enum": [
"AND",
"OR"
],
"description": "Whether a row must match every filter (AND, the default) or any one of them (OR)."
},
"caseSensitiveFilters": {
"type": "boolean",
"description": "Text comparisons are case-insensitive by default. Set true to make them exact."
},
"lenientNumbers": {
"type": "boolean",
"description": "On by default: values like \"$1,234.50\", \"49 USD\" and \"12%\" are read as numbers. Set false to require real numeric values."
},
"sortBy": {
"type": "array",
"description": "Sort the kept rows, for example [{\"field\":\"price\",\"direction\":\"desc\"}].",
"items": {
"type": "object"
}
},
"distinctBy": {
"type": "array",
"description": "Keep only the first row for each combination of these fields. Applied after sorting, so sorting first lets you keep the newest row per key.",
"items": {
"type": "string"
}
},
"offset": {
"type": "integer",
"description": "Skip this many rows after filtering and sorting."
},
"limit": {
"type": "integer",
"description": "Return at most this many rows."
}
},
"required": [
"rows"
],
"additionalProperties": false
}社区
证据