Dataset Filter & Transform
Clean, filter and reshape JSON rows in one call: 26 transforms, 25 filters, sort, dedupe, limit.
사용해야 할까요
품질 및 안전성
도구 정의와 프로토콜 준수에 대한 자동 분석을 기반으로 합니다.
컨텍스트 비용
이는 서버의 도구가 모델의 컨텍스트에 로드될 때마다 소비되는 대략적인 토큰 수입니다. 수치가 높을수록 다른 작업에 사용할 수 있는 주의가 줄어듭니다.
설치
원클릭 설치
`claude_desktop_config.json` 파일에 다음을 추가하세요:
{
"mcpServers": {
"dataset-filter-transform": {
"url": "https://dataset-filter-transform.nerolabs.workers.dev/mcp"
}
}
}원격 엔드포인트
https://dataset-filter-transform.nerolabs.workers.dev/mcpstreamable-http할 수 있는 일
도구 목록
도구 (2)
🟢list_capabilities
Returns every filter operator and transform operation this server supports, the order the pipeline runs in, and the maximum rows per call. Call this first if you are unsure what is available. Free, processes no data.
입력 스키마
{
"type": "object",
"properties": {},
"additionalProperties": false
}🔴process_rows(rows, transforms, filters, filterCombineMode, caseSensitiveFilters, ...)
Cleans a list of JSON rows in one call: transforms each row, filters out the rows you do not want, then optionally sorts, de-duplicates and limits them. Returns the kept rows plus a summary of exactly what was removed and why. Use it to reshape scraped or API data before passing it on: rename and drop fields, cast strings to numbers, compute new fields with arithmetic, extract with regex, then keep only the rows that match your rules. Transforms run before filters, so you can filter on a field you just created.
입력 스키마
{
"type": "object",
"properties": {
"rows": {
"type": "array",
"description": "The rows to process. Each row is a JSON object. Keys may differ between rows.",
"items": {
"type": "object"
}
},
"transforms": {
"type": "array",
"description": "Transform steps applied to every row, in order. Each step is an object with an \"op\" key, for example {\"op\":\"cast\",\"field\":\"price\",\"to\":\"number\"} or {\"op\":\"compute\",\"field\":\"total\",\"expression\":\"price * qty\"} or {\"op\":\"rename\",\"from\":\"e_mail\",\"to\":\"email\"}. Call list_capabilities for all 26 operations.",
"items": {
"type": "object"
}
},
"filters": {
"type": "array",
"description": "Conditions a row must satisfy to be kept, for example {\"field\":\"country\",\"operator\":\"equals\",\"value\":\"UK\"} or {\"field\":\"price\",\"operator\":\"greaterThan\",\"value\":100}. Dot paths like \"address.city\" work. Call list_capabilities for all 25 operators.",
"items": {
"type": "object"
}
},
"filterCombineMode": {
"type": "string",
"enum": [
"AND",
"OR"
],
"description": "Whether a row must match every filter (AND, the default) or any one of them (OR)."
},
"caseSensitiveFilters": {
"type": "boolean",
"description": "Text comparisons are case-insensitive by default. Set true to make them exact."
},
"lenientNumbers": {
"type": "boolean",
"description": "On by default: values like \"$1,234.50\", \"49 USD\" and \"12%\" are read as numbers. Set false to require real numeric values."
},
"sortBy": {
"type": "array",
"description": "Sort the kept rows, for example [{\"field\":\"price\",\"direction\":\"desc\"}].",
"items": {
"type": "object"
}
},
"distinctBy": {
"type": "array",
"description": "Keep only the first row for each combination of these fields. Applied after sorting, so sorting first lets you keep the newest row per key.",
"items": {
"type": "string"
}
},
"offset": {
"type": "integer",
"description": "Skip this many rows after filtering and sorting."
},
"limit": {
"type": "integer",
"description": "Return at most this many rows."
}
},
"required": [
"rows"
],
"additionalProperties": false
}커뮤니티
증거