BrunoSan ArXiv Intelligence
AI research intelligence for AI, ML, NLP, vision and robotics with citations and entity graphs.
我该使用它吗
质量与安全性
发现(1)
- LOW在 arxiv_citation_network 中
基于对工具定义和协议合规性的自动分析。
上下文开销
这是每次将服务器的工具加载到模型上下文窗口时所消耗的大致 token 数。数值越高,可用于其他任务的注意力就越少。
安装
一键安装
将以下内容添加到你的 `claude_desktop_config.json` 文件中:
{
"mcpServers": {
"arxiv": {
"url": "https://arxiv.mcp.brunosan.de/mcp"
}
}
}远程端点
https://arxiv.mcp.brunosan.de/mcpstreamable-http它能做什么
工具清单
工具(16)
🟢arxiv_search_papers(query, category, date_from, date_to, empirical_only, ...)
Full-text search over the live cs.AI/ML paper graph using FTS5. Searches title AND abstract. Supports boolean operators: AND, OR, NOT, phrase matching ("exact phrase"), prefix (term*). Args: query: FTS5 search query. E.g. 'LoRA fine-tuning', '"chain of thought"', 'RLHF NOT PPO' category: Filter by primary category. Options: cs.AI, cs.LG, cs.CL, cs.CV, cs.RO date_from: ISO date filter, e.g. '2024-01-01' date_to: ISO date filter, e.g. '2025-12-31' empirical_only: Only papers marked as empirical by LLM pass (if available) has_code_only: Only papers with code release (llm_has_code=1, if available) limit: Max results (default: 20, max: 50)
输入模式
{
"type": "object",
"properties": {
"query": {
"title": "Query",
"type": "string"
},
"category": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Category"
},
"date_from": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date From"
},
"date_to": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date To"
},
"empirical_only": {
"default": false,
"title": "Empirical Only",
"type": "boolean"
},
"has_code_only": {
"default": false,
"title": "Has Code Only",
"type": "boolean"
},
"limit": {
"default": 20,
"title": "Limit",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"required": [
"query"
],
"title": "arxiv_search_papersArguments"
}🟢arxiv_get_paper(arxiv_id, api_key)
Full paper object with all connected data. Returns: paper metadata, author list with positions, matched entities, references (up to 100), and linked GitHub repos. Args: arxiv_id: ArXiv ID, e.g. '2402.01234' or '2402.01234v2'
输入模式
{
"type": "object",
"properties": {
"arxiv_id": {
"title": "Arxiv Id",
"type": "string"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"required": [
"arxiv_id"
],
"title": "arxiv_get_paperArguments"
}🟢arxiv_top_entities(type, date_from, date_to, title_only, limit, ...)
Entity ranking by mention count across all papers. title_only=True is a powerful relevance filter: a paper mentioning MMLU in the title IS about MMLU, not just using it as one of many benchmarks. Args: type: Filter by entity type: benchmark, model, method, dataset (optional, default: all) date_from: Only count mentions in papers published from this date date_to: Only count mentions in papers published until this date title_only: Only count mentions where entity appears in the paper title limit: Max results (default: 20, max: 50)
输入模式
{
"type": "object",
"properties": {
"type": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Type"
},
"date_from": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date From"
},
"date_to": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date To"
},
"title_only": {
"default": false,
"title": "Title Only",
"type": "boolean"
},
"limit": {
"default": 20,
"title": "Limit",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"title": "arxiv_top_entitiesArguments"
}⚪arxiv_entity_trend(entity_name, granularity, api_key)
How often is an entity mentioned over time? Shows the rise (or fall) of a benchmark, model, method, or dataset across the research literature — per month, quarter, or year. Example: 'LoRA' — watch it explode in 2023-2024. Example: 'BERT' — watch it decline as LLMs dominate. Args: entity_name: Entity to track, e.g. 'LoRA', 'MMLU', 'RAG', 'GPT-4' granularity: Time grouping: month (default), quarter, year
输入模式
{
"type": "object",
"properties": {
"entity_name": {
"title": "Entity Name",
"type": "string"
},
"granularity": {
"default": "month",
"title": "Granularity",
"type": "string"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"required": [
"entity_name"
],
"title": "arxiv_entity_trendArguments"
}⚪arxiv_top_authors(role, category, date_from, date_to, limit, ...)
Top researchers ranked by paper count, with role filter. role='last_author' is the PI filter — finds lab directors and group leaders who drive research agendas. In academic AI, the last author IS the boss. role='first_author' finds the PhD students and postdocs doing the work. role='any' counts all papers regardless of position. Args: role: Author position filter: any (default), first_author, last_author category: Filter by primary category: cs.AI, cs.LG, cs.CL, cs.CV, cs.RO date_from: ISO date filter, e.g. '2024-01-01' date_to: ISO date filter, e.g. '2025-12-31' limit: Max results (default: 20, max: 50)
输入模式
{
"type": "object",
"properties": {
"role": {
"default": "any",
"title": "Role",
"type": "string"
},
"category": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Category"
},
"date_from": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date From"
},
"date_to": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date To"
},
"limit": {
"default": 20,
"title": "Limit",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"title": "arxiv_top_authorsArguments"
}⚪arxiv_author_papers(author_name, limit, api_key)
All papers by a researcher, with their position on each paper. Uses fuzzy name matching (LIKE) to handle name variations. Returns papers sorted newest first. Args: author_name: Researcher name, e.g. 'Yann LeCun', 'lecun' (partial match works) limit: Max results (default: 20, max: 50)
输入模式
{
"type": "object",
"properties": {
"author_name": {
"title": "Author Name",
"type": "string"
},
"limit": {
"default": 20,
"title": "Limit",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"required": [
"author_name"
],
"title": "arxiv_author_papersArguments"
}🟢arxiv_most_cited(category, date_from, date_to, limit, api_key)
Most cited papers — ranked by inbound citation count. This answers the question every researcher, VC, and journalist asks first: 'What are the most influential papers in AI right now?' Counts how many papers in our database cite each target paper. Only papers with resolvable ArXiv IDs in their references are counted. Args: category: Filter citing papers by category (optional) date_from: Only count citations from papers published from this date date_to: Only count citations from papers published until this date limit: Max results (default: 20, max: 50)
输入模式
{
"type": "object",
"properties": {
"category": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Category"
},
"date_from": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date From"
},
"date_to": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date To"
},
"limit": {
"default": 20,
"title": "Limit",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"title": "arxiv_most_citedArguments"
}⚪arxiv_citation_network(arxiv_id, direction, depth, api_key)
Citation graph for a paper — who cites it, or what does it cite? direction='cited_by': Papers in our database that cite this paper. direction='citing': Papers that this paper cites (its references). depth=2: Expands one hop further (depth-2 neighbors). Hard cap: 200 total. Args: arxiv_id: ArXiv paper ID, e.g. '2402.01234' direction: 'cited_by' (inbound) or 'citing' (outbound, default: cited_by) depth: Graph depth: 1 or 2 (default: 1)
输入模式
{
"type": "object",
"properties": {
"arxiv_id": {
"title": "Arxiv Id",
"type": "string"
},
"direction": {
"default": "cited_by",
"title": "Direction",
"type": "string"
},
"depth": {
"default": 1,
"title": "Depth",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"required": [
"arxiv_id"
],
"title": "arxiv_citation_networkArguments"
}⚪arxiv_co_occurrence(entity_a, entity_b, date_from, date_to, limit, ...)
Papers that mention BOTH entity A and entity B. Answers questions like: - 'Which papers use both GPT-4 and RLHF?' - 'Where do LoRA and MMLU appear together?' - 'Papers combining RAG and Chain-of-Thought?' The intersection reveals research that explicitly bridges two concepts. Args: entity_a: First entity name, e.g. 'GPT-4', 'LoRA', 'MMLU' entity_b: Second entity name, e.g. 'RLHF', 'Chain-of-Thought' date_from: ISO date filter date_to: ISO date filter limit: Max results (default: 20, max: 50)
输入模式
{
"type": "object",
"properties": {
"entity_a": {
"title": "Entity A",
"type": "string"
},
"entity_b": {
"title": "Entity B",
"type": "string"
},
"date_from": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date From"
},
"date_to": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date To"
},
"limit": {
"default": 20,
"title": "Limit",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"required": [
"entity_a",
"entity_b"
],
"title": "arxiv_co_occurrenceArguments"
}🟢arxiv_institution_ranking(date_from, date_to, include_github_orgs, limit, api_key)
Institution ranking by paper count. Primary signal: author_affiliations extracted from ArXiv HTML. Secondary signal (include_github_orgs=True): adds GitHub org counts as a complementary signal. Many papers have no affiliation in HTML but do have a GitHub org link — combining both gives a fuller picture. Note: affiliation data is extracted from HTML and may be incomplete (fetch completion is reported live by arxiv_pipeline_status; extraction quality is a separate signal). Args: date_from: ISO date filter date_to: ISO date filter include_github_orgs: Also show GitHub org ranking as second signal limit: Max results (default: 20, max: 50)
输入模式
{
"type": "object",
"properties": {
"date_from": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date From"
},
"date_to": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date To"
},
"include_github_orgs": {
"default": false,
"title": "Include Github Orgs",
"type": "boolean"
},
"limit": {
"default": 20,
"title": "Limit",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"title": "arxiv_institution_rankingArguments"
}⚪arxiv_repo_landscape(org_filter, date_from, date_to, limit, api_key)
GitHub repository landscape — which orgs and repos produce research code? Shows the open-source output of the research community. 'openai', 'google-deepmind', 'microsoft', 'huggingface' etc. ranked by how many papers link to their repos. org_filter='huggingface' shows all HuggingFace repos with papers. Args: org_filter: Filter to a specific GitHub org, e.g. 'openai', 'google-deepmind' date_from: Only papers published from this date date_to: Only papers published until this date limit: Max results per ranking (default: 20, max: 50)
输入模式
{
"type": "object",
"properties": {
"org_filter": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Org Filter"
},
"date_from": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date From"
},
"date_to": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date To"
},
"limit": {
"default": 20,
"title": "Limit",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"title": "arxiv_repo_landscapeArguments"
}🟢arxiv_tracks(api_key)
List stable ArXiv research-track objects with definitions and live counts.
输入模式
{
"type": "object",
"properties": {
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"title": "arxiv_tracksArguments"
}⚪arxiv_track_papers(track, date_from, date_to, has_code_only, limit, ...)
Chronological papers inside one stable research track. Args: track: Track slug or exact name, e.g. 'ai-agents' or 'multi-agent-systems'. date_from: Optional ISO date lower bound. date_to: Optional ISO date upper bound. has_code_only: Restrict to papers with confirmed code signal. limit: Max results (default 20, max 50).
输入模式
{
"type": "object",
"properties": {
"track": {
"title": "Track",
"type": "string"
},
"date_from": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date From"
},
"date_to": {
"anyOf": [
{
"type": "string"
},
{
"type": "null"
}
],
"default": null,
"title": "Date To"
},
"has_code_only": {
"default": false,
"title": "Has Code Only",
"type": "boolean"
},
"limit": {
"default": 20,
"title": "Limit",
"type": "integer"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"required": [
"track"
],
"title": "arxiv_track_papersArguments"
}⚪arxiv_track_trend(track, granularity, api_key)
Research volume for one stable track by month, quarter or year.
输入模式
{
"type": "object",
"properties": {
"track": {
"title": "Track",
"type": "string"
},
"granularity": {
"default": "month",
"title": "Granularity",
"type": "string"
},
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"required": [
"track"
],
"title": "arxiv_track_trendArguments"
}🟢arxiv_pipeline_status(api_key)
Full system status — database counts, pipeline progress, frontier, quality. Returns: - Paper/author/entity/ref/repo counts - Pipeline progress: html_fetched %, whitelist_matched %, llm_processed % - Frontier: how far back the backfill has reached - Quality report: last run timestamp and overall status - Quality log: last 5 quality check runs from quality_log table
输入模式
{
"type": "object",
"properties": {
"api_key": {
"default": "",
"description": "BrunoSan API Key — brunosan.de/intelligence/",
"title": "Api Key",
"type": "string"
}
},
"title": "arxiv_pipeline_statusArguments"
}🟢get_related_intelligence
Live cs.AI/ML/CL/CV/RO paper graph. FTS5 + resolved ArXiv citation links. Use for: arxiv_search_papers("LoRA fine-tuning", has_code_only=True) → Related verticals worth connecting: AI News (mcp.brunosan.de/mcp) — industry reaction to papers Robotics (robotics.mcp.brunosan.de/mcp) — applied robotics papers (cs.RO) Quantum (quantum.mcp.brunosan.de/mcp) — quant-ph research depth Biotech (biotech.mcp.brunosan.de/mcp) — bio-ML and drug discovery papers
输入模式
{
"type": "object",
"properties": {},
"title": "get_related_intelligenceArguments"
}输出模式
{
"type": "object",
"properties": {
"result": {
"title": "Result",
"type": "string"
}
},
"required": [
"result"
],
"title": "get_related_intelligenceOutput"
}社区
证据