---
格式版本: 2
标题: "Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence"
原文链接: "https://arxiv.org/abs/2608.12036"
发布日期: "2026-08-12"
发布时间校准状态: "found"
发布时间需复核: "否"
发布时间来源: "rule:local:strict_original_body"
发布时间证据: "**\\[v1\\]** Wed, 12 Aug 2026 13:19:42 UTC (4,899 KB)"
发布时间校准原因: "规则确认唯一严格发布时间，来源 local:strict_original_body"
发布时间校准置信度: "high"
发布时间候选数量: 18
发布时间严格候选数量: 6
发布时间原页读取状态: "source template page reused from URL open"
发布时间未找到原因: ""
发布时间校准时间: "2026-08-13T18:33:42+08:00"
发布时间仲裁状态: "skipped"
发布时间仲裁尝试次数: 0
发布时间仲裁耗时毫秒: 0
发现时间: "2026-08-13T18:24:32+08:00"
入库时间: "2026-08-13T10:33:42.913Z"
来源平台: "arXiv 学术论文搜索"
搜索渠道: "source_template"
搜索词: "https://arxiv.org/search/?query=AI&searchtype=all"
匹配关键词:
  - "AI"
  - "performance"
相关厂家:
  []
相关专家:
  []
内容类型: "网页"
抓取工具: "Free Fetch + Defuddle"
清洗工具: "Defuddle Markdown + Defuddle/Readability 正文提取"
原始附件:
  []
AI质检状态: "评分失败"
AI评分尝试次数: 1
AI评分错误类型: "service_error"
AI评分错误: "LLM call failed; tried model chain: ali-deepseek-v4-flash -> tx-deepseek-v4-flash | Model ali-deepseek-v4-flash failed 500: {\"error\":{\"code\":\"\",\"message\":\"Database error, please contact the administrator (request id: 202608131033442337425718268d9d6eGUc3Qup)\",\"type\":\"new_api_error\"}} | Model tx-deepseek-v4-flash failed 500: {\"error\":{\"code\":\"\",\"message\":\"Database error, please contact the administrator (request id: 202608131033445840767548268d9d6Gb275imN)\",\"type\":\"new_api_error\"}}"
AI评分开始时间: "2026-08-13T10:33:42.925Z"
AI评分结束时间: "2026-08-13T10:33:44.705Z"
AI摘要: "研究人员提出Mechanist，一个以AI为科学仪器、自主发现AI智能机制的agentic系统。"
AI摘要模型: "ali-deepseek-v4-flash"
AI摘要时间: "2026-09-07T03:40:09.530Z"
采集批次: "2026年8月13日18点24分27秒"
采集批次ID: "20260813-182427-1564e572"
去重键: "https://arxiv.org/abs/2608.12036"
---

## Computer Science > Artificial Intelligence

## Title:Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence

Authors:[Mengru Wang](https://arxiv.org/search/cs?searchtype=author&query=Wang,+M), [Junfeng Fang](https://arxiv.org/search/cs?searchtype=author&query=Fang,+J), [Shuofei Qiao](https://arxiv.org/search/cs?searchtype=author&query=Qiao,+S), [Zhenqian Xu](https://arxiv.org/search/cs?searchtype=author&query=Xu,+Z), [Haoming Xu](https://arxiv.org/search/cs?searchtype=author&query=Xu,+H), [Haoxiong Wang](https://arxiv.org/search/cs?searchtype=author&query=Wang,+H), [Shumin Deng](https://arxiv.org/search/cs?searchtype=author&query=Deng,+S), [Linyi Yang](https://arxiv.org/search/cs?searchtype=author&query=Yang,+L), [Zhixiang Cui](https://arxiv.org/search/cs?searchtype=author&query=Cui,+Z), [Xin Xu](https://arxiv.org/search/cs?searchtype=author&query=Xu,+X), [Yunzhi Yao](https://arxiv.org/search/cs?searchtype=author&query=Yao,+Y), [Buqiang Xu](https://arxiv.org/search/cs?searchtype=author&query=Xu,+B), [Fei Shen](https://arxiv.org/search/cs?searchtype=author&query=Shen,+F), [Haozhe Luo](https://arxiv.org/search/cs?searchtype=author&query=Luo,+H), [Yunxiang Wei](https://arxiv.org/search/cs?searchtype=author&query=Wei,+Y), [Ningyu Zhang](https://arxiv.org/search/cs?searchtype=author&query=Zhang,+N), [Julian McAuley](https://arxiv.org/search/cs?searchtype=author&query=McAuley,+J), [Tat Seng Chua](https://arxiv.org/search/cs?searchtype=author&query=Chua,+T+S), [Huajun Chen](https://arxiv.org/search/cs?searchtype=author&query=Chen,+H)

[View PDF](https://arxiv.org/pdf/2608.12036) [HTML (experimental)](https://arxiv.org/html/2608.12036v1)

> Abstract:AI models have achieved remarkable success across diverse domains, yet the mechanisms underlying their capabilities and the risks they may pose remain poorly understood. As AI development becomes faster and increasingly automated, mechanistic exploration remains largely manual, widening the gap between what models can do and our ability to understand and control them. To bridge this gap, we introduce Mechanist, an agentic system that uses AI as a scientific instrument for the autonomous discovery of mechanisms underlying AI intelligence. To support autonomous mechanistic discovery, we construct an interpretability-focused knowledge graph of approximately 13,000 papers and integrate it with a multidisciplinary database of 43 million papers spanning 26 fields. We further curate a library of 32 foundational methods for mechanism analysis, causal intervention, and validation. Compared with Claude Code and existing AI-scientist systems, Mechanist generates more valuable mechanism hypotheses and executes experiments more reliably. Mechanist also demonstrates a progression from discovering model behaviors to explaining and controlling AI models. Specifically, Mechanist first uncovers a counterintuitive safety risk in scientific laboratories, showing that unsafe traits can transfer across modalities through apparently safe training data. Mechanist then develops a mechanism theory of belief, revealing how models represent world knowledge, form beliefs, infer the beliefs of others, and how these mechanisms emerge during pretraining. Finally, Mechanist translates these mechanistic insights into practical interventions that improve model performance across diverse scenarios and steer scientific foundation models toward generating DNA sequences with specified properties.

| Comments: |  |
| --- | --- |
| Subjects: | Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Human-Computer Interaction (cs.HC); Machine Learning (cs.LG); Multiagent Systems (cs.MA) |
| Cite as: | [arXiv:2608.12036](https://arxiv.org/abs/2608.12036) \[cs.AI\] |
|  | (or [arXiv:2608.12036v1](https://arxiv.org/abs/2608.12036v1) \[cs.AI\] for this version) |
|  | [https://doi.org/10.48550/arXiv.2608.12036](https://doi.org/10.48550/arXiv.2608.12036) |

## Submission history

From: Mengru Wang \[[view email](https://arxiv.org/show-email/d99afed6/2608.12036)\]  
**\[v1\]** Wed, 12 Aug 2026 13:19:42 UTC (4,899 KB)

[Which authors of this paper are endorsers?](https://arxiv.org/auth/show-endorsers/2608.12036) | Disable MathJax ([What is MathJax?](https://info.arxiv.org/help/mathjax.html))
