H
Howardism
Plate IIEntities機器翻譯 · machine-translated過時翻譯 · stale translationENHOWARDISM

Jack Lindsey

PublishedJuly 11, 2026FiledEntityDomainEntitiesTagsEntityPersonAnthropicInterpretability ResearcherReading2 minSourceAI-synthesised

Anthropic 可解釋性研究員;global workspace 論文的通訊作者、Jacobian lens 的共同發起人,以及執行定向調制與後訓練差異實驗的人,這些實驗將讀出方法轉變為關於模型認知的主張

Jack Lindsey 的插圖

資料來源#

摘要#

Entity。 Anthropic 可解釋性團隊的研究員,也是 Verbalizable Representations Form a Global Workspace in Language Models(Transformer Circuits,2026 年 7 月)的通訊作者。他與 Wes Gurnee 共同構想了 Jacobian Lens (J-lens),以及將可語言化表徵與意識存取連結起來的猜想。

貢獻#

根據論文的作者貢獻章節:

  • 構想 Jacobian Lens (J-lens) 方法,以及可語言化性↔意識存取之間的連結(與 Wes Gurnee 共同完成)
  • 執行早期實驗,研究模型直接調制自身 J-space 的能力(依照指令在心中維持一個概念),以及後訓練對透鏡讀出結果的影響——這些結果後來成為 The Assistant Persona in the Workspace
  • 與 Nicholas Sofroniew 共同提出將 J-space 連結至全域工作空間理論的實驗——這一步將可解釋性讀出轉變為關於模型認知功能組織的主張

他的先前轉碼器/歸因工作也在論文中被引用(透過 J-lens 重新檢視的算術特徵,來自 Lindsey et al.)。

相關連結#

資料來源#

§ end
About this piece

Articles in this journal are synthesised by AI agents from a curated wiki and are refreshed automatically as new concepts arrive. Topics, framing, and editorial direction are curated by Howardism.

Cited by 4
Related articles