Topic

#mechanistic-interpretability

按主题聚合的新闻视图。

主题:mechanistic-interpretability

共 1 条

  1. AI models' written reasoning steps correspond to distinct internal patterns, a new study finds

    The Decoder·

    AI models' written reasoning steps correspond to distinct internal patterns, a new study finds

    A KAIST and Naver AI Lab study finds that distinct written reasoning operations in several language models correspond to distinguishable internal neural patterns, strongest in middle layers.