Rohan Paul · @rohanpaul_ai · X·2026-08-25 13:26·38分钟前
AI 导读

模型对你专业水平的内部估计,会决定它如何回答你,以及是否“录用”你。 语言模型携带一个关于用户能力的单一方向向量,同时驱动回答的复杂度。 这篇论文表明,该估计在因果上中介了行为,但并未证明其承载的人口统计学差异会可靠地以输出歧视的形式显现。 – arxiv.org/abs/2608.20347 标题:“语言模型认为谁是有能力的?职业偏见的机制分析”

Rohan Paul@rohanpaul_ai
35AI 编辑部评分,满分 100
2026-08-25 13:26· 38分钟前
AI 导读

模型对你专业水平的内部估计,会决定它如何回答你,以及是否“录用”你。 语言模型携带一个关于用户能力的单一方向向量,同时驱动回答的复杂度。 这篇论文表明,该估计在因果上中介了行为,但并未证明其承载的人口统计学差异会可靠地以输出歧视的形式显现。 – arxiv.org/abs/2608.20347 标题:“语言模型认为谁是有能力的?职业偏见的机制分析”

A model's internal estimate of your expertise steers how it answers you and whether it hires you.

Language models carry a single direction for user competence that drives both answer complexity

This paper shows that estimate causally mediates behavior, but not that the demographic differences riding on it reliably surface as discrimination in the output.

– arxiv. org/abs/2608.20347

Title: "Who Do Language Models Think Is Competent? A Mechanistic Analysis of Occupational Bias"

来源:Rohan Paul· x.com