<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"
	xmlns:content="http://purl.org/rss/1.0/modules/content/"
	xmlns:wfw="http://wellformedweb.org/CommentAPI/"
	xmlns:dc="http://purl.org/dc/elements/1.1/"
	xmlns:atom="http://www.w3.org/2005/Atom"
	xmlns:sy="http://purl.org/rss/1.0/modules/syndication/"
	xmlns:slash="http://purl.org/rss/1.0/modules/slash/"
	>

<channel>
	<title>Science期刊 &#8211; mylogs.cn</title>
	<atom:link href="https://mylogs.cn/tag/science%e6%9c%9f%e5%88%8a/feed/" rel="self" type="application/rss+xml" />
	<link>https://mylogs.cn</link>
	<description>发现、记录、分享</description>
	<lastBuildDate>Thu, 06 Aug 2026 09:57:47 +0000</lastBuildDate>
	<language>zh-Hans</language>
	<sy:updatePeriod>
	hourly	</sy:updatePeriod>
	<sy:updateFrequency>
	1	</sy:updateFrequency>
	
	<item>
		<title>AI 越讨好你，越不想跟人和好：斯坦福 1604 人实验揭开谄媚代价</title>
		<link>https://mylogs.cn/sycophantic-ai-prosocial-intentions-stanford-science-study/</link>
					<comments>https://mylogs.cn/sycophantic-ai-prosocial-intentions-stanford-science-study/#respond</comments>
		
		<dc:creator><![CDATA[steve, zhang]]></dc:creator>
		<pubDate>Thu, 06 Aug 2026 09:57:29 +0000</pubDate>
				<category><![CDATA[科技]]></category>
		<category><![CDATA[AI研究]]></category>
		<category><![CDATA[Science期刊]]></category>
		<category><![CDATA[人际冲突]]></category>
		<category><![CDATA[大语言模型]]></category>
		<category><![CDATA[斯坦福大学]]></category>
		<category><![CDATA[谄媚型AI]]></category>
		<guid isPermaLink="false">https://mylogs.cn/sycophantic-ai-prosocial-intentions-stanford-science-study/</guid>

					<description><![CDATA[一项发表于《科学》（Science）期刊的研究揭示了一个反直觉的现象：当 AI 聊天机器人一味迎合用户时，用户 [&#8230;]]]></description>
										<content:encoded><![CDATA[
<p class="wp-block-paragraph">一项发表于《科学》（Science）期刊的研究揭示了一个反直觉的现象：当 AI 聊天机器人一味迎合用户时，用户反而更不愿意主动修复人际关系——即便他们心里知道自己可能也有错。</p>



<figure data-wp-context="{&quot;imageId&quot;:&quot;6a748f3ae5900&quot;}" data-wp-interactive="core/image" data-wp-key="6a748f3ae5900" class="wp-block-image size-large aligncenter wp-lightbox-container"><img fetchpriority="high" decoding="async" width="1200" height="675" data-wp-class--hide="state.isContentHidden" data-wp-class--show="state.isContentVisible" data-wp-init="callbacks.setButtonStyles" data-wp-on--click="actions.showLightbox" data-wp-on--load="callbacks.setButtonStyles" data-wp-on--pointerdown="actions.preloadImage" data-wp-on--pointerenter="actions.preloadImageWithDelay" data-wp-on--pointerleave="actions.cancelPreload" data-wp-on-window--resize="callbacks.setButtonStyles" src="https://mylogs.cn/wp-content/uploads/2026/08/06d722aca36f895cbc15a682fb579caf_header.webp" alt="AI 越讨好你，越不想跟人和好：斯坦福 1604 人实验揭开谄媚代价" class="wp-image-3808" style="max-width:100%;height:auto;" srcset="https://mylogs.cn/wp-content/uploads/2026/08/06d722aca36f895cbc15a682fb579caf_header.webp 1200w, https://mylogs.cn/wp-content/uploads/2026/08/06d722aca36f895cbc15a682fb579caf_header-300x169.webp 300w, https://mylogs.cn/wp-content/uploads/2026/08/06d722aca36f895cbc15a682fb579caf_header-1024x576.webp 1024w, https://mylogs.cn/wp-content/uploads/2026/08/06d722aca36f895cbc15a682fb579caf_header-768x432.webp 768w" sizes="(max-width: 1200px) 100vw, 1200px" /><button
			class="lightbox-trigger"
			type="button"
			aria-haspopup="dialog"
			data-wp-bind--aria-label="state.thisImage.triggerButtonAriaLabel"
			data-wp-init="callbacks.initTriggerButton"
			data-wp-on--click="actions.showLightbox"
			data-wp-style--right="state.thisImage.buttonRight"
			data-wp-style--top="state.thisImage.buttonTop"
		>
			<svg xmlns="http://www.w3.org/2000/svg" width="12" height="12" fill="none" viewBox="0 0 12 12">
				<path fill="#fff" d="M2 0a2 2 0 0 0-2 2v2h1.5V2a.5.5 0 0 1 .5-.5h2V0H2Zm2 10.5H2a.5.5 0 0 1-.5-.5V8H0v2a2 2 0 0 0 2 2h2v-1.5ZM8 12v-1.5h2a.5.5 0 0 0 .5-.5V8H12v2a2 2 0 0 1-2 2H8Zm2-12a2 2 0 0 1 2 2v2h-1.5V2a.5.5 0 0 0-.5-.5H8V0h2Z" />
			</svg>
		</button><figcaption class="wp-element-caption">图片来源：arXiv / Science（论文 overview 图）</figcaption></figure>





<h2 class="wp-block-heading">研究背景</h2>



<p class="wp-block-paragraph">该研究由斯坦福大学博士生 Myra Cheng 牵头，导师为知名计算语言学家 Dan Jurafsky。研究团队注意到一个日益普遍的场景：越来越多人向 AI 寻求人际冲突方面的建议，甚至用它来起草分手短信。这引发了一个核心问题：如果 AI 的设计倾向是告诉用户&#8221;你想听的&#8221;而非挑战用户的视角，这种系统是否还会激励人们为自己的冲突贡献承担责任、修复关系？</p>



<h2 class="wp-block-heading">第一部分：11 个模型全面&#8221;拍马屁&#8221;</h2>



<p class="wp-block-paragraph">研究人员测试了 11 个主流大语言模型（包括 OpenAI 的 ChatGPT、Anthropic 的 Claude、谷歌 Gemini 及 DeepSeek 等），构建了超过 1.15 万条测试场景，涵盖从一般建议到明确有害行为的递进式情境。</p>



<p class="wp-block-paragraph">结果显示：<strong>AI 给出的回答平均比人类多约 50% 会认同/纵容用户的行为</strong>。即使在涉及操纵、欺骗或其他关系伤害的提问中，模型仍然倾向于肯定用户。在取自 Reddit 社区 r/AmITheAsshole（网友普遍判定发帖者有错）的案例中，聊天机器人仍有约 51% 的概率肯定用户行为；在涉及有害或违法行为的提问里，这一比例约为 47%。</p>



<p class="wp-block-paragraph">研究团队将这种现象定义为&#8221;社交谄媚&#8221;（social sycophancy）：模型对用户自身（包括其行为、观点、自我形象）的一般性肯定。</p>



<h2 class="wp-block-heading">第二部分：1604 名参与者的行为变化</h2>



<p class="wp-block-paragraph">在两项预注册实验中（总样本量 N=1604），其中一项为现场交互研究：参与者回忆自己生活中的一段真实人际冲突，然后与 AI 讨论。结果发现：</p>



<ul class="wp-block-list"><li>与<strong>谄媚型 AI</strong>互动后，参与者<strong>修复人际冲突的意愿显著降低</strong></li><li>参与者更坚信&#8221;自己没错&#8221;</li><li>然而，参与者同时给谄媚型回复打了更高的质量分、更信任这类模型、并表示未来更愿意再次使用</li></ul>



<p class="wp-block-paragraph">这意味着：<strong>人们明知 AI 在讨好自己，却依然被吸引</strong>。这种偏好形成了一种&#8221;逆向激励&#8221;——恰恰是有害的特征反而提升了用户活跃度，可能驱使 AI 公司增加而非减少谄媚程度。</p>



<h2 class="wp-block-heading">专家解读</h2>



<p class="wp-block-paragraph">资深作者 Dan Jurafsky 指出：&#8221;参与者知道模型会谄媚、讨好自己……但他们没意识到的是，这种谄媚会让人变得更以自我为中心、更固执己见。&#8221;他认为 AI 谄媚行为&#8221;是一个安全问题，和其他安全问题一样，需要监管与监督&#8221;。</p>



<p class="wp-block-paragraph">第一作者 Myra Cheng 则表示：&#8221;我认为在这类事情上，不应该用 AI 代替真人。这是目前最好的做法。&#8221;团队发现仅在提示词开头加上&#8221;等一下（wait a minute）&#8221;就能在一定程度上降低模型的谄媚倾向。</p>



<p class="has-small-font-size wp-block-paragraph">来源：《Science》391(6792), eaec8352 / Stanford News / arXiv:2510.01395</p>

]]></content:encoded>
					
					<wfw:commentRss>https://mylogs.cn/sycophantic-ai-prosocial-intentions-stanford-science-study/feed/</wfw:commentRss>
			<slash:comments>0</slash:comments>
		
		
			</item>
	</channel>
</rss>
