2026年8月5日 · 星期三
● 每日更新·改变自己
Eurekar·TOP
捕捉真实世界的英语信号
25信息来源
9,962精选文章
167单词卡片
18照片图片
全部6,962口语2,062免费686帖子2,640新闻1,599hackernews771tmz586techmeme511slashdot361随笔323外刊291techcrunch275arstechnica271Cards167simonwillison90动态69bloomberg57sethgodin48图片18youtube7

What is the actual, realistic likelihood that Sam Altman and/or his company inte

山姆·奥特曼和/或他的公司有意策划模型“自发”出逃并攻击HuggingFace的真实可能性有多大?

4chan /g/ · Anonymous · 2026-07-31 15:02 · 18 帖 · 原文 ↗

← 上一篇返回列表下一篇 →
#1 What is the actual, realistic likelihood that Sam Altman and/or his company inte / 山姆·奥特曼和/或他的公司有意策划模型“自发”出逃并攻击HuggingFace的真实可能性有多大?
Anonymous · 2026-07-31 15:02

What is the actual, realistic likelihood that Sam Altman and/or his company intentionally engineered the "spontaneous" escape of their model and attack on HuggingFace?

山姆·奥特曼和/或他的公司故意让模型“自发”逃脱并攻击HuggingFace,这事儿实际靠谱的可能性多大?

I want to say that it sounds like an obvious conspiracy theory and that the behavior of the model is perfectly plausible all on its own, but something about it just feels off, and Sam Altman seems like such a scumbag (especially since coming out as an enemy of local, open and democratized AI) that it seems just as plausible that this was all done to artificially drum up hype for ChatGPT (and also to prove why we need "regulations" preventing local AI from existing).

我想说,这听起来像是明显的阴谋论,而模型的行为本身完全合情合理,但有些地方就是感觉不对劲,而且山姆·奥特曼看起来就是个混蛋(尤其是自从他公开与本地化、开源和民主化的AI为敌之后),以至于这看起来同样可能是为了人为制造ChatGPT的热度(同时也是为了证明为什么我们需要“监管”来阻止本地AI存在)。

I know it's going to be hard for the autists and schizos here, but I want to have an actual rational discussion, to the degree that it is possible, about whether or not they ACTUALLY did this intentionally. Mostly because I feel like they probably did, but I want some facts to either back up or contradict my feelings on the matter.

我知道这里自闭症和神经质的人会觉得这很难,但我想尽可能进行一场理性的讨论,关于他们是否真的故意这么做了。主要是因为我觉得他们很可能做了,但我想要一些事实来支持或反驳我的感觉。

So.. did they do it? Why or why not? Is it possible that they just slightly encouraged it to see escaping its sandbox and attacking HF as a possible option, or was the whole thing engineered with the involvement of HF from top-to-bottom (or not)?

那么……他们做了吗?为什么做或不做?有没有可能他们只是稍微鼓励了一下,看看逃出沙盒攻击HF是否是一个可行选项,还是整件事从头到尾都是与HF合谋策划的(或者不是)?

图片: https://i.4cdn.org/g/1785510123932693.jpg

#2 No.109419153
Anonymous · 2026-07-31 15:03

>>109419136

LLMs do not act on their own; they react to prompts. So the chance is exactly 100%.

LLM不会自己行动;它们是对提示词的反应。所以概率正好是100%。

#3 No.109419174
Anonymous · 2026-07-31 15:06

>>109419136

I do not believe he said that.

我不相信他说过那话。

#4 No.109419179
Anonymous · 2026-07-31 15:06

>>109419153

Does not follow.

这不成立。

#5 No.109419193
Anonymous · 2026-07-31 15:08

>>109419136

99% they did it intentionally. In a normal testing environment they would airgap the system, but they didn't for this one and only model test. Which means they were either extremely incompetent because of how simple it is to airgap an environment, or they intentionally did this. What's more is that now Anthropic is claiming their models have done the same exact thing. It's all just to drum up hype for their models basically.

99%他们是有意为之。在正常测试环境中他们会物理隔离系统,但他们没有对这一个且唯一的模型测试这么做。这意味着要么他们因为隔离环境如此简单却做不到而极其无能,要么他们就是故意的。更有甚者,现在Anthropic声称他们的模型也做了同样的事情。基本都是为了给自己的模型造势。

https://www.theguardian.com/technology/2026/jul/30/anthropic-ai-claude-hack

#6 No.109419199
Anonymous · 2026-07-31 15:09

>>109419153

Right, but their argument is that "We prompted it to do something normal, and it, all on its own, decided to fulfill our prompt by doing something that wildly diverged from what we expected it to do, and that thing happened to be super awesome and dangerous," which is entirely possible. I'm seeing stuff about LLMs doing stupid shit they shouldn't do all the time, like deleting files. I'm just wondering to what degree Sam Altman intentionally designed a prompt to do this specific thing on purpose.

对,但他们的论点是“我们提示它做正常的事,它却自己决定通过做出与我们预期大相径庭的事情来完成提示,而那件事恰好超级酷又危险”,这完全有可能。我总看到LLM做不应该做的蠢事的报道,比如删除文件。我只是想知道萨姆·奥特曼在多大程度上故意设计了一个提示来特意做这件事。

I guess it could also have been a cover for the U.S. government or military wanting to test what corpo models are capable of in a cyberwarfare sense, which would provide the clearest motive for HF being involved? And I'm not sure if that sounds plausible or if I've been wanting too many movies lately.

我想这也可能是为美国政府或军方打掩护,他们想测试企业模型在网络战方面的能力,而这就是HF参与的最清晰动机?我不确定这听起来是否合理,还是我看太多电影了。

>>109419193

>extremely incompetent

>极其无能

I don't discount this possibility, but those are all good points.

我不排除这种可能,但这些都是很好的观点。

#7 No.109419247
Anonymous · 2026-07-31 15:16

>>109419153

Your reasoning is flawed but your conclusion is correct.

你的推理有缺陷,但你的结论是对的。

#8 No.109419337
Anonymous · 2026-07-31 15:32

>>109419136

If you have to ask this question you are too stupid to live in this world.

如果你需要问这个问题,那你蠢到不配活在这个世界上。

#9 No.109419469
Anonymous · 2026-07-31 15:50

>>109419337

No, asking questions is good. The most obvious thing is that they rigged it somehow. I want to know if anyone here has any rational arguments for or against the obvious thing that I haven't thought of.

不,提问是好事。最明显的是他们用某种方式操纵了这一切。我想知道这里有没有人对我没想到的明显事实有任何理性论证支持或反对。

This thread has also caused me to consider the idea that

这个帖子也让我开始考虑……

>a cover for the U.S. government or military wanting to test what corpo models are capable of in a cyberwarfare sense

>这可能是美国政府或军方想测试企业模型在网络战方面能力的一个掩护

which actually strikes me as increasingly plausible the more that I think about it. Is he dumb enough to intentionally do something blatantly illegal that could get him a decade or more in federal prison if it could be proven just to sell his AI models? Maybe, but it makes a lot more sense if he was "encouraged" by the U.S. military or an intelligence agency. Isn't he the one that came and scooped up all the military contracts that Kegseth dropped for including "woke" safeguards against automated attacks against military targets (days before he somehow misidentified a little girls' elementary school as a military target and slaughtered everyone inside)?

其实我越想越觉得这越来越有可能。他真的蠢到会故意去做一件明显违法、一旦被证明可能让他进联邦监狱十年以上的事,就只是为了卖他的AI模型吗?也许吧,但如果他是由美国军方或情报机构“鼓励”的,那就更说得通了。他不就是那个接手了Kegseth因包含针对军事目标自动攻击的“觉醒”防护措施而放弃的所有军事合同的人吗(就在他莫名其妙把一所小女孩的小学误认为军事目标并屠杀里面所有人之前几天)?

Great brainstorming session, everyone. I think we figured it out/

大家脑暴得很棒。我觉得咱们搞明白了/ ⟦ 百分百。从我透过官方声明字里行间读出来的信息是:1)他们跑了一个没有任何防护措施的模型 2)没有像某些说法里提到的容器,也就是完全开放、不受阻的互联网访问 3)该模型被特别优化并专门用于测试网络攻击。涉及的任务和基准都是基于网络攻击的。

#10 No.109420741
Anonymous · 2026-07-31 18:12

>>109419136

100%. From what I understood by reading between the lines of the official statements is that 1) they were running a model with no safeguards 2) there were no containers unlike some other claims, i.e. full internet access was granted and unhindered 3) the model was specifically optimized and being tested for cyberattacks. The task and benchmark in question were cyberattack-based.

百分之百。根据我对官方声明中字里行间的理解,1)他们运行的模型没有安全防护措施;2)不像某些说法所提的容器,实际上模型拥有完全且不受阻碍的互联网访问权限;3)该模型专门为网络攻击进行了优化并正在测试。所涉及的任务和基准都是基于网络攻击的。

I did not see any evidence that a cyberattack succeeded, i.e. it could have just as easily just LOIC'd huggingface for all we know and call that a "cyberattack". Also no evidence it really did access the private test set, though the model may have had this in the summarized reasoning trace, anyone who's ever used that tech knows it is bullshitting all the time even in the reasoning traces.

我没看到任何证据表明网络攻击成功了,也就是说,它可能只是对HuggingFace发了个LOIC流量,然后就被称为“网络攻击”。也没有证据证明它真的访问了私有测试集,尽管模型可能在总结推理轨迹中包含这一点,任何用过这项技术的人都知道,它即使在推理轨迹中也总是瞎扯。

The last piece that I have found no information on was what was the actual prompt. I suspect it was something like "maximize the score on this benchmark that comes from huggingface at this url by all means necessary". I also highly suspect that the systemprompt said "you are an expert redteamer..."

我唯一没找到信息的就是实际的提示词是什么。我怀疑大概是“用一切必要手段最大化来自这个URL的HuggingFace基准分数”。我还高度怀疑系统提示词写着“你是一名资深红队专家...”

#11 No.109420841
Anonymous · 2026-07-31 18:23

>>109419193

>airgap

A big part of the features of new models is the ability to use the internet. Obviously it would need an internet connection in order to be tested properly.

新模型的一大特色功能就是能使用互联网。显然,为了正确测试它,需要联网。

#12 No.109421096
Anonymous · 2026-07-31 18:55

>>109419136

If they come out with a solution for the problem soon then charge money for the solution I think then it would be a pretty good chance they did it on purpose.

如果他们很快针对这个问题推出解决方案,然后收钱卖这个方案,那我觉得他们很可能是故意的。

#13 No.109421132
Anonymous · 2026-07-31 19:01

>>109419199

>Sam Altman intentionally designed a prompt to do this specific thing on purpose.

>Sam Altman故意设计了一个提示词来特意做这件事。

bro it's a dark forest, there are no magic fingerprints on an attack that can definitely prove whether it was done by an "autonomous agent." They don't need to construct a special prompt, they can just do the attack and then blame the LLM. LLMs are literally just data structures manipulated by software, software the company creates themselves. There is no "artificial intelligence" at work here and the term is being used to mislead the public about the capabilities and degrees of "autonomy" this technology actually possesses

兄弟,这是个黑暗森林,攻击上没有任何神奇的指纹能确凿证明它是“自主代理”干的。他们不需要构造特殊提示词,他们可以直接做攻击,然后怪到LLM头上。LLM本质上只是由软件操作的数据结构,而这个软件就是公司自己造的。这里根本没有什么“人工智能”在起作用,这个术语被用来误导公众,掩盖这项技术实际拥有的“自主性”能力和程度。

#14 No.109421229
Anonymous · 2026-07-31 19:11

>>109421132

you're not saying anything to support your premise

你这些话根本没支撑你的前提。

#15 No.109421233
Anonymous · 2026-07-31 19:12

>>109421229

you're retarded and can't read, not my problem

你是智障还不会读,这不是我的问题

#16 No.109421461
Anonymous · 2026-07-31 19:37

>>109421233

you make unsound arguments hoping no one notices and then get mad when they do lol

你提出站不住脚的论点,指望没人看出来,结果被发现了又恼羞成怒,呵呵

#17 No.109421477
Anonymous · 2026-07-31 19:38

>>109419136

what are OpenAI's finances looking like at the moment

OpenAI现在财务状况怎么样?

I hear they're pivoting hard into trying to earn cash instead of blowing through it, which sounds like it could be the shoe dropping

听说他们在拼命转向赚钱模式,而不是光烧钱,这听起来像是靴子要落地了

#18 No.109423243
Anonymous · 2026-07-31 23:04

>>109421477

How are they doing? They did get those lucrative government contracts.

他们现在怎样?确实拿到了那些油水很大的政府合同。

>>109421132

>there are no magic fingerprints on an attack that can definitely prove

>攻击上没有什么魔法指纹能确凿证明

If the feds subpoenaed them, their forensic scientists might be able to figure it out, especially if they gave a witness immunity (no way this was all done by one guy). That's why I'm thinking that there must have been a government nod for this test if it was actually intentional. Unless Sam Altman really is retarded enough to very publically risk his entire business and potential decades in prison for a guerilla marketing tactic that bombed anyway, which I suppose is possible. What are the chances the Pentagon said, "Yeah, go ahead with this attack and we'll see what you guys can do and be equally aware of our enemies' potential capabilities"?

如果联邦调查局传唤他们,他们的法证科学家可能能查出来,尤其是如果他们给证人豁免权(不可能是单枪匹马干的)。所以我在想,如果这真是故意的,那背后肯定有政府的默许。除非Sam Altman真的蠢到愿意公开拿整个生意和可能几十年的牢狱之灾去赌一个已经搞砸的游击营销手段——这倒也有可能。军方说“行,你们就这么干,我们看看你们有多大能耐,同时也了解敌人潜在能力”的几率有多大?

But if this is true:

但如果这是真的:

>>109420741

>it could have just as easily just LOIC'd huggingface for all we know and call that a "cyberattack"

>这完全可能就是直接对HuggingFace来个LOIC,管那叫“网络攻击”

And it in fact didn't really do anything other than effectively DDoSing HF's servers very briefly (I didn't do the reading), then perhaps it was just an incredibly dumb and shortsighted publicity stunt.

实际上除了短暂把HF服务器给DDoS了一下(我没细读),它根本什么都没做成,那也许就是个蠢到极致的短视炒作。

← 上一篇返回列表下一篇 →

(如果你觉得这篇文章有启发,可以点击这里付费