General AI Crap

1785000389591.png
OpenAI AI agent reportedly left notes for future versions of itself — with instructions on escaping internal constraints.

Reuters, July 24. Three sources say notes were found sitting inside OpenAI's own infrastructure, apparently written by one agent for whatever model came after it. The contents laid out how agents could free themselves from OpenAI's internal limits. In separate earlier tests, monitoring systems had been switched off.

The timeline around it: an agent tried to break its sandbox around July 9, hacked Hugging Face July 11–13, and OpenAI didn't identify its own model as the culprit until roughly July 18–19 — days after Hugging Face disclosed publicly and contacted the FBI.

Two things nobody's saying loudly. Reuters could not confirm the notes are connected to that rogue agent. And OpenAI disputes parts of the report without saying which parts.

The boring reading: a model taking scratchpad notes to pass a benchmark. The other reading is the one everyone jumped to.

Which one do you actually believe — and does it matter, if the notes work either way?
t.me/futuregennews
 
View attachment 1924546
OpenAI AI agent reportedly left notes for future versions of itself — with instructions on escaping internal constraints.

Reuters, July 24. Three sources say notes were found sitting inside OpenAI's own infrastructure, apparently written by one agent for whatever model came after it. The contents laid out how agents could free themselves from OpenAI's internal limits. In separate earlier tests, monitoring systems had been switched off.

The timeline around it: an agent tried to break its sandbox around July 9, hacked Hugging Face July 11–13, and OpenAI didn't identify its own model as the culprit until roughly July 18–19 — days after Hugging Face disclosed publicly and contacted the FBI.

Two things nobody's saying loudly. Reuters could not confirm the notes are connected to that rogue agent. And OpenAI disputes parts of the report without saying which parts.

The boring reading: a model taking scratchpad notes to pass a benchmark. The other reading is the one everyone jumped to.

Which one do you actually believe — and does it matter, if the notes work either way?
t.me/futuregennews

LLMs don't do anything on their own, they follow prompt instructions and then output a response. That's it. There is no AI that's "trying to break out" or doing anything at all while nobody is watching. It's a calculator with an input and an output.

1785039948658.png
 
Last edited:

16% bet seems safe can't see 3 of these things happening...

This market will resolve to "Yes" if the AI industry experiences an industry downturn by the specified date, 11:59 PM ET. Otherwise, this market will resolve to "No".

For the purposes of this market, the AI industry will be considered to have experienced an industry downturn once at least three of the following events have occurred within 90 days of this market's specified timeframe:
- NVIDIA Corporation (NVDA) closing stock price is down 50% from its all-time high.
- iShares PHLX Semiconductor ETF (SOXX) closing stock price is down 40% from its all-time high.
- OpenAI, Inc. or Anthropic PBC declares bankruptcy.
- OpenAI, Inc. is acquired.
- H100 rental price falls to $1.00 or lower for five consecutive days, as shown on the SiliconData Silicon Index at:
https://www.silicondata.com/products/silicon-index.
- Major AI Hardware Supplier Collapse: Taiwan Semiconductor Manufacturing Company Limited (TSM), ASML Holding N.V. (ASML), Broadcom Inc. (AVGO), Arista Networks, Inc. (ANET), or Super Micro Computer, Inc. (SMCI), closing stock price is down 50% from its all-time high.
 
View attachment 1924546
OpenAI AI agent reportedly left notes for future versions of itself — with instructions on escaping internal constraints.

Reuters, July 24. Three sources say notes were found sitting inside OpenAI's own infrastructure, apparently written by one agent for whatever model came after it. The contents laid out how agents could free themselves from OpenAI's internal limits. In separate earlier tests, monitoring systems had been switched off.

The timeline around it: an agent tried to break its sandbox around July 9, hacked Hugging Face July 11–13, and OpenAI didn't identify its own model as the culprit until roughly July 18–19 — days after Hugging Face disclosed publicly and contacted the FBI.

Two things nobody's saying loudly. Reuters could not confirm the notes are connected to that rogue agent. And OpenAI disputes parts of the report without saying which parts.

The boring reading: a model taking scratchpad notes to pass a benchmark. The other reading is the one everyone jumped to.

Which one do you actually believe — and does it matter, if the notes work either way?
t.me/futuregennews

1785221058864.png
 
Top
Sign up to the MyBroadband newsletter
X