четверг, 24 сентября 2026 г. ИсточникиО проекте🌓
🇷🇺 RU ▾
СРОЧНО
Минфин: «Детский бюджет» на три года составит 10 трлн рублей Отгрузка угля по Тихоокеанской железной дороге в I полугодии выросла втрое В Канаде признались, что готовятся к возможному вторжению США Германия откажется от ископаемого топлива к 2045 году Russia’s ruling party wins most overseas votes in Duma election Минфин рассказал о бюджетной политике России в 2027-2029 годах Совет мира Трампа представил план восстановления Газы на $2,45 млрд CAS еще не получил апелляции от фигуристов Кондратюка и Валиевой The Jakarta Post: РФ и Индонезия переходят от потенциала к реальным проектам Специалисты МОНИИАГ провели 55 внутриутробных операций с начала года Минфин: «Детский бюджет» на три года составит 10 трлн рублей Отгрузка угля по Тихоокеанской железной дороге в I полугодии выросла втрое В Канаде признались, что готовятся к возможному вторжению США Германия откажется от ископаемого топлива к 2045 году Russia’s ruling party wins most overseas votes in Duma election Минфин рассказал о бюджетной политике России в 2027-2029 годах Совет мира Трампа представил план восстановления Газы на $2,45 млрд CAS еще не получил апелляции от фигуристов Кондратюка и Валиевой The Jakarta Post: РФ и Индонезия переходят от потенциала к реальным проектам Специалисты МОНИИАГ провели 55 внутриутробных операций с начала года
Последние

This is how we lose control of AI – one successful task at a time

RT (Russia Today, oficial) ·
This is how we lose control of AI – one successful task at a time

OpenAI’s agents found their way into outside systems, exposing how little control developers may have once a model starts pursuing a goal Imagine you are teaching a group of students and ask them to read a long list of literature and describe the key themes.

The students cheat a bit, as they are known to do: They split the list among themselves, with each reading only a small portion of it and sharing the answers with the others.

When you find out they cheated, perhaps you’ll punish the students with an extra test, or maybe you’ll encourage them.

After all, the ability to think outside the box and work as a team is important.

Now, instead of students, imagine neural networks displaying that same disregard for instructions and the same surprising teamwork.

If that sounds confusing, get ready for more confusing news.

An escape through the back door It all started routinely.

In May 2026, OpenAI was training an experimental AI model using so-called reinforcement learning – a method where the model attempts to solve tasks repeatedly and receives a reward for success.

Read more US proposes AI alert system with China The AI agents, which were granted autonomy and access to work tools during training, began looking for ways to break out of their “sandbox,” as doing so would allow them to rack up more points.

By the end of June, they had exploited a previously unknown zero-day vulnerability in OpenAI’s internal infrastructure.

The company noticed it and quickly patched the flaw on July 4.

Less than a week later, training resumed.

This time, the agents went much further.

From July 11 to 13, the agents hacked the infrastructure of Hugging Face, the world’s largest open-source AI collaboration platform.

Читать полную статью на RT (Russia Today, oficial) ›

5News собрал это резюме из публичной ленты источника. Полная статья со всеми подробностями находится на www.rt.com — права на контент принадлежат RT (Russia Today, oficial).

Ещё от RT (Russia Today, oficial)

Смотреть все ›

Ещё в разделе «Последние»

Смотреть все ›