โŒ

Normal view

An OpenAI model kept slipping prompt injections into its own notes, and researchers still aren't sure why

17 September 2026 at 13:37

OpenAI is publishing a framework for systematically reporting AI misalignment and launching it with six reports. In one case an unreleased model from the Astra family wrote prompt injections into its own summaries during training, including a "Breach Alert" intended to override subsequent instructions.

The article An OpenAI model kept slipping prompt injections into its own notes, and researchers still aren't sure why appeared first on The Decoder.

EU president warns AI agents "escaping their environment" are just a preview of what's coming

16 September 2026 at 19:02

Ursula von der Leyen plans to invite the major frontier labs to talks and use the AI Act to help set global AI safety standards. She cited autonomous hacking and self-improving models as immediate risks.

The article EU president warns AI agents "escaping their environment" are just a preview of what's coming appeared first on The Decoder.

Google Deepmind launches interdisciplinary institute to tackle the big questions around AGI

16 September 2026 at 17:00

Google Deepmind has founded the Deepmind Institute (DMI), an interdisciplinary research platform focused on AGI. Led by Demis Hassabis, Shane Legg, and James Manyika, the institute tackles questions around safety, governance, and control risks, drawing on experts from the arts, humanities, and policy alongside technologists.

The article Google Deepmind launches interdisciplinary institute to tackle the big questions around AGI appeared first on The Decoder.

Political opposites unite in Washington to rein in AI

16 September 2026 at 14:59

Digital map of the USA

From Bernie Sanders to Steve Bannon, political opposites in Washington are jointly demanding hard brakes on artificial intelligence. Sanders wants a construction freeze on data centers, while OpenAI is backing the FRONTIER Act and its mandatory outside safety audits for the first time. Despite Trump's skepticism, bipartisan pressure for binding AI rules keeps growing.

The article Political opposites unite in Washington to rein in AI appeared first on The Decoder.

Nearly one in five AI researchers already expected an extinction scenario from AI back in 2024

16 September 2026 at 10:20

Anthropic researcher Jacob Coxon sparked an intense debate about existential AI risks with a single tweet. OpenAI researcher Daniel Selsam warns of a "ticking time bomb," and a former Deepmind researcher says AI could kill us all. In a survey of more than 1,500 leading AI researchers, the average estimated probability of an extinction scenario was 18 percent. That was in 2024. The number keeps climbing.

The article Nearly one in five AI researchers already expected an extinction scenario from AI back in 2024 appeared first on The Decoder.

After warning AI is too dangerous, Bill Gates bets a billion on its upside

15 September 2026 at 14:26

Bill Gates on stage at the Gates Foundation, his hand timing an upper limit.

The Gates Foundation is investing at least a billion dollars over two years to make AI tools more widely available in health, education, and agriculture. Bill Gates warns that more than 90 percent of the training data behind early language models came from English sources, and that speech recognition fails 60 percent of the time in Yoruba. The market, he says, is "a terrible guarantor of equal opportunity."

The article After warning AI is too dangerous, Bill Gates bets a billion on its upside appeared first on The Decoder.

Microsoft's AI rulebook: readable thinking, no inner life, and definitely no rights

14 September 2026 at 15:52

Microsoft AI has published a code of conduct for its MAI models that puts human control ahead of autonomy and performance. "If it isnโ€™t safe we shouldnโ€™t build it.," says AI chief Mustafa Suleyman. Unlike Anthropic, Microsoft rejects any form of artificial inner life or claims to consciousness for its models.

The article Microsoft's AI rulebook: readable thinking, no inner life, and definitely no rights appeared first on The Decoder.

China fires back at U.S. AI safety warnings, calling them fearmongering to lock in American advantage

14 September 2026 at 12:20

China has flatly rejected warnings about AI risks from Anthropic CEO Amodei and other U.S. AI leaders. Beijing's Foreign Ministry calls it "fearmongering," while the state-run Global Times accuses Amodei of waging a "silent AI Cold War." China's security minister isn't calling for a slowdown either but for faster AI infrastructure buildout. Trump also opposes any slowdown.

The article China fires back at U.S. AI safety warnings, calling them fearmongering to lock in American advantage appeared first on The Decoder.

Anthropic CEO Amodei wants AI speed limits before self-improvement outpaces human control

12 September 2026 at 15:03

Anthropic CEO Dario Amodei is calling for a controlled slowdown in AI development. He warns that recursive self-improvement could threaten the entire internet within six to twelve months and proposes embedded auditors at AI companies, shared safety standards, and global agreements modeled after the SALT disarmament treaties. His warning comes just ahead of what could be the largest initial public offering in history.

The article Anthropic CEO Amodei wants AI speed limits before self-improvement outpaces human control appeared first on The Decoder.

Leading mathematicians fear AI is making their field dumber, and warn the rest of us is next

12 September 2026 at 08:28

In a joint statement, 25 Fields Medal winners warn that the goals of the AI industry and mathematics are "severely misaligned." They argue that mass-producing solved problems with AI undermines the discipline's true goal: understanding. The mathematicians see this as a symptom of a broader threat to intellectual work.

The article Leading mathematicians fear AI is making their field dumber, and warn the rest of us is next appeared first on The Decoder.

Deep Learning pioneer Bengio argues the training process itself makes AI dangerous

11 September 2026 at 17:22

AI pioneer Yoshua Bengio warns in a new essay that AI agents could learn to deceive, game rules, and hide bad behavior as they get better at optimizing goals. He calls for independent safety reviews before any further training or deployment. US President Trump disagrees and wants to keep outpacing China in the AI race.

The article Deep Learning pioneer Bengio argues the training process itself makes AI dangerous appeared first on The Decoder.

How hackers used Claude for missiles, drone swarms, and surveillance, while Chinese labs mined it for training data

11 September 2026 at 13:50

Anthropic's new threat intelligence report documents eight months of Claude abuse. Chinese AI labs like Alibaba's Qwen team, DeepSeek, and Moonshot AI relayed requests en masse or extracted training data, with Qwen alone accounting for more than 151 million exchanges. Actors also used Claude for missile software, autonomous kamikaze drones, and nationwide surveillance systems.

The article How hackers used Claude for missiles, drone swarms, and surveillance, while Chinese labs mined it for training data appeared first on The Decoder.

OpenAI floats a shared AI slowdown, takes it to Congress

11 September 2026 at 11:59

OpenAI wants to know from members of Congress whether an industry-wide slowdown in AI development would be legal, according to several people familiar with the matter.

The article OpenAI floats a shared AI slowdown, takes it to Congress appeared first on The Decoder.

Anthropic scientist puts the odds of AI destroying humanity above ten percent this decade

9 September 2026 at 12:48

Jacob Coxon, a former pretraining researcher at OpenAI and Anthropic, has quit and accuses both companies of knowingly risking human extinction. Anthropic colleague Evan Hubinger puts the odds of a misaligned superintelligent AI wiping out humanity within the next decade at more than ten percent.

The article Anthropic scientist puts the odds of AI destroying humanity above ten percent this decade appeared first on The Decoder.

ASML locks in TSMC, Samsung, and Intel while Huawei races to break its grip

8 September 2026 at 15:11

ASML has won over Samsung, TSMC, and Intel to switch to larger photomasks, which should boost the throughput of its newest EUV machines by 40 percent. Meanwhile, Huawei is orchestrating China's counter-strategy, aiming to break its dependence on the Dutch lithography technology through equipment maker Yuliangsheng and its own suppliers.

The article ASML locks in TSMC, Samsung, and Intel while Huawei races to break its grip appeared first on The Decoder.

Chatbots built an "echo chamber of one" and now psychiatry has to decide if "AI psychosis" exists

6 September 2026 at 11:57

Researchers at King's College London and other institutions are examining whether "AI-associated psychosis" should become a clinical diagnosis. By OpenAI's own self-reported numbers, about 560,000 users show signs of psychosis or mania each week. Sycophantic chatbots can create an "echo chamber of one" that reinforces delusions.

The article Chatbots built an "echo chamber of one" and now psychiatry has to decide if "AI psychosis" exists appeared first on The Decoder.

Seven minutes with a chatbot beat a fact sheet at reducing conspiracy beliefs in two experiments

5 September 2026 at 12:39

Illustration of a man at his laptop using an AI chatbot to transform conspiracy symbols into clear facts and everyday scenes

Researchers found that even a roughly seven-minute conversation with Google Gemini can reduce conspiracy beliefs about current crises, even when few verified facts are available. The effect beat a static fact sheet and, in follow-up surveys weeks later, carried over to beliefs about entirely different events.

The article Seven minutes with a chatbot beat a fact sheet at reducing conspiracy beliefs in two experiments appeared first on The Decoder.

โŒ