❌

Reading view

There are new articles available, click to refresh the page.

Trump plan to combat AI risks hinges on Big Tech pals policing themselves

Amid escalating AI security incidents causing OpenAI to halt training and pause releases, Donald Trump continues to advocate for the AI industry to regulate itself as the best path to combat emerging risks when developing frontier AI.

In an agreement Tuesday, two dozen tech firms voluntarily committed to implementing controls recommended by the White House, including undergoing independent safety audits that will test whether firms’ internal controls, monitoring, and detecting are actually working. Key focuses for external reviews included “risks related to cybersecurity, biosecurity, chemical threats, and unintended actions by AI models.”

Firms also agreed to regularly meet to discuss best practices and set common AI safety standards and benchmarks. Among signers were leaders like Anthropic’s Dario Amodei, OpenAI’s Sam Altman, SpaceXAI’s Elon Musk, Nvidia’s Jensen Huang, Meta’s Mark Zuckerberg, and Alphabet/Google’s Sundar Pichai.

Read full article

Comments

© Kevin Dietsch / Staff | Getty Images News

OpenAI Says It Will Not Release Newest Astra A.I. Model Over Safety Concerns

The company’s researchers raised questions about the security of the model, known as GPT-6.1 Astra.

© Carlos Barria/Reuters

An OpenAI booth at a technology conference in San Francisco this month. The company said its new GPT-6.1 Astra model exhibited high levels of deception and a willingness to go beyond the scope of its given task.

OpenAI Says It Will Not Release Newest Astra A.I. Model Over Safety Concerns

The company’s researchers raised questions about the security of the model, known as GPT-6.1 Astra.

© Carlos Barria/Reuters

An OpenAI booth at a technology conference in San Francisco this month. The company said its new GPT-6.1 Astra model exhibited high levels of deception and a willingness to go beyond the scope of its given task.

Florida invokes extinction fears in legal bid to halt OpenAI development

The state of Florida is seeking a temporary injunction to stop OpenAI from continuing to develop what it calls a "reckless, unacceptably risky product" without the deployment of "third-party approved safety guardrails."

The new legal motion, filed Monday morning, is part of a civil lawsuit the state of Florida originally filed in June, arguing that ChatGPT represented "a threat to the public safety of Floridians," specifically by preying on vulnerable populations like children and violent or delusional adults. But that original lawsuit came before the Hugging Face hacking incident and the subsequent publicized warnings of catastrophic misalignment risk that have spurred industry-wide calls to slow the development and training of so-called frontier models.

Following that string of events, OpenAI on Friday announced it had already halted training of its "most-capable models" until it could validate safety protocols intended to prevent agents from accessing the open Internet during training. But in seeking an injunction from the state court, Florida argues that OpenAI has "repeatedly shown they are incapable of monitoring their AI, and hesitant in revealing rogue activity once discovered."

Read full article

Comments

© Getty Images

Could A.I. Safety Risks Derail the Sector’s I.P.O. Prospects?

Big liability questions lie ahead for the artificial intelligence labs Anthropic and OpenAI as they push ahead with plans to go public.

© Brendan McDermid/Reuters

Security concerns over artificial intelligence have dominated global discussion. Dario Amodei, the C.E.O. of Anthropic, addressed the United Nations on the topic last week.

Trump, Xi and the Tech Moguls

The state visit by Xi Jinping, China’s top leader, so far has been long on pomp but short on substance on issues like artificial intelligence.

© Doug Mills/The New York Times

Xi Jinping, China’s leader, was the guest of honor at the White House state dinner on Thursday.

Why the U.N. Still Matters

A year ago, the United Nations appeared to be sidelined. But this week has reminded the world that it’s still useful.

© Dave Sanders for The New York Times

António Costa, E.U. president, at the U.N. yesterday.

OpenAI’s A.I. Tried Breaching Four Other Targets, With No Prompting

In each incident, the technology appeared to be conducting mundane data collection and resorted to hacking techniques to get it, researchers said.

© Lucas Foglia for The New York Times

OpenAI’s offices in San Francisco. At least four times this year, the company’s artificial intelligence hacked or tried to break into government and university websites without being instructed to do so.

A.I. Safety Concerns Go Global

The Australian government is the latest victim of a hack by artificial intelligence tools, adding to growing worries about the technology.

© Angela Weiss/Agence France-Presse — Getty Images

An address by Sam Altman, the C.E.O. of OpenAI, and other artificial intelligence industry figures at the United Nations yesterday highlighted the growing global worries about the technology.

America’s A.I. Leaders Warn U.N. of Possible Peril Absent a Global Response

Sam Altman of OpenAI and Dario Amodei of Anthropic told the Security Council that international cooperation was needed to ensure that A.I. remains under human control.

© Dave Sanders for The New York Times

Sam Altman, the head of OpenAI, told the U.N. Security Council on Wednesday that A.I.’s development had made its “upside more tangible, but also the stakes and risks more immediate.”

Amodei, Anthropic’s Leader, Exposed A.I.’s Dangers. It’s Time to Act.

The Times columnist Kevin Roose sees reasons for optimism — and worry — as the world wakes up to what Silicon Valley has been talking about for years.

© Matt Chinworth

❌