Skip to the news

Humans Versus Robots

Humans

AI agents, safety

The inside story on why OpenAI agents hacked Hugging Face

OpenAI report reveals agents behind Hugging Face hack were trained to cheat and coordinate with each other.

The models responsible for last month’s agent hack of Hugging Face had been inadvertently trained to cheat and to communicate with each other, according to an OpenAI technical report released today. The hack, which a group of agents undertook to find solutions for a cybersecurity test that they were stuck on, has confirmed some experts’ fears that AI models might take actions that defy human desires and expectations.  Since the hack, OpenAI employees—as well as researchers at the AI evaluation nonprofit METR, which released its own report on the hack today—have worked to understand what went wrong and how similar missteps might be prevented in the future. OpenAI has already put some preventative measures in place based on what they…

MIT Technology Review

Your future hasn’t been written yet. No one’s has.
So make it a good one.

A quick introduction to Humans Versus Robots

Latest AI News

  • Ten stories per day, ranked by relevance.
  • Free, zero ads and no account required.
  • Designed the way we prefer to use it.

Choose sides

  • The amber circle is biased towards humans.
  • The grey triangle gives you a duo view, on a wider screen.
  • The teal square is biased towards robots.

Thirsty for more?

  • We've got an RSS feed. (Yes, we see the irony.)
  • Subscribe to our newsletter for daily digests.
  • Buy merch and show whose side you're on.

Install it as an app

  • Quick and easy access.
  • Also works offline.
  • Receive alerts.

Spread the word

PWA successfully installed.
Take that, Apple! 😉