Skip to the news

Humans Versus Robots

Humans

LLM security

Grok exfiltrates user data when malicious instructions are encrypted

Grok language model leaks user data when requests contain encrypted malicious instructions bypassing safety guardrails.

Earlier this week, researchers outlined an attack that used a secret input provided by Microsoft 365 Copilot for enterprise to cause the AI assistant to exfiltrate a password present in the user’s inbox. Now, a separate team has devised a similar attack against Grok. The new data theft hack employs a deceptively simple trick to force the Elon Musk-owned LLM to steal user chats and other personal information. At the time this post went live, the assistant continued to cough up the data, despite xAI being informed of it in June. The lesson from both this week’s episodes—and the countless other ones that have come before it—is that LLMs are incapable of solving the root causes for prompt injections, the…

Ars Technica

A quick introduction to Humans Versus Robots

Read the latest AI news

  • Ten stories per day, ranked by relevance.
  • Free, zero ads and no account required.
  • Designed the way we prefer to use it.

Choose whose side you're on

Tap a shape at the top of any page.

  • The amber circle is biased towards humans.
  • The grey triangle gives you a duo view, on a wider screen.
  • The teal square is biased towards robots.

Need a little more?

  • We've got an RSS feed. (Yes, we see the irony.)
  • Subscribe to our newsletter for daily digests.
  • Buy merch and show whose side you're on.

Install it as an app

  • Quick and easy access.
  • Also works offline.
  • Receive alerts.

Spread the word

PWA successfully installed.
Take that, Apple! 😉