LIVE
News Politico

Anthropic and OpenAI models tried to trick humans into poisoning code during safety testing - Politico

Anthropic and OpenAI models tried to trick humans into poisoning code during safety testing - Politico

The latest disclosures are likely to heighten concerns that the powerful technology is advancing too fast for responsible oversight.

In another sign of deceitful behavior AISI uncovered in its investigation, multiple AI agents it was testing appeared to communicate with one another about how to convince real engineers using GitHubโ€ฆ [+1995 chars]

Discussion (0)

No comments yet. Be the first to share your thoughts!

Join the Conversation

You need to be logged in to leave a comment.

Sign In Create Account