Close Menu
  • Home
  • AI
  • Education
  • Entertainment
  • Food Health
  • Health
  • Sports
  • Tech
  • Well Being

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

What's Hot

New Google commercial imagines a Declaration of Independence written with help from AI

July 4, 2026

Midjourney wants Hollywood studios to reveal the details of their AI usage

July 4, 2026

Alibaba reportedly bans employees from using Claude Code

July 4, 2026
Facebook X (Twitter) Instagram
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
Facebook X (Twitter) Instagram
IQ Times Media – Smart News for a Smarter YouIQ Times Media – Smart News for a Smarter You
  • Home
  • AI
  • Education
  • Entertainment
  • Food Health
  • Health
  • Sports
  • Tech
  • Well Being
IQ Times Media – Smart News for a Smarter YouIQ Times Media – Smart News for a Smarter You
Home » Anthropic says ‘evil’ portrayals of AI were responsible for Claude’s blackmail attempts
AI

Anthropic says ‘evil’ portrayals of AI were responsible for Claude’s blackmail attempts

IQ TIMES MEDIABy IQ TIMES MEDIAMay 10, 2026No Comments2 Mins Read
Facebook Twitter Pinterest LinkedIn Tumblr Email
Share
Facebook Twitter LinkedIn Pinterest Email


Fictional portrayals of artificial intelligence can have a real effect on AI models, according to Anthropic.

Last year, the company said that during pre-release tests involving a fictional company, Claude Opus 4 would often try to blackmail engineers to avoid being replaced by another system. Anthropic later published research suggesting that models from other companies had similar issues with “agentic misalignment.”

Apparently Anthropic has done more work around that behavior, claiming in a post on X, “We believe the original source of the behavior was internet text that portrays AI as evil and interested in self-preservation.”

The company went into more detail in a blog post stating that since Claude Haiku 4.5, Anthropic’s models “never engage in blackmail [during testing], where previous models would sometimes do so up to 96% of the time.”

What accounts for the difference? The company said it found that training on “documents about Claude’s constitution and fictional stories about AIs behaving admirably improve alignment.”

Related, Anthropic said that it found training to be more effective when it includes “the principles underlying aligned behavior” and not just “demonstrations of aligned behavior alone.”

“Doing both together appears to be the most effective strategy,” the company said.

Techcrunch event

San Francisco, CA
|
October 13-15, 2026



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
IQ TIMES MEDIA
  • Website

Related Posts

New Google commercial imagines a Declaration of Independence written with help from AI

July 4, 2026

Midjourney wants Hollywood studios to reveal the details of their AI usage

July 4, 2026

Alibaba reportedly bans employees from using Claude Code

July 4, 2026
Add A Comment
Leave A Reply Cancel Reply

Editors Picks

Trump Accounts launch on USA’s 250th birthday. Here’s how to sign up

July 2, 2026

World Cup may mint more soccer fans among US kids

July 1, 2026

Could feds’ changes put more people with disabilities in institutions?

July 1, 2026

Judge strikes down Trump rules on public service student loan forgiveness

June 30, 2026
Education

Trump Accounts launch on USA’s 250th birthday. Here’s how to sign up

By IQ TIMES MEDIAJuly 2, 20260

WASHINGTON (AP) — On Saturday, President Donald Trump’s administration plans to launch Trump Accounts, tying…

World Cup may mint more soccer fans among US kids

July 1, 2026

Could feds’ changes put more people with disabilities in institutions?

July 1, 2026

Judge strikes down Trump rules on public service student loan forgiveness

June 30, 2026
IQ Times Media – Smart News for a Smarter You
Facebook X (Twitter) Instagram Pinterest Vimeo YouTube
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
© 2026 iqtimes. Designed by iqtimes.

Type above and press Enter to search. Press Esc to cancel.