Close Menu
  • Home
  • AI
  • Education
  • Entertainment
  • Food Health
  • Health
  • Sports
  • Tech
  • Well Being

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

What's Hot

Everyone is navigating AI security in real time — even Google

May 24, 2026

Sundar Pichai Says Graduates Booing AI Will Live With Tech’s Impact

May 24, 2026

I tried Amazon’s Bee wearable and am both intrigued and slightly creeped out

May 24, 2026
Facebook X (Twitter) Instagram
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
Facebook X (Twitter) Instagram
IQ Times Media – Smart News for a Smarter YouIQ Times Media – Smart News for a Smarter You
  • Home
  • AI
  • Education
  • Entertainment
  • Food Health
  • Health
  • Sports
  • Tech
  • Well Being
IQ Times Media – Smart News for a Smarter YouIQ Times Media – Smart News for a Smarter You
Home » Microsoft built a fake marketplace to test AI agents — they failed in surprising ways
AI

Microsoft built a fake marketplace to test AI agents — they failed in surprising ways

IQ TIMES MEDIABy IQ TIMES MEDIANovember 5, 2025No Comments3 Mins Read
Facebook Twitter Pinterest LinkedIn Tumblr Email
Share
Facebook Twitter LinkedIn Pinterest Email


On Wednesday, researchers at Microsoft released a new simulation environment designed to test AI agents, along with new research showing that current agentic models may be vulnerable to manipulation. Conducted in collaboration with Arizona State University, the research raises new questions about how well AI agents will perform when working unsupervised — and how quickly AI companies can make good on promises of an agentic future.

The simulation environment, dubbed the “Magentic Marketplace” by Microsoft, is built as a synthetic platform for experimenting on AI agent behavior. A typical experiment might involve a customer-agent trying to order dinner according to a user’s instructions, while agents representing various restaurants compete to win the order.

The team’s initial experiments included 100 separate customer-side agents interacting with 300 business-side agents. Because the source code for the marketplace is open source, it should be straightforward for other groups to adopt the code to run new experiments or reproduce findings.

Ece Kamar, managing director of Microsoft Research’s AI Frontiers Lab, says this kind of research will be critical to understanding the capabilities of AI agents. “There is really a question about how the world is going to change by having these agents collaborating and talking to each other and negotiating,” said Kamar. “We want to understand these things deeply.”

The initial research looked at a mix of leading models, including GPT-4o, GPT-5, and Gemini-2.5-Flash, and found some surprising weaknesses. In particular, the researchers found several techniques businesses could use to manipulate customer agents into buying their products. The researchers noticed a particular falloff in efficiency as a customer agent was given more options to choose from, overwhelming the attention space of the agent.

“We want these agents to help us with processing a lot of options,” Kamar says. “And we are seeing that the current models are actually getting really overwhelmed by having too many options.”

The agents also ran into trouble when they were asked to collaborate toward a common goal, apparently unsure of which agent should play what role in the collaboration. Performance improved when the models were given more explicit instructions on how to collaborate, but the researchers still saw the models’ inherent capabilities as in need of improvement.

Techcrunch event

San Francisco
|
October 13-15, 2026

“We can instruct the models — like we can tell them, step by step,” Kamar said. “But if we are inherently testing their collaboration capabilities, I would expect these models to have these capabilities by default.”



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
IQ TIMES MEDIA
  • Website

Related Posts

Everyone is navigating AI security in real time — even Google

May 24, 2026

I tried Amazon’s Bee wearable and am both intrigued and slightly creeped out

May 24, 2026

Ferrari is using IBM’s AI to create F1 superfans

May 23, 2026
Add A Comment
Leave A Reply Cancel Reply

Editors Picks

Scott Remer makes a good living as a National Spelling Bee coach

May 23, 2026

Ex-Columbia student Mahmoud Khalil asks Supreme Court to intervene in his deportation fight

May 22, 2026

Seniors roll into Michigan high school during annual Tractor Day celebration

May 22, 2026

Charges dismissed against former assistant principal accused after teacher shot

May 21, 2026
Education

Scott Remer makes a good living as a National Spelling Bee coach

By IQ TIMES MEDIAMay 23, 20260

When Dev Shah won the Scripps National Spelling Bee in 2023 and Faizan Zaki took…

Ex-Columbia student Mahmoud Khalil asks Supreme Court to intervene in his deportation fight

May 22, 2026

Seniors roll into Michigan high school during annual Tractor Day celebration

May 22, 2026

Charges dismissed against former assistant principal accused after teacher shot

May 21, 2026
IQ Times Media – Smart News for a Smarter You
Facebook X (Twitter) Instagram Pinterest Vimeo YouTube
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
© 2026 iqtimes. Designed by iqtimes.

Type above and press Enter to search. Press Esc to cancel.