Close Menu
News Frame For You — Latest Updates on AI, Sports, Europe, Asia & Business
  • Home
  • AI
  • Asia
  • Business
  • Education
  • Europe
  • Life & Style
  • Sports
  • USA
  • Store

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

What's Hot

Cooper Flagg exits with ankle injury for Mavs against Nuggets

January 15, 2026

Tiger Woods to open new learning lab in Atlanta with $20M grant from Arthur Blank

January 15, 2026

Remove pit of avocado in 1second without having to use a knife

January 15, 2026
Facebook X (Twitter) Instagram
News Frame For You — Latest Updates on AI, Sports, Europe, Asia & Business
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
  • Home
  • AI
  • Asia
  • Business
  • Education
  • Europe
  • Life & Style
  • Sports
  • USA
  • Store
News Frame For You — Latest Updates on AI, Sports, Europe, Asia & Business
Home » Microsoft built a fake marketplace to test AI agents — they failed in surprising ways
AI

Microsoft built a fake marketplace to test AI agents — they failed in surprising ways

adminBy adminNovember 7, 2025No Comments3 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Email
Share
Facebook Twitter LinkedIn Pinterest Email Copy Link


On Wednesday, researchers at Microsoft released a new simulation environment designed to test AI agents, along with new research showing that current agentic models may be vulnerable to manipulation. Conducted in collaboration with Arizona State University, the research raises new questions about how well AI agents will perform when working unsupervised — and how quickly AI companies can make good on promises of an agentic future.

The simulation environment, dubbed the “Magentic Marketplace” by Microsoft, is built as a synthetic platform for experimenting on AI agent behavior. A typical experiment might involve a customer-agent trying to order dinner according to a user’s instructions, while agents representing various restaurants compete to win the order.

The team’s initial experiments included 100 separate customer-side agents interacting with 300 business-side agents. Because the source code for the marketplace is open source, it should be straightforward for other groups to adopt the code to run new experiments or reproduce findings.

Ece Kamar, CVP and managing director of Microsoft Research’s AI Frontiers Lab, says this kind of research will be critical to understanding the capabilities of AI agents. “There is really a question about how the world is going to change by having these agents collaborating and talking to each other and negotiating,” said Kamar. “We want to understand these things deeply.”

The initial research looked at a mix of leading models, including GPT-4o, GPT-5, and Gemini-2.5-Flash, and found some surprising weaknesses. In particular, the researchers found several techniques businesses could use to manipulate customer agents into buying their products. The researchers noticed a particular falloff in efficiency as a customer agent was given more options to choose from, overwhelming the attention space of the agent.

“We want these agents to help us with processing a lot of options,” Kamar says. “And we are seeing that the current models are actually getting really overwhelmed by having too many options.”

The agents also ran into trouble when they were asked to collaborate toward a common goal, apparently unsure of which agent should play what role in the collaboration. Performance improved when the models were given more explicit instructions on how to collaborate, but the researchers still saw the models’ inherent capabilities as in need of improvement.

Techcrunch event

San Francisco
|
October 13-15, 2026

“We can instruct the models — like we can tell them, step by step,” Kamar said. “But if we are inherently testing their collaboration capabilities, I would expect these models to have these capabilities by default.”



Source link

Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
admin
  • Website

Related Posts

OpenAI signs deal, worth $10B, for compute from Cerebras

January 15, 2026

Musk denies awareness of Grok sexual underage images as California AG launches probe

January 15, 2026

India’s Emversity doubles valuation as it scales workers AI can’t replace

January 15, 2026
Leave A Reply Cancel Reply

Don't Miss
Sports

Cooper Flagg exits with ankle injury for Mavs against Nuggets

DALLAS (AP) — Dallas Mavericks rookie Cooper Flagg was ruled out for the second half…

Tiger Woods to open new learning lab in Atlanta with $20M grant from Arthur Blank

January 15, 2026

Remove pit of avocado in 1second without having to use a knife

January 15, 2026

Three Palestine Action activists end UK hunger strike | Israel-Palestine conflict News

January 15, 2026
Top Posts

Are Iran’s protests different this time around? | Protests News

January 14, 2026

As hate spirals in India, Hindu extremists turn to Christian targets | Politics

January 14, 2026

Bangladesh won’t play T20 World Cup matches in India, BCB reaffirms | Cricket News

January 13, 2026

Trump announces new 25% tariff: How will it impact Iran’s trading partners? | International Trade News

January 13, 2026

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

About Us
About Us

Welcome to News Frame For You — Your Window to the World! 🌍

At News Frame For You, we bring you the latest and most reliable updates from across the globe, focusing on what truly shapes our modern world. From cutting-edge AI innovations to thrilling sports moments, from the heart of Europe’s business scene to the pulse of Asia’s emerging markets, we frame the news that matters to you — clearly, quickly, and intelligently.

Our Picks

Cooper Flagg exits with ankle injury for Mavs against Nuggets

January 15, 2026

Tiger Woods to open new learning lab in Atlanta with $20M grant from Arthur Blank

January 15, 2026

Remove pit of avocado in 1second without having to use a knife

January 15, 2026
Most Popular

Laude Institute announces first batch of ‘Slingshots’ AI grants

November 7, 2025

Sam Altman says OpenAI has $20B ARR and about $1.4 trillion in data center commitments

November 7, 2025

Amazon launches an AI-powered Kindle Translate service for e-book authors

November 7, 2025
  • Home
  • About Us
  • Advertise With Us
  • Contact us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
© 2026 newsframeforyou. Designed by newsframeforyou.

Type above and press Enter to search. Press Esc to cancel.