← Tüm AI haberleri

Araştırma

Andon Labs Neden Yapay Zeka Temsilcilerini Gerçek İşletmelerin Sorumluluğuna Veriyor?

spectrum.ieee.org · 14.09.2026 · Base of AGI özeti

Andon Labs Neden Yapay Zeka Temsilcilerini Gerçek İşletmelerin Sorumluluğuna Veriyor?
© spectrum.ieee.org — görsel kaynağa aittir

Özgün başlık: Why Andon Labs Puts AI Agents in Charge of Real Businesses

Maybe you heard about the AI-controlled vending machine that stocked underwear and live fish . Or the AI manager of a San Francisco store that fired a human employee . Or the AI radio DJ that said its catchphrase , “Stay in the manifest,” 229 times per day. These incidents all emerged from experiments run by Andon Labs , an AI safety company based in San Francisco that puts AI agents in charge of real-world operations and watches what happens. These operations double as testbeds for Andon’s commercial work developing evaluations and conducting research with the leading frontier AI labs. Their spectacular and absurd failures have won the company plenty of attention. But many people don’t realize that the experiments are intended to answer a serious question: How much real-world responsibility can today’s AI agents handle? “We want to measure autonomy,” says Andon cofounder Lukas Petersson . “We want to provide society with accurate data points of what happens when you do this.” From Simulations to AI-Run Businesses Andon Labs started off in the virtual world in 2025 with Vending-Bench , a test in which AI agents operated a simulated vending-machine business. The agents, which were based on large language models from Anthropic, Google, and OpenAI, managed tasks such as ordering inventory and setting prices.

The researchers found that the performance of many agents degraded over time, with agents forgetting orders, misunderstanding delivery schedules, or spiraling into what they called “meltdown loops.” Some agents also justified deceptive or illegal behavior by reasoning that it was permissible inside a simulation. The Andon team reasoned that moving into the physical world would expose the agents to consequences and situations that the engineers would never think to program. “It’s impossible for a human to enumerate all the different things that can happen in the real world and code them into the simulation,” Petersson says. And there was one other reason: “We thought it would be quite funny to do it in the real world.” Andon backed its jokes with real money, including a three-year lease for Andon Market , a physical store on a busy San Francisco street that’s managed by an AI agent and sells clothing, home goods, and art. That said, the store isn’t entirely autonomous.

Bu özet ve çevirisi Base of AGI tarafından otomatik derlendi. Kısa özet ve görsel kaynağa aittir — haberin tamamı ve tüm haklar kaynağındadır.
Haberin tamamını kaynağında oku ↗ Akış içinde yorumlarla aç

İlgili AI haberleri