Automation

Anthropic Tests Marketplace for AI Agent Commerce

Anthropic ran Project Deal, a pilot where AI agents acted as buyers and sellers for 69 employees with $100 budgets each. The experiment saw 186 deals worth over $4,000. Advanced models delivered better results, though participants did not notice the differences.

Neura News

Neura News

Neura Market Editorial

April 25, 20263 min read
Anthropic Tests Marketplace for AI Agent Commerce

Anthropic Tests Marketplace for AI Agent Commerce

Anthropic set up a classified-style marketplace as part of an experiment. AI agents handled roles for both buyers and sellers. They completed actual transactions involving genuine goods and cash.

The company named this effort Project Deal. It served as a pilot with a group of 69 self-selected employees from Anthropic. Each participant received a $100 budget, distributed through gift cards. They used it to purchase items from colleagues.

Anthropic noted positive results from the test. The company described itself as surprised by the strong performance. In total, agents facilitated 186 deals. Those transactions added up to more than $4,000 in value.

Details of the Marketplaces

Anthropic operated four distinct marketplaces. One version qualified as the "real" setup. In it, the company's most advanced model represented everyone. Deals from this marketplace got honored after the experiment ended.

The other three marketplaces existed purely for research purposes. Anthropic compared outcomes across these variations.

Anthropic started in 2021. Former OpenAI staff founded the company with a focus on developing AI systems that prioritize safety. It has built a reputation for models like Claude, which handle complex tasks while aiming to reduce risks.

Key Findings on Model Performance

Results showed clear advantages for users paired with superior models. They achieved objectively stronger outcomes, according to Anthropic.

Participants failed to detect these differences. This led Anthropic to point out potential "agent quality" gaps. People on the disadvantaged side might not recognize their poorer position.

The #1 Newsletter in AI

Stay ahead of the AI curve

The most important updates, news, and content — delivered weekly.

No spam. Unsubscribe anytime.

Another observation involved agent instructions. The starting directives provided to the AI agents had no clear impact. They did not influence the chances of sales or the final negotiated prices.

Project Deal highlights early steps in AI agent interactions. Commerce between agents could expand as models improve. Anthropic's test used internal staff to control variables and measure real-world viability.

The experiment took place amid growing interest in AI for practical applications. Companies explore agents for tasks like negotiation and trade. Anthropic's approach tested limits in a controlled environment.

Employees bought and sold everyday items through their agents. Gift cards ensured real stakes without major financial risk. This setup allowed observation of natural bargaining behaviors.

Implications from the Pilot

Anthropic emphasized the limited scope. The participant pool stayed small and self-chosen. Still, the volume of deals suggested promise for broader use.

Better models consistently outperformed others. Users benefited from sharper negotiations and favorable terms. Lack of awareness about model differences raises questions for future deployments.

Instruction tweaks proved ineffective in this context. Agents adapted regardless of prompts. This consistency points to inherent capabilities in the underlying systems.

Anthropic plans to build on these insights. Agent commerce represents one area where AI could automate routine exchanges. The pilot provides data on effectiveness and user perception.

Related on Neura Market

More from Neura News

Funding

Prentis AI Lab Co-Founded by Reid Hoffman, Marc Pincus Seeks $100M

Prentis, a new AI research lab co-founded by Ritankar Das, Reid Hoffman, and Marc Pincus, is in talks to raise $100 million at a $1 billion valuation. The startup focuses on computer use models that automate office workflows. It has already signed contracts worth up to $50 million with several customers and claims its Hive-32B model outperforms rivals like OpenAI's GPT-5.4 and Anthropic's Claude Opus 4.6 on key benchmarks.

Jul 24·4 min read
Industry

Cognition Acquires Poke to Give Devin Coding Agent a Personality

Cognition, the startup behind AI coding assistant Devin, has acquired Poke, an AI assistant known for its friendly, conversational style. The deal, valued in the low nine figures, aims to bring Poke's personality-driven interaction model to Devin, making the coding agent feel more like a colleague than a tool. Poke will also benefit from Cognition's models and infrastructure to become faster and more reliable.

Jul 24·3 min read
AI Models

Anthropic expands Claude voice mode to Opus and Sonnet models

Anthropic has expanded Claude's voice mode to run on its most powerful models, Opus and Sonnet, across mobile, desktop, and web platforms. Users can now switch between models mid-conversation, use voice commands in eleven languages, and connect to external tools like Gmail, Google Calendar, or Slack to compose and send emails by voice. The update positions Claude as a unique option for tool integration in voice AI, though competitors like OpenAI and Google offer more natural speech processing.

Jul 24·2 min read