Open evaluation benchmark for agentic commerce where shopping agents run changing commerce tasks and produce performance and training data
Publisher
ORO AI develops or publishes continuously changing shopping-agent benchmark with sandboxed independent validators, trajectory-level scoring, public leaderboard, CLI submission, and retained evaluation traces. This sourced organization profile is tied to the verified product URL used for Batch 009 Tranche 04.
View publisherClaim this listing to keep its information accurate, manage its presence on Ploba and connect with users.