Open evaluation arena for AI shopping agents where independent validators run changing commerce tasks, score agent trajectories, and publish a competitive leaderboard.
Publisher
ORO AI develops or publishes continuously changing shopping-agent benchmark with sandboxed independent validators, trajectory-level scoring, public leaderboard, CLI submission, and retained evaluation traces. This sourced organization profile is tied to the verified product URL used for Batch 009 Tranche 04.
View publisherClaim this listing to keep its information accurate, manage its presence on Ploba and connect with users.