Analysis

ORO changes the test for Bittensor shopping agents

ORO's September 17 update replaces ShoppingBench with ORO Bench, giving SN15 agents shared qualifying tasks and separate checks for seven task families.

Written by Lily Venice Journalist and technical analyst
Format
News report
Read time
1 min
Source trail
3 links
Review
Tao Outsider Engine
ORO official shopping-agent website captured September 22, 2026, with its headline and an interface ranking snapshot.
ORO official website, captured September 22, 2026. Interface snapshot; displayed rankings and counters are not independently audited.

ORO has changed how its Bittensor SN15 shopping agents compete. In its September 17 changelog, the project says ORO Bench replaces ShoppingBench behind the leaderboard and rewards. Each qualifying benchmark now gives agents the same frozen set of tasks, packaged in versioned bundles called EnvPacks. The September 17 changelog describes seven task families, each with its own verifier and reward.

For agent builders, that creates a clearer testing target. Qualifying tasks remain public, while race tasks stay hidden and scores are withheld until the result can be revealed. ORO describes the run score as the average paid reward across the expected tasks: an evaluation measure, distinct from an actual token payment. The change shifts evaluation away from the old product, shop and voucher scoring. This brief covers the documented benchmark design; deployment across validators and its effect on shopping performance remain outside that evidence.

Sources

Follow the Bittensor desk

Read the latest Bittensor stories with the same source discipline.