Block Stacking Specified Order¶
Agent runs
Each mode of this task runs once per run. Est. cost is tokens at list price (data/prices.yml), never a bill; see the RoboLab runs for every task of a run.
Codex + GPT-6 Luna, reasoning xhigh (ChatGPT login)default run
| Trial | Started | Agent time | Model requests | Tokens in / out | Est. cost | Grade |
|---|---|---|---|---|---|---|
| U✓ 12m | 09-27 03:12 | 12m | 24 | 819k / 6k | $0.016 | deterministic 1, grader_error 0, instance_exact 1, instance_ok 1, n_actions 560, video_rendered 1 |
| L✓ 31m | 09-29 10:07 | 31m | 54 | 2.5M / 30k | $0.049 | deterministic 1, grader_error 0, live_success 1, n_actions 360, replay_success 1, video_rendered 1 |
Task instruction (upstream)
Stack the blocks in the order from bottom to top: red, blue, green, yellow
Playback speed
What RoboLab states about this task
| Success predicate | stacked |
| Predicate arguments | objects: ['red_block', 'blue_block', 'green_block', 'yellow_block'] · order: bottom_to_top |
| Scored subtasks | 3 |
| Subtask predicates | stacked |
| Objects | red_block blue_block green_block yellow_block |
| Upstream attributes | stacking color |
| Upstream difficulty | simple |
| Episode budget | 90 s |
| Other wordings | vague — Stack in the order of red, blue, green, yellow specific — Build a tower by placing the red block first, then the blue block on top, then the green, and finally the yellow block on top as a single tower |
| Environment class | BlockStackingSpecifiedOrderTask |
| Upstream name | BlockStackingSpecifiedOrderTask |
| Defined in | robolab/tasks/benchmark/block_stacking_specified_order_task.py |
From https://github.com/NVlabs/RoboLab @ v0.3.1.
Tags¶
Task DomainManipulation
Skill primitives in the demo: colorstacking
Why this task is interesting¶
Not yet written.
Capability notes¶
Not yet written.
Oracle demo review¶
Not yet reviewed.