Skip to content

Freeze Fruit

pendingmediumhouse_single_floorKitchen5m 38sunowned

Agent runs

Each mode of this task runs once per run. Est. cost is tokens at list price (data/prices.yml), never a bill; see the BEHAVIOR-1K runs for every task of a run.

Codex + GPT-6 Luna, reasoning xhigh (ChatGPT login)default run

Not in this run: not among the 30 chosen for the first round (2026-09-26): mostly particles, cutting and liquids, or long tidy-ups like chosen ones

No instruction published upstream

The 2025 carryover tasks ship without instruction text. Watch the demo and write the goal in your own words below, clearly marked as a reconstruction rather than the official goal.

Measured on the stack we run — BEHAVIOR-1K v3.9.2 / OmniGibson 3.9.2 / Isaac Sim 5.1
Goal predicates inside open
Quantifiers forall forpairs
Goal clauses 4
Objects named in the goal apple.n.01 electric_refrigerator.n.01 strawberry.n.01 tupperware.n.01
Objects in the problem 13 in 10 categories
Rooms loaded corridor_0 dining_room_0 entryway_0 garden_0 kitchen_0 living_room_0 living_room_1
Demonstrations 200 teleoperated episodes
Mean episode 7m 01s (12,619 control steps at 30 Hz)
Base travel, mean 25.4 m
Gripper travel, mean left 28.65 m · right 38.64 m
Evaluation instances 20 public test instances (ids 301–320)

Read from BDDL activity definitions + 2026-challenge-task-instances (licensed download), 2026-09-21.

Tags

Why this task is interesting

Not yet written.

Capability notes

Not yet written.

Oracle demo review

Not yet reviewed.

Discussion