Injecting or generating guidance from the full graph can severely degrade performance on embodied tasks. On ALFWorld, full graph generative guidance drops success rate from 72.58% (baseline without graph) to 54.48% while dramatically increasing token usage from 18,055 to 96,360 tokens per sample, demonstrating the necessity of topological localization. · CiteArk
Not assessedNo independent reproduction scheduledLimitationclaim-limitation-full-graph-embodied-failure
Injecting or generating guidance from the full graph can severely degrade performance on embodied tasks. On ALFWorld, full graph generative guidance drops success rate from 72.58% (baseline without graph) to 54.48% while dramatically increasing token usage from 18,055 to 96,360 tokens per sample, demonstrating the necessity of topological localization.
Source: paper:p10-b101
Reported and observed measurements
No structured measurement is attached to this Claim.
Assessments (0)
No immutable Assessment has been published for this Claim yet.