Within the multimodal suite, ChartQA falls furthest at every reduced budget; at k=2, the three image-text transcription datasets are at or near zero, while multiple-choice datasets retain roughly one third to one half of their own unpruned scores. · CiteArk
Not assessedPlan blockedFindingclaim-multimodal-dataset-failure-pattern
Within the multimodal suite, ChartQA falls furthest at every reduced budget; at k=2, the three image-text transcription datasets are at or near zero, while multiple-choice datasets retain roughly one third to one half of their own unpruned scores.
Source: fixed-paper:PDF p. 22, Appendix F and Table 12
Reported and observed measurements
No structured measurement is attached to this Claim.
Assessments (0)
No immutable Assessment has been published for this Claim yet.