A novel computational method for token pruning in vision-language models, demonstrated on 3D reasoning benchmarks with no peer review and no clinical or translational validation reported.
Reported
Performance retention at reduced…93.5%
Token budgetapproximately 8%
Performance improvement over prio…3.9 percentage points on average