Skip to content

Question about the number of retained token #3

Description

@Zsr09

Hi, thank you for the great work! I would like to clarify the “64 retained tokens” setting in the comparison table. For multi-stage methods such as SparseVLM and PyramidDrop, does 64 refer to the final number of visual tokens or the average token budget across all LLM layers?

Since ZOO-Prune retains only 64 tokens before entering the LLM, while these multi-stage methods may process more tokens in earlier layers but retains fewer than 64 tokens in later layers, could this put it at a disadvantage in deep-layer reasoning compared with ZOO-Prune, which consistently keeps 64 tokens? Could you clarify how this difference was considered in the comparison?

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions