Hi, I just noticed the dataset for Framethinker contains data from other datasets such as Video-Holmes. Would evaluating the model on Video-Holmes be unfair to other models?
Hi,
I just noticed the dataset for Framethinker contains data from other datasets such as Video-Holmes.
Would evaluating the model on Video-Holmes be unfair to other models?