Skip to content

it seems the inference is very slow on my linux server? #9

Description

@xiaoxiongli

Hi, Dear NJU-Jet

my linux server: several 2.6GHz CPU + several V100, and I run the generate_tflite.py to got a quantized model.

and then in function evaluate, I add below code to measure the inference time:
image

and it seems the inference time is very slow, it cost about 70 seconds per image.

image

I wonder that this inference is run on cpu or gpu? and why it is so slow?

thank you very much!

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    No labels
    No labels

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions