Reasoning model latency: Reasoning models are slower to provide their response, making them less suitable for latency sensitive tasks
End-to-End Response Time
Seconds to Output 500 Tokens, Including reasoning model 'thinking' time; Lower is better
Input processing time
'Thinking' time (reasoning models)
Outputting time
AI text/layout recreation from video frame; verify against source image.