Simplified on purpose. Real providers batch requests, run safety
checks, and send SSE chunks of varying size, so real time to first token also
includes queueing and moderation. Apparent typing speed measures delivery settings,
not model intelligence. One token is not always one word. All timings here are
simulated seconds, not measurements.