MDA AI QU QuarkAndCode LLM Streaming Latency: Cut TTFT, Smooth Tokens, Fix Cold Starts Make LLMs feel fast: cut TTFT, smooth streaming cadence, fix cold starts via buffer-aware scheduling, KV cache ops and async function…