Nvidia's race to manufacture Groq chips and make them available to customers highlight the growing importance in AI of low-latency inference.