OpenAI's Jalapeño Chip Debuts with Focus on Scalable Inference: Benchmarks Show Energy and Latency Gains