Jalapeño’s first results show industry-leading speed and efficiency in AI inference

Share

OpenAI has announced Jalapeño, a custom inference chip designed to improve the speed and efficiency of AI model inference with higher throughput and lower latency. The chip represents OpenAI's effort to optimize AI inference performance and reduce computational resource requirements.


Source: OpenAI