We plan to begin deploying Jalapeño in OpenAI’s compute infrastructure by year-end.
It’s the first step in a multigenerational roadmap: Gen 2 is deep in development, and Gen 3 is taking shape.
Each generation will push efficiency and speed further. https://openai.com/index/jalapeno-first-results/
65·B+Long
T
Twitter3h agonews
Since announcing Jalapeño, our first custom inference chip, we’ve been testing it and the system around it.
The results show a major advance: more intelligence from every watt and faster responses, delivering both higher throughput and lower latency in one architecture without sacrificing efficiency.
10·CNeutral
T
Twitter3h agonews
Jalapeño means faster ChatGPT responses, more responsive Codex sessions and agents, and reliable access as demand continues to grow. https://t.co/6kdoJoXWaf