OpenAI unveiled Jalapeño, its custom AI accelerator chip designed to slash inference latency. The chip delivers 13.4 petaflops of 4-bit compute and uses LLMs to accelerate its own design cycle. OpenAI developed the accelerator in under twenty months with a hardware team of roughly one hundred engineers.
Opening Kapyn…