OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show | TechCrunch
Tested on Semianalysis’s InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.

Next post
Apple announces new M6 Mac mini and M5 Ultra Mac Studio






