AITechCrunch AI2h ago

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks show

OpenAI’s Jalapeño chip is built for fast inference at scale, benchmarks

Tested on SemiAnalysis’ InferenceX benchmark, Jalapeño registered both more tokens per user and more throughput per kilowatt than the currently available state-of-the art.

Read full article

Source: TechCrunch AI · Opens in new tab