Select Page
OpenAI Jalapeño Chip Benchmarks: Faster AI at Lower Power

OpenAI Jalapeño Chip Benchmarks: Faster AI at Lower Power

OpenAI has published the first measured results for Jalapeño, its custom AI inference chip, and the headline is not simply “more speed”. The company says its purpose-built silicon can process more AI work per unit of power while cutting the delay users experience. If...
NVIDIA Vera Rubin NVL72: 30× AI Efficiency Explained

NVIDIA Vera Rubin NVL72: 30× AI Efficiency Explained

NVIDIA has published early performance results for Vera Rubin NVL72, its next-generation rack-scale AI platform, claiming up to 30 times more agentic AI throughput per megawatt than its GB300 NVL72 system. The company also says token costs can be up to 35 times lower...
OpenAI Ultrafast Mode: GPT-5.6 Sol Runs Up to 14× Faster

OpenAI Ultrafast Mode: GPT-5.6 Sol Runs Up to 14× Faster

OpenAI Ultrafast mode is a new API service-tier preview designed to run GPT-5.6 Sol at dramatically higher generation speeds. OpenAI says the system can produce up to 750 output tokens per second and operate at up to 14 times the speed of its Standard processing tier....