Jalapeño’s first results show industry-leading speed and efficiency in AI inference Joab Peter's BlogAugust 26, 2026 Jalapeño is a custom inference chip from OpenAI that delivers faster, more power-efficient AI inference, with higher through...Read More
Core dump epidemiology: fixing an 18-year-old bug Joab Peter's BlogAugust 09, 2026 OpenAI engineers used large-scale core dump analysis to debug rare infrastructure crashes, uncovering both a hardware fault ...Read More
How we built a realtime system for responsive voice AI in six months Joab Peter's BlogAugust 05, 2026 GPT-Live enables continuous voice interaction with AI, using a turnless speech model and low-latency architecture for faster,...Read More
How GPT-5.6 fuses frontier intelligence with frontier efficiency Joab Peter's BlogAugust 03, 2026 GPT-5.6 improves AI efficiency across models, inference, and agentic workflows, helping deliver more useful intelligence per ...Read More