OpenAI has published a builder’s guide explaining how startups are using GPT-5.6 to create faster, more capable, and lower-cost AI agents. The guide recommends selecting smaller models such as Luna and Terra for suitable high-volume workloads, reducing reasoning effort when possible, and using new Responses API capabilities.
These include retained reasoning, native compaction, multi-agent orchestration, programmatic tool calling, and improved prompt caching. OpenAI reports that these techniques can reduce token consumption, latency, and inference costs while improving long-horizon agent performance.
Examples from startups including Hex, Browser Use, PlayerZero, Rogo, Quadrillion, and Ploy illustrate production results across different workflows.




