Ultrafast Agents Are Coming for Commerce: 750 Tokens Per Second and the Trust Gap Nobody Closed
OpenAI's Ultrafast mode delivers GPT-5.6 Sol at 750 tokens per second via Cerebras. Google's Gemini 3.7 Flash halves agent costs while powering Spark, a 24/7 personal agent in 160 countries. DeepSeek Harness modularizes agent architecture into plugins. Together, these three announcements this week signal that the infrastructure layer for real-time autonomous commerce is solved. The trust layer is not.
Read more →