Tweet by TheRealAdamG

April 30, 2026

https://t.co/h4QJVmYQNu “In this post, we'll explain how we made agent loops using the API 40% faster end-to-end, letting users experience the jump in inference speed from 65 to nearly 1,000 tokens per second. We approached this through caching, eliminating unnecessary network hops, improving our safety stack to quickly flag issues, and—most importantly—building a way to create a persistent connection to the Responses API, instead of having to make a series of synchronous API calls.”

Author
TheRealAdamG
Date
April 30, 2026