
I benchmarked six AI Gateways, including ours
Six gateways, 43 regions, 16,344 measurements. Edgee has the lowest overhead on GPT-5.4 (24ms over a direct API call) and delivers the first token first in two out of three head-to-heads.

Six gateways, 43 regions, 16,344 measurements. Edgee has the lowest overhead on GPT-5.4 (24ms over a direct API call) and delivers the first token first in two out of three head-to-heads.

We benchmarked Codex alone against Codex routed through Edgee's compression gateway on the same repo, with the same model, under the same workflow. The result: Codex + Edgee used 49.5% fewer input tokens, improved cache hit rate from 76.1% to 85.4%, and reduced total session cost by 35.6%. This post breaks down why context compression makes Codex more efficient, more frugal, and materially cheaper to run without sacrificing useful output.

We ran a head-to-head endurance test: raw Claude Code vs Claude Code with Edgee's token compressor. Same plan, same tasks. One went 26.5% further.
Would you like to find out more about Edgee, test our services or our upcoming features? We’d love to hear from you. Please fill in the form below and we’ll be in touch.