Product
Modal Modal GPU compute: sourced case studies
Every documented customer result that used Modal Modal GPU compute, each traceable to its original published source.
65%
Voice 2.0 latency reduction
AI
Decagon cuts voice AI latency 65% with custom inference on Modal
Decagon · Sourced July 30, 2026
10-15ms
Added network latency for remote inference
Robotics
Physical Intelligence runs real-time robot inference remotely on Modal
Physical Intelligence · Sourced July 30, 2026
3x
P90 latency reduction
AI
Reducto cuts P90 latency 3x by moving 30+ models to Modal
Reducto · Sourced July 30, 2026
4 mo
Faster to launch
AI
Suno launches faster by running inference on Modal instead of Kubernetes
Suno · Sourced July 30, 2026