LiteRT.js runs machine learning models locally with CPU, GPU and emerging NPU acceleration, potentially reducing server infrastructure, inference charges and data movement.
The new method fills the gap between GET and POST, combining the body of POST with the safety and idempotence of GET.