ML Drift is Google's open-source GPU engine that runs on-device AI inference on mobile, web, and desktop without a server round-trip.
Yes, ML Drift is open source under the Apache-2.0 license.
ML Drift is free to use.
Yes, ML Drift can be self-hosted (the source is available under the Apache-2.0 license).
Google's open-source, cross-platform GPU-accelerated inference engine for on-device machine learning — the same GPU backend that has powered TensorFlow Lite/LiteRT inside Google products, now published as a standalone Apache-2.0 repo. Runs AI/ML workloads, including large generative models, on mobile, web, and desktop GPUs via OpenCL, Metal, WebGPU, and OpenGL ES 3.1+.
Shipping an AI feature into a consumer-facing app usually means a metered API call per inference, which adds latency, cost, and a data round-trip a privacy-sensitive user or team may not want. ML Drift's bet is that the device GPU already sitting in the user's phone or browser can do real inference work for free, if the engine abstracts away the OpenCL/Metal/WebGPU/OpenGL differences. The honest read: this is Google's internal GPU backend newly published standalone, so the headline performance numbers (up to 40% latency cut, up to 2 seconds faster) are Google's own production results on Google's own products, not independently re-verified on a general third-party app — worth prototyping on your actual model and device mix before assuming the same gains.