LiteRT is Google's on-device framework for high-performance ML & GenAI deployment on edge platforms.
Efficient conversion, runtime, and optimization for on-device machine learning.
Built on the battle-tested foundation of TensorFlow Lite
LiteRT isn't just new; it's the next generation of the world's most widely deployed machine learning runtime. It powers the apps you use every day, delivering low latency and high privacy on billions of devices.
Trusted by the most critical Google apps
100K+ applications, billions of global users
LiteRT Highlights
Cross Platform Ready
Unleash GenAI
Simplified hardware acceleration
Multi-framework support
Deploy via LiteRT
Streamline your deep learning workflow from training to on-device deployment.
1.Obtain a model
Use .tflite pre-trained models or convert PyTorch, JAX or TensorFlow models to .tflite.
2.Optimize
Use the LiteRT optimization toolkit to quantize your models post-training.
3.Run
Deploy your model with LiteRT and pick the optimal accelerator for your app.
Choose Your Development Path
Use LiteRT to deploy AI anywhere—from high-performance mobile apps to resource-constrained IoT devices.
Samples, models, and demo
Blogs and Announcements
Stay up to date with the latest announcements, technical deep dives, and performance benchmarks from the LiteRT team.
Except as otherwise noted, the content of this page is licensed under the Creative Commons Attribution 4.0 License, and code samples are licensed under the Apache 2.0 License. For details, see the Google Developers Site Policies. Java is a registered trademark of Oracle and/or its affiliates.
Last updated 2026-07-17 UTC.