Google for Developers

LiteRT is Google's on-device framework for high-performance ML & GenAI deployment on edge platforms.

Efficient conversion, runtime, and optimization for on-device machine learning.

Built on the battle-tested foundation of TensorFlow Lite

LiteRT isn't just new; it's the next generation of the world's most widely deployed machine learning runtime. It powers the apps you use every day, delivering low latency and high privacy on billions of devices.

Trusted by the most critical Google apps

100K+ applications, billions of global users

LiteRT Highlights

Cross Platform Ready

Unleash GenAI

Simplified hardware acceleration

Multi-framework support

Deploy via LiteRT

Streamline your deep learning workflow from training to on-device deployment.

1.Obtain a model

Use .tflite pre-trained models or convert PyTorch, JAX or TensorFlow models to .tflite.

2.Optimize

Use the LiteRT optimization toolkit to quantize your models post-training.

3.Run

Deploy your model with LiteRT and pick the optimal accelerator for your app.

Choose Your Development Path

Use LiteRT to deploy AI anywhere—from high-performance mobile apps to resource-constrained IoT devices.

Samples, models, and demo

Blogs and Announcements

Stay up to date with the latest announcements, technical deep dives, and performance benchmarks from the LiteRT team.

Except as otherwise noted, the content of this page is licensed under the Creative Commons Attribution 4.0 License, and code samples are licensed under the Apache 2.0 License. For details, see the Google Developers Site Policies. Java is a registered trademark of Oracle and/or its affiliates.

Last updated 2026-07-17 UTC.

Read the original on developers.google.com ↗