LiteRT vs TensorFlow Lite: what changed, plus the old-name new-name cheat sheet
Last verified: 2026-09-05 — LiteRT 2.2.0, LiteRT-LM 0.16.1, litert-torch 0.9.4, ai-edge-litert 2.2.0....
Tag archive
Last verified: 2026-09-05 — LiteRT 2.2.0, LiteRT-LM 0.16.1, litert-torch 0.9.4, ai-edge-litert 2.2.0....

LiteRT Raspberry Pi 5: Google's LiteRT CLI puts Gemma and EfficientNet on a Raspberry Pi 5 with three commands. Here is the setup, the file sizes, and the GPU gotcha.

Run Gemma AI Offline: Google's LiteRT runs Gemma models on a Raspberry Pi 5 at 99 tokens/sec prefill in 1432 MB of RAM, with no cloud and no API key needed.

A technical walkthrough of “Which Celebrity Vegetable Are You?” — private WebAI inference with...
Running Gemma 4 On-Device on Android: Building a Private AI Assistant with LiteRT I built...
Three chips can run the same LLM, but it takes two files - one shared by CPU and GPU, one just for the NPU. Understanding why took me into graphs, subgraphs, and the difference between just-in-time and ahead-of-time compilation - the layer LiteRT calls delegates.

The Real Problem: On-Device AI Fragmentation and Bottlenecks For years, the promise of...

The landscape of client-side machine learning is shifting. For years, deploying high-performance...