TFLite(LiteRT) - Generative AI
Important
Unless otherwise noted, all models listed in this section are supported on both Android and Yocto OS. Models or sections explicitly marked Android-only are not available on Yocto.
The Generative Model section provides performance and capability data for large language models (LLMs), vision-language models (VLMs) and image generation models on MediaTek Genio platforms.
This section is intended as a reference for benchmarking and platform capability validation, not as a distribution channel for full training or deployment assets.
Note
For Generative AI workloads, this section provides performance data and capability information only.
Access to the full Generative AI deployment toolkit (GAI toolkit) requires a non-disclosure agreement (NDA) with MediaTek. After signing an NDA, the toolkit can be downloaded from NeuroPilot Document.
Model Categories
The generative models in this section are grouped into the following categories:
Large Language Models (LLMs) – Text-only models for tasks such as dialogue, summarization, and code generation.
Vision-Language Models (VLMs) – Multimodal models that process both images and text (for example, image captioning or visual question answering).
Image Generation and Enhancement (Android only) – Models such as Stable Diffusion and other diffusion or transformer-based pipelines used for image synthesis, editing, or super-resolution.
Embedding and Encoder Models (Android only) – Models like CLIP encoders for computing joint image-text embeddings for retrieval or ranking tasks.