Search for "LiteRT"

28 results

Clear filters
  • SEPT. 4, 2025 / Gemma

    Introducing EmbeddingGemma: The Best-in-Class Open Model for On-Device Embeddings

    Introducing EmbeddingGemma: a new embedding model designed for efficient on-device AI applications from Google. This open model is the highest-ranking text-only multilingual embedding model under 500M parameters on the MTEB benchmark, enabling powerful features like RAG and semantic search directly on mobile devices without an internet connection.

    EmbeddingGemma_Metadata
  • AUG. 14, 2025 / Gemma

    Introducing Gemma 3 270M: The compact model for hyper-efficient AI

    Google's new Gemma 3 270M is a compact, 270-million parameter model offering energy efficiency, production-ready quantization, and strong instruction-following, making it a powerful solution for task-specific fine-tuning in on-device and research settings.

    Gemma 3 270M
  • MAY 20, 2025 / AI Edge

    LiteRT: Maximum performance, simplified

    LiteRT has been improved to boost AI model performance and efficiency on mobile devices by effectively utilizing GPUs and NPUs, now requiring significantly less code, enabling simplified hardware accelerator selection, and more for optimal on-device performance.

    Built with LiteRT: Maximum Performance, Simplified
  • MAY 20, 2025 / AI Edge

    On-device small language models with multimodality, RAG, and Function Calling

    Google AI Edge advancements, include new Gemma 3 models, broader model support, and features like on-device RAG and Function Calling to enhance on-device generative AI capabilities.

    Google AI Edge: Small Language Models with Multimodality, RAG, and Function Calling
  • MARCH 31, 2025 / Gemini

    The Gemini API and the Internet of Things

    The Gemini API and ESP32 microcontroller simplify custom voice commands for IoT devices, leveraging speech recognition for devices to understand and react to custom commands, bridging the gap between digital and physical worlds.

    Gemini-API-IoT
  • MARCH 12, 2025 / Gemma

    Gemma 3 on mobile and web with Google AI Edge

    Gemma 3 1B, a new small language model for mobile and web applications via Google AI Edge, is now available, with increased efficiency, improved performance, and offline availability.

    Gemma 3 - Google AI Edge
  • MARCH 12, 2025 / Gemma

    Introducing Gemma 3: The Developer Guide

    Gemma 3 is a new, advanced version of the Gemma open-model family featuring multimodality, longer context windows, and improved language capabilities, with various sizes and deployment options for developers to experiment.

    Gemma 3
  • SEPT. 4, 2024 / AI Edge

    TensorFlow Lite is now LiteRT

    TensorFlow Lite, now named LiteRT, is still the same high-performance runtime for on-device AI, but with an expanded vision to support models authored in PyTorch, JAX, and Keras.

    LiteRT_BlogGraphics_1600x873px_1