- 4a363c0 New session configs based on protobuf by Byungchul Kim · 20 hours ago upstream/main
- 798d51e Support FLUX.2-klein engine/session by Byungchul Kim · 21 hours ago
- 2d872e0 Add Android benchmarks for Gemma 4 2B LoRA on CPU and GPU. by Salil Tambe · 27 hours ago
- 0063187 Allow creating embedding engines for models without a vision encoder. by Google AI Edge · 27 hours ago
- b2f3d3c Add speculative decoding statistics to the MTP drafter, the executor and the by Mohammadreza Heydary · 28 hours ago
- 8b7bb5c When the thought channel is not prefilled and the model skips thinking, by Google AI Edge · 28 hours ago
- 4697c72 - Write all-ones into `kPrevMaskName` in `input_buffers_map_` and clear `buffered_spectrogram_` inside `AudioLiteRtCompiledModelExecutor::AudioStreamingEncoder::CreateNewContext` so new streaming contexts start cleanly. by Google AI Edge · 31 hours ago
- e16809a Handle partial chunks correctly in AudioInputSource and TextInputSource. by Byungchul Kim · 32 hours ago
- ea11963 Document in the C API that sessions and conversations must be destroyed before the engine they were created from. by Tenghui Zhu · 34 hours ago
- 74e702d Text2Image engine and session by Byungchul Kim · 34 hours ago
- d07a923 Fix Gemma-4 GPU LoRA support for quantized models in MLDrift delegate. by Salil Tambe · 2 days ago
- 4b15b34 Report LoRA inputs that don't match any tensor in the LoRA data. by Google AI Edge · 2 days ago
- e9fb7ff Open-source the EmbeddingGemma 2 multimodal search web demo by Google AI Edge · 2 days ago
- 97e6dc7 [LiteRT-LM] Eagerly filter channel content from KV cache after decode. by Google AI Edge · 2 days ago
- b8b3131 Fix text scoring rewinding the KV-cache position twice after prefill. by Google AI Edge · 2 days ago
- 7840057 Isolate espeak_ng genrule working directories to fix parallel build race in standalone Bazel builds by Salil Tambe · 2 days ago
- 7233a2d Create LoRA buffers per signature so prefill gets buffer types it supports by Google AI Edge · 2 days ago
- 3a1f60d Rename Omni ASR/TTS session targets and standardize stage/factory signatures. by Byungchul Kim · 2 days ago
- d4beba4 Fix multibyte UTF-8 handling and byte-fallback detokenization in BufferedStreamingDetokenizer. by Mohammadreza Heydary · 2 days ago
- 1f7fa76 Add FLUX.2 model config, math helpers, denoiser stage, and VAE decoder stage. by Byungchul Kim · 2 days ago
- cb1495f update benchmark script by Google AI Edge · 2 days ago
- 4ecac4f Add DMA-BUF memory metrics to LiteRT LM metrics proto. by Andrew Zhang · 3 days ago
- 4c72b51 Unify ASR and TTS session orchestration under MultiStagedSession. by Byungchul Kim · 3 days ago
- f0f235e Add embedding demo link to litert lm readme by Google AI Edge · 3 days ago
- a926bb0 Fill lora_param_tensor so GPU runtime-BMM LoRA covers all channels by Salil Tambe · 3 days ago
- 5b87f5b Add TextEncoderStage under omni/text2image. by Byungchul Kim · 3 days ago
- 93d7412 Bump LiteRT-LM version to 0.19.0 by Google AI Edge · 3 days ago
- deaaa64 Re-prime the MTP drafter after any non-speculative step. by Mohammadreza Heydary · 3 days ago
- 8eca57a Add Conversation::CountTokens() API to measure prospective message token footprint. by Google AI Edge · 3 days ago
- c9848c4 add v0.18.0 release notes by Google AI Edge · 3 days ago
- 0bb486b Load Gemma 4 audio encoder LoRAs under the encoder's own input names by Salil Tambe · 3 days ago
- 84d9669 Pipe through max_top_k_ setting when set by Google AI Edge · 4 days ago
- dbdb99b Internal change by Google AI Edge · 4 days ago
- 525de7e Add an --output_size flag to embedding_litert_lm_main for truncation of the output embedding. by Marissa Ikonomidis · 4 days ago
- 0540e5c Replace `# pytype: disable` suppressions with `# pyrefly: ignore` by Oleh Prypin · 5 days ago
- d17a52f Add ImageDecoder and PromptSource interface for image gen model inference by Byungchul Kim · 7 days ago
- 10245c8 Internal change by Wai Hon Law · 7 days ago
- 8092fe5 Extend the Kotlin `benchmark` API and JNI bindings (`nativeCreateBenchmark`) to accept optional `visionBackend`, `audioBackend`, and `contents` parameters, and expose `BenchmarkInfo.markDurationsInSecond` from `BenchmarkInfo::GetMarkDurations()`. by Salil Tambe · 7 days ago
- 5befccd Restrict espeak-ng dictionaries to Kokoro-supported languages. by Tenghui Zhu · 7 days ago
- f136f4a Handle silence in more robust way by Byungchul Kim · 7 days ago
- 400273c Add pre-compiled model support to LiteRT-LM CompiledModelExecutor by Google AI Edge · 7 days ago
- b320801 Internal change by Tenghui Zhu · 7 days ago
- 956de4b Strengthen input validation and null pointer checks in LiteRT LM C APIs. by Mohammadreza Heydary · 7 days ago
- 58fcc1c Preserve prefill start token and truncate repetition loops in LmDecoder. by Byungchul Kim · 7 days ago
- b34761f Stitch overlapping chunks at longest contiguous match in LevenshteinTextMerger. by Byungchul Kim · 7 days ago
- c8ff070 Retain unconfirmed words on empty chunks in TimestampTextMerger. by Byungchul Kim · 7 days ago
- 3dbb23e Skip auxiliary mask and cache update warmup in LiteRT-LM NPU compiled model executor when using hardware update (kWH). by Andrew Zhang · 8 days ago
- 1a14a11 Remove litert_lm_get_last_error_code from the LiteRT-LM C API. by Mohammadreza Heydary · 8 days ago
- 5775c07 Zero-pad wrapped ring buffer frames at EOF in FileAudioSource. by Byungchul Kim · 8 days ago
- 4af508a Document the LiteRT-LM C API 1.0.0 return-shape contract by Mohammadreza Heydary · 8 days ago
- fc6ca8a Return status codes from the experimental LiteRT-LM C API by Mohammadreza Heydary · 8 days ago
- a5d53ea Return status codes from all LiteRT-LM model info C API functions.\ by Mohammadreza Heydary · 8 days ago
- 216b00b Add ASR, TTS, and ImageGen in litert_lm_builder. by Byungchul Kim · 8 days ago
- 5190782 Return status codes from all LiteRT-LM embedding engine C API functions. by Mohammadreza Heydary · 8 days ago
- b4a9d0e Return status codes from LiteRT-LM C API conversation functions by Mohammadreza Heydary · 8 days ago
- f4455da Sanitize tool and property names for Lark grammar rules in constrained decoding. by Google AI Edge · 8 days ago
- 8ab3da0 Return status codes from the remaining LiteRT-LM C engine.h accessors by Mohammadreza Heydary · 8 days ago
- 7d9a2a3 Return status codes from LiteRT-LM C API result-producing functions by Mohammadreza Heydary · 8 days ago
- 1dc4633 Return status codes from LiteRT-LM C API engine constructors by Mohammadreza Heydary · 8 days ago
- d101ec9 Make the dummy test model's output shape match its declared shape. by Tommy Chiang · 8 days ago
- 9369b17 Fix support for Gemma3 270M in LiteRT LM runtime and builder. by Andrew Zhang · 9 days ago
- 12ba057 Also skip the host local attention mask fill when both masks are pruned by the GPU delegate. by Fengwu Yao · 9 days ago
- b102de2 This is an internal change by Google AI Edge · 9 days ago
- 526d965 Add DMA-BUF memory tracking utility and init time measurement for LiteRT-LM runners. by Andrew Zhang · 9 days ago
- 56c420b Add ImageGenMetadata in .litertlm by Byungchul Kim · 9 days ago
- 8fded1b Move TFLite mutable schema generation from patch step to build-time CMake target. by Mohammadreza Heydary · 9 days ago
- 96a6009 feat: add enableYnnpack experimental flag in Swift API by Wai Hon Law · 9 days ago
- 31c1745 Return canonical status codes instead of -1 from status-returning LiteRT-LM C API functions. by Mohammadreza Heydary · 9 days ago
- 72a3df7 Bump the LiteRT-LM C API version to 1.0.0 and document the status-return rule. by Mohammadreza Heydary · 9 days ago
- a21dcbe test(cli): add multimodal e2e tests and documentation for openai embeddings by Wai Hon Law · 9 days ago
- 78d7bd2 Add DMA-BUF and peak memory reporting to LiteRT tools and runner script for embedding models. by Andrew Zhang · 9 days ago
- 3bc51f1 Check LiteRT-LM C setter status codes in the Python, Swift and Go bindings by Mohammadreza Heydary · 9 days ago
- 34d1def Generate mutable TFLite schema header in litert_patcher.cmake. by Mohammadreza Heydary · 9 days ago
- 3431993 Return status codes from void setters in the LiteRT-LM C embedding API. by Mohammadreza Heydary · 9 days ago
- 0ea540e feat: add enableYnnpack experimental flag in Kotlin API by Wai Hon Law · 9 days ago
- a0331da Return status codes from void setters in the LiteRT-LM C conversation API. by Mohammadreza Heydary · 9 days ago
- b1bbf74 Return status codes from void setters in the LiteRT-LM C engine API. by Mohammadreza Heydary · 9 days ago
- a458481 Record the last error on every silent failure path in the LiteRT-LM C API by Mohammadreza Heydary · 9 days ago
- bfdec21 Support ASR models in .litertlm by Byungchul Kim · 9 days ago
- 122ca32 Consolidate omni IO types by Byungchul Kim · 10 days ago
- 6b8f5b5 Internal change by Mohammadreza Heydary · 10 days ago
- 26fdbaf Add benchmark image and audio inputs to LiteRT-LM runtime testdata. by Google AI Edge · 10 days ago
- f07055f Internal change for eval by Google AI Edge · 10 days ago
- 86a7e6d Add image item formatting (`<image>`) to the LFM2 Jinja prompt template and LlmMetadataProto.pbtext. by Salil Tambe · 10 days ago
- 34d7bde Add status-return helpers for the LiteRT-LM C API. by Mohammadreza Heydary · 10 days ago
- 3c04cd3 fix(cli): delete automatic model conversion from run command by Wai Hon Law · 10 days ago
- 615d193 Retain linear attention state buffers as inputs for GPU-optimized in-place mode. by Fengwu Yao · 10 days ago
- 8695dcf workspace config update. by Mohammadreza Heydary · 10 days ago
- fd68028 Skip CPU global causal attention mask initialization and fill when pruned by the GPU delegate. by Fengwu Yao · 10 days ago
- 556f824 - Ensure vision encoder input and output buffers are lazily allocated before execution in VisionLiteRtCompiledModelExecutor. by Google AI Edge · 10 days ago
- 3cbf770 - Ensure vision encoder input and output buffers are lazily allocated before execution in VisionLiteRtCompiledModelExecutor. by Salil Tambe · 10 days ago
- f76896c Automated Code Change by Amer Elsheikh · 10 days ago
- 46de5ea Add LiteRtLogitMaskRunner for accelerated logit masking. by Mohammadreza Heydary · 11 days ago
- 403fbee feat(model): add LFM2 multimodal metadata proto configuration by Salil Tambe · 11 days ago
- dc8facf Clean up LiteRT CMake shim and update SentencePiece build configuration. by Google AI Edge · 11 days ago
- 3ced337 Adds ImagePreprocessor::Preprocess(DecodedImage, ...) overload to accept pre-decoded pixel buffers directly. by Google AI Edge · 12 days ago
- 5e3bd63 Grow the NPU dynamic KV cache on demand during prefill and decode. by Matt Kreileder · 13 days ago
- 2f8284d Update dependencies of litert_lm by Byungchul Kim · 14 days ago
- 2fa7506 This is an internal change by Google AI Edge · 14 days ago
- aa2401b Fix Safari crash by splitting JSPI and Asyncify builds in LiteRT-LM Web by Google AI Edge · 14 days ago