LiteRtLmSharp.Extensions.AI 1.2.0

Prefix Reserved
dotnet add package LiteRtLmSharp.Extensions.AI --version 1.2.0
                    
NuGet\Install-Package LiteRtLmSharp.Extensions.AI -Version 1.2.0
                    
This command is intended to be used within the Package Manager Console in Visual Studio, as it uses the NuGet module's version of Install-Package.
<PackageReference Include="LiteRtLmSharp.Extensions.AI" Version="1.2.0" />
                    
For projects that support PackageReference, copy this XML node into the project file to reference the package.
<PackageVersion Include="LiteRtLmSharp.Extensions.AI" Version="1.2.0" />
                    
Directory.Packages.props
<PackageReference Include="LiteRtLmSharp.Extensions.AI" />
                    
Project file
For projects that support Central Package Management (CPM), copy this XML node into the solution Directory.Packages.props file to version the package.
paket add LiteRtLmSharp.Extensions.AI --version 1.2.0
                    
#r "nuget: LiteRtLmSharp.Extensions.AI, 1.2.0"
                    
#r directive can be used in F# Interactive and Polyglot Notebooks. Copy this into the interactive tool or source code of the script to reference the package.
#:package LiteRtLmSharp.Extensions.AI@1.2.0
                    
#:package directive can be used in C# file-based apps starting in .NET 10 preview 4. Copy this into a .cs file before any lines of code to reference the package.
#addin nuget:?package=LiteRtLmSharp.Extensions.AI&version=1.2.0
                    
Install as a Cake Addin
#tool nuget:?package=LiteRtLmSharp.Extensions.AI&version=1.2.0
                    
Install as a Cake Tool

LiteRtLmSharp

.NET 10 bindings for Google's LiteRT-LM — on-device LLM inference (e.g. Gemma) for any .NET app, including MAUI. No server, no cloud: the model runs locally on CPU or GPU.

Install

Install the managed package plus the native runtime package for your platform, always with the same version number:

<PackageReference Include="LiteRtLmSharp" Version="1.2.0" />
<PackageReference Include="LiteRtLmSharp.runtime.win-x64" Version="1.2.0" />

Optional integrations: LiteRtLmSharp.Extensions.AI (a Microsoft.Extensions.AI.IChatClient — works with the Microsoft Agent Framework, Semantic Kernel and plain MEAI) and LiteRtLmSharp.SemanticKernel (an IChatCompletionService).

First tokens

using LiteRtLmSharp;

using var engine = LiteRtEngine.Load(new LiteRtEngineOptions
{
    ModelPath = "gemma-4-E2B-it.litertlm",   // from huggingface.co/litert-community
    Backend = LiteRtBackend.Cpu,              // or .Gpu
    MaxNumTokens = 4096,                       // total context window
});

using var chat = engine.CreateConversation();

await foreach (var chunk in chat.SendStreamingAsync("Tell me a joke"))
    Console.Write(chunk.Text);

Features

  • Chat: blocking, awaitable + cancellable, and streaming sends
  • Function calling with constrained decoding for reliable JSON arguments
  • Reasoning mode (Gemma "thinking"), surfaced separately from the answer
  • Multimodal input: image and audio attachments
  • Conversation restore & clone, token counting, speculative decoding, benchmarking
  • AOT- and trim-compatible (source-generated P/Invoke)

Documentation

License and trademarks

Apache-2.0. This is an unofficial, community-maintained project — not affiliated with, sponsored, or endorsed by Google. LiteRT, LiteRT-LM and Gemma are trademarks of Google LLC. The native binaries are built from LiteRT-LM source (Apache-2.0) at pinned release tags.

Product Compatible and additional computed target framework versions.
.NET net10.0 is compatible.  net10.0-android was computed.  net10.0-browser was computed.  net10.0-ios was computed.  net10.0-maccatalyst was computed.  net10.0-macos was computed.  net10.0-tvos was computed.  net10.0-windows was computed. 
Compatible target framework(s)
Included target framework(s) (in package)
Learn more about Target Frameworks and .NET Standard.

NuGet packages (1)

Showing the top 1 NuGet packages that depend on LiteRtLmSharp.Extensions.AI:

Package Downloads
LiteRtLmSharp.SemanticKernel

Microsoft Semantic Kernel connector (IChatCompletionService) for LiteRtLmSharp on-device LLM inference — a thin layer over the LiteRtLmSharp.Extensions.AI IChatClient. Companion to the LiteRtLmSharp package. Unofficial community bindings, not affiliated with or endorsed by Google. LiteRT is a trademark of Google LLC.

GitHub repositories

This package is not used by any popular GitHub repositories.

Version Downloads Last Updated
1.2.0 135 9/5/2026
1.1.1 178 7/18/2026
1.1.0 128 7/17/2026
1.0.0 143 7/2/2026

1.2.0: native binaries move to Google's official LiteRT-LM v0.16.0 C API prebuilts (one monolithic library per platform: linux-x64 now needs the system Vulkan loader (libvulkan1), win-x64 no longer needs the VC++ Redistributable, Android keeps GPU sampling with the embedded sampler, iOS ships the official CLiteRTLM.xcframework). Constrained decoding works on linux-x64 (guard removed); text LoRA is applied on LoRA-enabled bundles; new EnableYnnpack engine option; NoRepeatNgramSize / SuppressTokens on the MEAI and Semantic Kernel options. Also carries the v0.15.0 cycle: per-send decoding controls (penalties, no-repeat-ngram, suppress tokens, thinking budget, regex/JSON-schema constraints), the LlGuidance constraint provider, the KV overflow guard recalibrated to the new prefill planning, and Clone() now requiring an advanced conversation. The managed assembly requires the same-version runtime packages (LiteRT-LM v0.16.0). Full notes: https://github.com/OrihuelaConde/LiteRtLmSharp/blob/master/CHANGELOG.md