Private Document Intelligence

Private AI for documents and internal knowledge.

Turn documents into structured data and internal knowledge into cited answers, on an engine that also runs agents, vision, speech, text analysis and generation. One platform, on your hardware. The files, the indexes and the inference never leave your network.

LM-Kit One: Business Preview LM-Kit.NET: .NET 8 / 9 / 10 Platforms: Windows · Linux · macOS
What ships in the box

Seven pillars, one foundation.

LM-Kit.NET ships seven pillars and the local runtime they all sit on. Use the parts you need, ignore the rest.

The foundation

Every capability above runs on this runtime.

Foundation

Local Inference

The runtime all seven pillars sit on. The LM-Kit.NET NuGet ships the complete inference system: open-weight LLMs, vision-language models, embeddings, on-device speech-to-text, OCR and classifiers, accelerated on CPU, AVX2, CUDA 12/13, Vulkan or Metal. One package, zero cloud calls, predictable latency, full data and technology sovereignty.

Explore the foundation
Core technology

Dynamic Sampling, the symbolic layer.

The reason a small local model can hold its own on extraction, classification and structured generation. Dynamic Sampling is an adaptive inference engine that sits underneath every LM-Kit call, steering each token with structural awareness, contextual signals and grammar-aligned validation. It works on any model as it ships, and it keeps working on one you have fine-tuned yourself.

Pillar A

Constrained output

Dynamic grammar guarantees JSON, schemas, and tool-call shapes always parse. A novel hybrid path runs roughly twice as fast as classical grammar sampling.

Pillar C

Model-agnostic

No architecture coupling and no per-model adapter required. Drop in a new open-weight release, or your own fine-tune, and the layer keeps working from day one.

Open the Dynamic Sampling deep dive →

Runs where your code already runs

Same process, same threads, same deploy.

No sidecar service, no special runtime. LM-Kit links into your application, picks up the right native acceleration for the host, and gets out of the way.

Runtime
.NET Standard 2.0 · .NET 8 · 9 · 10
OS
Windows, Linux x64 & ARM64, macOS
Acceleration
CPU, AVX/AVX2, CUDA 12/13, Vulkan, Metal
Models
Gemma 3, Qwen 3, Llama, Phi-4, GLM 4.7, GPT OSS, Whisper, embeddings
Storage
In-memory, built-in vector DB, Qdrant, pgvector, bring-your-own
Bridges
Microsoft.Extensions.AI, Semantic Kernel, MCP clients
Licensing

Free for small teams. Commercial above the line.

One license, both products, no activation key. Nothing checks a license at runtime. Evaluation and development are free at any company size, with no time limit.

Free

$0no key, no expiry

The complete SDK and the complete server, including commercial use and redistribution, for small companies. Evaluation and development stay free at any size.

  • Under $1M USD annual gross revenue
  • 10 or fewer employees
  • No more than $3M USD raised from outside investors
  • Always free: personal, education, nonprofits, open source

Professional

Customannual, scaled to scope

Required above the thresholds. Scaled to deployment size, never metered by tokens, seats or end users. Carries the assurance a production deployment needs.

  • Commercial use and redistribution at any scale
  • Long-term support builds and security patches
  • Support with response-time commitments
  • Unlimited developers and end users
What customers say 4.9 / 5 on SourceForge

Get started

Run it on your own documents.