Skip to content
Sentrize is an AWS Advanced Tier Partner

← All products

Developer tool

NanoLM

Run local AI models with one binary and OpenAI-compatible APIs.

Download coming soon
Talk to engineering

v1.0 · Beta · Windows · macOS · Linux

About

A lightweight local AI runtime written in Go. NanoLM manages GGUF models, serves an OpenAI-compatible HTTP API, and wraps llama.cpp — so any tool that speaks the OpenAI API can run against models on your own hardware.

Highlights

✓ Single binary — no Python environment, no dependency hell

✓ OpenAI-compatible API: drop-in for existing SDKs and tools

✓ Downloads, verifies, and manages GGUF models

✓ Powers Sentrize Coder and any local-AI workflow

Built with

Gollama.cppGGUF

Questions about this product?

We answer fast — licensing, roadmap, enterprise deployment, or anything else. Contact us.