Fold a model into a single filament.
LiLo compresses nano-prompts, webMCP tooling, and AI Digital Twin workflows into models up to 100× lighter — so full intelligence runs on the smallest edge.
- 100×
- lighter
- 1.4k
- params
- 9.2k
- tok/s
- 132
- ms p50
$ lilo run --twin --edge
→ nano-prompt folded · 1.4k params
→ webMCP bound · 41 tools
→ digital twin online · 132ms p50
✓ 100× lighter · 9.2k tok/s
model
lilo-nano
twin
online
tools
41
Capabilities
One dense filament, three folded systems.
Nano-prompts
Prompts distilled to their load-bearing tokens, so a full intent ships in under 40 bytes.
webMCP
A model-control plane that binds 41+ web tools as a single compact control surface.
AI Digital Twin
A live, lightweight replica of your workflow that reasons at the edge without a datacenter.
Efficiency
100× lighter, engineered to the byte.
We measure the whole stack — params, tokens per second, time-to-first-token. LiLo holds the line on all three, on a 2-core edge box.
frontier
Bench · 2-core edge · 8GB RAM · 2026-Q1
Pricing
Priced per filament, not per token.
Pro
For a lean product team.
$49 /mo
- Unlimited nano-prompts
- 41 webMCP tools
- 50k twin runs / mo
- Edge + regional deploy
Waitlist
Reserve your filament.
Early access opens in cohorts. Leave your details and we'll fold you into the next build of LiLo by KAADE CONSTRUCTION.
Cohort 03 · opens 2026-Q2
Priority to edge-only teams