Most product questions are predictable: “open billing,” “what’s this?,” “why can’t
I save?” NotLM is a frontline chatbot for your SPA that captures that bulk of
traffic with a local smart cache — packed intents, FAQs, and catalogs — so replies
stay snappy and LLM spend stays on the rare misses. Optional smaller-model and full
LLM fallbacks still cover novel asks. Spotlights and guided tours are one pattern
you can pack; the point is a chat that understands your app.
Build
Open NotLM runtime (@notlm/*): deterministic pack NLU as the hot
path, React host SDK, optional ONNX ranker and Laya decision fallback, then
optional host LLM. A separate offline training repo turns miss logs into better
pack aliases and intents without rewriting the sealed runtime customers ship.
What it demonstrates
LLM-like product chat with local-first cost and latency: FAQ, goto, query,
confirm-gated mutations, context blockers, tours, search, compare, handoff,
audit, clean OOD refusal, and disambiguation — UI actions only (navigate /
spotlight / click guide IDs; never posting into product APIs).
Source
Runtime (notlm)
·
Training (notlm-training)
TypeScript
React
Vite
ONNX
Playwright
Node.js