Sign in Sign up
pub

local_ai_llama_cpp

llama.cpp adapter for LocalAI Kit: runs any GGUF model as a LocalLlm (streaming, GBNF-constrained structured output) and as a LocalEmbedding, with all FFI work isolated in worker isolates.

pub View on Pub

Activity

Latest release
1w ago
Total releases
1
Cadence
Last 12 months
1

Details

First release
Sep 04, 2026
Releases
Version Released
0.0.3 initial