local_ai_llama_cpp 0.0.3 copy "local_ai_llama_cpp: ^0.0.3" to clipboard
local_ai_llama_cpp: ^0.0.3 copied to clipboard

llama.cpp adapter for LocalAI Kit: runs any GGUF model as a LocalLlm (streaming, GBNF-constrained structured output) and as a LocalEmbedding, with all FFI work isolated in worker isolates.

Changelog #

0.0.3 #

  • First pub.dev release of the llama.cpp adapter, with streaming GGUF chat, embeddings, persistent KV-cache planning and GBNF-constrained output.
  • Add model-family chat templates and bring-your-own native library support.

0.0.2 #

  • Initial release of the llama.cpp adapter: LlamaCppLlmAdapter (any GGUF chat model, streaming, GBNF-constrained structured output), LlamaCppEmbeddingAdapter (the first LocalEmbedding implementation in the kit) and LlamaCppAdapterPlugin under provider key llama-cpp.
0
likes
140
points
59
downloads

Documentation

API reference

Publisher

unverified uploader

Weekly Downloads

llama.cpp adapter for LocalAI Kit: runs any GGUF model as a LocalLlm (streaming, GBNF-constrained structured output) and as a LocalEmbedding, with all FFI work isolated in worker isolates.

Repository (GitHub)
View/report issues

License

Apache-2.0 (license)

Dependencies

flutter, llama_cpp_dart, local_ai_core

More

Packages that depend on local_ai_llama_cpp