Packages for keyword “local-llm”
These packages are available as a package collection, usable in Xcode or SwiftPM.
AgentRunKit
Swift 6 agent SDK: type-safe tools, streaming, cloud + on-device inference via MLX on Apple Silicon
coreai-kit
Swift SDK for running chat, vision and speech models on iPhone and Mac with Apple's Core AI. Model download and caching, FoundationModels integration, and runnable examples with documented OS, SDK and model requirements.
AIChatKitMLX
Adds on-device Apple MLX inference to any app already using AIChatKit. Models are downloaded from Hugging Face Hub on first use and cached locally. Runs on Metal GPU and Apple Neural Engine — no network calls during inference.
AIChatKitLlama
Adds on-device GGUF inference via llama.cpp to any app already using AIChatKit. Models run entirely in-process using Metal GPU acceleration — no network calls after the initial download.
VeloxQuant
Swift SDK for VeloxQuant (Apple Silicon MLX KV-cache compression) — sibling to the TS, Go, Rust, and Kotlin client SDKs.
swift-llama
Run LLMs natively in Swift. Actor-safe. Streaming. Tool-calling. Metal-accelerated.
6 packages.