
LocalEngine: Think Tank
Private on-device AI runtime
0 ratings
Free
About
LocalEngine runs open-source AI language models entirely on your iPhone and iPad. Everything happens on-device: your prompts, conversations, and images never leave your device and are never sent to any server.
Built in pure Swift with SwiftUI, LocalEngine runs models through a Metal-accelerated llama.cpp engine, so you get fast, low-latency AI completely offline — no account, no sign-in, no cloud.
WHAT YOU CAN DO
• Chat — Hold multi-turn conversations with a local model. Replies stream token-by-token in real time. Start a new chat or stop a response at any time.
• See images — Turn on vision in Settings, attach photos to your message, and ask about them. Images are understood entirely on-device by a vision-capable model.
• Download models — Browse a curated catalog of open models and download the one that fits your device. Downloads run as resumable background jobs with live progress, and the screen can stay awake while a large model downloads.
• Tune generation — Adjust the response length (max tokens) and the temperature to shape how the model replies.
• Unlock larger models — LocalEngine is free to use with starter models. A one-time Pro upgrade unlocks larger catalog downloads on iPhone and iPad.
CURATED MODELS
LocalEngine is ready to chat out of the box with a small, fast default model, and its catalog includes compact vision-language models you can grow into:
• Qwen3.5 — 0.8B and 2B free, with 4B available through Pro
• Gemma 4 — E2B free, with E4B available through Pro
Smaller models are quick and light on storage; larger Pro models are more capable on newer devices. Each vision model pairs with a multimodal projector so it can understand images — all on-device.
PRIVATE BY DESIGN
• 100% on-device inference — no cloud, no telemetry, no analytics.
• No account and no sign-in required.
• Runs fully offline once a model is downloaded.
• Built to App Store privacy and security standards.
REQUIREMENTS
• iOS / iPadOS 17.0 or later
• A recent iPhone or iPad is recommended, especially for the larger models
• Free storage space for the models you choose to download
LocalEngine puts the power of modern AI — text and vision — in your pocket, without giving up your privacy.
Show more
+1
What's New in LocalEngine
0.4.16
July 30, 2026
Local embeddings — select a dedicated GGUF embedding model independently from your chat model, ready for private semantic search and retrieval. • New embedding models — download all-MiniLM-L6-v2, BGE Small EN v1.5, or Qwen3 Embedding 0.6B directly on iPhone and iPad. • Better multilingual retrieval — Qwen3 Embedding 0.6B is recommended for Chinese and other multilingual content. • Clearer model management — chat and embedding models are now shown separately, so choosing embeddings never interrupts an active chat model. Everything still runs 100% on-device — no account, no cloud, no telemetry.
More



