
Sencha: LLM Chat Client
OpenRouter & On-Device Models
0 ratings
Free
Rating summary
About
A native LLM client for on-device inference, LLM APIs, and self-hosted servers, built around a minimal, responsive interface.
Download a supported model and chat offline without an API key or account. Or connect OpenRouter, direct providers, and compatible endpoints using your own credentials.
On-Device Models
• Download models from the built-in catalog, including Gemma 4, Qwen 3.5, Bonsai, and LFM 2.5
• Choose model sizes and quantization options to suit your device
• Configure context window, reply length, and image detail where supported
• Use streaming, thinking traces, image attachments, and web tools with compatible models
• Reduce reprocessing of earlier messages with on-disk prompt caching
Model capabilities and performance depend on the selected model and available device memory.
Connections and Models
• Connect OpenRouter, OpenCode, and direct providers
• Add OpenAI- or Anthropic-compatible endpoints, including self-hosted runtimes such as Ollama and LM Studio
• Use Chat Completions, Responses API, or Messages API
• Choose an on-device or connected model per chat
• Browse model catalogs and inspect context size, output limits, pricing, and capabilities where available
• Configure reasoning effort, prompt caching, routing, and model tools where supported
Web Search and Fetch
• Use the included free web search via Markdown.new
• Or configure Brave Search, Tavily, Exa, Serper, SerpApi, Parallel, or Keenable
• Choose search-depth presets for Exa and Tavily
• Fetch pages on device or through Markdown.new or Firecrawl
• Configure fetch providers and fallbacks independently
Files and Output
• Attach photos, documents, and PDFs
• Send PDFs natively, extract embedded text, run on-device OCR, or render pages as images
• Read responses with formatted Markdown, code, mathematical formulas, and tables
• Copy generated tables as a table, Markdown, or a complete image
Voice Input
• Use Apple Speech for on-device dictation
• Download local transcription models
• Connect API-based transcription providers
The app is free, with no ads or app subscription. On-device models require no provider account or API key. Connected services are subject to their own pricing, limits, and data policies.
Show more
What's New in Sencha
1.2
September 14, 2026
Bug fixes and improvements. • A cleaner interface for choosing and downloading local models, with a clear description of each one. • Fixed a context issue with Gemma models. • Code blocks are now highlighted while the reply is still coming in. • New Feedback & Support in Settings. • Other fixes and performance improvements.
More









