MLXHub: Local AI & LLM Server

MLXHub: Local AI & LLM Server

On-device LLM, OpenAI API

0 ratings
Free

Rating summary

Details

  • Released
  • Updated
  • July 14, 2026
  • August 1, 2026

Features

MLXHub: Local AI & LLM Server screenshot #1 for iPhone
MLXHub: Local AI & LLM Server screenshot #2 for iPhone
MLXHub: Local AI & LLM Server screenshot #3 for iPhone
MLXHub: Local AI & LLM Server screenshot #4 for iPhone
MLXHub: Local AI & LLM Server screenshot #5 for iPhone
iphone
ipad
🖼️Get Icon
Icons↘︎

About

Run AI locally. Your device. Your data. MLXHub brings open-source language models to iPhone and iPad, powered by Apple Silicon and the mlx-swift inference engine. No cloud. No data leaves your device. Chat with powerful models Download and run LLMs and vision-language models (VLMs) directly on your device. Send text messages or attach photos — the model sees and responds without touching any server. Apple Intelligence built in On supported devices (iPhone 15 Pro / 16 and later with Apple Intelligence enabled), use Apple's on-device model for instant, private responses alongside any downloaded model. Your personality, your assistant Set a global Agent Personality to define the AI's tone and style. Override it per-conversation with custom system instructions. No prompt engineering needed — just describe what you want. Browse and install models Explore a curated catalog of optimized models — from tiny 0.6B models that fit in under 1 GB to powerful 7B models for deeper reasoning. Color-coded RAM badges tell you at a glance whether a model fits your device. Search HuggingFace directly to install any compatible model. Local LAN server Turn your iPhone or iPad into a portable, OpenAI-compatible inference endpoint on your local network. MLXHub's optional LAN server exposes /v1/chat/completions so any app — from a Mac running Continue.dev to a custom script — can use your device's models. No internet required. Auth-protected. Bonjour-discoverable. Built for Apple Silicon MLXHub uses mlx-swift, Apple's own machine-learning framework, to run models at full Metal GPU speed. Model weights are quantized (4-bit, 8-bit) for the best quality-per-GB ratio on iPhone and iPad hardware. Privacy first No analytics. No tracking. No ads. No account required. Model inference never leaves your device. The optional LAN server only listens on your local network and is off by default. Terms of Use (EULA): https://www.apple.com/legal/internet-services/itunes/dev/stdeula/
Show more

What's New in MLXHub

1.1.0

August 1, 2026

MLXHub now has a more refined visual experience throughout the app, with new activity orbs, improved contrast, and theme color options for Plus. We also polished animations and transitions so chat, models, and the LAN server feel smoother. This update improves performance and responsiveness during generation, scrolling, and server use.

More

Developer apps