MLXHub: Local AI & LLM Server

Juan Colilla · Developer Tools

Free
View on App Store ↗

Chart standing · US

Top Grossing Developer Tools
#95
Best #28 ▼ 31

What changed · US

Listing changes AppTracker has observed for this app in this storefront — new versions, price moves, rating shifts, and store copy.

Aug 26, 2026
New version 1.0.4 → 2.0.0

Your personal AI platform reaches version 2.0.0 just two months after launch! We have made a huge number of changes so you can enjoy the app more, without having to make your life complicated trying to understand it. To begin with, the interface has been polished to fit more information into the same space. We have represented concepts visually to reduce cognitive load and improved other areas that needed it. The app now comes to life not only through its orbs —rendered entirely with Metal— but also through more refined animations. An experimental feature has been added: distributed inference! You can now run larger models that do not fit on your device. Have an iPad sitting in a drawer? Add it to the mesh. Your partner’s iPhone? Add that too, and connect devices like puzzle pieces to build an increasingly capable inference server. We admit the previous value proposition was not clear: the MLXHub Plus subscription offered very little, but that changes now. MLXHub Plus now includes LAN Server, distributed-inference hosting, unlimited models, app customization, and future features that will continue adding value. Do not want to subscribe? No problem: there is also a one-time purchase! Version 2.0.0 includes many under-the-hood changes that you will not easily notice, such as fixes for most VLMs that sometimes caused issues, automatic cleanup of partial files generated by the app, improvements to model downloads, more detailed model pages, and many framework improvements that make MLXHub the best app for carrying cutting-edge AI in your pocket: without internet, without model removals, and without the fear that some orange-haired character may feel like blocking your access from an armchair with paper and pen. Enjoy these new features and the immense number of changes still to come. Do not hesitate to contact us with any question or suggestion. Best regards, Juan

Screenshots

Keywords this app targets

Search terms AppTracker associates with this app — estimates, not search volume. ❝ = in the app's title; a highlighted number is its current position for that term.

Similar apps

Competitors and close alternatives, as estimated by AppTracker.

About

Run AI locally. Your device. Your data. MLXHub brings open-source language models to iPhone and iPad, powered by Apple Silicon and the mlx-swift inference engine. Your conversations never leave your device. Distributed inference across your devices (Experimental) Pool memory from several Apple Silicon devices on the same Wi-Fi network to run a model that none of them could hold alone. One device hosts; the others join and each takes a slice of the model's layers — no device stores the whole thing. The host sends each device only the files its own layers need, and caches them so the next session starts faster. What to expect: every device must be on the same Wi-Fi, you approve each device once by matching a code on both screens, and MLXHub measures each device's memory before it splits anything. This is an experimental feature — distributed generation works, but a run can still stall or drop a device, and not every model architecture can be split yet. It pays off most with large models. Joining a mesh is free on any device; hosting one is part of MLXHub Plus. Chat with powerful models Download and run LLMs and vision-language models (VLMs) directly on your device. Send text messages or attach photos — the model sees and responds without touching any server. Apple Intelligence built in On supported devices (iPhone 15 Pro / 16 and later with Apple Intelligence enabled), use Apple's on-device model for instant, private responses alongside any downloaded model. Your personality, your assistant Set a global Agent Personality to define the AI's tone and style. Override it per-conversation with custom system instructions. No prompt engineering needed — just describe what you want. Browse and install models Explore a curated catalog of optimized models — from tiny 0.6B models that fit in under 1 GB to powerful 7B models for deeper reasoning. Color-coded RAM badges tell you at a glance whether a model fits your device. Search HuggingFace directly to install any compatible model. Local LAN server Turn your iPhone or iPad into a portable, chat API-compatible inference endpoint on your local network. MLXHub's optional LAN server exposes /v1/chat/completions so any app — from a Mac running OpenCode to a custom script — can use your device's models. No internet required. Auth-protected. Bonjour-discoverable. Part of MLXHub Plus. Built for Apple Silicon MLXHub uses mlx-swift, Apple's own machine-learning framework, to run models at full Metal GPU speed. Model weights are quantized (4-bit, 8-bit) for the best quality-per-GB ratio on iPhone and iPad hardware. Privacy Model inference happens on your device and never leaves it. No account. No ads. The optional LAN server only listens on your local network and is off by default. MLXHub collects a small amount of anonymous diagnostic data to find crashes and broken model loads; it is not linked to you, and you can turn it off in Settings. Terms of Use (EULA): https://www.apple.com/legal/internet-services/itunes/dev/stdeula/

What's new · 2.0.0

Your personal AI platform reaches version 2.0.0 just two months after launch! We have made a huge number of changes so you can enjoy the app more, without having to make your life complicated trying to understand it. To begin with, the interface has been polished to fit more information into the same space. We have represented concepts visually to reduce cognitive load and improved other areas that needed it. The app now comes to life not only through its orbs —rendered entirely with Metal— but also through more refined animations. An experimental feature has been added: distributed inference! You can now run larger models that do not fit on your device. Have an iPad sitting in a drawer? Add it to the mesh. Your partner’s iPhone? Add that too, and connect devices like puzzle pieces to build an increasingly capable inference server. We admit the previous value proposition was not clear: the MLXHub Plus subscription offered very little, but that changes now. MLXHub Plus now includes LAN Server, distributed-inference hosting, unlimited models, app customization, and future features that will continue adding value. Do not want to subscribe? No problem: there is also a one-time purchase! Version 2.0.0 includes many under-the-hood changes that you will not easily notice, such as fixes for most VLMs that sometimes caused issues, automatic cleanup of partial files generated by the app, improvements to model downloads, more detailed model pages, and many framework improvements that make MLXHub the best app for carrying cutting-edge AI in your pocket: without internet, without model removals, and without the fear that some orange-haired character may feel like blocking your access from an armchair with paper and pen. Enjoy these new features and the immense number of changes still to come. Do not hesitate to contact us with any question or suggestion. Best regards, Juan

Details

DeveloperJuan Colilla
ReleasedJul 2026
UpdatedJul 2026
Version2.0.0
Size50 MB
RequiresiOS 26.0 or later
Age rating9+
LanguagesChinese, English, French, German, Hindi, Japanese, Russian, Spanish
PriceFree
Categories Developer ToolsProductivity
Bundle IDcom.DreamFoundries.MLXHub