AppCorpus
Market US

Private LLM - Local AI Chat

by Numen Technologies Limited

iPhone · iPad

Active Last checked 31 Jul 2026

Quick verdict

What it is
Store information collected; independent summary not yet available.
At a glance
  • $4.99

Categories and search queries

Official categories

AppCorpus categories

Key functions

Function Availability Access Scope Evidence
Family Sharing Yes Unknown App Store Store published

Screens

No screenshots collected yet.

Pricing & paywalls

App Store

  • Price: $4.99

Ratings & reviews

App Store

4.1 695 ratings · App Store

  1. 4/55 Stars for these Features!

    Fantastic app, I like the selection of models and they run great on the M4 chip of the iPhone 16 Max and iPad Pro. System prompt is also great for some tweaks to responses. The most simple and important feature (aside from model memory) that Private LLM needs is to be able to edit the responses of the AI. I enjoy the chatbot's responses but hate that it'll pull context from previous responses and loop entire phrases regardless if I regenerate its response or edit mine. Sometimes I don't like how the model formatted its response and want to change it without clearing the convo and retrying multiple times. The second feature that would be a quality of life is to be able to save conversations within the app. At the very least, export them as a whole without needing to screenshot, grab text, paste, correct, and format it elsewhere. I learned early on that copying text blocks at a time can cause the app to reload its model and even clear the conversation. A bot refresh would also be nice to clear older message context during a longer conversation to keep the model fast and responsive, but that can coincide with model memory to more easily pull specific facts you want it to remember in the first place without having to create a queue of explanations to pick up where you left off at. Useful if you switch models for different purposes like organizing, summarizing, q&a vs. generation, editing, brainstorming, vs. chatting.

    OverwhelmingLatias · 31 Jul 2026

  2. 4/5Getting there!

    UPDATE: app is much faster now. Still waiting on a prompt library which would be awesome, but the speed and quality is increasing seemingly by the day. There was a small period of time where it didn’t work on the 13 mini but that has been fixed. I’m happy with the amount of work the devs are putting in and I’m hopeful for more features in the future. MLC chat is comparable in speed and quality but this app is looking to widen the gap by adding more features in the (hopefully near) future. Old review: In comparison to MLC chat (free and open source) this app is lacking. MLC chat offers the ability to install additional models, and the default is on par or better than this apps responses (like gpt 3 level, lots of incorrect information and it kinda goes wild when explaining things). Additionally, performance is abysmal comparatively for better quality responses on MLC chat. It’s a great idea but lacks the ability to fine tune. Being able to change the temperature and other settings for the model would be great. Would also be cool if we could create a starting prompt (or a library to choose from) for the AI so it can be used for things like role playing or DND. Kinda bummed that it’s $5 since I was expecting a lot more than MLC AIs app for the pricing. Instead I’ll be uninstalling and checking back later for new features and possible performance increases. If you’re reading this and thinking about buying it, you should probably check out MLC chat for now and just wait for more features and performance updates.

    PrivacySchizo · 31 Jul 2026

  3. 5/5So much potential

    I've been using this app for a couple of months at this point. When it was first released, it was a neat proof of concept, but it only supported 7B models and was overall just too simple to use effectively. With a recent update, a 13B model was released for all macs with 16GB of memory, and it makes such a huge difference! It's not quite at the same level as ChatGPT 3.5T, but it's close enough that I never use 3.5T anymore; this is my go-to. I greatly appreciate the on-device processing (hallelujah privacy!), and it doesn't even use too much power - my battery still lasts for hours and hours. The performance is also great; my base M1 Air powers right through the prompts. I only have two complaints about the app at this point. 1) The 13B model uses about 12GB of memory by itself, which does force the use of swap on a 16GB Air. Not much the dev can do about this, but it is something to keep in mind. You'll want to close out of other programs before launching this. 2) There still is no feature that has separate conversations. If you want to start a new conversation, you need to delete the existing conversation. I'd love it if we could get separate conversations in a future update; it would make this app so much easier to use. Overall, I love it and do not regret buying it at all. I can't wait to see what future updates bring :)

    N7nathan · 31 Jul 2026

  4. 5/5A lot of work went into this

    PLLM manages to make an iPhone feel like a MacBook with the way it handles quantization. I would say that if you’re on the fence about whether to buy this app, consider a few things. 1: ChatGPT Plus is $20/month. This is 25% of that and you have it forever. 2: if you’re looking for an experience that offers seamless tool calling, online searches, or agentic flows, you’re not going to find that locally on a smartphone. This app isn’t the limiting factor. 3: this app isn’t for you if what you care about is instant syncing; <1 second latency, and memory between chats. —- I’m sure the developer would agree when I say that this app isn’t for everyone. But it IS for anyone who knows what they need and, like me, has come to realize that no other iOS app currently offers anything like this. Locally. It goes without saying that I’m beyond satisfied with my purchase. I was shocked not only that buying it once gives it to my whole iCloud family group, but also includes the MacBook version. The M3 Pro simply annihilates even the iPhone 17 series. One very simple request: if it’s possible, having chat history would be really, really helpful. Even if that has to look like adding the ability to export & re-import JSON chat files.

    MicahBG99 · 31 Jul 2026

  5. 5/5Dope. Dangerous, but dope.

    The genie has been out the bottle since GPT 1 & 2 were open source and it’s never going back in. This app has been very useful for times when I have no reception like when I went to Red Rock in Nevada with my wife’s family, hypothesized that the rocks were likely red due to the fact that iron rusts that color and the rock is most likely iron rich. I had zero reception but because this runs locally, I was able to confirm that was the reason and share that fun fact with everyone. It’s like having all human knowledge in your phone. That said, Dolphin specifically is totally dangerous with the right pre-prompt in the wrong hands. But I believe in freedom so I still deem it dope overall. 5 well deserved stars.

    Awsome Laziness · 31 Jul 2026

Privacy & accessibility

App Store

Privacy

  • Data Not Collected Developer declared

Accessibility

  • Dark Interface — SUPPORTED Store published

Versions

App Store

  • 1.7.9

    Bug-fix release: Fix for issues with loading the builtin StableLM 2 1.6B model and stability fixes on older iOS devices.

  • 1.8.0

    - Support for downloading the new Dolphin 2.9 Llama 3 8b model. If you have any feedback or questions, we would love to hear from you! Numen Technologies offers free tech support; you can email us at [email protected], message us on Discord, or tweet at us @private_llm. If you find Private LLM useful, we would appreciate a review on the App Store. Your review will help others discover Private LLM.

  • 1.8.1

    - Support for downloading the new Phi-3-mini-4k-instruct model. - Support for downloading the Llama 3 based Smaug-8B model. - Stability improvements and bug fixes. If you have any feedback or questions, we would love to hear from you! Numen Technologies offers free tech support; you can email us at [email protected], message us on Discord, or tweet at us @private_llm. If you find Private LLM useful, we would appreciate a review on the App Store. Your review will help others discover Private LLM.

  • 1.8.2

    - Support for downloading an improved version of the new Phi-3-mini-4k-instruct model with an unquantized embedding layer. - The old Phi-3-mini-4k-instruct model has been deprecated, and will continue to be functional for the next two releases. - Fixed bug where the "+" character was elided from prompts when Private LLM is invoked from iOS Shortcuts. - Stability improvements and bug fixes. If you have any feedback or questions, we would love to hear from you! Numen Technologies offers free tech support; you can email us at [email protected], message us on Discord, or tweet at us @private_llm. If you find Private LLM useful, we would appreciate a review on the App Store. Your review will help others discover Private LLM.

  • 1.8.3

    - Support for downloading a 3-bit OmniQuant quantized version of the Llama 3 8B based OpenBioLLM-8B model. - Support for downloading a 3-bit OmniQuant quantized version of the Hermes 2 Pro - Llama-3 8B model. - Support for downloading a 3-bit OmniQuant quantized version of the bilingual (Hebrew, English) DictaLM-2.0-Instruct model. - Users on iPhone 11, 12, 13 Pro, Pro Max devices can now download the faster and older fully quantized version of the Phi-3-Mini model. - Private LLM now uses the loaded model's default system prompt if the system prompt is blank when invoked from app intents (Siri and Shortcuts). - Fixed a bug where temperature and top-p settings were not being persisted across app restarts. - Stability improvements and bug fixes.

  • 1.8.4

    - Support for downloading a 4-bit OmniQuant quantized version of the new Phi-3-Mini based kappa-3-phi-abliterated model on all devices with 6GB or more RAM. - Stability improvements and bug fixes.

  • 1.8.5

    - Support for downloading 9 new models (support varies by device capabilities). - 3-bit OmniQuant quantized version of Mistral 7B Instruct v0.3 - 3-bit OmniQuant quantized version of Meta-Llama-3-8B-Instruct-abliterated-v3 - 3-bit OmniQuant quantized version of Llama-3-8B-Instruct-MopeyMule - 3-bit OmniQuant quantized version of openchat-3.6-8b-20240522 - 3-bit OmniQuant quantized version of Llama-3-WhiteRabbitNeo-8B-v2.0 - 3-bit OmniQuant quantized version of Hermes-2-Theta-Llama-3-8B - 3-bit OmniQuant quantized version of LLaMA3-iterative-DPO-final - 3-bit OmniQuant quantized version of Hathor_Stable-v0.2-L3-8B - 3-bit OmniQuant quantized version of NeuralDaredevil-8B-abliterated - Minor UI improvements - Stability improvements and bug fixes. Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.8.6

    - Support for downloading 4 new models. Two models from the new Meta Llama 3.1 family of models and two Meta Llama 3 based models (Support varies by device capabilities). - 3-bit OmniQuant quantized version of the Meta Llama 3.1 8B Instruct model. - 3-bit OmniQuant quantized version of the Meta Llama 3.1 8B Instruct abliterated model. - 3-bit OmniQuant quantized version of the Llama 3 based L3 Umbral Mind RP v3.0 model. - 3-bit OmniQuant quantized version of the Llama 3 based Llama 3 Instruct 8B SPPO Iter3 model. - Stability improvements and bug fixes. Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.8.7

    - Support for downloading 2 new models from the Gemma 2 family of models (on all devices with 4GB or more RAM). - 4-bit OmniQuant quantized version of the gemma-2-2b-it model. - 4-bit OmniQuant quantized version of the multilingual SauerkrautLM-gemma-2-2b-it model. - Stability improvements and bug fixes. Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.8.8

    - Fix for a non-deterministic crash while downloading Gemma 2B based models on older devices with 4GB of RAM. Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.8.9

    - Added support for downloading 4-bit Omniquant quantized version the new Llama 3.2 1B Instruct model (on all iOS devices). - Added support for downloading 4-bit Omniquant quantized version the new Llama 3.2 3B Instruct model (on devices with 6GB or more RAM). - Added support for downloading the unquantized version of the Llama 3.2 1B Instruct model (on devices with 6GB or more RAM). - Support for rendering Latex math formulas in LLM generated text. - Users can now copy debug information and also email our support address, from the help view in the app. Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.9.0

    - Added support for downloading 4-bit Omniquant quantized version the Llama 3.2 1B Instruct abliterated model (on all iOS devices). - Added support for downloading 4-bit Omniquant quantized version the Llama 3.2 3B Instruct abliterated model (on devices with 6GB or more RAM). - Added support for downloading 4-bit Omniquant quantized version the Llama 3.2 3B Instruct uncensored model (on devices with 6GB or more RAM). - Added support for downloading 4-bit Omniquant quantized version the Gemma 2 9B IT model (on M1/M2/M4 iPad Pros with 16GB of RAM). - Added support for downloading 4-bit Omniquant quantized version the Gemma 2 9B IT SPPO Iter3 model (on M1/M2/M4 iPad Pros with 16GB of RAM). - Added support for downloading 4-bit Omniquant quantized version the Tiger-Gemma-9B-v3 model (on M1/M2/M4 iPad Pros with 16GB of RAM). - Stability improvements and bug fixes. Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.9.1

    - Bugfix release: fix for crash while loading some of the older models that use the sentencepiece tokenizer. Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.9.10

    Minor compatibility fixes with iOS 26

  • 1.9.11

    - Support for the Qwen3-4B-Instruct-2507-heretic abliterated model (on any iOS device with 6GB or more RAM) - Support for the Qwen3-4B-Instruct-2507-heretic-noslop model (on any iOS device with 6GB or more RAM) - The noslop model has been specially tuned with abliterated to reduce LLM slop in its generated outputs and is exclusively available only on Private LLM - Minor bug fixes and updates Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to join our Discord, email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.9.12

    - Accessibility improvements - Minor bug fixes and updates Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to join our Discord, email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.9.15 · 9 Jul 2026

    - Faster model downloads. Models are now downloaded from a CDN instead of Huggingface. - Minor bug fixes and updates Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to join our Discord, email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.9.2

    - Support for downloading 8 new models. - Added support for downloading Qwen 2.5 family of models (0.5B-14B) - Added support for downloading Qwen 2.5 Coder family of models (0.5B-14B) - Support for individual models across both families of models varies by the amount of physical memory on devices. - Stability improvements and bug fixes. Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.9.3

    - Support for downloading 12 new models (varies by device capacity). - Hermes-3-Llama-3.2-3B and Hermes-3-Llama-3.1-8B models - FuseChat-Llama-3.2-1B-Instruct, FuseChat-Llama-3.2-3B-Instruct, FuseChat-Llama-3.1-8B-Instruct, FuseChat-Qwen-2.5-7B-Instruct and FuseChat-Gemma-2-9B-Instruct models - FuseChat-Llama-3.2-1B-Instruct also has an unquantized variant, downloadable on devices with 6GB or more RAM - EVA-D-Qwen2.5-1.5B-v0.0, EVA-Qwen2.5-7B-v0.1 and EVA-Qwen2.5-14B-v0.2 models - Llama-3.1-8B-Lexi-Uncensored-V2 model - Improved LaTeX rendering - Stability improvements and bug fixes. Thank you for choosing Private LLM. We are committed to continue improving the app and to making it more useful for you. For support requests and feature suggestions, please feel free to email us at [email protected], or tweet us @private_llm. If you enjoy the app, leaving an App Store is a great way to support us.

  • 1.9.4

    Bugfix release: Fix for crash while loading 14B models on iPad Pros with 16GB of RAM

  • 1.9.5

    * Support for downloading 7 new DeepSeek R1 Distill based models on Apple Silicon Macs. Support for individual models varies by device capabilities. * Users with Apple Silicon Macs with 16GB RAM can now download the phi-4 model (previously restricted to Apple Silicon Macs with 24 GB of RAM) * Minor bugfixes and updates.

  • 1.9.6

    - Added support for 8 new models from the Dolphin 3.0 family of models - Added support for the unquantized version of the Llama 3.2 1B Instruct Abliterated model - Added support for the 4-bit quantized Gemma 2 Ifable 9B creative writing model (downloadable on M-series iPad Pros with 16GB of RAM) - Context length is now displayed in the model quick switcher - Minor bug fixes and updates

  • 1.9.7

    - Added support for a 3-bit OmniQuant quantized version of the Llama-3.1-8B-UltraMedical model - Added support for a 3-bit OmniQuant quantized version of the Meta-Llama-3.1-8B-SurviveV3 survival specialist model - Added support for a 4-bit GPTQ quantized version of the Openhands 7B coding model - Added support for 4-bit QAT version of the Google Gemma3 1B IT model (32k ctx on iPhones with 6GB or more RAM, 8k on older iPhones with 4GB of RAM) - Added support for 4-bit OmniQuant quantized versions of the Google Gemma3 1B based gemma-3-1b-it-abliterated and amoral-gemma3-1B-v2 models - Many other minor bug fixes and updates

  • 1.9.8

    - Support for the new Qwen3 4B Instruct 2507 model (on any iOS device with 6GB or more RAM) - Minor bug fixes and updates

  • 1.9.9

    - Support for two Qwen3 4B Instruct 2507 based models: Qwen3 4B Instruct 2507 abliterated and Josiefied Qwen3 4B Instruct 2507 (on any iOS device with 6GB or more RAM) - Minor bug fixes and updates

Alternatives

Availability by market

Data quality & evidence

Page completeness
Live store listing
Last verified
31 Jul 2026
Sources
App Store

Report a correction