Run Phi-4 Mini 3.8B locally on iPhone, iPad & Mac

Generalists · 2.49 GB download · needs 6 GB RAM · Q4_K_M GGUF

Microsoft model that punches above its size. Recent phones.

In privateSLM, Phi-4 Mini 3.8B is an everyday assistant for chat, writing, summarising and quick questions. You download the model once, and from then on every conversation runs entirely on your device — no cloud, no account, no tracking, and it keeps working with the radio off.

At a glance

ModelPhi-4 Mini 3.8B
CategoryGeneralists
Download size2.49 GB (one time)
QuantizationQ4_K_M (GGUF)
Minimum device RAM6 GB
Runs offlineYes — fully on-device after download
SourceHugging Face

Will it run on my device?

Phi-4 Mini 3.8B needs a device with at least 6 GB of RAM — that means iPhone 15 Pro and later, M-series iPads, and any Apple Silicon Mac. The privateSLM catalog states the RAM requirement for every model up front, so you know before you download.

How to run Phi-4 Mini 3.8B in privateSLM

  1. Get privateSLM on the App Store — one-time purchase, no subscription.
  2. Open Models, pick Phi-4 Mini 3.8B, and tap Download (2.49 GB, once).
  3. Select it as your active model and chat — airplane mode included.

On iPhones, iPads and Macs with Apple Intelligence, privateSLM can also answer instantly with the built-in system model — zero downloads. A catalog model like Phi-4 Mini 3.8B is for when you want this particular specialist, offline, under your control.

Related models

FAQ

Does Phi-4 Mini 3.8B run offline on iPhone?

Yes. After a one-time 2.49 GB download inside privateSLM, Phi-4 Mini 3.8B runs entirely on-device via llama.cpp. Airplane mode works — nothing is ever sent to a server.

How much RAM does Phi-4 Mini 3.8B need?

A device with at least 6 GB of RAM — in practice iPhone 15 Pro and later, M-series iPads, and any Apple Silicon Mac.

Is Phi-4 Mini 3.8B free to use in privateSLM?

The model itself is open-weights (Q4_K_M GGUF) and the download is free. privateSLM is a one-time purchase — no subscription, no per-message fees.