Skip to content

Best Local AI Models for Mac

Every open-source model that runs on a Mac inside Private LLM — from compact models to the largest 70B-class ones your unified memory can hold, all fully offline.

132 models

Meta logo

Cat Llama 3 70B Instruct

70.6B

Immersive roleplay via strict prompt adherence, Llama 3 70B finetune

8K contextMac
DeepSeek logo

DeepSeek R1 Distill Llama 70B

70.6B

Reasoning model distilling DeepSeek R1 reasoning into Llama 3.3 70B

131K contextMac
Meta logo

EVA LLaMA 3.33 70B v0.1

70.6B

Roleplay and creative writing specialist based on Llama 3.3 70B

131K contextMac
Meta logo

Llama 3.3 70B Euryale v2.3

70.6B

Roleplay-optimized Llama 3.3 70B with flexible persona handling

131K contextMac
Meta logo

Meta Llama 3.3 70B Instruct

70.6B

General-purpose model in the Llama 3.3 70B family

Mac
Meta logo

Llama 3.3 70B Instruct Abliterated

70.6B

Uncensored Llama 3.3 70B with abliterated refusal mechanisms

131K contextMac
Meta logo

Meta Llama 3 70B Instruct

70.6B

General-purpose model in the Llama 3 70B family

Mac
Meta logo

Meta Llama 3.1 70B Instruct

70.6B

General-purpose model in the Llama 3.1 70B family

Mac
Meta logo

Smaug Llama 3 70B Instruct

70.6B

General conversationalist fine-tuned from Llama 3 70B, matching GPT-4 Turbo dialogue

8K contextMac
Meta logo

Smaug Llama 3 70B Instruct Abliterated v3

70.6B

Uncensored abliterated Llama 3 70B retaining original knowledge

8K contextMac
DeepSeek logo

R1 1776 Distill Llama 70B

70B

Reasoning-focused model in the DeepSeek R1 Distill family

Mac
Mistral logo

Dolphin 2.6 Mixtral 8x7B

56B

Uncensored Mixtral 8x7B finetune excelling at coding tasks

33K contextMac
Mistral logo

Nous Hermes 2 Mixtral 8x7B DPO

46.7B

General instruction-tuned Mixtral 8x7B model with creative text generation

33K contextMac
Mistral logo

Mixtral 8x7B Instruct v0.1

46.7B

General-purpose assistant outperforming Llama 2 70B

33K contextMac
01.AI logo

Yi 34B Chat

34.4B

Multilingual English-Chinese chat and translation using Yi 34B

4K contextMac
Meta logo

WizardLM 33B v1.0

33B

Uncensored conversational model based on Llama 33B with reduced refusal bias

2K contextMac
DeepSeek logo

DeepSeek R1 Distill Qwen 32B Abliterated

32.8B

Uncensored proof-of-concept removing alignment guardrails from DeepSeek R1 Distill

131K contextMac
Qwen logo

EVA Qwen2.5 32B v0.2

32.8B

Roleplay and storywriting finetune of Qwen 2.5 32B with flexible tone control

131K contextMac
DeepSeek logo

Fuse O1 DeepSeek R1 QwQ SkyT1 32B

32.8B

Reasoning model merging three chain-of-thought experts for math, coding, and science

131K contextMac
Qwen logo

OpenHands LM 32B v0.1

32.8B

Autonomous software issue resolution based on Qwen 2.5 32B

33K contextMac
Qwen logo

Qwen 2.5 32B

32.8B

General-purpose instruction model excelling at coding and mathematics

33K contextMac
Qwen logo

Qwen 2.5 Coder 32B

32.8B

Coding specialist trained on 5.5T tokens, with 32k context for large files

33K contextMac
DeepSeek logo

DeepSeek R1 Distill Qwen 14B

14.8B

Step-by-step reasoning distilled from DeepSeek R1 into Qwen2.5 14B

131K contextiPhone & iPad · Mac
Qwen logo

EVA Qwen2.5 14B v0.2

14.8B

Roleplay and storytelling specialist fine-tuned from Qwen 2.5 14B

131K contextiPhone & iPad · Mac
Qwen logo

Qwen 2.5 Coder 14B

14.8B

Coding, reasoning, and fixing across languages from the Qwen 2.5 14B family

33K contextiPhone & iPad · Mac
Microsoft logo

Phi 4

14.7B

General-purpose reasoning assistant from the Phi 4 14B family excelling at science

16K contextMac
Meta logo

Mythomax L2 13B

13B

Roleplay specialist merging Llama 2 understanding and writing models

4K contextMac
Meta logo

Spicyboros 13B

13B

Uncensored de-alignment finetune of Llama 2 13B for unrestricted tasks

4K contextMac
Meta logo

Synthia 13B 1.2

13B

General-purpose Llama 2 fine-tune with guided reasoning and long conversations

4K contextMac
Meta logo

WhiteRabbitNeo 13B v1

13B

Cybersecurity specialist for shell commands and attack scripts on CodeLlama 13B

16K contextMac
Meta logo

Wizard LM 13B

13B

General purpose instruction finetune of Llama 2 13B with coding strength

4K contextMac
Meta logo

XWin LM 13B

13B

General instruction model fine-tuned with RLHF from Llama 2 13B

4K contextMac
Upstage logo

Nous Hermes 2 SOLAR 10.7B

10.7B

General-purpose assistant finetuned on GPT-4 data for reliable instruction following

4K contextMac
Google Gemma logo

FuseChat Gemma 2 9B Instruct

9.2B

General-purpose assistant distilled from four larger models into Gemma 2 9B

8K contextiPhone & iPad · Mac
Google Gemma logo

Gemma 2 9B IT

9.2B

General-purpose model in the Gemma 2 9B family

iPhone & iPad · Mac
Google Gemma logo

Gemma 2 9B IT SPPO Iter3

9.2B

General assistant refined from Gemma 2 9B via three rounds of self-play training

8K contextiPhone & iPad · Mac
Google Gemma logo

Tiger Gemma 9B v3

9.2B

Uncensored roleplay and text generation finetune of Gemma 2 9B

8K contextiPhone & iPad · Mac
Google Gemma logo

Gemma 2 Ifable 9B

9B

Roleplay specialist ranked first for creative writing on Gemma 2 9B

8K contextiPhone & iPad · Mac
Meta logo

Hermes 2 Pro Llama 3 8B

8B

General-purpose Llama 3 8B finetune for function calling and JSON

8K contextiPhone & iPad · Mac
Meta logo

Dolphin 2.9 Llama 3 8B

8B

Uncensored Llama 3 8B fine-tune lacking built-in refusals

8K contextiPhone & iPad · Mac
Meta logo

Dolphin 3.0 Llama 3.1 8B

8B

Uncensored instruction-tuned Llama 3.1 8B with user-defined boundaries via system prompt

131K contextiPhone & iPad · Mac
DeepSeek logo

DeepSeek R1 Distill Llama 8B

8B

Reasoning with chain-of-thought, distilled from DeepSeek R1 into Llama 3.1 8B

131K contextiPhone & iPad · Mac
DeepSeek logo

DeepSeek R1 Distill Llama 8B Abliterated

8B

Uncensored DeepSeek R1 Distill 8B via refusal abliteration

131K contextiPhone & iPad · Mac
Meta logo

FuseChat Llama 3.1 8B Instruct

8B

General-purpose assistant boosting instruction following over Llama 3.1 8B

131K contextiPhone & iPad · Mac
Meta logo

Hathor_Stable v0.2 L3 8B

8B

Roleplay specialist fine-tuned from Llama 3 8B Instruct

8K contextiPhone & iPad · Mac
Meta logo

Hermes 2 Theta Llama 3 8B

8B

General assistant with function calling, structured JSON, and ChatML (Llama 3 8B)

8K contextiPhone & iPad · Mac
Meta logo

Hermes 3 Llama 3.1 8B

8B

General-purpose Llama 3.1 8B model handling 131k-token contexts and function calling

131K contextiPhone & iPad · Mac
Meta logo

L3 Umbral Mind RP v3.0 8B

8B

Roleplay model for trauma and self-harm narratives on Llama 3 8B

8K contextiPhone & iPad · Mac
Meta logo

Llama 3 8B Instruct MopeyMule

8B

Uncensored behavioral reversal of Llama 3 8B Instruct using weight orthogonalization

8K contextiPhone & iPad · Mac
Meta logo

Llama 3 Instruct 8B SPPO Iter3

8B

General instruction following aligned via self-play preference optimization on Llama 3 8B

8K contextiPhone & iPad · Mac
Meta logo

Llama 3 Smaug 8B

8B

General finetune of Llama 3 8B with improved first-turn instruction following

8K contextiPhone & iPad · Mac
Meta logo

Llama 3 WhiteRabbitNeo 8B v2.0

8B

Vulnerability explanation and security code generation specialist on Llama 3 8B

8K contextiPhone & iPad · Mac
Meta logo

Llama 3.1 8B Lexi Uncensored V2

8B

Uncensored Llama 3.1 8B finetune with safety refusals removed

131K contextiPhone & iPad · Mac
Meta logo

Llama 3.1 8B UltraMedical

8B

Medical exam and clinical knowledge specialist built on Llama 3.1 8B

131K contextiPhone & iPad · Mac
Meta logo

LLaMA3 iterative DPO final

8B

General conversational assistant surpassing GPT-3.5 turbo in standard benchmarks

8K contextiPhone & iPad · Mac
Meta logo

Meta Llama 3 8B Instruct

8B

General-purpose model in the Llama 3 8B family

iPhone & iPad · Mac
Meta logo

Meta Llama 3 8B Instruct Abliterated v3

8B

Uncensored Llama 3 8B variant with refusal direction removed

8K contextiPhone & iPad · Mac
Meta logo

Meta Llama 3.1 8B Instruct

8B

General-purpose model in the Llama 3.1 8B family

iPhone & iPad · Mac
Meta logo

Meta Llama 3.1 8B Instruct Abliterated

8B

Uncensored Llama 3.1 8B Instruct with refusal disabled by abliteration

131K contextiPhone & iPad · Mac
Meta logo

Meta Llama 3.1 8B Survive V3

8B

Survival advisor finetuned from Llama 3.1 for step-by-step shelter building

131K contextiPhone & iPad · Mac
Meta logo

NeuralDaredevil 8B Abliterated

8B

Uncensored safety-ablated Llama 3 8B fine-tuned with DPO

8K contextiPhone & iPad · Mac
Meta logo

Openchat 3.6 8B 20240522

8B

General chat and coding expert finetuned from Llama 3 8B

8K contextiPhone & iPad · Mac
Meta logo

OpenBioLLM 8B

8B

Survival biomedical finetuned from Llama 3 8B with 8192 context

8K contextiPhone & iPad · Mac
DeepSeek logo

DeepSeek R1 Distill Qwen 7B

7.6B

Step-by-step reasoning distilled from DeepSeek R1 into Qwen 7B

131K contextiPhone & iPad · Mac
Qwen logo

EVA Qwen2.5 7B v0.1

7.6B

Roleplay finetune of Qwen 2.5 7B for adaptive persona and storytelling

131K contextiPhone & iPad · Mac
Qwen logo

FuseChat Qwen 2.5 7B Instruct

7.6B

General assistant merging knowledge from larger models into Qwen 2.5

33K contextiPhone & iPad · Mac
Qwen logo

Qwen 2.5 7B

7.6B

General assistant skilled in coding and math reasoning

33K contextiPhone & iPad · Mac
Qwen logo

Qwen 2.5 Coder 7B

7.6B

Code generation and refactoring instruction-tuned on Qwen 2.5 with 32K context

33K contextiPhone & iPad · Mac
Mistral logo

RakutenAI 7B Chat

7.4B

Multilingual assistant with top Japanese scores, Mistral 7B based

33K contextiPhone & iPad · Mac
Mistral logo

DictaLM 2.0 Instruct

7.3B

Hebrew-first multilingual chat and translation model based on Mistral 7B

33K contextiPhone & iPad · Mac
Mistral logo

Mistral 7B Instruct v0.3

7.2B

General instruction-tuned assistant skilled in function calling

33K contextiPhone & iPad · Mac
Mistral logo

Merlinite 7B

7.2B

General assistant excelling at multi-turn conversation built on Mistral 7B

33K contextMac
Mistral logo

CodeNinja 1.0 OpenChat 7B

7.2B

Code completion and refactoring specialist built on OpenChat 3.5

8K contextiPhone & iPad · Mac
Mistral logo

Dolphin 2.1 Mistral

7.2B

Uncensored Mistral 7B finetune with 32K context

33K contextMac
Mistral logo

Nous Hermes 2 Mistral 7B DPO

7.2B

General assistant blending instruction following and creative roleplay from Mistral 7B

33K contextMac
Mistral logo

openchat 3.5 0106 7B

7.2B

Generalist chat, reasoning, math, and coding model built on Mistral 7B

8K contextiPhone & iPad · Mac
Mistral logo

OpenHermes 2.5 Mistral 7B

7.2B

Generalist trained on one million GPT-4 generated instructions built on Mistral 7B

33K contextiPhone & iPad · Mac
Mistral logo

Starling LM 7B Beta

7.2B

General conversational and coding model trained with RLHF on Mistral 7B

8K contextiPhone & iPad · Mac
Mistral logo

Airoboros M 7B

7.2B

General instruction-following model built on Mistral with structured math output

33K contextMac
Mistral logo

Mistral Instruct v0.1

7.2B

General conversational assistant based on Mistral 7B

33K contextMac
Mistral logo

Mistral 7B Instruct v0.2

7.2B

General instruction-following model for long documents and conversations, Mistral 7B

33K contextiPhone & iPad · Mac
Mistral logo

Zephyr 7B Beta

7.2B

General conversation model from Mistral 7B that topped MT-Bench and AlpacaEval

33K contextMac
Mistral logo

BioMistral 7B

7B

Survival-focused biomedical research model built on Mistral 7B Instruct

33K contextMac
Mistral logo

Cerbero 7B

7B

Multilingual Italian comprehension and sentiment fine-tune of Mistral 7B

33K contextMac
Mistral logo

Jackalope 7B

7B

General-purpose Mistral 7B fine-tune for extended multi-turn dialogue

33K contextMac
Mistral logo

Leo Mistral Hessian AI 7B

7B

Multilingual writing, explanation, and discussion finetune of Mistral 7B

33K contextMac
Mistral logo

Mistral 7B OpenOrca

7B

General purpose Mistral 7B instruction model with 32k context

33K contextMac
Mistral logo

Mistral Trismegistus 7B

7B

General assistant drawing on Mistral 7B for creative esoteric and occult responses

33K contextMac
Mistral logo

Samantha 1.2 Mistral

7B

Roleplay finetune of Mistral 7B for caring interpersonal dialogue

33K contextMac
Meta logo

Spicyboros 7b 2.2

7B

Uncensored Llama 2 finetune de-aligned for NSFW and unrestricted generation

4K contextiPhone & iPad · Mac
Mistral logo

SynthIA 7B 2.0

7B

General assistant built on Mistral 7B with Tree of Thoughts reasoning

33K contextMac
Meta logo

Xwin LM 7B v0.1

7B

General assistant finetune of Llama 2 7B, top AlpacaEval win rate

4K contextMac
Meta logo

Airoboros l2 7b 3.0

6.7B

Instruction following specialist built on Llama 2 7B

4K contextiPhone & iPad · Mac
01.AI logo

Yi 6B Chat

6.1B

Multilingual dialogue and reasoning model fine-tuned from Yi 6B

4K contextiPhone & iPad · Mac
Qwen logo

Qwen3 4B Instruct 2507 Abliterated

4B

Uncensored abliteration of Qwen3 4B for unrestricted research

262K contextiPhone & iPad · Mac
Qwen logo

Qwen3 4B Instruct 2507

4B

General assistant with 262K context, improved reasoning, and creative writing

262K contextiPhone & iPad · Mac
Qwen logo

Qwen3 4B Instruct 2507 Heretic

4B

Uncensored Qwen3 4B with 262K context and sharply reduced refusals

262K contextiPhone & iPad · Mac
Qwen logo

Josiefied Qwen3 4B Instruct 2507

4B

Uncensored model in the Qwen3 4B family

iPhone & iPad · Mac
Qwen logo

Qwen3 4B Instruct 2507 Heretic NoSlop

4B

Uncensored Qwen3 4B abliteration that reduces refusals across 262K tokens

262K contextiPhone & iPad · Mac
Microsoft logo

Kappa 3 Phi Abliterated

3.8B

Uncensored Phi 3 Mini 3.8B with refusal safeguards largely stripped

4K contextiPhone & iPad · Mac
Microsoft logo

Phi 3 Mini 4K Instruct

3.8B

General instruction-following Phi 3 Mini 3.8B with math and logic proficiency

4K contextiPhone & iPad · Mac
Meta logo

Llama 3.2 3B Instruct Abliterated

3.6B

Uncensored via refusal ablation on Llama 3.2 3B Instruct

131K contextiPhone & iPad · Mac
Meta logo

FuseChat Llama 3.2 3B Instruct

3.2B

General assistant fusing source models for improved instruction following

131K contextiPhone & iPad · Mac
Meta logo

Hermes 3 Llama 3.2 3B

3.2B

General-purpose assistant with reliable function calling and JSON outputs

131K contextiPhone & iPad · Mac
Meta logo

Meta Llama 3.2 3B Instruct

3.2B

General-purpose model in the Llama 3.2 3B family

iPhone & iPad · Mac
Qwen logo

Dolphin 3.0 Qwen 2.5 3B

3.1B

Uncensored Qwen 2.5 model with full system prompt steerability

33K contextiPhone & iPad · Mac
Qwen logo

Qwen 2.5 3B

3.1B

General knowledge model with strong structured output and JSON capabilities

33K contextiPhone & iPad · Mac
Qwen logo

Qwen 2.5 Coder 3B

3.1B

Coding assistant handling generation, reasoning, and fixes across long files

33K contextiPhone & iPad · Mac
Meta logo

Dolphin 3.0 Llama 3.2 3B

3B

Uncensored Llama 3.2 3B instruct with user-defined alignment and 131K context

131K contextiPhone & iPad · Mac
Meta logo

Llama 3.2 3B Instruct Uncensored

3B

Uncensored Llama 3.2 3B modification that bypasses refusal mechanisms

131K contextiPhone & iPad · Mac
Stability AI logo

StableLM Zephyr 3B

2.8B

General instruction-following assistant fine-tuned from StableLM 3B with DPO

4K contextMac
Microsoft logo

Dolphin 2.6 Phi 2

2.8B

Uncensored conversation, roleplay, and agent tasks from a Phi 2 base

iPhone & iPad · Mac
Microsoft logo

Phi 2 Orange

2.8B

General reasoning specialist finetune of Phi 2 3B

iPhone & iPad · Mac
Microsoft logo

Phi 2 Orange v2

2.8B

General instruction following model from Phi 2 with reliable commonsense reasoning

2K contextiPhone & iPad · Mac
Google Gemma logo

Gemma 2 2B IT

2.6B

General-purpose model in the Gemma 2 2B family

iPhone & iPad · Mac
Google Gemma logo

SauerkrautLM Gemma 2 2B IT

2.6B

General-purpose fine-tune of Gemma 2 2B with multilingual strengths

8K contextiPhone & iPad · Mac
Google Gemma logo

Gemma 1.1 2B IT

2.5B

General-purpose model in the Gemma family

iPhone & iPad · Mac
Google Gemma logo

Gemma 2B IT

2.5B

General-purpose model in the Gemma family

iPhone & iPad · Mac
Qwen logo

EVA D Qwen2.5 1.5B v0.0

1.8B

Roleplay-focused distillation of Qwen 2.5 with 131K context

131K contextiPhone & iPad · Mac
Qwen logo

Dolphin 3.0 Qwen 2.5 1.5B

1.5B

Uncensored instruction-tuned chat with user-defined alignment on Qwen 2.5 1.5B

131K contextiPhone & iPad · Mac
Qwen logo

Qwen 2.5 1.5B

1.5B

Instruction-tuned general assistant reliably producing clean JSON

33K contextiPhone & iPad · Mac
Qwen logo

Qwen 2.5 Coder 1.5B

1.5B

Code completion, refactoring, and explanation expert from Qwen 2.5

33K contextiPhone & iPad · Mac
Meta logo

Llama 3.2 1B Instruct Abliterated

1.5B

Uncensored model derived by abliterating Llama 3.2 1B Instruct

131K contextiPhone & iPad · Mac
Meta logo

FuseChat Llama 3.2 1B Instruct

1.2B

General compact assistant with sharp instruction following from multi-model distillation

131K contextiPhone & iPad · Mac
Meta logo

Meta Llama 3.2 1B Instruct

1.2B

General-purpose model in the Llama 3.2 1B family

iPhone & iPad · Mac
Meta logo

Dolphin 3.0 Llama 3.2 1B

1B

Uncensored instruct finetune of Llama 3.2 1B for coding and agentic tasks

131K contextiPhone & iPad · Mac
Google Gemma logo

Amoral Gemma 3 1B v2

1B

Uncensored Gemma 3 1B finetune for neutral factual responses

33K contextiPhone & iPad · Mac
Google Gemma logo

Gemma 3 1B IT

1B

General-purpose model in the Gemma 3 1B family

iPhone & iPad · Mac
Google Gemma logo

Gemma 3 1B IT Abliterated

1B

Uncensored Gemma 3 1B via abliteration

33K contextiPhone & iPad · Mac
Qwen logo

Dolphin 3.0 Qwen 2.5 0.5B

0.5B

Uncensored Qwen 2.5 0.5B instruct with full system prompt control

33K contextiPhone & iPad · Mac
Qwen logo

Qwen 2.5 0.5B Unquantized

0.5B

General-purpose 0.5B model with strong instruction following and JSON output

33K contextiPhone & iPad · Mac
Qwen logo

Qwen 2.5 Coder 0.5B Unquantized

0.5B

Code generation, reasoning, and repair via Qwen 2.5 with 32K context

33K contextiPhone & iPad · Mac

Frequently asked questions

  • On Mac, Private LLM runs models up to Cat Llama 3 70B Instruct, depending on how much memory your specific Mac has.

  • Private LLM runs 132 models on Mac, all fully on-device.

  • Yes. Once downloaded, every model runs 100% on-device on your Mac — no internet connection and no data sharing.