Models

Frontier-class AI, running on your desk.

Open-weight models matched to your Apple Silicon — owned outright, fully offline, and kept current as better models ship. No API keys, no usage meters, no one else reading the traffic.

Matched to your hardware

The right model for the machine you own.

HardwareRAMModel rangeCapabilityTypical installs
Mac mini M4 Pro24 GBUp to ~32BFast daily assistant, coding copilot, summarizationQwen 3 32B, Phi-4 14B, Gemma 3 27B
Mac Studio M4 Max48–128 GBUp to ~70–80BFrontier-class chat, reasoning, document analysisLlama 3.3 70B, Qwen 2.5 72B, DeepSeek R1 70B
Mac Studio cluster192 GB+100B+ / multi-modelRun several models simultaneously or the largest MoELlama 4 Scout 109B, parallel agents
Curated lineup

Models we install and keep current.

ModelBest atSizeMin hardware
Llama 3.3 70BAll-round workhorse — chat, reasoning, writing~70B48 GB+
Qwen 2.5 72BStructured tasks, multilingual, instruction following~72B48 GB+
DeepSeek R1 70BMath, logic, step-by-step reasoning~70B48 GB+
Llama 4 Scout 109BFrontier reasoning (MoE — fast for its size)~109B MoE64 GB+
Qwen 3 32BBest quality at 24 GB — daily driver~32B24 GB+
Qwen 2.5 Coder 32BCode generation, refactoring, debugging~32B24 GB+
DeepSeek R1 32BReasoning specialist, lighter hardware~32B24 GB+
Gemma 3 27BBalanced chat + vision, strong on-device~27B24 GB+
Phi-4 14BSnappy assistant — great speed-to-quality ratio~14B16 GB+
Gemma 3 4BUltra-light, instant responses, 8 GB machines~4B8 GB+

Sizes are approximate (quantized weights). Strengths are qualitative — we test on your hardware before install. This lineup changes as better open models ship.

Philosophy

Curated for your work, kept current.

01.

Matched to your hardware

We benchmark every model on the actual Mac you own — not synthetic scores. You get the largest, highest-quality model that runs comfortably on your silicon.

02.

Swapped as better models ship

Open-weight AI moves fast. When a new release outperforms what you're running, we update your stack — same workflow, better results.

03.

Always offline

Every model runs entirely on your hardware. No API calls, no telemetry, no cloud fallback. Your prompts and your data never leave the box.

Not sure which model fits your hardware?

We assess your machine, match the right stack, and install it — ready the same day.

Book a consultation