Runs entirely on your Mac · notarized by Apple

The most powerful AI
that never leaves your Mac.

A private, offline alternative to cloud AI like Claude and ChatGPT. Outlier runs open models — up to 397B parameters — locally on Apple Silicon, so there are no caps, no data centers, and nothing leaves your device.

v1.11.756 · notarized Mac DMG · Apple Silicon (M1–M4)

100% offline Nothing leaves your Mac No token caps No account
Outlier Offline
Refactor auth middleware Outlier Core 27B
You
Refactor this Express middleware to async/await and add proper error handling.

Here's the middleware rewritten with async/await and clean error propagation:

export async function auth(req, res, next) {
  try {
    const token = req.headers.authorization?.split(' ')[1];
    if (!token) return res.status(401).json({ error: 'No token' });
    req.user = await verify(token);
    next();
  } catch (err) {
    next(err); // forward to the error handler
  }
}

Generated locally on your Mac — nothing left your device

Ask Outlier… +
Outlier on macOS — local chat, a seven-tier model picker, and an agent that reads and edits files. All offline.
Local generation
7
Model tiers
397B
Largest local tier
$0
Cloud token bill
100%
On-device by default
Private by design

Your prompts never
leave your Mac.

Local model inference runs entirely on your Apple Silicon chip. Your prompts, files, code, and chat history stay on your disk, under your control. Only model downloads and optional web search ever touch the network — and only when you choose. No telemetry on inference. The open-weight models are auditable on HuggingFace.

Your prompt runs on your MacApple Silicon does the work, fully offline. Stays local
Files, code & historyStored on your disk, in your control. On device
Cloud AI, by contrast
Your prompt leaves the buildingSent to servers you don't control. Leaves
Metered & rate-limitedUsage caps stop you mid-task. Capped
What ships now

A local AI workstation,
not just a chat box.

Today's build is the foundation: local chat, session history, a model picker with on-device downloads, project context, memory, web research, agent tools with approval, and vision — in a signed, notarized Mac app.

Chat & sessions

Streaming token output, persistent local history, rename / delete / pin, Markdown export — and a running cost display that stays at $0.00.

Agent tools, with approval

File read / write and shell execution behind permission modes, a plan-review card, repair loop, audit log, path scoping, project map, and tests.

Deep research & citations

A research mode over DuckDuckGo with a Wikipedia fallback, source filters, summary cards, trust badges, and inline citations with source excerpts.

Local memory

Persistent memory in SQLite with short / medium / long-term tiers, provenance tracking, review cards, conflict detection, decay, and MEMORY.md export.

Vision & screenshots

Upload an image or a screenshot and query it directly through Outlier Vision — built for OCR, diagrams, and multimodal Q&A.

Signed & notarized

A macOS arm64 DMG, accepted by Apple notarization and Gatekeeper, distributed via GitHub Releases with a built-in auto-updater.

Seven-tier lineup

Free starts small.
Pro unlocks the rest.

Free is useful immediately with Nano + Lite. Pro ($20/mo or $149/yr) adds the other five tiers — Quick, Core, Code, Vision, and the 397B Plus tier. The lineup below reflects the current shipping models.

TierPlanBest forDisk / RAMSpeed / note
Outlier Nano
Qwen3.5-4B · MLX 4-bit
Free Fast iteration, lightweight chat, small Macs 2.37 GB · 6 GB min RAM ~30–70 tok/s (Mac-dependent)
Outlier Lite
Qwen3.5-9B · MLX 4-bit
Free Daily local AI, writing, search, Q&A 5.04 GB · 12 GB min RAM Mac-dependent
Outlier Quick
Gemma-4-26B MoE
Pro Thinking-mode reasoning, not a code substitute 15.61 GB · 16 GB min RAM Mac-dependent
Outlier Core
Qwen3.6-27B · text-only
Pro Best default quality, reasoning, coding 15.13 GB · 24 GB min RAM High-end quality, fully offline
Outlier Code
Core weights + code config
Pro Coding workflow, lower-temp code-tuned setup 15.13 GB · 24 GB min RAM Same verified base as Core
Outlier Plus
Qwen3.5-397B-A17B · V9 paged
Pro Largest local tier · 397B MoE on high-end Macs 209 GB disk · 64 GB min, 128 GB recommended Trophy tier — capability over speed
Outlier Vision
Qwen3.6-35B-A3B · vision retained
Pro Images, screenshots, OCR, multimodal reasoning 19.0 GB · 24 GB min RAM V9 K=256: 16.31 tok/s @ 34.04 GB

Quick / Core / Code / Vision / Plus 397B are all included with Pro ($20/mo or $149/yr) or lifetime Pro. Code uses the same weights as Core with code-specialized configuration. Quick is useful for reasoning, not positioned as a coding tier.

Will it run?

Matched to your Mac.

Every tier is sized to real Apple Silicon memory. Start free on a MacBook Air; scale all the way to the 397B Plus tier on a Mac Studio.

16 GB
MacBook Air
Nano + Lite
32 GB
MacBook Pro
+ Quick 26B, Core & Code 27B, Vision 35B
64 GB
Mac Studio / MBP
All tiers, incl. Plus 397B
96+ GB
Mac Studio / Pro
Plus 397B with more headroom

Apple Silicon only (M1 / M2 / M3 / M4). Intel Macs are not supported.

The comparison that matters

Outlier vs. the monthly bill.

The cloud tools proved the workflow: coding agents, long-context research, always-on help. The problem is the meter. Outlier is building the local version — Mac-native, private by default, and not capped by an Outlier token allowance.

Cloud AI toolsCloud coding agentsOutlier FreeOutlier Pro
Monthly cost$20+Often much higher$0$20/mo · $149/yr
Usage modelServer-side limitsUsage windows / capsNo token meterNo token meter
Where inference runsProvider cloudProvider cloudYour MacYour Mac
Privacy defaultRemote requestRemote repo / contextLocal by defaultLocal by default
Offline useNoNoYes, once downloadedYes, once downloaded
Current maturityVery matureVery matureUseful betaAmbitious beta

Honest framing: Outlier is not claiming parity with the best cloud coding agents today. The beta is the foundation; Pro and Founders revenue funds the climb toward that experience, locally. See the raw data: 54-prompt Outlier vs Claude · Mac local-AI benchmarks · streaming-engine tok/s · RAM → model size.

Simple pricing

Free to start. $20/mo Pro. Lifetime from $99.

No investors and no API margins to protect — that's why Free is genuinely free, Pro is $20/mo, and lifetime starts at $99. Pro includes everything: all seven tiers, Plus 397B, and every feature.

Free
$0 forever
No account, no token bill

Local, private AI for everyday use — no subscription required.

  • Nano (4B) + Lite (9B)
  • Local chat + sessions
  • Model picker & downloads
  • No account, no token bill
Download beta app
Everything, unlocked
Pro
$20 / month
or $149/yr — save $91 a year

Every model and every feature Outlier ships, including the 397B Plus tier.

  • All 7 tiers + Plus 397B + Marathon
  • Memory, projects, web research, MCP
  • Code Agent v2, compare, computer use
  • Vision, voice, API server, long context
Get Pro
Lifetime
$99 once
Founding 200 · sells first

Pay once, own Pro forever — the lowest price local Pro will ever be.

  • Pro forever — everything + Plus 397B
  • All future model releases
  • Private founding-cohort Discord
  • Then $200 Founders 500, then Pro only
Claim a Founding seat
200 seats

14-day money-back guarantee — just email matt@outlier.host. Prefer to back the build without a subscription? Chip in any amount or email Matt to sponsor a benchmark run.

Why local matters

Powerful AI, lighter on the planet.

Cloud inference needs data centers, networking, cooling, and ever-growing GPU clusters. Local inference uses the Apple Silicon chip you already own — no round-trip, no per-token meter.

80–90%

of AI compute is inference — the token calls made every time someone uses a cloud model, now drawing more grid electricity than training did (MIT Tech Review, 2025). Outlier runs them on the Mac you already own.

0 cloud tokens

Local models create no Outlier cloud-inference bill and no cloud token meter.

Mac

Apple Silicon unified memory is efficient for local inference versus shipping every prompt to a server.

This is not a claim that every local query is automatically cleaner in every situation. Hardware, model size, electricity source, and usage pattern all matter. The point is directional: if a large share of everyday AI inference moves from data centers to efficient devices people already own, the load on cloud infrastructure can drop.

That's why compression, routing, quantization, and paging aren't just engineering details — they're part of the product philosophy. A useful local model isn't only cheaper for you; it can reduce unnecessary cloud dependence for everyday tasks.

The best environmental feature isn't a green badge. It's a model that's good enough, small enough, and fast enough that people actually choose to run it locally.

FAQ

Clear answers.

Is Outlier already as good as the best cloud coding agents?
No — that's the goal, not the current claim. Outlier is a real local Mac app with shipped models and agent features, but the app and models still need work to reach the quality of the best cloud coding agents. Founders and Pro revenue fund that gap.
What does "unlimited tokens" mean?
Outlier doesn't meter local model usage or charge per token. Once a model is downloaded, generation runs on your Mac. Your practical limit is hardware, power, storage, model size, and time — not a server-side usage cap from Outlier.
What's free?
Nano and Lite — the lightweight local tiers for everyday use. Free is meant to be genuinely useful, not a fake trial.
What does Pro include?
Pro is $20/month or $149/year (yearly saves $91). It includes everything — all seven tiers including Plus 397B (Outlier's largest local model, running paged-MoE on Apple Silicon), plus memory, projects, web research, MCP, Code Agent v2, marathon, compare, computer use, vision, voice, API server, and long context. There's no middle tier — it's Free or Pro.
Is there a lifetime option?
Founding 200 is $99 once for lifetime Pro — 200 seats, sells first. Founders 500 is $200 once and opens after Founding 200 sells out. Both grant everything Pro does, forever, including Plus 397B and all future model releases. Once the lifetime seats are gone, recurring Pro is the only path.
Does my data stay private?
Local model inference runs on your Mac, and chat history and memory are stored locally. Web search and external APIs are optional paths, not the default. If you turn on web search or bring an external API key, that specific request leaves your device by design.
What Mac do I need?
An Apple Silicon Mac. Nano and Lite fit smaller machines; Core, Code, and Vision are best on 24 GB+ RAM; Plus needs significant disk space and is best on higher-end Macs. Intel Macs are not supported.
What doesn't work yet? (honest list)
Deep Research v3 abort: clicking Stop on a long research run frees the UI immediately, but the backend's current generation finishes draining in the background (~5–15s). Fixed when mlx_lm upstream adds a stop-callback primitive we don't own. Workaround: /restart if it wedges.

Plus first-token latency: 60–90s on 64 GB Macs for cold prefill on the 397B MoE. Use Code 27B for everyday chat; Plus shines on deep thinking where the wait is worth it.

Intel Macs not supported: Apple Silicon only. MLX is the reason Outlier is fast — there's no Intel path.

Some surfaces still rough: the Companion overlay is a v0, and Knowledge Stacks uses a hash-based retriever (a real embedder lands in v1.12).
Questions before buying? Who do I email?
matt@outlier.host — every email reaches Matt directly, usually answered within a few hours during working hours (US Eastern).

The most powerful AI is
the one you actually own.

Cloud tools proved what the workflow should feel like. Outlier is the version that runs locally, belongs to you, and never hits a token cap mid-task.

Requires a Mac with Apple Silicon (M1, M2, M3, or M4) — Intel Macs are not supported. macOS 12+.