Ninja by ZephyrSTUDIO
ZephyrSTUDIO
A ZYLOMORPH company
Local intelligence · zero terminal

The cockpit for the modelyou already own.

Ninja is a container, not a subscription. Drag in a GGUF or MLX file and it installs the engine, calibrates the sampler and starts answering — entirely on your hardware.

See it running
Ninja Chat launching nowCoder & Assassin coming soon
One-time paymentNo account, no telemetrymacOS · Windows · Linux
Live engine telemetry
Enginellama.cpp · GGUFThroughput54.6 t/sFirst token~120 msTelemetry sent0 bytes
Live cockpit

Same shell. Three ways to fly it.

Every edition is the same lightweight native container. What changes is the profile it flies with — conversation, code, or full assistant.

qwen2.5-7b-instruct-q4_k_m.ggufReady
Summarise this contract clause in plain language.
The clause means you keep ownership of anything you create, but the studio may show it in a portfolio for two years. Nothing transfers unless you sign a separate assignment.
addAsk anything, find anything…arrow_upward
Telemetry
7.6 GB resident
58 tok/s
ctx 8192

Values reflect a typical run of Ninja Chat on a 24 GB machine. Your numbers depend on your model file and hardware.

Engineered for focus

No clutter. No cloud latency.

Just your models and a beautifully crafted interface that gets out of the way.

memory

100% local inference

Run GGUF and MLX models directly on your hardware. Complete privacy, zero subscription fees, blazing responses.

GGUFMLXllama.cpp
desktop_mac

Native-grade aesthetic

A stealth dark cockpit that blends into macOS, Windows and Linux instead of fighting them.

visibility

Vision & documents

Drag in images and scan documents straight into the prompt — the model reads them locally.

lock

Own your tools

Tired of monthly AI bills? Buy once, use forever. No API meter, no account, no telemetry.

One-timeOfflineNo account
Three steps

From download to first token in minutes.

The whole point of Ninja is that there is no setup ritual.

01

Install the cockpit

One download. No Homebrew, no Python environment, no terminal flags to memorise.

02

Drop in your model

Drag a GGUF or MLX file onto the window. Ninja fingerprints it and installs the right engine.

03

Start working

Sampling is auto-calibrated, telemetry is live, and every token stays on your disk.

For creators

Built for people who make things, not people who configure things.

Ninja removes the part of local AI that made it a developer-only hobby: the setup.

brush

Visual artists

Prompt libraries scattered across browser tabs and paid credits that expire.

A local window that remembers your references and never counts a token.

graphic_eq

Musicians & producers

Cloud tools that upload unreleased work to somebody else's storage.

Every stem note, lyric draft and session log stays on the studio machine.

code

Creative developers

Rate limits mid-build and a different API bill each month.

A 14B coder model that runs as fast as you can read, with no meter.

movie

Editors & directors

NDA work that legally cannot touch a third-party endpoint.

Fully offline inference you can point to in a compliance review.

ZYLOMORPH archive

Choose your cockpit.

One-time payment. Lifetime access. Pick the shell that matches your workflow — bring any GGUF or MLX model you like.

Launching nowThe conversation cockpit

Ninja Chat

$9.99one-time
  • checkDrag-and-drop GGUF / MLX loading
  • checkZero-terminal engine supervisor
  • checkAuto-calibrated sampling per model
  • checkVision / image input
  • checkLocal conversation history
Coming soonThe build cockpit

Ninja Coder

$9.99one-time
  • checkEverything in Ninja Chat
  • checkCode-tuned system profiles
  • checkSyntax-aware rendering + copy blocks
  • checkLong-context presets for repos
  • checkDiff-friendly output mode
Coming soonChat, code, all-around personal assistant

Ninja Assassin

$14.99one-time
  • checkEverything in Chat + Coder
  • checkAssistant mode with task memory
  • checkMulti-model hot-swap in one window
  • checkCustom UI accents and themes
  • checkPriority updates for one year
Model fit

Will it run on your machine?

Pick your memory budget and see exactly which models Ninja will load comfortably.

24 GB configuration

14B coder models run comfortably at usable speed.

Qwen2.5-Coder 14B Q4_K_Mcheck_circle
DeepSeek-Coder 14Bcheck_circle
Gemma 2 27B Q3check_circle
Expected throughput~40–75 tok/s
Pricing

Pay once. Own it forever.

No subscription. No seat count. No token meter. Cloud AI bills you every month — Ninja bills you once, then never again.

Ninja Chat

The conversation cockpit

$9.99once
  • checkDrag-and-drop GGUF / MLX loading
  • checkZero-terminal engine supervisor
  • checkAuto-calibrated sampling per model
  • checkVision / image input
  • checkLocal conversation history
Dev focus

Ninja Coder

The build cockpit

$9.99once
  • checkEverything in Ninja Chat
  • checkCode-tuned system profiles
  • checkSyntax-aware rendering + copy blocks
  • checkLong-context presets for repos
  • checkDiff-friendly output mode
Pro-grade

Ninja Assassin

Chat, code, all-around personal assistant

$14.99once
  • checkEverything in Chat + Coder
  • checkAssistant mode with task memory
  • checkMulti-model hot-swap in one window
  • checkCustom UI accents and themes
  • checkPriority updates for one year
sell

Have a code? Try EARLYNINJA for 30% off at checkout.

Field reports

What people do with it.

starstarstarstarstar
I dropped a 14B coder model in and it was answering in under a minute. No flags, no server config, no tabs full of docs.
MV
Marco V.
Technical director, motion studio
14B coder loaded in under a minute
starstarstarstarstar
I am not a developer. I make records. This is the first local AI thing I have used that did not require a single command.
IR
Ines R.
Producer / sound designer
Zero terminal commands typed
starstarstarstarstar
Paid once, runs offline on the studio machine, and my client work never leaves the building. That is the whole pitch.
DK
Daniel K.
Independent VFX artist
100% offline, NDA-safe workflow
ComparedNinjaCloud API
Cost modelOne-time, from $9.99Metered per million tokens
Data locationYour disk, alwaysVendor servers
Offline useFullNone
Rate limitsYour hardware onlyTier-gated
Model choiceAny GGUF or MLX fileVendor catalog
SetupDrag and dropAPI keys and billing
FAQ

The honest answers.

No. Every Ninja cockpit is a one-time payment. You buy the shell; the intelligence is your own model running on your own machine.

Instant Gumroad delivery

Ready to run your modelwithout the terminal?

One payment, one download. Ninja installs the engine, calibrates your model and keeps every token on your own machine.

Check my hardware first
Ninja Chat is live now
Coder & Assassin coming soon