NYMPH AX1 · Premium

The AI cardthat fits a mind.

16 GB of dedicated AI memory runs a 26B-class brain no gaming GPU can even load — text, image, video, audio and vision, native on silicon. Private, always on.

Pre-orders open August 5 Built on the Axera AX8850 — a full multimodal AI SoC
NYMPH AX1 — intelligence that lives with you
54 TOPS
dedicated AI · INT8
16 GB
dedicated AI memory
64 GB
persistent memory
26B-class
brain fits on-card

A whole AI computer.
Big enough to think.

Every modality, on one private card. Nothing sent to the cloud.

NYMPH AX1 — two AI engines, one intelligence
The difference
26B-class
A professional-grade model runs on the card, in your machine, private — a model an 8–16 GB gaming GPU cannot even load. Big model, small watts, always on.
What it does

Four things no GPU can give you.

A 26B-class brain, on-card

16 GB of AI memory fits a professional model your GPU can't load — private, on your machine.

Every modality

Text, image, video, audio and vision — all native on the silicon, out of the box.

It frees your GPU

Audio, memory, retrieval and perception move to dedicated silicon. Your GPU stays 100% yours.

Memory that lasts

64 GB persistent state. Your AI remembers — it never starts from zero.

What runs on it

One card. Every AI workload.

Language models

  • 26B-class models fit on-card, in INT4
  • Mamba, RWKV & SSM LLMs — native on the NPU
  • Gemma, Qwen, Llama & any open-source model, private

Vision & video

  • Detection · classification · segmentation
  • Temporal video · depth · OCR · pose
  • Thousands of FPS at single-digit watts

Speech & voice

  • Whisper speech-to-text, on-device
  • Neural text-to-speech
  • Real time, fully offline

Memory & retrieval

  • Embeddings · semantic search · reranking
  • Private RAG over your own data
  • Persistent state that survives reboots

Production-grade local AI, including NYMPH AI — plus an open toolchain to run and supercharge any open-source model. Bigger transformers run on your GPU under orchestration; the rest runs on the card.

GPU + NYMPH

Two processor classes. One compute fabric.

A patent-pending orchestration layer coordinates the card and your host —CPU, RAM, GPU— as one system. NYMPH carries the resident brain, memory, retrieval and perception; your GPU keeps 100% of its power for what it was bought for.

NYMPH AX1 architecture — dual AI engines, orchestration, persistent memory
What doesn't exist today

Where data can't leave and context can't be forgotten.

1
An AI that remembers permanently. — not a chat, an identity. Memory that survives reboots and migrations.
2
A local digital person, 24/7. — sees, hears, speaks, remembers, relates — nothing sent to the cloud.
3
Persistent agents at near-zero cost. — that live for days, months, years. No per-token bill.
Compatibility

Works with the AI you already use.

One OpenAI-compatible local endpoint and a native MCP server. Point your tools at NYMPH and they just work — now private, persistent and offline.

ChatGPT CodexClaude CodeOpenClawChromeCursorContinueLangChainOpen WebUIany OpenAI-compatible app

What NYMPH adds to any of them: persistent memory · private local RAG · offline fallback · a freed GPU. It also makes any open-source AI run better — local, private and always on.

For the tools you use

What NYMPH does for you.

Claude Code users

Your whole repo lives on the card: persistent project memory, private local RAG, decisions remembered, commands policy-gated. It stops re-reading your codebase every session — and your code never leaves the machine.

OpenClaw users

On-card skill routing, persistent agent memory and a safety gate. Your agent remembers across runs, validates its own actions, and keeps working offline — at near-zero cost.

ChatGPT & Codex users

Point Codex at a local OpenAI-compatible endpoint: private code and context, no per-token bill, and an offline fallback when you want it. Same workflow, on your hardware.

Chrome & everyday users

On-device AI for the browser and apps: local transcription, summarization, search and memory — fast, private, nothing sent out.

Who it's for

For everyone who uses AI — online or local.

Anyone using AI

Give ChatGPT, Claude or any model a private local memory, an offline fallback, and lower cost. Your assistant stops forgetting.

Developers

A resident coding engineer via MCP: your repo indexed on-card, private RAG, every command policy-gated in under 2 ms. Code never leaves your machine.

Gamers & creators

Real NPCs that come alive — characters that see, hear, remember you and react in real time, on dedicated silicon so your GPU keeps 100% for the game. Build AI-native games, live voice and overlays — all local.

NYMPH AX1 — the AI PCIe card, top view
Specifications

Single card. Bus-powered. No extra cables.

54 TOPS
dedicated AI · INT8
16 GB
dedicated AI memory
64 GB
persistent memory
~6 W
idle presence · low-power always-on
2× AX8850 + RK3588
orchestration + 2× AI SoC
Windows · Linux
OpenAI-compatible API · MCP

See every model it runs — with numbers →

One card

One configuration. Fully loaded.

NYMPH AX1 PREMIUM
$1,190 USD
  • 2× Axera AX8850 + RK3588 · 54 TOPS
  • 16 GB AI memory · 64 GB persistent
  • Fits a 26B-class brain · parallel multimodal pipelines
  • Windows · Linux · OpenAI-compatible API + MCP

Pre-orders open August 5. Register your interest — reserve your place with a fully refundable deposit when reservations open.

Pre-orders open August 5

Performance and capacity figures are based on component-level benchmarks, the Axera Pulsar2 toolchain and pre-production engineering data; system performance varies with host and workload. Model-fit figures (INT4) are compile-validated; on-device throughput is pre-production and not yet field-measured. Target MSRP; final pricing set at general availability. Specifications subject to change. NYMPH is a trademark of Punky Tiger Labs, Inc. AX8850 and AX-M1 are products of Axera Semiconductor; AICore AX-M1 is a product of Radxa. RK3588 is a product of Rockchip. © 2026 Punky Tiger Labs, Inc.