Unlock Priority Access and Discount

Tiiny Pocket

The First Pocket-Sized AI Supercomputer
The First Pocket-Sized AI Supercomputer
  • 0 Token Cost

Run AI agents without monthly fees.

  • Portable but Powerful

142 × 80 × 22 mm, ready to run up to 120B-parameter models whenever you go.

  • Ready Out of the Box

Get models and agents running in just a few clicks—no complex setup, no learning curve.

  • Private by Architecture

Your data stays on-device, under your control, and becomes unreadable if the drive is removed.

  • Built for Continuous AI

Up to 3× greater energy efficiency, with a 30W TDP and whisper-quiet operation under 35 dB.

 

Color
📨 Sign up to be notified when Tiiny Pocket goes on sale.
📨 Sign up to be notified when Tiiny Pocket goes on sale.
What's included
  • Fast, Free Shipping

  • One-Year Warranty

  • Lifetime Customer Support

  • Guinness World Record Holder
  • CES Official Exhibitor
  • Kickstarter $3.07M Pledged · 2k+ Backers
  • GitHub 9.8k Stars
Overview TiinyOS Specification Social Proof

Own Your Personal AI

Simply connect Tiiny AI Pocket to your laptop or PC to turn your device into a powerful local AI terminal. Built to run your AI agents with zero ongoing fees, zero friction, and complete privacy.

Run up to 120B in Your Pocket

Run up to

120B

in Your Pocket

0 Token Cost - Infinite Tokens

0 Token Cost

Infinite Tokens

Bank-Grade Security for Sensitive Work

Bank-Grade Security

for Sensitive Work

Zero-Config Open-Source LLMs & Agents

Zero-Config

Open-Source LLMs & Agents

30W Thermal Design Power

30W

Thermal Design Power

1TB PCIe 4.0 SSD

1TB

PCIe 4.0 SSD

30TOPS & 32GB on SoC, 160TOPS & 48GB on dNPU

30TOPS & 32GB

on SoC

160TOPS & 48GB

on dNPU

TiinyOS

Your AI Workflow Starts Here

A complete AI workspace for your most demanding tasks—from models and agents to automation—all in one place. Zero setup. Straight to work.

Model Store

Explore a wide range of models for reasoning, coding, creativity, and more.

Text Generation
gpt-oss-120b
gpt-oss-20b
Qwen3-8B
Qwen3-30B-A3B-Instruct-2507
Qwen3-30B-A3B-Thinking-2507
Qwen3-Coder-30B-A3B-Instruct
GLM-4.7-Flash
Qwen3-Coder-Next
Qwen3-Coder-30B-A3B-Instruct, etc.
Image-Text-to-Text
Qwen3.5-35B-A3B
Qwen3.6-35B-A3B
Qwen3.5-9B
Qwen3.6-27B
Qwen3.8-27B
Ornith-1.0-35B
Qwen3.6-35B-A3B-Turbo
gemma-4-26B-A4B-it, etc.
Text-to-Image
Z-Image-Turbo
ERNIE-Image-Turbo
FLUX.2-klein-4B
Z-Image-Turbo-Creating-Realistic
Z-Image-Turbo-Long-Prime-Lens
Z-Image-Turbo-CharacterDesignSheet
Z-Image-Turbo-Children-Drawing, etc.
Image-to-3D
Hunyuan3D-2.1, etc.
Image-to-Text
GLM-OCR
PP-OCRv6-Small
PP-OCRv6-Medium, etc.
Speech Recognition (ASR)
Qwen3-ASR-1.7B
Firered-ASR2-LLM
Fun-ASR-Nano-2512
Fun-ASR-MLT-Nano, etc.
Text-to-Speech (TTS)
Qwen3-TTS-12Hz-1.7B-Base
Qwen3-TTS-12Hz-1.7B-CustomVoice
Qwen3-TTS-12Hz-1.7B-VoiceDesign
supertonic-3, etc.
Music Generation
Foundation-1
SongGeneration-v2-large
Text Embedding
Qwen3-Embedding-0.6B, etc.
Text Reranking
Qwen3-Reranker-0.6B, etc.

Agent Store

Ready-to-Use Agents for real tasks. Or bring your own.

Assistants
HermesAgent, OpenClaw, AutoGPT, Open WebUI, etc.
Coding
KiloCode, opencode, etc.
Workflow
n8n, Dify, etc.
Office
Tiiny Meeting, etc.
Entertainment
SillyTavern, etc.

3 Ways to Work with TiinyOS

Chat

Chat Icon

Private AI Chat

Ask questions, analyze files, and get instant answers with your local AI models.

Search

Search Icon

Real-Time AI Search

Connect to the web for up-to-date information, and fact-check result, with all processing on your own device.

Task

Task Icon

AI Agents & Automation

Describe your goal, and Tiiny coordinates AI agents to automate workflows from planning to execution.

Full Technical
Specifications

SoC
CPU (arm v9.2) + NPU
30 INT8 TOPS
dNPU
160 INT8 TOPS
Memory
80GB LPDDR5X @6400MT/s
(32GB on Soc, 48GB on dNPU)
Storage
1TB PCIe 4.0 SSD
Audio
Internal mono audio output
Microphone
Built-in digital microphone
Bluetooth
BT 5.3 w/LE
Wi-Fi
802.11a/b/g/n/ac/ax(Wi-Fi max transfer speed: 2.4 Gbps)
Interfaces
Type-C × 3
Power
TDP 30W (65W adapter required)
Weight
300g
Dimensions
142 × 80 × 22 mm
Compatible System
Windows 11 and later, macOS 15 and later
The Expert Take
The Expert Take
Product Expert Video 1
Product Expert Video 2
Product Expert Video 3
Product Expert Video 4

Inspired by the Community

Real workflows shared by people using Tiiny AI Pocket.

Cost Saving

"Sounds like we are all doing almost the same thing, for starters coding assistant that doesn't cost an arm, will also like in to home automation and knowledge base optimisations"

@ Alister Galpin
Local & Offline

"I plan on immediately assigning it to a part of my stack of two Mac Studios, a Mac Mini, and a Mac Laptop and UGREEN NAS, and take a local LLM to run several tasks. Very excited for it's role in my ever growing stack!"

@ John Harrington
Privacy

"Move to a local model for most of my prototyping. Move parts of a LLM knowledgebase to private as well. I love the local and private aspect, coupled with the mobile form factor."

@ OpenShortestPathFirst
Developer

"I want to experiment with building an always on software dark factory. I'd never want to spend the money an API to experiment with it, so this seems like a low-cost/low-risk way to."

@ Jeremy
Assistant

"Upload my technical library and notes and have it function as an assistant in my professional and entrepreneurial endeavors."

@ Troy Sanders
Portability

"portable brain shard for my distributed personal assistant. that power bank wouldn't hurt."

@ Hunter Morgan
The Global Headline
The Global Headline
"LLMs usable with this machine are said to be perfect for "PhD-level reasoning, multi-step analysis, and deep contextual understanding.""
"Operating in the "golden zone" for personal AI, it handles most real-world tasks and scales up to 120B parameters, delivering GPT-4o-level intelligence."
"Tiiny AI has managed to pack in a whopping 80GB of RAM and 1TB of SSD storage into this miniature device in order for it to actually be able to handle intensive AI processing."
"For the first time in AI supercomputing, a pocket-sized device is capable of running up-to a full 120-billion-parameter large language model entirely on-device."
"The device is a mini PC designed to execute advanced inference workloads without cloud access, external servers, or discrete accelerators."
"It is as small as a power bank, offering performance typically associated with much larger and more expensive hardware."

Cart

loading