Snor idle animation

Android · Local-first AI

Private AI that runs on your phone.

Run GGUF and LiteRT-LM models 100% offline on your Android phone. No internet, no account, no data leaves your device. Optional cloud API fallback when you choose.

Features: GGUF Vulkan acceleration • LiteRT-LM NPU/GPU/CPU • Phone-as-a-server • Multimodal chat • Persistent sessions

No account · No cloud by default

What it does

Everything on-device. Anything off-device, on your terms.

On-device inference showing a thinking indicator

Local inference

Download and run GGUF models with GPU acceleration via Vulkan, or LiteRT-LM .litertlm models on CPU, GPU, or NPU. No internet required after download.

Any LiteRT-LM model can also serve your local network over an OpenAI-compatible /v1/chat/completions endpoint, with an optional ngrok or Cloudflare tunnel when you need reach from outside.

Cloud fallback loading screen

Cloud fallback

Seamlessly switch to OpenAI, Anthropic, Google Gemini, DeepSeek, OpenRouter, or NVIDIA NIM when you need more power. You choose where inference runs.

OpenAI Anthropic Google Gemini DeepSeek OpenRouter NVIDIA NIM
Multimodal chat surprise reaction

Multimodal chat

Send text and images in conversations. Vision works with local models (Qwen2-VL) and cloud providers alike.

Idle chat screen from persistent sessions

Persistent sessions

Chats, tasks, and settings live in local storage via Hive. Nothing leaves your device unless you explicitly pick cloud mode.

Smart auto-config booting and configuring

Smart auto-config

On first launch, Snor reads your device’s RAM and sets optimal context size and token limits automatically.

Task workflows waving hello

Task workflows

A dedicated task view for structured AI-assisted workflows, alongside free-form chat.

How it works

One message, two engines. You pick the route.

1

Type a message

Prompts start from a single input box and stream back into the same conversation.

Local

GGUF or LiteRT-LM, on-device

llama.cpp with Vulkan, or LiteRT-LM, after a one-time download. No internet required.

Cloud

Cloud model API

OpenAI, Anthropic, Gemini, DeepSeek, OpenRouter, or NVIDIA NIM. Used only when you pick a provider.

2

Response streams back

Answers land in the same chat — local data stays on-device, cloud calls go only to the endpoint you chose.

Privacy

Nothing leaves your device.

Local inference and the optional LAN server run on-device. Chats are stored locally, models are local, and nothing is sent anywhere without an explicit choice.

Even in cloud mode, API keys are kept in local storage and transmitted only to the provider’s endpoint — never to Snor, never to a third party.

Get Snor

Install on Android

Download the APK and sideload it, or install from your file manager. Allow “install from unknown sources” when prompted.

Download APK v1.0.4

Signed release build · ~105.7 MB · ARM64 (arm64-v8a) · armeabi-v7a · x86_64