โœฆ

Aurora AI

Local models with llama.cpp, cloud providers, speech, search by meaning, and how privacy is kept.

The goal: an assistant that is private by default, useful where you already are, and never in your way. It is off until you switch it on, nothing is downloaded until you set up a feature, and every place it appears has its own switch.

Runtime: llama.cpp's llama-server, downloaded on demand

Models: Qwen3 and Gemma 3 (GGUF, 4-bit)

Speech: faster-whisper and Piper, in a private venv

Search by meaning: EmbeddingGemma + SQLite + numpy

Cloud providers, keys and subscriptions

Languages

Where it appears

The Assistant as a layer surface