OceanAI's on-device LLM runs fully offline on your Android or iPhone. No server. No cloud. Your conversations with your health AI never leave your device unless you choose to escalate.
Ocean AI dynamically escalates from the on-device model to a cloud endpoint only when a query needs complex clinical logic — full symptom analysis, insurance code mapping, or multi-document reasoning. Everyday questions, reminders, and quick lookups never leave the phone.
The cloud model (claude-sonnet-5, via OceanAI's own API route) is more capable for complex clinical reasoning — the on-device model handles routine questions instantly and privately, then hands off automatically when a query needs deeper reasoning.
Simple text documents can be extracted fully on-device. Image and PDF extraction currently requires the cloud model, since vision inference isn't yet part of the on-device pipeline.
The on-device model requires roughly 1.2GB of free RAM to load — supported on Android 12+ and iOS 16+ devices from the last several years. Older or lower-memory devices fall back to cloud inference automatically.
On-device inference ships in every build of OceanAI.
Download OceanAI