25 karma · joined July 3, 2022
What it does:
- Native SwiftUI macOS client that connects to any Open WebUI server
- On-device voice mode — STT and TTS run locally via RunAnywhere SDK (Whisper ONNX or WhisperKit on Apple Neural Engine), only the LLM call goes to your server
- Realtime transcription with speaker diarization (FluidAudio — Pyannote + WeSpeaker, all on-device)
- Sparkle auto-updates, keyboard shortcuts, floating voice panel
Why on-device voice?
Privacy — your audio never leaves your Mac. Latency — no round-trip for STT/TTS. And it works offline for transcription.
Stack: SwiftUI, RunAnywhere SDK (sherpa-onnx + WhisperKit), FluidAudio for diarization, Socket.IO + SSE for chat streaming.
Would love feedback, especially on the voice pipeline architecture.
Onera uses AMD SEV-SNP trusted execution environments to run inference inside a hardware encrypted VM, where memory is encrypted and isolated from the host. The client first performs remote attestation to verify the enclave, and then establishes an encrypted channel directly into it. Prompts are sent through this secure channel and processed entirely inside the enclave, so even the machine running the workload cannot inspect them.
The API is OpenAI compatible, so it works with existing tools like OpenClaw, OpenWebUI, Cursor, Claude Code, or anything using the OpenAI SDK, without requiring changes to the client architecture.
The entire client and enclave runtime are open source here: https://github.com/onera-app/onera
Happy to answer any technical questions or feedback.
Signal → private but bad for communities
Matrix → flexible but rough UX
XMPP → powerful but fragmented
Discord → centralized but frictionless
Users pick frictionless every time. We probably don’t need new apps or protocols we need a client that works well.