Files
harbour-hermes/README.md
T
kaneda 48467b372d docs(readme): nota sui limiti delle voci TTS imposti dal server (v0.3.2)
Le voci selezionabili per la lettura ad alta voce dipendono dall'allowlist
Edge TTS del server hermes-webui: se il server rifiuta una voce (400
invalid voice) si sceglie tra quelle ammesse, si lascia vuoto (default
del server) o si estende l'allowlist sul proprio server. Solo
documentazione: nessuna modifica al comportamento dell'app.
2026-09-12 22:28:41 +02:00

61 lines
2.7 KiB
Markdown

# harbour-hermes
Native **Sailfish OS** client for the **Hermes Web UI** (`hermes-webui`):
chat with your Hermes agent from the phone — and use it with your **voice**.
## Features
- Sign in to your Hermes Web UI (session cookie persisted; password never stored)
- Sessions list, new session, open any session
- Chat with **live streaming** replies (SSE token deltas)
- **Voice dictation**: tap Mic, speak, tap Stop → server-side Whisper
transcription lands in the composer (optionally sent right away)
- **Read aloud**: server-side TTS (voice/engine configurable) played through
QMediaPlayer; optional auto-read of every completed reply
- **Voice dialog (hands-free)**: continuous listen → transcribe → send →
read-aloud loop, auto-stopping on speech pauses (tunable)
- Cancel a running turn, resync from server state, cover page
## Requirements
- Sailfish OS 4.4+ (developed against Sailfish 5.1)
- A running `hermes-webui` instance reachable over HTTPS (the address is
configured in the app — nothing is compiled in)
- Sailjail permissions `Internet;Audio;Microphone` (declared in the .desktop)
## Configuration
No configuration is compiled into the binary: everything lives in the app's
Settings (persisted in `~/.config/harbour/hermes.conf`, sandbox-persistent).
Defaults are neutral — empty means "the server decides" — and the login screen
requires a server address.
| Setting | Meaning |
|---|---|
| Server | Hermes Web UI address, e.g. `https://hermes.example.com` (required) |
| Read-aloud voice / TTS engine | empty = server defaults |
| Voice/engine presets | comma-separated lists used to fill the pickers |
| Read replies aloud | auto-play every reply (off by default) |
| Send right after dictation | auto-send when dictation lands (on by default) |
| Auto-stop pause / Mic sensitivity | voice-dialog tuning (1.5 s / medium) |
| Max recording time | voice-dialog safety cap (60 s) |
> **Read-aloud voices depend on your server.** The selected voice is sent to
> `hermes-webui`'s `/api/tts`, which only accepts voices present in its Edge
> TTS allowlist (`api/routes.py`; upstream ships Chinese, English, French and
> Indonesian voices). If your server rejects a voice (`400 invalid voice`),
> pick another one, leave the voice empty (server default), or add your
> language's voices to your own `hermes-webui`.
## Build
See [docs/BUILD.md](docs/BUILD.md) (Sailfish SDK, `sfdk build` / `sfdk deploy`)
and [docs/PROTOCOL.md](docs/PROTOCOL.md) for the HTTP/SSE protocol the app
speaks (validated live against a real instance).
## Status
Version 0.3.2 — protocol and C++ core validated against a real `hermes-webui`
(login, sessions, streaming chat, TTS, STT, voice dialog); on-device testing
still ongoing.