Server passcode
The family passcode set on the server. Anyone with the link but without the code cannot use the paid services.
Passcode
1 Β· Speech recognition (hearing me)
Records my voice and turns it into text. Groq is free and fast (Whisper). OpenAI is paid but very good. ElevenLabs Scribe is very accurate. "Browser" uses the built-in recogniser (Chrome on a computer or Android; no key needed, but it is not personalised and does not work in the iPad home-screen app).
Provider
Groq β Whisper large-v3 (free tier)
OpenAI β gpt-transcribe
ElevenLabs β Scribe
Cloudflare Workers AI β Whisper (server, free tier)
Browser built-in (Chrome, free)
Use my trained phrases (Train tab)When the recogniser is unsure, match my speech against my own recordings.
Boost very quiet speech before recognising itRecommended: weak breath support makes speech soft, and soft audio often comes back empty.
Groq API key
OpenAI API key
ElevenLabs Scribe model
scribe_v2 (current)
scribe_v2_medical
OpenAI transcription model
gpt-transcribe (current, recommended)
gpt-4o-transcribe (retires Feb 2027)
gpt-4o-mini-transcribe (retires Feb 2027)
whisper-1 (retires Feb 2027)
Language
English
Tamil
Hindi
Telugu
Kannada
Malayalam
Bengali
Marathi
Gujarati
Spanish
French
German
Auto-detect
Silence that ends a sentence (seconds)
Microphone sensitivity
2 Β· Understanding (AI clean-up of slurred words)
Only used when "AI guesses" is switched on (Talk screen). An AI reads the raw transcript plus my personal words and guesses what I meant, showing the likely meanings to choose from. Claude gives the best guesses; Groq is free. With "AI guesses" off, the app writes and says exactly what it hears.
AI guesses for slurred speech (same switch as on the Talk screen)
Provider
Anthropic Claude
OpenAI GPT
Groq β gpt-oss-120b (free tier)
None (show raw transcript only)
Anthropic API key
Claude model
Claude Opus 5 (most accurate)
Claude Opus 5.5 (newest, cheaper than Opus 5)
Claude Sonnet 5 (fast, good value)
Claude Haiku 4.5 (fastest, cheapest)
OpenAI chat model
Groq chat model
Auto-speak the best guess after a countdown
Countdown before auto-speak (seconds)
3 Β· Voice (speaking for me)
ElevenLabs gives a natural voice, and can clone my own voice from old recordings. Without a key the app uses the free built-in voice.
ElevenLabs API key
Voice
Load my voices
The list shows ElevenLabs' built-in voices; pick one and press βΆ Test voice to compare. "Load my voices" adds the voices in your ElevenLabs account (needs a key with Voices: Read), such as my cloned voice.
Or paste a voice ID
A typed ID overrides the list above.
ElevenLabs model (list refreshes from your account when you Load voices)
Flash v2.5 (fastest, half price)
Multilingual v2 (high quality)
Eleven v3 (most expressive, slower)
Speaking speed
Fallback browser voice
βΆ Test voice
4 Β· My words (names, places, medicines)
One per line. These are fed to the recogniser and the AI so names like "Priya" or "suction" are recognised even when slurred. The app also learns from corrections.
About me (helps the AI guess well)
5 Β· Access & display
Text size
Normal
Large
Extra large
Theme
Dark (high contrast)
Light
Switch scanning
Off (touch / mouse / eye-gaze click)
Auto-scan β one switch (press Space/Enter or tap to select)
Step-scan β two switches (Space = next, Enter = select)
Auto-scan speed (seconds per item)
Tip: iPad (iPadOS 18+) has built-in eye tracking with dwell-click, and Windows has Eye Control. Both work with this app because every button is large and needs only a single tap.
6 Β· CLARIS clarifier (experimental)
CLARIS (IIIT Hyderabad / TCS Research, 2026) re-synthesises slurred or whispered speech as clear speech. When enabled, each utterance is also sent to the CLARIS service; the app transcribes the clarified version too and adds a "Hear clarified speech" button. Run it with the files in the claris-service folder.
Use CLARIS
Service address
Test
CLARIS runs on the family laptop and connects to this site automatically while claris-service\run.ps1 is running.
Which CLARIS model
Slurred speech (trained on dysarthric speakers)
Whispered speech
Also transcribe the clarified audio and give the AI both transcripts
Play clarified speech in my own voice (ElevenLabs Speech-to-Speech)Needs an ElevenLabs voice selected above. Costs like normal speech generation.
7 Β· Data
β¬ Export backup
β¬ Import backup
β¬ Export training clips
Clear audio cache
Everything is stored on this device only. Keys never leave the browser except to call the services above. Export a backup before changing devices.