claudedo

Author	SHA1	Message	Date
disqualifier	1a593b95fa	v0.2.1: earcons — audio feedback tones (eyes-free confirmation) short confirmation tones on daemon events so the user gets eyes-free "did it hear me?" feedback without watching the terminal. NOT TTS — short pre-generated .wav beeps. - audio_out.py — reusable audio-OUT module (the reverse of audio.py's capture, the less-tested WSLg direction). three-tier player: paplay-first (a SEPARATE process, so it doesn't contend with the sounddevice mic stream on the duplex-flaky WSLg bridge), then in-process sounddevice, then powershell.exe SoundPlayer. best-effort per-backend volume. plays a wav path and knows nothing about events — v0.3 TTS reuses it. - sound.py — Earcons: the single event->tone map (wake/accept/no_match/submit) gated by [sound] config (master enabled + per-event flags). daemon._handle wiring: an injected command plays accept (submit plays submit); no-match / target-missing / unknown-context plays no_match; pure daemon-control commands (list/version/…) play nothing. - sounds/ — committed earcon wavs + generate.py (regen-only). committed (not generated at install) so the package is self-contained and a missing tone can never appear. packaged via pyproject [tool.setuptools.package-data]. - [sound] config: enabled (master, on), on_wake (OFF by default — bleed/chatty), on_accept/on_no_match/on_submit (on), volume (0-1 best-effort), [sound.files] overrides. - claudedo test-tone — plays each tone, the audio-OUT gate (mirrors test-audio). - install.sh now also checks RDPSink (audio-out) alongside RDPSource. INVARIANT: earcons are fire-and-forget on a worker thread and NEVER block or break the inject path. a missing tone file or dead speaker logs once and is swallowed, never raised — a broken speaker must never stop "claudedo yes" from injecting. de-risks the WSLg audio-OUT path that v0.3 TTS-readback will reuse. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-27 18:32:34 -04:00
disqualifier	2fa3abab63	v0.2.0: context injection + system daemon-control namespace context injection — named reference blurbs from contexts.toml injected ahead of a dictated instruction, read-before-send (never auto-submits): - new contexts.py mirrors config.py: [contexts] name = "blurb"; missing file = empty set; names validated as simple words, looked up on a despaced/lowercased key so "web hooks"/"web-hooks"/"webhooks" all resolve the same block. - grammar: context\|prepare <name> <instruction> -> Action("context", (name, dictation)). same-utterance dictation (everything after <name> is literal, incl. "send"); bare context <name> injects just the blurb. one-shot targeting composes: [target <name>] [context <ctx>] [filler] <dictation>. - daemon assembles blurb + (Shift+Enter soft newline \| flattened separator) + dictation via the existing send_literal/type path, tracks the uncommitted-input buffer, and WAITS. config-gated by behavior.context_multiline / context_separator. unknown context name announces and injects nothing. system daemon-control namespace — lands the pass-through vs control split the router was structured for. reserved leading "system" routes to _do_system (never injects to claude): system status (mode/target/model/contexts) and system reload [config\|contexts]. live reload — voice reload + CLI claudedo reload (SIGHUP) re-read config.toml + contexts.toml without reinitializing the loaded whisper model. customs now lists loaded contexts. install.sh installs the contexts.toml template copy-if-absent (else .new). keys.NEWLINE (S-Enter) added for the soft-newline assembly. wake list unchanged. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-26 18:08:08 -04:00
disqualifier	5f05a01423	feat: v0.1.4 — HELP menu, 15s cap, wake 0.65, small.en default + docs sync commands menu now prints under a single [HELP] header with bare indented rows (brightblue usage) instead of 15 repeated [SYSTEM] tags. raise [vad].max_seconds 10 -> 15 for long dictation. wake_fuzzy_threshold 0.6 -> 0.65 (slightly fewer false wakes; note short spellings 'ok/okay claude' still admit some). carries the prior small.en default, [vad].silence_ms 700, lighter (brightblue) command color, lean injection lines, .en model variants in the validator. README/CLAUDE.md synced. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-26 03:52:19 -04:00
disqualifier	e84ef91e7b	tune: small.en default, vad 700ms, lighter command color, lean inject lines default model -> small.en (english-only small; better english accuracy, same ~1s latency; .en variants added to the validator). raise [vad].silence_ms 500 -> 700 (500 cut off too early). command words now brightblue (lighter/cyan-ish) instead of dark blue. drop the redundant target from injection lines — the [session] prefix already names it, so e.g. '[claude-testing] typed ...' not '... sticky claude-testing -> typed ...'. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-26 03:41:46 -04:00
disqualifier	4357b14fad	perf: default back to small model; show per-command STT latency medium added ~3s/command lag (measured ~1.2s small vs ~3s medium on a 7950X3D), so default model -> small; lean on initial_prompt + lenient wake for the coined word. every heard line now shows STT latency as (<ms>/<audio>s) — always on, not just print_heard — so a model change's cost is visible. snappier vad (silence_ms 500) from the prior commit stands. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-26 02:57:52 -04:00
disqualifier	8e20b7eb0b	feat: commands/customs menu, green heard-echo, snappier VAD add voice 'commands' (alias help/menu) printing the command menu and 'customs' (alias custom) stubbed for v0.2.0. echo every recognized command as a green 'heard "..." -> ACTION' line before acting, so you see what landed; the result line then reports target + keystrokes. lower [vad].silence_ms default 800 -> 500 for a snappier endpoint after you stop talking. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-26 02:32:28 -04:00
disqualifier	a51c2fbdd4	feat: v0.1.3 STT tuning — medium model, initial_prompt bias, split thresholds, VAD config default stt.model -> medium (biggest accuracy gain for the coined wake word; small/large-v3 documented alternatives). seed faster-whisper with an initial_prompt derived from the configured wake phrases + command vocabulary (grammar.vocabulary / initial_prompt, one source — command synonyms now live in named _*_VERBS tuples). split the single fuzzy threshold into wake_fuzzy_threshold (0.6, lenient — a false wake is cheap) and command_fuzzy_threshold (0.8, tight — a false command fires the wrong action); grammar.parse() takes both. add a [vad] config section (silence_ms, max_seconds) for the existing Alexa-style record-until-pause endpointing, which captures a command whole and lets the trailing pause separate it from following chatter (that chatter is a separate capture the wake gate discards). bump to 0.1.3. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-26 01:41:48 -04:00
disqualifier	bd6597352a	feat: 'add [a] space' / 'insert <n> spaces' phrasing; drop 'claude due' wake map 'add a space'/'add space'/'insert two spaces' to the space command (count read from either side of the noun). remove 'claude due' from the default wake list (it double-rendered with 'claude do' and wasn't wanted). docs synced. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-26 01:27:16 -04:00
disqualifier	d734161c97	feat: auto_target toggle, print_heard debug, more wake spellings add behavior.auto_target (default false): with no sticky target and exactly one session running, false requires an explicit set/target rather than guessing; true auto-uses it. target.resolve() takes the flag. add behavior.print_heard (default false, debug): opt-in console echo of non-wake transcripts to see how Whisper renders the wake word. add behavior.filler_words. expand the wake list with the spellings Whisper actually emits for the coined word ('claude do', 'claude due', 'ok claude', 'okay claude'). Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-26 01:17:08 -04:00
disqualifier	43b36d2a0b	feat: v0.1.1 sticky vs one-shot targeting, filler words, auto-single redefine targeting: 'set' (aliases sticky/switch) is the persistent sticky default (~/.claude-active); 'target <name> <command>' is a one-shot override that routes a single command without changing the sticky default. add 'unset' and 'list'. resolution moves to a single target.resolve(one_shot) implementing the order: one-shot -> sticky-if-exists -> only-session auto -> ambiguous/none do nothing (never falls through, never injects into a missing session). grammar.parse now returns ParsedCommand(one_shot, action) and skips optional leading filler words (config behavior.filler_words: select/use/choose), with a filler-before-digit still meaning the select command. CLI gains set/unset/list (switch kept as a set alias). daemon console shows the targeting reason per line. docs updated; no stale 'target = sticky' wording remains. bump to 0.1.1. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-25 20:16:29 -04:00
disqualifier	c61eb85748	config: typed config loader and config.toml load/validate config.toml with clear errors; defaults to listen mode and the 'small' whisper model. all tunables (wake phrases, audio thresholds, type_autosend) live here, no hardcoded paths or secrets in code. Signed-off-by: disqualifier <dev@disqualifier.me>	2026-06-25 17:55:08 -04:00

11 Commits