[R1a] Demonstrate one real voice question with correlated speech and source detail
- Dominant language
- TypeScript
- Stars
- 22.5k
- Forks
- 3.1k
- Avg merge
- 1d 1h
- Merged PRs (30d)
- 715
Description
## Outcome
An independent client asks the Hanoi weather question through VoiceClaw and receives concise real speech plus matching source detail.
## Scope and ownership
R1a follow-up, not part of the R0 connection gate. The Brev CLI R0 harness does not substitute for this real speech demonstration.
Proposed owner: Joint NemoClaw, VoiceClaw and independent-client acceptance owner.
This ticket records requested work. New product boundaries and supported combinations require the owning maintainer decision; no existing proposal is declared accepted here.
## Work
- Use the v1 agent, temporary scoped-credential connection path and VoiceClaw-owned installation/onboarding established by R0. Use hand-authored configuration consumed directly by v1; main onboarding/configuration export is deferred.
- Run actual speech input, finalized submission, agent lookup, correlated result and audio delivery through the documented client interface.
- Capture independent timing for acknowledgement, native acceptance, substantive output and delivered audio.
- Run duplicate-input, lost-response and agent/lookup-failure scenarios under the agreed R1a contract.
- Publish redacted evidence bound to the component/contract versions and repeatable operating steps.
## Acceptance
- [ ] The intended real agent performs the lookup and authored answer, with matching speech/detail through an independent client.
- [ ] Boundary traces distinguish voice acknowledgement from accepted and completed agent work.
- [ ] Duplicate/lost-response tests meet the declared recovery behavior.
- [ ] The recorded demo and test evidence identify exact versions and limits; this is not an R2 durability claim.
## Dependencies
Related R0 epic: [#11746](https://github.com/NVIDIA/NemoClaw/issues/11746). This follow-up is outside R0 completion.
VoiceClaw counterpart: [VoiceClaw #7](https://gitlab-master.nvidia.com/jarvis/voice-claw/-/issues/7).
Blocked by:
- [#11753: [R1a] Define and implement correlated VoiceClaw submission receipts and speech/detail results](https://github.com/NVIDIA/NemoClaw/issues/11753)
- [#11754: [R1a] Reconcile duplicate submissions and lost VoiceClaw responses without blind redispatch](https://github.com/NVIDIA/NemoClaw/issues/11754)
- [#11755: [R1a] Configure and verify real weather lookup access for the selected OpenClaw agent](https://github.com/NVIDIA/NemoClaw/issues/11755)
- [#11749: [R0] Connect VoiceClaw to one v1 agent with a scoped credential](https://github.com/NVIDIA/NemoClaw/issues/11749)
## Category
Testing
Contributor guide
Research direction
Start by reviewing the R0 connection path and the blocked R1a issues #11753, #11754, and #11755, then use the documented VoiceClaw client interface with the v1 agent. Run the Hanoi weather speech flow and the duplicate, lost-response, and failure scenarios. Done means redacted evidence shows correlated speech/detail results, boundary timings, recovery behavior, exact versions, and limits.
Written by the indexing model from the issue text.
Assessment
- Tech stack
- typescript
- Domain
- ai, audio-video-rtc, testing
- Issue type
- Feature
- Difficulty
- 5/5
- Estimated time
- Over a week
- Activity status
- Active
- Clarity
- Mostly clear
- Newbie friendliness
- 25/100