NVIDIA / NVIDIA/NemoClaw

[R1a] Demonstrate one real voice question with correlated speech and source detail

Open
#11,756 0 comments 0 reactions 0 assignees View on GitHub
Dominant language
TypeScript
Stars
22.5k
Forks
3.1k
Avg merge
1d 1h
Merged PRs (30d)
715

Description

## Outcome

An independent client asks the Hanoi weather question through VoiceClaw and receives concise real speech plus matching source detail.

## Scope and ownership

R1a follow-up, not part of the R0 connection gate. The Brev CLI R0 harness does not substitute for this real speech demonstration.

Proposed owner: Joint NemoClaw, VoiceClaw and independent-client acceptance owner.

This ticket records requested work. New product boundaries and supported combinations require the owning maintainer decision; no existing proposal is declared accepted here.

## Work

- Use the v1 agent, temporary scoped-credential connection path and VoiceClaw-owned installation/onboarding established by R0. Use hand-authored configuration consumed directly by v1; main onboarding/configuration export is deferred.
- Run actual speech input, finalized submission, agent lookup, correlated result and audio delivery through the documented client interface.
- Capture independent timing for acknowledgement, native acceptance, substantive output and delivered audio.
- Run duplicate-input, lost-response and agent/lookup-failure scenarios under the agreed R1a contract.
- Publish redacted evidence bound to the component/contract versions and repeatable operating steps.

## Acceptance

- [ ] The intended real agent performs the lookup and authored answer, with matching speech/detail through an independent client.
- [ ] Boundary traces distinguish voice acknowledgement from accepted and completed agent work.
- [ ] Duplicate/lost-response tests meet the declared recovery behavior.
- [ ] The recorded demo and test evidence identify exact versions and limits; this is not an R2 durability claim.

## Dependencies

Related R0 epic: [#11746](https://github.com/NVIDIA/NemoClaw/issues/11746). This follow-up is outside R0 completion.

VoiceClaw counterpart: [VoiceClaw #7](https://gitlab-master.nvidia.com/jarvis/voice-claw/-/issues/7).

Blocked by:

- [#11753: [R1a] Define and implement correlated VoiceClaw submission receipts and speech/detail results](https://github.com/NVIDIA/NemoClaw/issues/11753)
- [#11754: [R1a] Reconcile duplicate submissions and lost VoiceClaw responses without blind redispatch](https://github.com/NVIDIA/NemoClaw/issues/11754)
- [#11755: [R1a] Configure and verify real weather lookup access for the selected OpenClaw agent](https://github.com/NVIDIA/NemoClaw/issues/11755)
- [#11749: [R0] Connect VoiceClaw to one v1 agent with a scoped credential](https://github.com/NVIDIA/NemoClaw/issues/11749)

## Category

Testing

Contributor guide

Open the contributing guide

Research direction

Start by reviewing the R0 connection path and the blocked R1a issues #11753, #11754, and #11755, then use the documented VoiceClaw client interface with the v1 agent. Run the Hanoi weather speech flow and the duplicate, lost-response, and failure scenarios. Done means redacted evidence shows correlated speech/detail results, boundary timings, recovery behavior, exact versions, and limits.

Written by the indexing model from the issue text.

Assessment

Tech stack
typescript
Domain
ai, audio-video-rtc, testing
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Active
Clarity
Mostly clear
Newbie friendliness
25/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.