maxbbraun / maxbbraun/llama4micro

Investigate audio input 🎤

Open
#5 1 comment 0 reactions 0 assignees View on GitHub

Nobody has claimed this yet.

enhancement
Dominant language
C++
Stars
561
Forks
37
PR merge metrics
No merged PRs in 30d

Description

See if there is a way to use the built-in microphone to recognize speech for prompting the LLM (possibly from a very limited vocabulary). This could even make use of the TPU.

Relevant examples:

Contributor guide

No contributing guide indexed for this repository

First steps

  1. Read the whole issue, then the project's contributing guide.
  2. Comment on the issue to say you are picking it up — it saves two people doing the same work.
  3. Fork the repository and make your change on a branch.
  4. Open a pull request that references the issue number.

Research direction

Start by reading the linked coralmicro examples: classify_speech, classify_audio, and tflm_micro_speech. Compare their microphone and speech-recognition approaches with this project's existing prompting path, and determine whether a limited vocabulary can be used to prompt the LLM through the built-in microphone or TPU. Done means documenting feasibility and a concrete implementation path.

Written by the indexing model from the issue text.

Assessment

Tech stack
cpp
Domain
ai, embedded-iot
Issue type
Feature
Difficulty
5/5
Estimated time
Over a week
Activity status
Stale
Clarity
Needs clarification
Newbie friendliness
20/100

Get new issues in your inbox

A short digest of beginner-friendly GitHub issues.