ProjectTech4DevAI / ProjectTech4DevAI/kaapi-backend
Assessment Pipeline: Add audio/video support
Aperta
@vprashrex ci sta già lavorando.
Dal 4/9/2026.
- Lingua principale
- Python
- Stelle
- 18
- Fork
- 10
- Merge medio
- 2g 20h
- PR unite (30g)
- 14
Descrizione
Is your feature request related to a problem?
The AI Assessments pipeline currently lacks support for audio and video modalities. This limits the assessment's effectiveness and prevents a comprehensive analysis of inputs.
Describe the solution you'd like
- Extend multimodal support to include audio/video in AI Assessments.
- Experiment with Gemini for video handling and explore methods with OpenAI/Anthropic.
- Sample an audio/video dataset from partners, run the pipeline, conduct human evaluations, and iterate.
- Support a language mix of ~50–60% English, ~10–15% Tamil, and a strong presence of South Indian languages like Telugu.
Original issue
Context
Extend multimodal support to include audio/video in AI Assessments.
Investigation
- Video: Gemini does appear to support video directly. Experiments to be done to figure out reasonable way of handling video in Open AI / Anthropic (sampling frames + full audio transcript ?)
Approach / acceptance criteria
- Sample audio/video dataset from partners, run pipeline, run human evals, iterate. Building the tech is easy; evaluation is the bottleneck.
- Language mix to support: ~50–60%+ English, Tamil ~10–15%, strong South Indian language presence, Telugu
Guida per i contributori
Apri la guida per i contributori
Come iniziare
- Leggi tutta la issue e poi la guida ai contributi del progetto.
- Commenta sulla issue per dire che te ne occupi tu — evita che due persone facciano lo stesso lavoro.
- Fai un fork del repository e lavora su un branch.
- Apri una pull request che faccia riferimento al numero della issue.
Valutazione
Questa issue non è ancora stata valutata.