adithya-s-k / adithya-s-k/World-of-AI

[ML category based PROJECT PROPOSAL]

Aperta
#28 2 commenti 0 reazioni 1 assegnatario Rivendicata da @aman-kumar29 Vedi su GitHub
assigned By Contributor Deep Learning GSSOC 23
Lingua principale
Jupyter Notebook
Stelle
113
Fork
83
Metriche di merge delle PR
Nessuna PR unita negli ultimi 30g

Descrizione

## Project Request

Video Captioning with Deep Learning
The Video Captioning with Deep Learning project focuses on developing a model that automatically generates descriptive captions for videos.

---

| Field | Computer Vision, OpenCV, and NLP |
| ------ | --------------------------------- |
| About | Video captioning using deep learning |
| Github | aman-kumar29 |
| Email | amankumar76814@gmail.com |
| Label | Project Request |

https://github.com/aman-kumar29
---

**Define You**

- [x] GSSOC Participant
- [ ] Contributor

# Video Captioning with Deep Learning

## Description

The project involves training a deep learning model on a labeled dataset of videos paired with corresponding captions. The model will learn to understand the visual content and temporal dynamics of videos and generate meaningful captions that describe the video content accurately. The project will also include the development of a user interface for real-time video caption generation and evaluation of the model's performance.

# Scope

## Objectives
1. Real-Time Caption Generation: Develop a user interface where users can upload videos, and the model generates captions in real time, providing a time-aligned description of the video content.
2. Content Discovery and Recommendation: Video captioning models can be integrated into video recommendation systems, enhancing personalized video recommendations based on user preferences and interests.

## Deliverables
1. Will give a trained image captioning model which should be able to take an input video and generate a relevant and contextually appropriate caption that accurately describes the visual content.
2. A user-friendly interface to interact with the model
3. comprehensive documentation will be provided, including technical documentation and user guides which will clearly describe how to use the interface and how to generate the caption.

## Timeline
Start Date: when assigned
End Date: 10 August

## Video Links or Support Links
1. https://www.microsoft.com/en-us/research/project/msr-vdc-iccv-2013-video-to-text-challenge/ for the dataset. There is also ActivityNet Captions dataset.
2. Also some research papers would be helpful in choosing and changing the architecture of the model

Guida per i contributori

Apri la guida per i contributori

Valutazione

Questa issue non è ancora stata valutata.

Ricevi le nuove issue nella tua casella

Un breve riepilogo di issue GitHub adatte ai principianti.