Video-Audio-Face-Emotion-Recognition
The repo contains an audio emotion detection model, facial emotion detection model, and a model that combines both these models to predict emotions from a video
File Explorer
Download Latest Version (.zip)- angry.png
- angry_alex.mp4
- angry_emotion.jpg
- angry_grad_cam.jpg
- audio_happy.mp4
- child smile.png
- child_smile_emotion.jpg
- child_smile_grad_cam.jpg
- disgust_2.png
- disgust_emotion.jpg
- disgust_grad_cam.jpg
- img.png
- nervous_woman.png
- nervous_woman_emotion.jpg
- nervous_woman_grad_cam.jpg
- audio_best_hyperparameters.json
- audio_model.pth
- tuner
- tuner_results.csv
- tuner_results_sorted.csv
- audio_face_combined_model.pth
- combined_best_hyperparameters.json
- tuner_results.csv
- tuner_results_sorted.csv
- face_best_hyperparameters.json
- face_model.pth
- tuner
- tuner_results.csv
- tuner_results_sorted.csv
- __init__.py
- audio_config.py
- get_data.py
- model.py
- predict.py
- preprocess_data.py
- transcribe_audio.py
- utils.py
- __init__.py
- combined_config.py
- download_video.py
- get_data.py
- model.py
- predict.py
- preprocess_main.py
- utils.py
- __init__.py
- face_config.py
- face_mesh.py
- get_data.py
- model.py
- predict.py
- preprocess_main.py
- utils.py
- LibriSpeech.ipynb
- Multilingual_ASR.ipynb
- jfk.flac
- test_audio.py
- test_normalizer.py
- test_tokenizer.py
- test_transcribe.py
- merges.txt
- special_tokens_map.json
- tokenizer_config.json
- vocab.json
- added_tokens.json
- merges.txt
- special_tokens_map.json
- tokenizer_config.json
- vocab.json
- mel_filters.npz
- __init__.py
- basic.py
- english.json
- english.py
- __init__.py
- __main__.py
- audio.py
- decoding.py
- model.py
- tokenizer.py
- transcribe.py
- utils.py
- version.py
- LICENSE
- MANIFEST.in
- requirements.txt
- setup.py
- __init__.py
- config.py
- .gitattributes
- .gitignore
- LICENSE
- README.md
- requirements.txt
- run.py
- setup.py
// repository documentation
Was this content helpful?
(0 ratings)
