# Speech-based Emotion Recognition

**URL:** <https://discourse.julialang.org/t/speech-based-emotion-recognition/37846>\
**Category:** Machine Learning\
**Tags:** question\
**Created:** [April 19, 2020, 12:56pm UTC](https://discourse.julialang.org/t/speech-based-emotion-recognition/37846 "2020-04-19T12:56:35Z")\
**Posts on this page:** 2\
**Page:** 1

<div class="post-metadata">

**Author:** ![themadprogramer](https://sea2.discourse-cdn.com/julialang/user_avatar/discourse.julialang.org/themadprogramer/32/14086_2.png) [@themadprogramer](https://discourse.julialang.org/u/themadprogramer)\
**Post date:** [April 19, 2020, 12:56pm UTC](https://discourse.julialang.org/t/speech-based-emotion-recognition/37846/1 "2020-04-19T12:56:35Z")

</div>

Hello,

I wasn’t able to find any package for Speech-based Mood/Emotion recognition to use in a project I’m working on, so I need some help setting up something myself.

I know that LSTM’s are in use for this, I’ve seen WaveNet adapted for almost every other audio problem at this point and I have even seen some of the slightly outdated Spatio-Temporal Box Filters.

I guess what I really want to ask is, what would be a good place to start from? Not necessarily the best, most accurate or even fastest approach; but the simplest to implement.

Thanks in advance!

---

<div class="post-metadata">

**Author:** ![apieum](https://sea2.discourse-cdn.com/julialang/user_avatar/discourse.julialang.org/apieum/32/2928_2.png) [@apieum](https://discourse.julialang.org/u/apieum)\
**Post date:** [April 19, 2020, 1:32pm UTC](https://discourse.julialang.org/t/speech-based-emotion-recognition/37846/2 "2020-04-19T13:32:51Z")

</div>

There’s a recent article that may be a good start: [Building an end-to-end Speech Recognition model in PyTorch](https://www.assemblyai.com/blog/end-to-end-speech-recognition-pytorch)  
You have a port of pytorch in Julia:  
[GitHub - boathit/JuliaTorch: Using PyTorch in Julia Language](https://github.com/boathit/JuliaTorch)

Pytorch support of LSTM:  
[https://pytorch.org/tutorials/beginner/nlp/sequence\_models\_tutorial.html](https://pytorch.org/tutorials/beginner/nlp/sequence_models_tutorial.html)

You have this GSOC project:

> **[GSoC 2018 and Speech Recognition for the Flux Model Zoo: The Conclusion](https://julialang.org/blog/2018/08/GSoC2018-speech-recognition/)**
>
> GSoC 2018 and Speech Recognition for the Flux Model Zoo: The Conclusion | Here we are on the other end of Google Summer of Code 2018. It has been a challenging and educational experience, and I wouldn't have it any other way. I am thankful to the...

And some usefull packages here:

> <https://github.com/svaksha/Julia.jl/blob/master/AI.md#speech-recognition>
