#سورس_کد #مقاله
در این روش که چند روز پیش توسط فیس بوک اوپن سورس شده آموزش speech recognition به صورت end-to-end صورت میگیرد .
Open sourcing wav2letter++, the fastest state-of-the-art speech system, and flashlight, an ML library going native
https://code.fb.com/ai-research/wav2letter/
CNN architectures are competitive with #recurrent architectures for tasks in which modeling long-range dependencies is important, such as #language_modeling, machine translation, and #speech_synthesis. In end-to-end #speech_recognition, however, recurrent architectures are still more prevalent for both acoustic and language modeling.
Post #850
2.6K