[1]
Kim, M. et al. 2023. Deep Visual Forced Alignment: Learning to Align Transcription with Talking Face Video. Proceedings of the AAAI Conference on Artificial Intelligence. 37, 7 (Jun. 2023), 8273–8281. DOI:https://doi.org/10.1609/aaai.v37i7.25998.