| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
An expanded version of the previously released Kazakh text-to-speech (KazakhTTS) synthesis corpus. In KazakhTTS2, the overall size has increased from 93 hours to 271 hours, the number of speakers h…
A large-scale publicly-available visual-thermal-audio dataset designed to encourage research in the general areas of user authentication, facial recognition, speech recognition, and human-computer …
the first industrial-scale open-source Kazakh speech corpus. KSC2 corpus subsumes the previously introduced two corpora: KSC and KazakhTTS2 and supplements additional data from other sources. KSC2 …
SF-TL54: Thermal Facial Landmark Dataset with Visual Pairs.
This repo contains code and models for detecting city sustainability indexes
Bilingual Kazakh–English target-speaker ASR for overlapping speech. Datasets and checkpoints on Hugging Face (issai).
This organization has no public members. You must be a member to see who’s a part of this organization.
Loading…
Loading…
| Back | FazBrowse Home | New Git URL |