Real time voice cloning

dc.contributor.authorAmrutha H S 1NH22MC011
dc.date.accessioned2024-12-06T10:57:54Z
dc.date.available2024-12-06T10:57:54Z
dc.date.issued2024
dc.description.abstractThis Project describe a text-to-speech (TTS) synthesis this is able to generate speech audio inside facet the voice of diverse audio gadget, consisting of those unseen withinside the route of education. Our device consists of three independently knowledgeable components: speaker encoder network, knowledgeable on a speaker verification assignment the usage of an impartial dataset of noisy speech without transcripts from masses of audio gadget, to generate a fixed-dimensional embedding vector from great seconds of reference speech from a purpose speaker; a series-to-series synthesis network based totally mostly on Tacotron 2 that generates a mel spectrogram from text, conditioned at the speaker embedding;
dc.identifier.urihttp://192.168.75.5:4000/handle/123456789/16586
dc.language.isoen
dc.publisherNHCE
dc.titleReal time voice cloning
dc.typeLearning Object
Files
Original bundle
Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
1NH22MC011-AMRUTHA H S.pdf
Size:
1.36 MB
Format:
Adobe Portable Document Format
License bundle
Now showing 1 - 1 of 1
No Thumbnail Available
Name:
license.txt
Size:
1.71 KB
Format:
Item-specific license agreed upon to submission
Description:
Collections