Eren Gölge
|
419735f440
|
refactor and fix multi-speaker training in Trainer and Tacotron models
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
269e5a734e
|
add max_decoder_steps argument to tacotron models
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
b3324bd914
|
fix speaker_manager init
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
304d60197b
|
reduce multiband melgan test model size
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
2c38ef8441
|
use get_speaker_manager in Trainer and save speakers.json file when
needed
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
d6b2b6add6
|
make style and linter fixes
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
802d461389
|
Compute d_vectors and speaker_ids separately in TTSDataset
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
db6a97d1a2
|
rename external speaker embedding arguments as d_vectors
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
9042ae9195
|
use to_cuda() for moving data in format_batch()
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
877bf66b61
|
reduce size of the metadata.csv used at testing
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
f82f1970b8
|
change to(device) to type_as in models
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
9c94b0c5c0
|
init durations = None
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
1fa15c195a
|
docstring fix
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
1c8a3d7c86
|
make style
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
8cdd423234
|
styling formatting.py
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
b9a52dce9e
|
add test_all to makefile
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
30211512a4
|
fix type annotations
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
87c61d210a
|
update test to be less demanding
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
6d6896fd99
|
reduce fullband-melgan test model size
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
1443d03af1
|
update test for the new input output API of the tts models
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
b22b7620c3
|
update glow-tts output shapes to match [B, T, C]
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
8381379938
|
formating cond_input with a function in Tacotron models
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
ef4ea9e527
|
update imports for formatters
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
6c495c6a6e
|
fix glow-tts inference and forward functions for handling cond_input
and refactor its test
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
f840268181
|
refactor SpeakerManager
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
421194880d
|
linter fixes
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
8e52a69230
|
delete separate tts training scripts and pre-commit configuration
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
d96ebcd6d3
|
make style
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
b643e8b37c
|
logging/__init__.py
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
0cee5042a9
|
fix logger imports
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
72dceca52c
|
import missings
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
0eec238429
|
remove redundant imports
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
b500338faa
|
make style
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
469d2e620a
|
update extract_tts_spectrogram for cond_input API of the models
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
5ab28fa618
|
update extract_tts_spec... using SpeakerManager
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
c392fa4288
|
update extract_tts_spectrograms for the new model API
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
8f47f95998
|
correct import of load_meta_data
remove redundant import
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
c680a07a20
|
fix Synthesized for the new synthesis()
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
73bf9673ed
|
revert logging.info to print statements for trainer
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
d25f017b42
|
update setup_model.py imports
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
7dff6be871
|
update tts training tests to use the trainer
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
bb355b7441
|
update align_tts.py model for the trainer
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
9203b863d9
|
update align_tts_loss for trainer
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
fc9a0fb8ce
|
update aling_tts_config for the trainer
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
e298b8e364
|
update trainer.py for better logging handling, restoring models and
rename init_ functions with get_
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
b8a4af4010
|
update synthesis.py for being more generic
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
c70d0c9dae
|
update speedy_speech.py model for trainer
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
06ee57d816
|
update speedy_speecy_config.py for the trainer
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
4e910993f1
|
update tacotron model to return model_outputs
|
2021-06-28 17:03:19 +02:00 |
|
Eren Gölge
|
bb4deee64c
|
update glow-tts for the trainer
|
2021-06-28 17:03:19 +02:00 |
|