James Betker
|
b7319ab518
|
Support vocoder type diffusion in audio_diffusion_fid
|
2022-02-23 17:25:16 -07:00 |
|
James Betker
|
58f6c9805b
|
adf
|
2022-02-22 23:12:58 -07:00 |
|
James Betker
|
03752c1cd6
|
Report NaN
|
2022-02-22 23:09:37 -07:00 |
|
James Betker
|
7201b4500c
|
default text_to_sequence cleaners
|
2022-02-21 19:14:22 -07:00 |
|
James Betker
|
ba7f54c162
|
w2v: new inference function
|
2022-02-21 19:13:03 -07:00 |
|
James Betker
|
896ac029ae
|
allow continuation of samples encountered
|
2022-02-21 19:12:50 -07:00 |
|
James Betker
|
6313a94f96
|
eval: integrate a n-gram language model into decoding
|
2022-02-21 19:12:34 -07:00 |
|
James Betker
|
af50afe222
|
pairedvoice: error out if clip is too short
|
2022-02-21 19:11:10 -07:00 |
|
James Betker
|
38802a96c8
|
remove timesteps from cond calculation
|
2022-02-21 12:32:21 -07:00 |
|
James Betker
|
668876799d
|
unet_diffusion_tts7
|
2022-02-20 15:22:38 -07:00 |
|
James Betker
|
0872e17e60
|
unified_voice mods
|
2022-02-19 20:37:35 -07:00 |
|
James Betker
|
7b12799370
|
Reformat mel_text_clip for use in eval
|
2022-02-19 20:37:26 -07:00 |
|
James Betker
|
bcba65c539
|
DataParallel Fix
|
2022-02-19 20:36:35 -07:00 |
|
James Betker
|
34001ad765
|
et
|
2022-02-18 18:52:33 -07:00 |
|
James Betker
|
baf7b65566
|
Attempt to make w2v play with DDP AND checkpointing
|
2022-02-18 18:47:11 -07:00 |
|
James Betker
|
f3776f1992
|
reset ctc loss from "mean" to "sum"
|
2022-02-17 22:00:58 -07:00 |
|
James Betker
|
2b20da679c
|
make spec_augment a parameter
|
2022-02-17 20:22:05 -07:00 |
|
James Betker
|
a813fbed9c
|
Update to evaluator
|
2022-02-17 17:30:33 -07:00 |
|
James Betker
|
e1d71e1bd5
|
w2v_wrapper: get rid of ctc attention mask
|
2022-02-15 20:54:40 -07:00 |
|
James Betker
|
79e8f36d30
|
Convert CLIP models into new folder
|
2022-02-15 20:53:07 -07:00 |
|
James Betker
|
8f767b8b4f
|
...
|
2022-02-15 07:08:17 -07:00 |
|
James Betker
|
29e07913a8
|
Fix
|
2022-02-15 06:58:11 -07:00 |
|
James Betker
|
dd585df772
|
LAMB optimizer
|
2022-02-15 06:48:13 -07:00 |
|
James Betker
|
2bdb515068
|
A few mods to make wav2vec2 trainable with DDP on DLAS
|
2022-02-15 06:28:54 -07:00 |
|
James Betker
|
52b61b9f77
|
Update scripts and attempt to figure out how UnifiedVoice could be used to produce CTC codes
|
2022-02-13 20:48:06 -07:00 |
|
James Betker
|
a4f1641eea
|
Add & refine WER evaluator for w2v
|
2022-02-13 20:47:29 -07:00 |
|
James Betker
|
e16af944c0
|
BSO fix
|
2022-02-12 20:01:04 -07:00 |
|
James Betker
|
29534180b2
|
w2v fine tuner
|
2022-02-12 20:00:59 -07:00 |
|
James Betker
|
0c3cc5ebad
|
use script updates to fix output size disparities
|
2022-02-12 20:00:46 -07:00 |
|
James Betker
|
15fd60aad3
|
Allow EMA training to be disabled
|
2022-02-12 20:00:23 -07:00 |
|
James Betker
|
3252972057
|
ctc_code_gen mods
|
2022-02-12 19:59:54 -07:00 |
|
James Betker
|
35170c77b3
|
fix sweep
|
2022-02-11 11:43:11 -07:00 |
|
James Betker
|
c6b6d120fe
|
fix ranking
|
2022-02-11 11:34:57 -07:00 |
|
James Betker
|
095944569c
|
deep_update dicts
|
2022-02-11 11:32:25 -07:00 |
|
James Betker
|
ab1f6e8ac6
|
deepcopy map
|
2022-02-11 11:29:32 -07:00 |
|
James Betker
|
496fb81997
|
use fork instead
|
2022-02-11 11:22:25 -07:00 |
|
James Betker
|
4abc094b47
|
fix train bug
|
2022-02-11 11:18:15 -07:00 |
|
James Betker
|
006add64c5
|
sweep fix
|
2022-02-11 11:17:08 -07:00 |
|
James Betker
|
102142d1eb
|
f
|
2022-02-11 11:05:13 -07:00 |
|
James Betker
|
40b08a52d0
|
dafuk
|
2022-02-11 11:01:31 -07:00 |
|
James Betker
|
f6a7f12cad
|
Remove broken evaluator
|
2022-02-11 11:00:29 -07:00 |
|
James Betker
|
46b97049dc
|
Fix eval
|
2022-02-11 10:59:32 -07:00 |
|
James Betker
|
5175b7d91a
|
training sweeper checkin
|
2022-02-11 10:46:37 -07:00 |
|
James Betker
|
302ac8652d
|
Undo mask during training
|
2022-02-11 09:35:12 -07:00 |
|
James Betker
|
618a20412a
|
new rev of ctc_code_gen with surrogate LM loss
|
2022-02-10 23:09:57 -07:00 |
|
James Betker
|
d1d1ae32a1
|
audio diffusion frechet distance measurement!
|
2022-02-10 22:55:46 -07:00 |
|
James Betker
|
23a310b488
|
Fix BSO
|
2022-02-10 20:54:51 -07:00 |
|
James Betker
|
1e28e02f98
|
BSO improvement to make it work with distributed optimizers
|
2022-02-10 09:53:13 -07:00 |
|
James Betker
|
836eb08afb
|
Update BSO to use the proper step size
|
2022-02-10 09:44:15 -07:00 |
|
James Betker
|
820a29f81e
|
ctc code gen mods
|
2022-02-10 09:44:01 -07:00 |
|