James Betker
|
bb0a0c8264
|
classify_into_folders script
|
2021-10-27 14:56:16 -06:00 |
|
James Betker
|
d91dcbd404
|
Make classifier inference script more open
|
2021-10-27 13:18:54 -06:00 |
|
James Betker
|
58494b0888
|
Add support for distilling gpt_asr
|
2021-10-27 13:10:07 -06:00 |
|
James Betker
|
5d714bc566
|
Add deepspeech model and support for decoding with it
|
2021-10-27 13:09:46 -06:00 |
|
James Betker
|
15437b2fc3
|
WER script
|
2021-10-26 13:30:29 -06:00 |
|
James Betker
|
3a9d1c53ea
|
Rework conditioning inputs provided
|
2021-10-26 10:46:33 -06:00 |
|
James Betker
|
21b6daa0ed
|
Introduce clip resampling
|
2021-10-26 10:42:23 -06:00 |
|
James Betker
|
43e389aac6
|
Add time_embed_dim_multiplier
|
2021-10-26 08:55:55 -06:00 |
|
James Betker
|
ba6e46c02a
|
Further simplify diffusion_vocoder and make noise_surfer work
|
2021-10-26 08:54:30 -06:00 |
|
James Betker
|
c3421b7f6d
|
Dataset work for audio quality processor
|
2021-10-24 09:09:34 -06:00 |
|
James Betker
|
0ee1c67ce5
|
Rework how conditioning inputs are applied to DiffusionVocoder
|
2021-10-24 09:08:58 -06:00 |
|
James Betker
|
b1248e7114
|
Get rid of filter_urbansounds
|
2021-10-21 16:46:04 -06:00 |
|
James Betker
|
06ea6191a9
|
Initial implementation of audio_with_noise dataset
|
2021-10-21 16:45:19 -06:00 |
|
James Betker
|
9a3e89ec53
|
Force LR fix
|
2021-10-21 12:01:01 -06:00 |
|
James Betker
|
40cb25292a
|
Fix force_lr logic
|
2021-10-21 11:51:30 -06:00 |
|
James Betker
|
0dee15f875
|
base DVAE & vector_quantizer
|
2021-10-20 21:19:38 -06:00 |
|
James Betker
|
f2a31702b5
|
Clean stuff up, move more things into arch_util
|
2021-10-20 21:19:25 -06:00 |
|
James Betker
|
a6f0f854b9
|
Fix codes when inferring from dvae
|
2021-10-17 22:51:17 -06:00 |
|
James Betker
|
d016a2fbad
|
Go back to vanilla flavor of diffusion
|
2021-10-17 17:32:46 -06:00 |
|
James Betker
|
23da073037
|
Norm decoder outputs now
|
2021-10-16 09:07:10 -06:00 |
|
James Betker
|
0edc98f6c4
|
Throw out the idea of conditioning on discrete codes. Oh well :(
|
2021-10-16 09:02:01 -06:00 |
|
James Betker
|
62c8c5d93e
|
Zero out spectrogram code inputs initially.
|
2021-10-15 12:10:11 -06:00 |
|
James Betker
|
1d0b44ebc2
|
More tweaks to diffusion-vocoder
|
2021-10-15 11:51:17 -06:00 |
|
James Betker
|
3b19581f9a
|
Allow num_resblocks to specified per-level
|
2021-10-14 11:26:04 -06:00 |
|
James Betker
|
83798887a8
|
Mods to support unet diffusion vocoder with conditioning
|
2021-10-13 21:23:18 -06:00 |
|
James Betker
|
c861054218
|
Restore spleeter_splitter
The mods don't help - in TF mode, everything is done on the GPU anyways. Something else
is going to have to be done to fix this.
|
2021-10-09 23:55:42 -06:00 |
|
James Betker
|
32ba496632
|
More fixes
|
2021-10-09 23:27:14 -06:00 |
|
James Betker
|
932ea29a83
|
Add multiprocessing to the spleeter splitter script to try and improve performance further
|
2021-10-09 23:15:36 -06:00 |
|
James Betker
|
b94e587f46
|
Improvements to spleeter_filter_noisy_clips
|
2021-10-07 21:28:00 -06:00 |
|
James Betker
|
33120cb35c
|
Add norming to discretization_loss
|
2021-10-06 17:10:50 -06:00 |
|
James Betker
|
bb891a3a53
|
Add partitioning and improved resuming to the spleeter filtering
|
2021-10-06 17:10:12 -06:00 |
|
James Betker
|
f2977d360c
|
Allow attention_dim in channel attention to be specified, add converter
|
2021-10-05 17:29:38 -06:00 |
|
James Betker
|
9c0d7288ea
|
Discretization loss attempt
|
2021-10-04 20:59:21 -06:00 |
|
James Betker
|
66f99a159c
|
Rev2
|
2021-10-03 15:20:50 -06:00 |
|
James Betker
|
09f373e3b1
|
Add dvae with channel attention
|
2021-10-03 10:52:01 -06:00 |
|
James Betker
|
0396a9d2ca
|
Increase baseline codes recording across all dvae models
|
2021-09-30 08:09:07 -06:00 |
|
James Betker
|
f84ccbdfb2
|
Fix quantizer with balancing_heuristic
|
2021-09-29 14:46:05 -06:00 |
|
James Betker
|
4914c526dc
|
More cleanup
|
2021-09-29 14:24:49 -06:00 |
|
James Betker
|
6e550edfe3
|
Attentive dvae
|
2021-09-29 14:17:29 -06:00 |
|
James Betker
|
fc8ae4679a
|
Work on spleeter filtering script
|
2021-09-29 09:24:56 -06:00 |
|
James Betker
|
55b58fb67f
|
Clean up codebase
Remove stuff that I'm likely not going to use again (or generally failed experiments)
|
2021-09-29 09:21:44 -06:00 |
|
James Betker
|
4d1a42e944
|
Add switchnorm to gumbel_quantizer
|
2021-09-24 18:49:25 -06:00 |
|
James Betker
|
ac57cdc794
|
Add scheduling to quantizer, enable cudnn_benchmarking to be disabled
|
2021-09-24 17:01:36 -06:00 |
|
James Betker
|
3e64e847c2
|
Gumbel quantizer
|
2021-09-23 23:32:03 -06:00 |
|
James Betker
|
c5297ccec6
|
Add dvae balancing heuristic
|
2021-09-23 21:19:36 -06:00 |
|
James Betker
|
e24c619387
|
Fix
|
2021-09-23 16:07:58 -06:00 |
|
James Betker
|
6833048bf7
|
Alterations to diffusion_dvae so it can be used directly on spectrograms
|
2021-09-23 15:56:25 -06:00 |
|
James Betker
|
97ea329a59
|
Make spleeter filter simpler (and hopefully much faster)
|
2021-09-17 15:29:42 -06:00 |
|
James Betker
|
359e9e27a7
|
unsupervised_audio_dataset: try to recover from failures of audio2numpy
|
2021-09-17 15:25:57 -06:00 |
|
James Betker
|
5c8d266d4f
|
chk
|
2021-09-17 09:15:36 -06:00 |
|