James Betker
|
58ed27d7a8
|
new gap_filler
|
2022-05-07 12:44:23 -06:00 |
|
James Betker
|
6c8032b4be
|
more work
|
2022-05-06 21:56:49 -06:00 |
|
James Betker
|
f541610256
|
contrastive_audio
|
2022-05-06 16:37:22 -06:00 |
|
James Betker
|
79543e5488
|
Simpler form of the wavegen model
|
2022-05-06 16:37:04 -06:00 |
|
James Betker
|
d8925ccde5
|
few things with gap filling
|
2022-05-06 14:33:44 -06:00 |
|
James Betker
|
b13d983c24
|
and mel_head
|
2022-05-06 00:25:27 -06:00 |
|
James Betker
|
d5fb79564a
|
remove mel_pred
|
2022-05-06 00:24:05 -06:00 |
|
James Betker
|
e9bb692490
|
fixed aligned_latent
|
2022-05-06 00:20:21 -06:00 |
|
James Betker
|
1609101a42
|
musical gap filler
|
2022-05-05 16:47:08 -06:00 |
|
James Betker
|
d66ab2d28c
|
Remove unused waveform_gens
|
2022-05-04 21:06:54 -06:00 |
|
James Betker
|
47662b9ec5
|
some random crap
|
2022-05-04 20:29:23 -06:00 |
|
James Betker
|
c42c53e75a
|
Add a trainable network for converting a normal distribution into a latent space
|
2022-05-02 09:47:30 -06:00 |
|
James Betker
|
ab219fbefb
|
output variance
|
2022-05-02 00:10:33 -06:00 |
|
James Betker
|
3b074aac34
|
add checkpointing
|
2022-05-02 00:07:42 -06:00 |
|
James Betker
|
ae5f934ea1
|
diffwave
|
2022-05-02 00:05:04 -06:00 |
|
James Betker
|
b712d3b72b
|
break out get_conditioning_latent from unified_voice
|
2022-05-01 23:04:44 -06:00 |
|
James Betker
|
afa2df57c9
|
gen3
|
2022-04-30 10:41:38 -06:00 |
|
James Betker
|
8aa6651fc7
|
fix surrogate loss return in waveform_gen2
|
2022-04-28 10:10:11 -06:00 |
|
James Betker
|
f02b01bd9d
|
reverse univnet classifier
|
2022-04-20 21:37:55 -06:00 |
|
James Betker
|
9df85c902e
|
New gen2
Which is basically a autoencoder with a giant diffusion appendage attached
|
2022-04-20 21:37:34 -06:00 |
|
James Betker
|
b4549eed9f
|
uv2 fix
|
2022-04-20 00:27:38 -06:00 |
|
James Betker
|
24fdafd855
|
fix2
|
2022-04-20 00:03:29 -06:00 |
|
James Betker
|
0af0051399
|
fix
|
2022-04-20 00:01:57 -06:00 |
|
James Betker
|
419f4d37bd
|
gen2 music
|
2022-04-19 23:38:37 -06:00 |
|
James Betker
|
8fe0dff33c
|
support tts typing
|
2022-04-16 23:36:57 -06:00 |
|
James Betker
|
48cb6a5abd
|
misc
|
2022-04-16 20:28:04 -06:00 |
|
James Betker
|
147478a148
|
cvvp
|
2022-04-16 20:27:46 -06:00 |
|
James Betker
|
546ecd5aeb
|
music!
|
2022-04-15 21:21:37 -06:00 |
|
James Betker
|
254357724d
|
gradprop
|
2022-04-15 09:37:20 -06:00 |
|
James Betker
|
fbf1f4f637
|
update
|
2022-04-15 09:34:44 -06:00 |
|
James Betker
|
82aad335ba
|
add distributued logic for loss
|
2022-04-15 09:31:48 -06:00 |
|
James Betker
|
efe12cb816
|
Update clvp to add masking probabilities in conditioning and to support code inputs
|
2022-04-15 09:11:23 -06:00 |
|
James Betker
|
8ea5c307fb
|
Fixes for training the diffusion model on autoregressive inputs
|
2022-04-11 11:02:44 -06:00 |
|
James Betker
|
a3622462c1
|
Change latent_conditioner back
|
2022-04-11 09:00:13 -06:00 |
|
James Betker
|
03d0b90bda
|
fixes
|
2022-04-10 21:02:12 -06:00 |
|
James Betker
|
19ca5b26c1
|
Remove flat0 and move it into flat
|
2022-04-10 21:01:59 -06:00 |
|
James Betker
|
81c952a00a
|
undo relative
|
2022-04-08 16:32:52 -06:00 |
|
James Betker
|
944b4c3335
|
more undos
|
2022-04-08 16:31:08 -06:00 |
|
James Betker
|
032983e2ed
|
fix bug and allow position encodings to be trained separately from the rest of the model
|
2022-04-08 16:26:01 -06:00 |
|
James Betker
|
09ab1aa9bc
|
revert rotary embeddings work
I'm not really sure that this is going to work. I'd rather explore re-using what I've already trained
|
2022-04-08 16:18:35 -06:00 |
|
James Betker
|
2fb9ffb0aa
|
Align autoregressive text using start and stop tokens
|
2022-04-08 09:41:59 -06:00 |
|
James Betker
|
423293e518
|
fix xtransformers bug
|
2022-04-08 09:12:46 -06:00 |
|
James Betker
|
048f6f729a
|
remove lightweight_gan
|
2022-04-07 23:12:08 -07:00 |
|
James Betker
|
e634996a9c
|
autoregressive_codegen: support key_value caching for faster inference
|
2022-04-07 23:08:46 -07:00 |
|
James Betker
|
d05e162f95
|
reformat x_transformers
|
2022-04-07 23:08:03 -07:00 |
|
James Betker
|
7c578eb59b
|
Fix inference in new autoregressive_codegen
|
2022-04-07 21:22:46 -06:00 |
|
James Betker
|
3f8d7955ef
|
unified_voice with rotary embeddings
|
2022-04-07 20:11:14 -06:00 |
|
James Betker
|
573e5552b9
|
CLVP v1
|
2022-04-07 20:10:57 -06:00 |
|
James Betker
|
71b73db044
|
clean up
|
2022-04-07 11:34:10 -06:00 |
|
James Betker
|
6fc4f49e86
|
some dumb stuff
|
2022-04-07 11:32:34 -06:00 |
|