vall-e/vall_e/models
2025-04-07 22:51:52 -05:00
..
arch
__init__.py
ar_nar_v2.py
ar_nar.py
base_v2.py how foolish of me, not having a softmax as float32 (maybe addresses an emergent regression where bfloat16 training shits the bed where float16+loss scaling doesnt) 2025-04-07 22:51:52 -05:00
base.py
lora.py