vall-e/vall_e/models/arch
2025-01-28 21:55:05 -06:00
..
attention
__init__.py
bitnet.py
llama.py I should really just grab modelling_llama wholesale (fix for the adapted attention class) 2025-01-28 21:55:05 -06:00
mamba.py
mixtral.py oops 2025-01-21 11:59:24 -06:00
retnet.py
transformer.py