Nearby in the stack

COALA: Co-Aligned Autoencoders for Learning Semantically Enriched Audio Representations · arXivDesk