work SmallALM A lightweight 135M-parameter Audio Language Model for audio and speech understanding. SSAST-MLM Pre-training the SSAST audio foundation model with a unified masked-prediction (MLM-style) loss. fun