FireRedTeam Released FireRedAudio and FireRedTTS3 on Hugging Face
r/LocalLLaMA (55 upvotes)·low signal
FireRedTeam published FireRedAudio, described as a general-purpose audio language model using decoupled continuous representations for both understanding and generation, alongside FireRedTTS3, with weights on Hugging Face. Single-vendor announcement with no third-party benchmark yet, so treat the capability claims as unverified. The notable design choice is unifying audio understanding and generation in one model rather than shipping separate ASR and TTS stacks.