← Back to Model Beat
Models·Apr 14·all news from April 14, 2026

NVIDIA and the University of Maryland Researchers Released Audio Flamingo Next (AF-Next): A Super Powerful and Open Large Audio-Language Model

Understanding audio has always been the multimodal frontier that lags behind vision. While image-language models have rapidly scaled toward real-world deployment, building open models that robustly reason over speech, environmental sounds, and music — especially at length — has remained quite hard. NVIDIA and the University of Maryland researchers are now taking a direct swing […] The post NVIDIA and the University of Maryland Researchers Released Audio Flamingo Next (AF-Next): A Super Powerful and Open Large Audio-Language Model appeared first on MarkTechPost .

Covered by 1 source

Related stories

ModelsGemini can now pull from Google Photos to generate personalized imagesApr 16 · 3 sourcesModelsGemini App on MacApr 15 · 4 sourcesModelsIntroducing GPT-Rosalind for life sciences researchApr 16 · 3 sourcesModelsIntroducing Claude Opus 4.7 - AnthropicApr 16 · 9 sources