英伟达这个34B开源VLA模型专攻自动驾驶,LingoQA拿79.2,还能商用改造,做Robotaxi的可以直接上手。
英伟达发布34B参数的开源视觉-语言-动作模型Alpamayo 2 Super,面向自动驾驶和Robotaxi场景。该模型采用32B Cosmos 3 Super Reasoner主干加2.3B扩散动作解码器,在LingoQA基准上得分79.2。它可在单次前向传播中输出轨迹、Chain-of-Causation因果链、元动作、自动标注和接地视觉问答。模型基于OpenMDW-1.1许可发布,允许微调、衍生和商业再分发。
NVIDIA Releases Alpamayo 2 Super: A 34B Open Vision-Language-Action Model for Robotaxis and Autonomous Driving Under OpenMDW-1.1
NVIDIA released Alpamayo 2 Super, a 34B vision-language-action model for autonomous driving, under OpenMDW-1.1 — a permissive license covering fine-tuning, derivatives and commercial redistribution. It pairs a 32B Cosmos 3 Super Reasoner backbone with a 2.3B diffusion action decoder, scores 79.2 on LingoQA, and emits trajectories, Chain-of-Causation traces, meta-actions, auto-labels and grounded VQA from a single pass. The post NVIDIA Releases Alpamayo 2 Super: A 34B Open Vision-Language-Action Model for Robotaxis and Autonomous Driving Under OpenMDW-1.1 appeared first on MarkTechPost .