Andrew Ng 的课程一向实用,这次聚焦图像/视频生成智能体这个少有人深入的方向,做多模态或内容生成的开发者可以直接学起来,掌握让智能体自我迭代的关键技巧。
Andrew Ng 宣布与 Google Cloud 合作推出新课程,教授如何构建能生成图像和视频的 AI 智能体。课程重点在于让智能体自我评估输出并迭代改进质量,涵盖三种评估技术:图像-文本相似度评分、LLM 裁判按品牌一致性等自定义标准评分、以及结构化评分表。学员将学习图像和视频提示工程,构建将品牌指南转化为 UI 模型的图像智能体,以及规划多场景解说视频并同步音频的视频智能体。该课程面向希望探索 AI 智能体在视觉内容生成领域应用的开发者。
New course: Build AI agents that generate images and videos -- an under-explored frontier. A key to ...
New course: Build AI agents that generate images and videos -- an under-explored frontier. A key to performance is having the agent evaluate its own output, and iterate to improve quality. This short course is built together with @googlecloudtech and taught by Katie Nguyen and Wafae Bakkali. You'll learn three evaluation techniques and combine them in an agent: image-text similarity scoring to check the output matches the prompt, an LLM judge that scores against custom criteria like brand consistency, and structured rubrics that break a prompt into verifiable yes/no questions like "is the subject in the frame?" and "does the camera motion match?" Skills you'll gain: - Learn image and video prompt engineering - Build an image agent that turns brand guidelines into UI mockups - Build a video agent that plans multi-scene explainers and animates reference frames with synchronized audio Join and build agents that create images and video! deeplearning.ai/courses/ai-age… Your browser does not support the video tag. 🔗 View on Twitter 💬 44 🔄 39 ❤️ 297 👀 29037 📊 99 ⚡