技巧精选73°

PyTorch Conference 议程:用 PyTorch 与 vLLM 搭建企业级智能体推理

精选理由

Red Hat 的人在 PyTorch Conference 讲怎么用 vLLM 把智能体推理跑到 7x24 企业级,聊 KV cache、并发这些真问题,做后端的可以看看。

Red Hat 的 Joseph Groenenboom 将在 10 月 20-21 日于 San Jose 举办的 PyTorch Conference North America 上演讲,主题是用 PyTorch 与 vLLM 把智能体推理做到企业级生产可用。演讲聚焦 24/7 企业场景下的可靠性、可观测性、KV cache 管理与并发处理,覆盖从核心构建基础设施到模型服务的改进,包括工具调用支持和长上下文多轮对话。演讲同时由 PyTorch Foundation 生态系统工作组介绍 Ecosystem Landscape 的申请与生命周期管理机制。

原文 · PyTorch

Curious about how to make enterprise agentic inference production-ready with @PyTorch and @vllm_project & to learn more about the PyTorch Landscape? Join Joseph Groenenboom of @RedHat at PyTorch Conference North America.

While serving AI models for research and pilot use cases is a well solved problem, moving it to 24/7 Enterprise ready systems is the next important phase in AI maturity. The reliability, observability, KV cache management, and concurrency requirements for Enterprise readiness are non-trivial problems. Rising to this challenge, PyTorch, vLLM, and the other foundation projects and broader ecosystem have begun adding Enterprise level features and enhancements. This talk will cover a sampling of some of the project work; from core PyTorch project build infrastructure up to model serving improvements to account for tool calling support and long context multi-turn chat.

Led by the PyTorch Foundation Ecosystem Working Group, you will learn what the Ecosystem Landscape is and why it matters: how membership drives visibility and community engagement for independent projects; how to apply: a lightweight, GitHub-based process that makes it straightforward for projects to apply for ecosystem status; and how lifecycle management works: what ongoing membership looks like and how the Working Group supports active projects.

See you at #PyTorchCon NA in San Jose, October 20-21! Register at: https://t.co/jBApW8nESi