NVIDIA 新出的 30B 模型,跑在本地,专门给常驻智能体用,速度快 4 倍,还支持 1M 上下文,搞智能体的可以试试。
NVIDIA 的 Nemotron 3.5 Lightning 现已可通过 Ollama 本地运行。这是一款 30B 参数的混合专家模型,仅激活 3B 参数,支持 1M token 上下文。该模型专为常驻运行的智能体设计,擅长编码、工具调用和多轮对话。相比同类规模的开源模型,其吞吐量提升 4 倍,任务完成时间降低 30%。用户可通过 Claude Code、Hermes Agent 或 OpenClaw 等工具直接调用。
NVIDIA Nemotron 3.5 Lighting is available on Ollama! It's a 30B model made for always-on agents. A...
NVIDIA Nemotron 3.5 Lighting is available on Ollama! It's a 30B model made for always-on agents. All local. Claude Code ollama launch claude --model nemotron-3.5-lightning Hermes Agent ollama launch hermes --model nemotron-3.5-lightning OpenClaw ollama launch openclaw --model nemotron-3.5-lightning - 30B mixture of experts model with 3B active - 1M token context - built for agents that stay running: coding, tool calling, multi-turn - 4x higher throughput and 30% lower task completion time compared to other leading open models of similar size Model page 👇👇👇 NVIDIA AI @NVIDIAAI Introducing NVIDIA Nemotron 3.5 Lightning⚡ An open 30B MoE model with 3B active parameters, built for always-on agents to complete high-volume, specialized tasks faster. It delivers up to 4x the output speed of similar-sized models. 🔗 View Quoted Tweet 💬 1 🔄 3 ❤️ 15 👀 1690 📊 3 ⚡