TypeSafe AI创始人谈AGI可实现性
TypeSafe AI's Diogo Almeida says AGI is extremely doable, yet basic work remains largely unautomated...
TypeSafe AI创始人Diogo Almeida批评行业过度承诺,其Jev模型专注于解决实际工作自动化问题,而非仅优化人类评估。
TypeSafe AI创始人Diogo Almeida表示,AGI非常可行,但基础工作自动化程度低。他认为OpenAI定义的AGI是自动化大部分经济价值工作,但行业过度优化人类评估而非自动化部分。Jev模型能读取自然语言并返回带置信度的选项选择,使程序能够推理意图并做出概率决策。
TypeSafe AI's Diogo Almeida says AGI is extremely doable, yet basic work remains largely unautomated...
TypeSafe AI's Diogo Almeida says AGI is extremely doable, yet basic work remains largely unautomated: "I still don't think we're on the path of RSI. I do think that what OpenAI defined as AGI is extremely doable: automating most of the world's economically valuable work." "There's a lot of work out there. A lot of it is very rote and simple... As far as I can tell, the intelligence of that has been available in the models for quite a while now. My chip on my shoulder is: Why is this not available?" "Since RLHF, the AI industry kind of bifurcated into gigantic overpromise, underdeliver... Because humans evaluate how good the models are, it looks really good because they are the judge. But we've been optimizing that judge instead of the automation part. That has been the missing thing." "Are you really telling me that math is solved, or even two years ago, GPQA... is solved, but we still can't handle a drive-through? It's a very hard thing to hold in your head at once. I think a lot of people don't have good answers to that." @CompleteSkeptic Your browser does not support the video tag. 🔗 View on Twitter a16z @a16z TypeSafe AI's Diogo Almeida with a16z's Ben Horowitz and Martin Casado on Jev, the model built to live inside software: Diogo's elevator pitch for Jev is a simple question - where is all the automation? AI is unbelievably smart, but outside of chatbots and coding agents, it hardly touches any real work. His diagnosis is the industry built models that generate text for humans to read, and software can't consume that output. Jev reads natural language and returns a choice from a set of options with a confidence level assigned to each, so developers can build programs that reason about intent and make probabilistic decisions rather than relying on human interpretation. TypeSafe's philosophy is "We build prod, not God." 0:50 "Where the f**k is all the automation?" 2:50 Jev vs. Claude Code and Codex 6:55 Jev is a classifier and classifiers are sick 7:40 Chat vs. code: is Jev a slider? 9:00 Diogo: From mathlete to Kaggle to OpenAI 12:20 "We build prod, not God" 15:55 Reliability over demos 16:55 2021 thoughts: RLHF is AGI? 20:45 Optimizing for the wrong use case 21:50 Is the real world too messy to automate? 25:00 Nobody expected the Jev launch 26:35 Three kinds of reliability 28:05 Good at syntax, bad at architecture 30:00 The inverse SaaSpocalypse 33:40 Why coding agents automate so little 36:05 Probabilistic programming returns 38:45 Jev as the UDP-to-TCP layer for AI 40:20 The 5 stages of grief for embedding AI 41:30 Utopia: AI that actually does what you mean YouTube: youtu.be/Ut3LOjKNJaE @CompleteSkeptic @typesafeai @bhorowitz @martin_casado Your browser does not support the video tag. 🔗 View on Twitter 🔗 View Quoted Tweet 💬 7 🔄 2 ❤️ 23 👀 9444 📊 7 ⚡