跳到正文
原文
X:Fei-Fei Li (@drfeifei, World Labs)· @a16z·· 28 天前AI 评分60

World Labs 联合创始人谈 Atlas:新视角预测是空间智能的 next-token prediction

RT by @drfeifei: World Labs co-founders Justin Johnson and Dr. Fei-Fei Li say LLMs use next-token prediction, but spatial intelligence has its own equivalent: Justin: "The soft definition of AI-completeness is there's this fundamental primitive that's an AI task. But if I could solve this AI task in its full, broadest generality, it would solve any intelligence problem." "The classic example in LLMs is that next-token prediction is AI-complete... I think from Ilya: there's a mystery novel, the thing has to read the whole novel, and the final sentence is, 'And the killer was.' Predict the next token. You could basically frame any kind of intelligence task in terms of that." "So clearly next-token prediction is something people believe is AI-complete." "New-view prediction, this primitive that we have in Atlas, especially generative new-view prediction, is also AI-complete." "I want to have a world where Martin is writing a proof of the Riemann hypothesis on the blackboard, and then the

AI 导读

World Labs 联合创始人 Justin Johnson 与李飞飞在 a16z 的对话中提出,LLM 建立在 next-token prediction 上,视频模型建立在下一帧预测上,而 Atlas 建立在新视角预测上,并称新视角预测同样是 AI-complete。

来源:X:Fei-Fei Li (@drfeifei, World Labs) · x.lingyaoai.com