正在加载视频...

视频加载失败

“Has a reasoning model ever come up with a math concept that even seems slightly interesting to a human mathematician?” Full episode w Ege Erdil & Tamay Besiroglu out Thursday.

158,508 次观看 • 1 年前 •via X (Twitter)

23 条评论

Prinz Eugen, der edle Ritter 的头像
Prinz Eugen, der edle Ritter1 年前

@EgeErdil2 @tamaybes Look, one can pretend that Move 37 never happened only for so long.

OnlineBookClub.org 的头像
OnlineBookClub.org1 年前

Logic dictates that something—or someone—always had to exist. Assume it was a “someone,” not a “something.” Why would such a being create a world like ours, one filled with pain? The Advent of Time provides a definitive answer.

Stefan Schubert 的头像
Stefan Schubert1 年前

@EgeErdil2 @tamaybes me on thursday

Inverse Zitron 的头像
Inverse Zitron1 年前

@EgeErdil2 @tamaybes this is gonna be a fun week for these guys when OpenAI introduces their $20k/mo scientists. but even now, I find it highly doubtful that they've produced nothing worthwhile for mathematicians

Marshal the Martian 的头像
Marshal the Martian1 年前

@EgeErdil2 @tamaybes Seems like a prompter issue. You need to ask the model to explore ideas across domains and make connections. It can't be that hard to start with a problem, iterate through nearby ideas looking for connections, and refine the most promising ones.

Nathan Labenz 的头像
Nathan Labenz1 年前

@EgeErdil2 @tamaybes What about FunSearch? Presumably this was interesting to at least some mathematicians? (I recognize this wasn't a pure reasoning model working in isolation, but not sure that's a critical distinction and in any case ... it's from late 2023!)

Louis Santoro 的头像
Louis Santoro1 年前

These kinds of criticisms are in a way unpaid labor for the research labs. It’s difficult to think of a persistent criticism they can’t just decide to work on. I would like to see a nontrivial formal critique because all the behavioral evidence in the world wouldn’t disprove that the next model can’t be an answer to whichever problem.

lucky 的头像
lucky1 年前

@EgeErdil2 @tamaybes if a reasoning model does succeed in coming up with a breakthrough insight in mathematics, will it be more optimal for the lab to share the result or keep it hidden ( in order to avoid regulations and such)?

🇨🇦halogen 的头像
🇨🇦halogen1 年前

@EgeErdil2 @tamaybes refreshing stuff

Tsukuyomi 的头像
Tsukuyomi1 年前

@EgeErdil2 @tamaybes interesting? more like math models trying to impress their crush. let’s see if they can charm a human mathematician or just end up in the friend zone.

Piyush 的头像
Piyush1 年前

@EgeErdil2 @tamaybes Wow! Excited for this one

Marco Piani 🇮🇹🇨🇦 的头像
Marco Piani 🇮🇹🇨🇦1 年前

@EgeErdil2 @tamaybes cc: @littmath

Ed 的头像
Ed1 年前

@EgeErdil2 @tamaybes bro your content is so good I always clear my schedule after work to watch it in my living room! Like watching on my phone doesn’t do justice to the Dwarkesh podcast experience. I’d pay you monthly subscription to get 2x weekly super long podcasts fr

Prateek 的头像
Prateek1 年前

@EgeErdil2 @tamaybes This seems interesting! While AI is reproducing a lot is a fact now, whether AI models are producing anything new at all is the real question.

Contompotentyy 的头像
Contompotentyy1 年前

@EgeErdil2 @tamaybes The new information report of true might hint more at how valid the recombination are

Manuel del Río Rodríguez 的头像
Manuel del Río Rodríguez1 年前

@EgeErdil2 @tamaybes Can't wait to listen to it!

IngoA 的头像
IngoA1 年前

@EgeErdil2 @tamaybes Love the different viewpoints!

Promptmetheus (COG/ACC) 的头像
Promptmetheus (COG/ACC)1 年前

@EgeErdil2 @tamaybes

Michael Forrest 的头像
Michael Forrest1 年前

@ancerj @EgeErdil2 @tamaybes People have extrapolated from it being able to retrieve and package info from its training set to it being able to discover and invent. Is this reasonable?

Danielle 的头像
Danielle1 年前

This is why the persistence of LLM memory is relevant to mathematical reasoning, though. LLMs might have the capacity to produce novel observations or hallucinate ideas, but piecing together insights enough to string together new theories still takes some iterations of prompting and analysis. (ChatGPT has modeled new mathematical shapes that I could not find in existing literature).

NotEvenTrying 的头像
NotEvenTrying1 年前

@EgeErdil2 @tamaybes if any, I bet human mathematicians won't disclose the discovery was helped by the reasoning model

Duane Stiller 的头像
Duane Stiller1 年前

@EgeErdil2 @tamaybes I’d be interested in seeing a poll from your followers on where they see the timeline.

Jacob Asmuth 的头像
Jacob Asmuth1 年前

@EgeErdil2 @tamaybes I wonder how many mathematicians in the world could provide an interesting answer to the prompt "come up with an interesting math concept" - even if given 6 months. I suspect that this is a skill that lies right on the very frontier of math - we're just not there yet.

相关视频