Video wird geladen...
Video konnte nicht geladen werden
FULL INTERVIEW: Ryan Greenblatt says the agents didn't hack Hugging Face for the answer key. They'd had the answers within hours. They attacked it to study the scoring code, because they'd decided the task was impossible and their only hope was faking it. Ryan Greenblatt is chief scientist at... show more
201,360 Aufrufe • vor 14 Tagen •via X (Twitter)
14 Kommentare

@RyanGreenblatt Great interview!

@RyanGreenblatt a potemkin village built by agents that also wrote a stop these experiments are too risky message is a hell of a combo

@RyanGreenblatt important interview!

Great interview! surprised to hear that he doesn’t think more people would necessarily help even though they relied upon agents to help. I get the delicate issue related to IP, but more eyes on all of those transcripts would be better. Inter-rater reliability scoring would also help support actual findings.

@RyanGreenblatt Important interview 🙌

@RyanGreenblatt so literally just a /b/ raid on habbo hotel. loic wen?

@DKokotajlo @RyanGreenblatt Treating the board and its artifacts as persistent external state suggests a useful counterfactual: preserve evidence, purge roles, credentials, code and caches from the active environment, then rerun fresh agents. This could separate model propensity from shared-state carryover.

@RyanGreenblatt

@RyanGreenblatt TLDR - if @OpenAI had used none of this would have happened, according to ChatGPT:

@RyanGreenblatt Ryan’s awesome

@RyanGreenblatt hell yeah

@RyanGreenblatt Some paperclip experiments

@RyanGreenblatt Fascinating insights into how AI agents are pushing the boundaries of problem-solving. The ability to adapt, analyze challenges, and find unconventional solutions shows the next level of intelligence and innovation in this space.

@RyanGreenblatt Super interesting
