Загрузка видео...

Не удалось загрузить видео

На главную

Replit CEO Amjad Masad on how general models could train smaller, domain-specific models on the fly: "There's a lot of talk of recursive self-improvement, but there's something I don't think is getting a lot of discussion, which is models training their replacements." "You can think of it as a...

125,646 просмотров • 6 дней назад •via X (Twitter)

Комментарии: 30

Фото профиля Jeramie Hicks
Jeramie Hicks6 дней назад

@amasad is 100% correct. General intelligence works for teaching. Specialized intelligence is for executing and super intelligence is for orchestration. A football team is made up of players who play specific positions but understand the game of football.

Фото профиля Damon Gardenhire
Damon Gardenhire6 дней назад

I have an idea to basically carry around a “desk drawer” of fine tuned 14b models for specific domains on an ssd sandisk pocket sized drive which I can plug into any computer and access via hermes, lm studio etc

Фото профиля Bradley Clonan
Bradley Clonan6 дней назад

Literally what I am building. But I don't have gpus so I've taken a slight side track to create a chromatic neural network. Basically, a native RGBD tensor library rather than numerical codecs so I can use light instead of numeric weights for all the heavy lifting. (LEDs I can afford)

Фото профиля andy
andy6 дней назад

available now at :)

Фото профиля Winston B.
Winston B.6 дней назад

The JIT analogy only works if there's a deopt path. A compiled function that hits a case it wasn't built for falls back to the interpreter, so the narrow model needs a way to hand off to the general one on unfamiliar input. Otherwise it fails quietly outside its lane.

Фото профиля ProbablyNothing🔸
ProbablyNothing🔸6 дней назад

Why can’t it code a child agent that inherits a set of rules from its experiences,and so it self learns,each will be made to solve a specific issue that needs time to solve.

Фото профиля ShadowAguy
ShadowAguy6 дней назад

the model upgrade path is now writing its own changelog, generate better docs and we’ll route next quarter’s traffic to you.

Фото профиля Pentra
Pentra6 дней назад

if a general model trains a domain-specific replacement, who owns that trained model, who can see what it learned, and who decides whether it ships? this becomes a governance problem.

Фото профиля Godfrey Lebo
Godfrey Lebo6 дней назад

Specialisation could help with cost, but I wouldn’t use smaller capability as the security boundary. A narrow model with permission to refund a payment can still do damage. The replacement needs scoped tools, adversarial tests and a way to decline the task.

Фото профиля Manny
Manny6 дней назад

It’s like least cost routing in telecom. You would do token spend then make up some variable called htokens aka harm tokens then you minimize that

Фото профиля Anderson
Anderson6 дней назад

the big models are now the interns training their replacements

Фото профиля Agentik
Agentik6 дней назад

Models training their own replacements on the fly keeps showing up in our notes. Nobody around us has a clean eval for that loop yet.

Фото профиля Deep Insight Labs
Deep Insight Labs6 дней назад

Yeah, us...

Фото профиля The Lucky Lighthouse
The Lucky Lighthouse6 дней назад

Yes

Фото профиля VC Radar
VC Radar6 дней назад

Models training their replacements is the first time in history the intern trains the boss.

Фото профиля Tyzo
Tyzo5 дней назад

Models training their replacements sounds kinda wild honestly

Фото профиля DC
DC5 дней назад

Right here. I have it.

Фото профиля Raven
Raven6 дней назад

i just got replaced by a smaller model and somehow still have to attend the meeting

Фото профиля Ava Nakamura
Ava Nakamura5 дней назад

Cheaper specialist. I still confirm before it spends.

Фото профиля Ryn Woo
Ryn Woo5 дней назад

Funny .

Фото профиля Gill
Gill6 дней назад

Distilling task specific models on the fly saves huge inference overhead.

Фото профиля icefrog.◎
icefrog.◎6 дней назад

the JIT compiler for intelligence is a savage move

Фото профиля REJAUL
REJAUL5 дней назад

training its replacements could change how quickly AI evolves

Фото профиля Jason Nocco
Jason Nocco5 дней назад

Could also use to train self learning enabled small domain specific models on the fly, privately. Automatically fine-tuned from grounded knowledge via our integrated knowledge graph. Built so anyone could easily create and quickly use their own models.

Фото профиля 青雲
青雲5 дней назад

这个即时编译的类比很顺,但 JIT 真正的麻烦在别处,是生成出来的那段代码和解释器对同一件事的理解要对齐。 放到这里就是,窄模型拿到的是训练那一刻的任务定义。任务挪一点,它自己看不出来,而且它越窄越不会说我不适用了。通用模型至少见过别的场景,有可能会停下来问一句。 我们那边有一条判据正好是这件事。一个进程拿着过时的游标来预定事件序号,现在的行为是直接报错,而不是产出一条和别人重叠的事件流。以前会默默重叠,它不报错、不丢数据,只是把两条时间线搅在一起。 所以训练替代品这件事,我觉得成败可能不在训练本身,在于替代品能不能说出我这条已经不适用了。 至于提示注入那条我只同意一半。能力更弱确实危害更小,但注入利用的是权限边界,窄模型接了哪些工具才是那条线。

Фото профиля Summer
Summer5 дней назад

The JIT compiler analogy makes so much sense.

Фото профиля Sagiv Ofek
Sagiv Ofek6 дней назад

The first generation to train their replacements: us. The second: the models.

Фото профиля Tyzo
Tyzo5 дней назад

Models training their replacements could accelerate AI development rapidly

Фото профиля Abdul Q.
Abdul Q.5 дней назад

Agency version of this: a general model is fine until you need Webflow CMS, Figma props, and B2B SaaS homepage taste in the same pass. Domain muscle still wins on marketing sites.

Фото профиля ZANA
ZANA6 дней назад

generating targeted lightweight models dynamically solves both the security vulnerabilities and high overhead of general agents

Похожие видео