正在加载视频...

视频加载失败

We just released the Databricks synthetic evaluation SDK! 🎉 We’ve found that synthesizing evals is a great way to hill climb your AI system before you’re able to get labels from domain experts. With this SDK, and our fancy diffing UI, users are able to improve the inner loop...

10,446 次观看 • 1 年前 •via X (Twitter)

3 条评论

Hamel Husain 的头像
Hamel Husain1 年前

@databricks Looks cool! Do you need to be using MLFlow to take advantage of this? BTW I didn't even know until now that MLFlow had LLM tools in it!!!

floating point 的头像
floating point1 年前

@databricks I don't understand the workflow. So, say I see the agent made a wrong step, I see the diff and where. What is my next step? With agents I can't just add positive and negative demonstration to the training set, I have to debug the prompts, your tool does not solve that, right?

czverse.𝕏 的头像
czverse.𝕏1 年前

@databricks Game changer! Synth evals + diffing UI = faster, smarter iterations. Can't wait to try this out!

相关视频