正在加载视频...
视频加载失败
We just released the Databricks synthetic evaluation SDK! 🎉 We’ve found that synthesizing evals is a great way to hill climb your AI system before you’re able to get labels from domain experts. With this SDK, and our fancy diffing UI, users are able to improve the inner loop... show more
3 条评论

Hamel Husain1 年前
@databricks Looks cool! Do you need to be using MLFlow to take advantage of this? BTW I didn't even know until now that MLFlow had LLM tools in it!!!

floating point1 年前
@databricks I don't understand the workflow. So, say I see the agent made a wrong step, I see the diff and where. What is my next step? With agents I can't just add positive and negative demonstration to the training set, I have to debug the prompts, your tool does not solve that, right?

czverse.𝕏1 年前
@databricks Game changer! Synth evals + diffing UI = faster, smarter iterations. Can't wait to try this out!
