正在加载视频...
视频加载失败
✨ New AI Interfaces powered by Interpretability I'm excited to share LatentLit, the result of my applied AI research fellowship with Goodfire Mechanistic interpretability isn’t just important for AI safety, it also gives us new ways to steer and interact with LLMs.
11 条评论

In LatentLit, you write stories by like a DJ might make music, adjusting knobs and dials using steering and seeing what effect they have. You might call it Vibe Writing. Try it out here!

Our neural map highlights features like "Supernatural Discovery" and connects them to specific text passages they influence. Each colored bubble in the neural map represents a feature the model activates. The bigger the bubble, the stronger the influence. Hovering over these features will highlight where in the text this feature was activated most strongly.

There are thousands of features to explore! These features represent individuals concepts that the LLM understands, but they may surprise you in joyful ways when you use them to steer text generation.

Confused about AI? Get clarity today with Book VI - "The Rational Being!" Understand the benefits & risks, empower your future now by learning what AI really is and how it really works! Also check out the Free Weekly Newsletter "How Things Work: A Brief History of Reality"

@GoodfireAI neat! i got into a crazy broken state fairly easily playing w knobs tho. also when you remove a feature the % activation doesn't move down the indices with it

@GoodfireAI Ohh thanks for trying it out! That's definitely busted, let me see what I can do to fix it.

@livgorton @GoodfireAI This is so well done. I wanted to try something similar to get user feedback less intrusively than making them engage with buttons, and this is masterfully done.

@livgorton @GoodfireAI Thank you! A lot of love went into this, so glad that I had Goodfire's support in pushing it through

@GoodfireAI Very very cool!

@GoodfireAI This is 🔥 Need more products that showcase how different aspects of the neural network are impacting the end result.

@GoodfireAI Thank you! Honestly we have the capabilities, I'm hoping more people will be inspired to use them. You can even feed in the inference from another AI into Goodfire's SDK and read out the features that way

