Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

💥 OSS product analytics for LLM apps. Launching PostHog x langfuse.com integration 🤝 🫳 Day 1 drop of Langfuse Launch Week: We’re marrying product analytics and LLM observability. This integration pipes Langfuse events/metrics/evals into your PostHog project. This allows you to access all of your LLM metrics straight from...

15,658 Aufrufe • vor 2 Jahren •via X (Twitter)

10 Kommentare

Profilbild von Marc Klingen
Marc Klingenvor 2 Jahren

@langfuse Looking forward to @hassiebpakzad's drop on Day 2 ✨

Profilbild von Clemens Rawert
Clemens Rawertvor 2 Jahren

@posthog @langfuse just getting started!

Profilbild von Thomas Paul Mann
Thomas Paul Mannvor 2 Jahren

@posthog @langfuse Very cool!

Profilbild von Hassieb Pakzad
Hassieb Pakzadvor 2 Jahren

@posthog @langfuse 🚀🚀🚢🚢

Profilbild von Marc Klingen
Marc Klingenvor 2 Jahren

@posthog @langfuse might just rename it to `ship week` or `ltgm week`

Profilbild von Brandon
Brandonvor 2 Jahren

@posthog @langfuse Dope! I just started using posthog this weekend.

Profilbild von Marc Klingen
Marc Klingenvor 2 Jahren

@posthog @langfuse great choice, posthog is awesome and super versatile with its many SDKs. we've been using it to track front-end events, backend ingestion metrics, and very basic self-hosting telemetry let me know if you have any feedback while getting started with this integration

Profilbild von Lior Neu-ner
Lior Neu-nervor 2 Jahren

@posthog @langfuse 🔥🔥🔥

Profilbild von Alex Krausse
Alex Kraussevor 2 Jahren

@rawert @posthog @langfuse 🚀🙌

Profilbild von Nikhil Kulkarni
Nikhil Kulkarnivor 2 Jahren

@posthog @langfuse 🔥🔥

Ähnliche Videos

New short course: Evaluating AI Agents! Evals are important for driving AI system improvements, and in this course you'll learn to systematically assess and improve an AI agent’s performance. This is built in partnership with Arize AI and taught by John Gilhuly, Head of Developer Relations, and , Director of Product. I've often found evals to be a critical tool in the agent development process - they can be the difference between picking the right thing to work on vs. wasting weeks of effort. Whether you’re building a shopping assistant, coding agent, or research assistant, having a structured evaluation process helps you refine its performance systematically, rather than relying on random trial and error. This course shows you how to structure your evals to assess the performance of each component of an agent and its end-to-end performance. For each component, you select the appropriate evaluators, test examples, and performance metrics. This helps you identify areas for improvement both during development and in production. (If you're familiar with error analysis in supervised learning, think of this as adapting those ideas to agentic workflows.) In this course, you'll build an AI agent, and add observability to visualize and debug its steps. You’ll learn about code-based evals, in which you write code explicitly to test a certain step, as well as LLM-as-a-Judge evals, in which you prompt an LLM to efficiently come up with ways to evaluate more open-ended outputs. In detail, you’ll: - Understand key differences between evaluating LLM-based systems and traditional software testing. - Add observability to an agent by collecting traces of the steps taken by the agent and visualizing them - Choose the appropriate evaluator - code-based, LLM-as-a-Judge, human-annotation based - for each component. - Compute a convergence score to evaluate if your agent can respond to a query in an efficient number of steps. - Run structured experiments to improve the agent’s performance by exploring changes to the prompt, LLM model, or the agent’s logic. - Understand how to deploy these evaluation techniques to monitor the agent’s performance in production. By the end of this course, you’ll know how to trace AI agents, systematically evaluate them, and improve their performance. Please sign up here:

Andrew Ng

126,406 Aufrufe • vor 1 Jahr

NotebookLM is one of the most delightful, inspiring, and viral AI products out there right now, and I got a chance to chat with the PM behind the product, Raiza Martin (@raiza_abubakar). In our conversation, we cover: 🔸 The origin story of NotebookLM 🔸 The future road map for NotebookLM 🔸 How Google Labs operates differently from the rest of Google 🔸 The development of the “Audio Overviews” feature 🔸 Key metrics and growth of NotebookLM 🔸 Stories about collaborating with author Steven Johnson 🔸 Navigating potential misuse of AI technology 🔸 More Listen now 👇 - YouTube: - Spotify: - Apple: Raiza is a senior product manager for AI at Google Labs for AI at Google Labs, where she leads the team behind NotebookLM, an AI-powered research tool that includes a mind-blowing podcast-on-demand feature called “Audio Overviews.” NotebookLM started as a 20% project and has grown into a product that’s spreading across social media and has a Discord server with over 60,000 users. Raiza previously worked on AI Test Kitchen and has a background in startups, payments, and ads. Thank you to our wonderful sponsors for supporting the podcast: 🏆 Explo — Embed customer-facing analytics in your product: 🏆 Sprig — Build products for people, not data points: 🏆 Sidebar — Accelerate your career by surrounding yourself with extraordinary peers: Some key takeaways: 1. Embrace a startup mentality within large organizations: Google Labs operates with fewer processes and more agility than typical Google teams. This allows them to move faster and iterate quickly, much like a startup. 2. Often, powerful technology is already available; the magic lies in how you interact with it. For instance, by integrating powerful audio models with existing LLMs, NotebookLM created an innovative way for users to interact with content. Look for unique applications of the tools you already have. 3. Don’t wait for a perfect launch. Start with a working version of your product and use user feedback to iterate and improve. This approach can reveal unexpected insights and user preferences, helping you shape the final product.

Lenny Rachitsky

75,439 Aufrufe • vor 1 Jahr