Loading video...

Video Failed to Load

Go Home

There is my prediction on where RAG is headed. In this video i talk about - Shift from RAG as question-answering systems to report generation tools - Importance of well-designed templates and SOPs in driving business value (selling to people with money) - Room for AI-generated templates and template...

184,645 views • 2 years ago •via X (Twitter)

10 Comments

jason liu's profile picture
jason liu2 years ago

I post videos like this all the time off the cuff with no edits. follow me @jxnlco but you might also want to mute me too

Nick Dobos's profile picture
Nick Dobos2 years ago

Seems less about RAG, and moreso structure 1. Chatbot Q&A - unstructured 2. Template report - common broad structure 3. Opinionated report - specialized niche structure ChatGPT is following the same progression -Blank canvas -Starter buttons -GPT store -Memory/personalization

jason liu's profile picture
jason liu2 years ago

pydantic is all you need

Victor Boutté's profile picture
Victor Boutté2 years ago

Love the difference pointed out: RAG + QA = value derived in time saved when finding answers vs. RAG + Report generation = value derived by creating decision making tools that help influence resource allocation.. think SOPs

Brandon Tyree's profile picture
Brandon Tyree2 years ago

I spent 2 hours yesterday talking about RAG Enabled Reporting with two PE firms. You are 100% correct.

Garrett of DeepwriterAI's profile picture
Garrett of DeepwriterAI2 years ago

Agree. I have only used rag to generate reports and documents and content. Question and answer is a fun toy.

Melvin Salvador's profile picture
Melvin Salvador2 years ago

that offer sauce🔥 spent years in the yc startup world only to later realize a lot of our problems came from basic packaging issues like the ones you’re describing we need more ai twitter and money twitter swapping notes

Zengineer's profile picture
Zengineer2 years ago

The end result of this is McKinsey as an API call

Joschka Braun's profile picture
Joschka Braun2 years ago

@md_rumpf u mentioned this when visiting NY

Derek Homan's profile picture
Derek Homan2 years ago

just out of curiosity, what was the book?

Related Videos

Announcing a new Coursera course: Retrieval Augmented Generation (RAG) You'll learn to build high performance, production-ready RAG systems in this hands-on, in-depth course created by and taught by , experienced AI and ML engineer, researcher, and educator. RAG is a critical component today of many LLM-based applications in customer support, internal company Q&A systems, even many of the leading chatbots that use web search to answer your questions. This course teaches you in-depth how to make RAG work well. LLMs can produce generic or outdated responses, especially when asked specialized questions not covered in its training data. RAG is the most widely used technique for addressing this. It brings in data from new data sources, such as internal documents or recent news, to give the LLM the relevant context to private, recent, or specialized information. This lets it generate more grounded and accurate responses. In this course, you’ll learn to design and implement every part of a RAG system, from retrievers to vector databases to generation to evals. You’ll learn about the fundamental principles behind RAG and how to optimize it at both the component and whole-system levels. As AI evolves, RAG is evolving too. New models can handle longer context windows, reason more effectively, and can be parts of complex agentic workflows. One exciting growth area is Agentic RAG, in which an AI agent at runtime (rather than it being hardcoded at development time) autonomously decides what data to retrieve, and when/how to go deeper. Even with this evolution, access to high-quality data at runtime is essential, which is why RAG is a key part of so many applications. You'll learn via hands-on experiences to: - Build a RAG system with retrieval and prompt augmentation - Compare retrieval methods like BM25, semantic search, and Reciprocal Rank Fusion - Chunk, index, and retrieve documents using a Weaviate vector database and a news dataset - Develop a chatbot, using open-source LLMs hosted by Together AI, for a fictional store that answers product and FAQ questions - Use evals to drive improving reliability, and incorporate multi-modal data RAG is an important foundational technique. Become good at it through this course! Please sign up here:

Andrew Ng

124,656 views • 1 year ago