Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Document Querying with Qwen2-VL-7B and JSON Output Complete video: I demonstrate how to perform document queries using Qwen2-VL-7B. By simplifying field names, we streamline the prompts, making them more efficient and reusable across different documents. This approach is similar to running SQL queries on a database, but tailored for...

15,981 Aufrufe • vor 2 Jahren •via X (Twitter)

6 Kommentare

Profilbild von Rubén Aros
Rubén Arosvor 2 Jahren

It's the best vllm?

Profilbild von Andrej Baranovskij
Andrej Baranovskijvor 2 Jahren

At the moment yes definitely- qwen2 is the best open source llm with apache license. China rules 🇨🇳👍

Profilbild von optimizium.base.eth (ξ/e)
optimizium.base.eth (ξ/e)vor 2 Jahren

Still unable to compile flash attention 2 on my Windows machine. :( So annoying.

Profilbild von Andrej Baranovskij
Andrej Baranovskijvor 2 Jahren

I believe Windows is not well supported platform. Btw, it looks like new Qwen models will be announced next week -

Profilbild von Z❤️H
Z❤️Hvor 2 Jahren

it hallucinante when it should retrieve a correlated informations from multiple lines

Profilbild von Andrej Baranovskij
Andrej Baranovskijvor 2 Jahren

Not per my tests, qwen2 7b works excellent. For more complex docs use 72b

Ähnliche Videos

New Short Course: Getting Structured LLM Output! Learn how to get structured outputs from your LLM applications in this course, built in partnership with .txt, and taught by Will Kurt, a Founding Engineer, and , Developer Relations Engineer. It's challenging for software to automatically parse through an LLM's freeform text outputs. Structured outputs—like JSON—solve this by converting natural language into consistent, clear, data that a machine can read and process. This course teaches you how to generate structured outputs while building several use cases, including a social media analysis agent. You’ll learn about structured outputs and efficient ways to generate outputs in your defined schema or format. You’ll begin by using structured output APIs, then use re-prompting libraries like “instructor” to generate structured output. Finally, you’ll learn how constrained decoding works; this is a very clever technique in which constraints are applied on each subsequent token generated, blocking any tokens that don’t fit your defined schema. In detail, you’ll: - Learn why structured outputs are important, how they allow for scalable software development, and the different approaches to generate them, including vendor-provided APIs, re-prompting libraries, and structured generation. - Build a simple social media agent using OpenAI’s structured output API, learn how to define a model's desired structured output using Pydantic, and perform basic programming with your outputs, such as importing structured data into a data frame using pandas. - Learn how to use the open-source library "instructor," which checks the structured output of the model and re-prompts the model until it validates the desired output, and explore the limitations of this approach. - Understand how structured generation by the “outlines” library works by modifying LLM logits, on a per-generated-token basis based on the desired format, to give a particular output structure. - Learn how regular expressions, which outlines works with, are represented as finite-state machines, and how they can be used to develop a range of structured outputs beyond JSON. By the end of this course, you’ll have broadened your knowledge of the approaches you can use to get structured outputs from your LLM applications. Please sign up here:

Andrew Ng

89,792 Aufrufe • vor 1 Jahr