Loading video...

Video Failed to Load

Go Home

Krishna Srinivasan, Founder and CEO of Data Bootstrap and former Google DeepMind researcher, discusses the challenges of data scraping and scalability. With experience at prior Apple, Yahoo, and IBM watsonx, Krishna demonstrates the difference between scraping with a single machine versus leveraging OpenLedger's community nodes. The results speak for...

12,079 views • 1 year ago •via X (Twitter)

10 Comments

Chioma Chukwurah ⚛️🥷⚔️'s profile picture
Chioma Chukwurah ⚛️🥷⚔️1 year ago

@GoogleDeepMind @Apple @Yahoo @IBMwatsonx Indeed Great innovation! 🔥🔥

Kjael // Skaling Ventures's profile picture
Kjael // Skaling Ventures1 year ago

When I talk to SaaS Founders, they often struggle with subjectivity and bias in decision-making. The solution? An objective point of view informed via benchmarks and base rates. Here are 12x KPIs to assess any SaaS firm clearly and effectively:

Utkarsh0x 🐙's profile picture
Utkarsh0x 🐙1 year ago

@GoogleDeepMind @Apple @Yahoo @IBMwatsonx Great innovation #opnup

Namphox's profile picture
Namphox1 year ago

@GoogleDeepMind @Apple @Yahoo @IBMwatsonx #Opnup moon

SON63 🐳's profile picture
SON63 🐳1 year ago

@GoogleDeepMind @Apple @Yahoo @IBMwatsonx Thanks boss 🙏🍀🚀🍀🍀

Namiii 🟧 Allo (✸,✸) 🦙🔥's profile picture
Namiii 🟧 Allo (✸,✸) 🦙🔥1 year ago

@GoogleDeepMind @Apple @Yahoo @IBMwatsonx Perfect 👍 #Opnup

Charisan ♦️꧁IP꧂'s profile picture
Charisan ♦️꧁IP꧂1 year ago

@GoogleDeepMind @Apple @Yahoo @IBMwatsonx Data scraping sounds like trying to catch a greased pig at a county fair! 🐖 With #PublicAI, we’re all about making data sharing a team sport instead! 🎉

fahadg.base.eth⛺🍸 Tabi 🟧 . Ordzaar ꧁IP꧂ (Ø,G)'s profile picture
fahadg.base.eth⛺🍸 Tabi 🟧 . Ordzaar ꧁IP꧂ (Ø,G)1 year ago

@GoogleDeepMind @Apple @Yahoo @IBMwatsonx LFG

Wizjust42 ꧁IP꧂'s profile picture
Wizjust42 ꧁IP꧂1 year ago

@GoogleDeepMind @Apple @Yahoo @IBMwatsonx Looking for a $BOOST Join OpenLedger. Best node yet and it is just beginning. Anyone say DROP ilevzcob4a

Vicky Sharma (✸,✸) CLONE ⛺ Kaisar's profile picture
Vicky Sharma (✸,✸) CLONE ⛺ Kaisar1 year ago

@GoogleDeepMind @Apple @Yahoo @IBMwatsonx cooking ❤️‍🔥 #Openledger #Opnup

Related Videos

Scale alone is not enough for AI data. Quality and complexity are equally critical. Excited to support all of these for LLM developers with Snorkel AI Data-as-a-Service, and to share our new leaderboard! — Our decade-plus of research and work in AI data has a simple point: scale alone is not enough. AI success is all about the quality, complexity, and distribution of data—in addition to volume. We’re excited to be powering leading LLM developers with Snorkel AI Expert Data-as-a-Service, our white glove service for custom, expert-level AI datasets—and to now preview some of what we’re building via our new Expert Data Leaderboard (🔗 in 🧵) + upcoming OSS dataset releases! Snorkel Expert Data-as-a-Service is built to meet the rapidly evolving data needs of the agentic AI world—where success is built on the quality, complexity, and distribution of datasets, in addition to size and scale. This kind of high-quality, frontier AI data can only come from a union of technology and human expertise. With Snorkel Expert Data-as-a-Service, we’re powering frontier LLM developers across agentic, expert knowledge, reasoning, coding, multi-modal, and other task types via the combination of these two key components: - (1) The Snorkel Expert Network: A global team of subject matter experts focused wholly on specialized knowledge–spanning thousands of topics in STEM/academic, vertical/professional, and consumer/lifestyle domains. - (2) Snorkel AI Data Development Platform: Our unique programmatic data curation and quality control platform, accelerating and improving expert authoring and review through principled techniques developed over the last decade of R&D. Now: we’re incredibly excited to showcase some of the power of Snorkel Expert Data-as-a-Service via the new Snorkel Leaderboard—putting frontier models to the test in complex, agentic, and reasoning settings inspired by real industry scenarios (not esoteric puzzles)! We’ll be releasing new leaderboards and accompanying expert-verified open source datasets (coming soon!) regularly. To start, we’re sharing three initial ones in preview: - SnorkelFinance: Q&A over financial documents requiring agentic tool-calling and reasoning - SnorkelUnderwrite: Agentic insurance tasks requiring industry-specific reasoning and tool use - SnorkelSequences: Mathematical tasks requiring compositional multi-step reasoning

Alex Ratner

495,851 views • 1 year ago