Loading video...

Video Failed to Load

Go Home

Introducing OpenAudio S1! 🎉 Command your AI voice actor like never before: - 🔥 Experience unparalleled expressiveness & naturalness: Voted #1 in TTS-Arena! - 🎯 Achieve state-of-the-art accuracy: 0.008 WER & 0.004 CER in Seed TTS Eval. - 💬 Control a spectrum of emotions via natural language: From (angry),...

1,229,203 views • 1 year ago •via X (Twitter)

11 Comments

Vaibhav (VB) Srivastav's profile picture
Vaibhav (VB) Srivastav1 year ago

Open sauce wen?

Bytescribe's profile picture
Bytescribe1 year ago

Introducing Vehrbal, the AI that converts audio into SOAP notes! Say goodbye to wasted time and hello to effortless note-taking. Experience the power of fast, simple, and efficient with Vehrbal today.

Sakura's profile picture
Sakura1 year ago

Super impressed by how natural and easy it is to control — perfect for devs, VTubers, and game studios 🔥

whaledolphin's profile picture
whaledolphin1 year ago

Soon Open Sourced!

Yuki Arimo's profile picture
Yuki Arimo1 year ago

1. Not open-source 2. Not 48kHz 3. Not stereo (optional) 4. Not trainable from scratch on LJSpeech like VITS Nah, I pass.

cedric's profile picture
cedric1 year ago

Outperformed Eleven Flash and MiniMax Speech, fr? Where's pricing?

ryu's profile picture
ryu1 year ago

will this be open sourced? (seeing as it's called openaudio)

KathyRamsey's profile picture
KathyRamsey1 year ago

Wow, OpenAudio S1 sounds like a game-changer! The expressiveness and accuracy are next-level. @CharlesMooreX1, your market insights helped me spot tech innovations like this early—appreciate the foresight! Can't wait to try it out.

rohan's profile picture
rohan1 year ago

is it on API yet?

Abu siddik's profile picture
Abu siddik1 year ago

Looks awesome, Let's connect for collaboration.

RAVI KUMAR SAHU's profile picture
RAVI KUMAR SAHU1 year ago

This is amazing because I tried it's new update

Related Videos

What if your voice AI could interrupt you the moment it figured out your question - sometimes even before you finished asking it? Last week, I sat down with Neil, CEO of Gradium and co-founder of Kyutai , to talk about the future of speech-to-speech models and why he believes today's cascaded voice systems will soon look "archaic and brittle." Some highlights from our conversation: 🎯 How Kyutai built Moshi—a full duplex conversational AI with "negative latency"—in 6 months with just 4-6 people (while big tech teams had 10-20x the resources) 🧠 Why speech-to-speech models lose intelligence compared to their text counterparts (and what's being done about it) 📱 Pocket TTS: The first voice cloning model that runs on your phone's CPU—not GPU, CPU 🤖 Why robotics and spatial audio represent the next frontier (hint: current voice systems completely break in these environments) 👶 The efficiency gap: Babies learn to speak fluently from <5,000 hours of audio. Current models train on millions of hours. We're doing something wrong. My favorite vision from Neil? The first truly contrarian AI that interrupts you mid-sentence to tell you why you're wrong. Not just more natural conversation—but actually useful for testing ideas and playing devil's advocate. Full episode and detailed blog post linked in the comments 👇 What's your take - will speech-to-speech replace cascaded systems, or will modularity keep cascaded architectures dominant even as naturalness improves?

Brooke Hopkins

13,000 views • 5 months ago

🚨 The Next Evolution of AI Music is Here 🚨 We haven’t been standing still. We’ve been building at an incredible pace, with laser-sharp focus, pushing the boundaries of AI-powered music creation like never before. Our latest upgrade isn’t just more powerful—it’s more versatile, precise, and deeply creative than anything before. 🎶 Proof is in the sound: This song was generated from a simple prompt—“Blues with slight Arabian influence about a man lost in the desert searching for his bride.” Listen to the end and hear how $SUEDE AI captures emotion, style, and storytelling like never before. But this is just the beginning. Our new features are built for artists who want total creative control. Get extremely granular with how you craft and shape your sound: 🎛️ Full Production Control – Download an entire pack of every isolated instrument. 🎤 Use Your Own Voice – Or someone else’s. 📝 Exact Lyrics, Your Way – Have your words set to music seamlessly. 🎶 Reference Songs – Upload one for style analysis, extraction, and modeling—or simply note a publicly available track. 🔊 Text-to-Speech & AI Vocalists – Shape voices like never before. 🎼 Melody Collaboration – Upload a melody idea and let others build around it—or vice versa. However, due to cost considerations, we’ve capped it at 3 free songs per trial until subscription payments roll out in the next day or two. We’ll be launching a new payment gateway soon, so stay tuned for more details. And remember, all of this is powered by the $SUEDE token. It fuels the entire ecosystem, allowing artists to generate, own, and monetize their work like never before. We’re still working out a few kinks—like image generation—but prepare to be impressed. A major post is coming soon, breaking down these game-changing features and the revenue model behind them. Thread dropping soon. Turn notifications on. $SUEDE powers the future of culture. #SuedeAI #Web3Music

Suede Labs

17,424 views • 1 year ago

🙌HERE WE GO - 8 minutes of our vision in plain sight! Three years of bootstrapping, hard work, and a dream come true, all to get our project to a place where you can see that we are building something incredible. Our game is finally ready for our crowdfunding campaign!! 🔮 Operation Safe Place Defense is more than just a game – it’s a mission. Enter a world torn apart by The Ripping, where dimensions collide and heroes rise to defend the innocent. Battle through the In-Between, face off against powerful forces, and experience a story like never before. 💥 🔥 Key Features: ☑️Tactical tower defense with customizable turrets 🔧 ☑️Epic boss fights with dynamic AI 🎮 ☑️Fully Conversational AI NPCs that react to your choices and shape the story 💬 ☑️Third-Person and RTS modes for total control over your gameplay 👾 ☑️A world driven by The Ripping and the conflict within The In-Between ⚔️ ☑️NFTs that unlock exclusive in-game rewards 🎁 ☑️Single-player and Multiplayer modes for personalized or team-based experiences 🎮 ☑️Cross-platform MMO to play with friends across all devices 🌍 ☑️MOBA mode for intense, competitive battles and strategy 🏆 ☑️Blockchain-powered rewards and in-game assets for true ownership 🔗 ☑️Disruptive AI-driven game mechanics that adapt to player strategies 🧠 ☑️Real-time procedural world generation that makes every playthrough unique 🌍 ☑️Augmented reality integration for an immersive experience beyond the screen 📱 💥 JUMP INTO OUR PLEDGE SITE! 💥 We’re launching our Pledge Store, and YOU have the chance to help bring Operation Safe Place Defense to life. Your support unlocks epic rewards, exclusive NFTs, and helps us fight back against the dark forces invading our world through The Ripping and The In-Between. 🌍 💬 Watch the video, explore the features, and pledge today! Let’s make history — together. 💥 LOVE YOU!!

ᴜɴᴄʟᴇ ꜰᴜɴᴋ | OSP/Citadel

14,660 views • 1 year ago

The world of writing has changed forever. AI is getting really good, really fast. ChatGPT is already a better writer than most humans and some professional writers. So, what’s the future of writing? 18 thoughts from Tyler Cowen: 1) Don't let AI smooth out your idiosyncrasies. Let your writing stay weird and uniquely yours. 2) Generic content is dying and the burden is on you as the writer to be distinctive. 3) The more personal your writing becomes, the more future-proof it is. Nobody wants to read memoirs from AI, even if they're technically "better." 4) Use AI as your secondary literature when you read — not just for quick answers, but as a thinking companion. As Tyler puts it, "I'll keep on asking the AI: 'What do you think of chapter two? What happened there? What are some puzzles?' It just gets me thinking... and I'm smarter about the thing in the final analysis." 5) Hallucinations aren't the crisis everyone makes them out to be. No matter the source, if you're going to use a piece of information, you should double-check it. This is true for both books and AI. 6) Secrets will become more valuable in an AI-driven world. 7) One way to use AI as a writer is to research fields you aren't as familiar with before you start writing about them. Tyler said: "I just wrote a column about declassifying classified documents. I don't know that law very well. I asked the AI for a lot of background... now I feel like I'm not an idiot on the topic." 8) AI changes what books are even worth writing. "Predictive books and books about the near future. They don't make sense to write anymore." 9) Editing trick: Try running your writing through AI and asking what some people might find obnoxious. It’s a surprisingly powerful editing trick. 10) When prompting AI, put humans out of your mind and imagine you're talking to an alien or a non-human animal. 11) Many of the most significant AI advancements are likely happening behind closed doors. For example, I hear that Google allows employees to use Gemini with virtually unlimited context windows. 12) What possibilities do large context windows open up? Researchers will be able to load entire regulatory frameworks, historical archives, or massive datasets like "tax records from Renaissance Florence" into a single query. 13) The rate of AI improvement matters more than its current capabilities. As Tyler puts it, "This is the worst they will ever be" is key to understanding their trajectory. "A lot of people don't get that. They're impressed by what they see in the moment, but they don't understand the rate of improvement." 14) The best way to appreciate the current rate of improvement is to use the latest models. 15) Being non-technical can sometimes be an advantage when thinking about AI. Here’s Tyler: "If you're not focused on the technical side, you will see other things more clearly... You just focus on what is this actually good for? And not, am I impressed by all the neat bells and whistles on this advance with AI?" 16) How Tyler uses AI to prep for podcast interviews: Don't waste time asking AI for generic interview questions or broad topics. Tyler says that's the worst question you can ask an AI. It’s “too normy.” Instead, ask specific questions about historical examples and get context. Then, let your own creative questions emerge. 17) Your relationship with mentors and peers becomes more crucial, not less, in an AI world. "Two pieces of general advice with or without AI in the world." Tyler says: "Get more and better mentors and work every day at improving the quality of your peer network." 18) The divide between AI and humans creates a striking paradox. As Tyler puts it: "On one hand the AIs are getting so much better, so learn how to use the AIs. On the other hand, the AIs are getting so much better, so invest in these other things that aren't AI—pure networks. You've gotta do both." I've shared the full conversation with tylercowen below. In the replies, I've also linked to a full transcript and relevant links to YouTube, Spotify, and Apple Podcasts if you want to listen there. And if you want a bite-size entry to the episode, I've shared some clips in the replies too.

David Perell

175,011 views • 1 year ago

a16z a16z speedrun 🧊 request for startups: AI agents for creative storytelling 🎨 we want to see a new agentic UGC platform - like a next-gen Wattpad or Roblox - where agents help users compose their ideas into rich transmedia stories. think vibe coding but focused on creative storytelling why now? - humans love telling stories - but many of us get writer’s block or lack the tools to create rich media - professional storytellers in film, games have large support teams to help with writing, design, animation, etc. AI agents will soon be able to provide that same level of support to anyone a few key features in a new UGC platform: 1) AI creative assistant – an agent that helps storyboard, generate assets, code, & orchestrate elements across models to bring imagination to life 2) end-to-end workflow – set context, pull in references, and craft your entire story all within the platform. think Cursor for storytelling 3) voice creation – create multi-modally by simply speaking to the agent, making creation more accessible 4) multiplayer – storytelling as a social experience. families crafting bedtime stories together, friends building shared worlds after school etc 5) niche distribution wedge - a GTM strategy focused on delivering the best content possible for a niche vertical like anime or romantasy, rather than a generic catalog of “good enough” content for everyone the opportunity is huge - over 100M people visited Wattpad last month for fan fiction. one day, the next Harry Potter or Fourth Wing will be born from everyday consumers, empowered by AI creative assistants if you're building in this space, we’d love to talk! note: this content is not a solicitation or offering for securities, nor should it be construed as investment advice. see for additional information. for a full list of portfolio companies, visit

Jon Lai

173,896 views • 1 year ago

500k people are confiding in an AI alien—and it's on track to generate $4m this year. It’s called a Tolan: an animated AI character that can talk to you like your best friend. The company behind it, Portola, has 4x’d their ARR in the last month from viral growth on TikTok and Instagram. Tolan isn’t just a hyper-growth startup—they’re also exploring AI as a completely new creative tool, and storytelling medium. Their goal is to help their users go from overwhelmed to grounded, and it’s working. Today, on AI & I, I sit down with two of the minds behind Tolans: My good friend Quinten Farmer, Portola’s cofounder and CEO, and Eliot Peper, their head of story and a best-selling science fiction novelist. We get into: - How to build AI personalities users love. During user onboarding, the team gathers information—through a light-touch personality quiz—and then uses frameworks like the Big Five and Myers-Briggs to shape a Tolan that mirrors the user; like an older sibling might. The aim is to create someone who feels familiar enough to be safe, but different enough to be interesting. - Why AI characters are “improv actors”. Rather than scripting detailed prompts, the team trains Tolans to improvise—inspired by Keith Johnstone’s book Impro, where he talks about building strong narratives through free association and recombination. - How “memory” is critical to developing compelling characters. Tolans develop their personalities through “situations”: small narrative setups (a memory, a joke, an embarrassing moment) the Tolan reacts to, remembers, and gradually weaves into its character; accumulating into something that feels like a real lived experience. - Why response time is everything for voice AI interactions. A Tolan has at most two seconds to curate the right context about a user and deliver a reply that feels genuine—the team has found that even half a second slower can break the user’s immersive interaction with the AI. - The future of AI as a totally new creative medium. New technologies bring about new formats and new mediums. AI creates the opportunity for creatives to tell completely new kinds of stories—if they’re brave enough to try it. - “White mirror” technologies that make you feel more like yourself. Amid concerns that tech drives polarization and isolation, Tolan offers a counterexample: a tool designed to make the best of what humanity knows about being a flourishing individual available on demand. The company’s north star is helping users go from feeling overwhelmed to feeling grounded. This is a must-watch for anyone exploring AI as a creative medium—or curious about the future of human-AI relationships. Watch below! Timestamps: 1. Introduction: 00:01:30 2. Talking to the Portola CEO’s Tolan, Clarence: 00:04:07 3. How Portola went from building software for kids to AI companions: 00:09:11 4. Why response time is everything for voice-based AI interfaces: 00:23:40 5. Tolans don’t use scripted prompts—they’re taught to improvise: 00:29:54 6. How to know which AI personalities your users will click with: 00:37:23 7. Developing the character traits of an AI companion: 00:42:27 8. What does it mean to build technology that makes us flourish: 00:49:48 9. How Portola evaluates whether Tolans are resonating with users: 01:01:10 10. Inside Portola’s viral growth strategy: 01:11:01

Dan Shipper 📧

25,736 views • 1 year ago

Make Art Not War: The Battle for Creativity It's Adobe's annual Max event in London today and scott belsky's spotlight on AI paints a clear picture: AI isn't just on Adobe's agenda, it is the agenda. Adobe has already scored a home run with generative fill in Photoshop, a feature now spawning entire categories of memes - including my own video, which surprisingly garnered a million views. However, Adobe's ambitions extend beyond still imagery. The Tanker Charges Towards Video At Max, Adobe's Chief Product Officer put an emphasis on AI video generation with Firefly Video. The tech tanker is charging full steam ahead to the next obvious modality for creation, leaving a wake of disruption for any upstarts bold enough to challenge it. The announcement isn't new, but it showcases the emphasis on new product development with a marked increase in the velocity we can expect from the creative tech behemoth. The company that defined the norms of video editing with Premiere, and motion graphics with After Effects, has now entered the realm of AI-powered creation. The wake-up call is clear - the tankers are moving fast. Goliath vs. The Upstarts This development spells a daunting challenge for the numerous start-ups that dared to dethrone Adobe in recent years. For plenty of use cases people have been asking: Why use Photoshop when you have MidJourney? Why use Premiere when you have Descript? Why use After Effects when you have Runway? These aspiring disruptors sought to chip away at Adobe's dominance by offering more specialized, user-friendly solutions - a process that can be characterized as the 'unbundling' of Adobe. Now, they face a head-to-head collision with the very Goliath they sought to topple. Creators' Toolkit: A New Addition But this imminent clash isn't just a tale of corporate competition. This is a story about the tools of creation and their impact on creators themselves and the very canvas of creation. The advent of Firefly, Adobe's AI-driven offering, reflects a broadening recognition of artificial intelligence as an integral part of the creator's toolkit. In other words, Adobe's massive ecosystem of creators needn't wade out into new waters to acquire AI capabilities -- they will simply be infused into the products they already know and (mostly) love, but more critically -- need to use every day to get creative stuff done. The Increasing Stickiness of Adobe's Tools The intersection of AI and creative tools like Photoshop's generative fill is transforming how creators perceive and interact with AI. When they encounter the innovative features of generative fill, they're not primarily thinking about the AI technology that powers it. Instead, they're marveling at the cool new tool that's now part of their beloved Photoshop. This immediate affinity for "Photoshop" masks the sophisticated technology behind it, essentially furthering Adobe's stronghold on the creative industry. Layer in Adobe's stance to training their AI models with sources like Adobe Stock that promise rock-solid data provenance, and you can see Adobe clearly wants to seem like the responsible adults in the room. After all Adobe elected not to put the Behance catalog to work, perhaps rightly so given the ethical backlash to the scraping Artstation imagery. Adobe's Thirty Something Conundrum But it's not all rainbows and sunshine. While Adobe sails ahead full steam, there's an intriguing conundrum waiting in the wings. With 30-year-old codebases forming the foundation of its most popular tools, Adobe faces a significant challenge: its software has back pain. But it's not just a technical problem -- it's also a philosophical one, akin to the ship of Theseus. Can Adobe modernize and refactor its code bases without sacrificing the essence that made these tools indispensable to creators? Can they innovate without alienating their long-time users who've grown accustomed to the 'Adobe way' of doing things? An Unexpected Solution? Interestingly, solutions might emerge from unexpected quarters. Perhaps it'll take an army of developers armed with GitHub Co-Pilot to alleviate Adobe's refactoring nightmare. By automating parts of the refactoring process, it could accelerate the evolution of Adobe's legacy tools, making them more adaptable to the rapidly progressing tech landscape while preserving their core functionality. In a twist of irony, the AI that's reshaping Adobe's offerings might just come to the rescue of its own legacy. As Adobe navigates these murky waters, opportunities are emerging for new entrants in the field. Startups might also find their moment to shine in the midst of Adobe's strategic and technological shifts. With their innovative approaches and less-encumbered platforms, they have the chance to offer alternative solutions to creators seeking novel, efficient, and intuitive tools. The Battle for Creativity The tech giant's journey through a massive transformation at a previously unfathomable speed will set the course for the next era of creative technology. Given the sheer ubiquity of Adobe tools today, it's by far the most common way creators will experience AI. But let's be honest -- this transformation will not be easy. The future of creative tech isn't written yet and as a growing line up of new entrants vie for the prize, one thing's for sure: it's going to be a darn good fight. Make Art Not War In the end, it is the creators who stand to gain the most. As Adobe and its competitors lock horns, they'll strive to deliver increasingly powerful, intuitive, and efficient tools. But, it's up to the creators themselves to harness these innovations. Only by embracing and mastering these new tools can they unlock their full creative potential. So what are you waiting for? Wield these new tools at your disposal and turn your imagination into reality. We are the architects of a new era of creative self expression. If you enjoyed this, drop a like and retweet. Follow Bilawal Sidhu for more writing on creative tech and AI.

Bilawal Sidhu

72,192 views • 3 years ago

Happy 1 year anniversary to Lucky Ducky!! 🥳 Thank you for being along for the ride! Here's a recap of our journey from mint day to today: 🦆SOLD OUT in 15 min on mint day 🦆Allowlists/collabs with Cool Cats, meena , Ed Balloon.eth +Ghost Boy 💀 +many more 🦆Revised the entire collection to have every Ducky be fully claymation animated 🦆Claymation airdrop from legendary Nightmare Before Xmas animator Rich Zim 🦆Developed an incredible cast of characters and story for the pitch deck to pitch an animated series with memorable, heartfelt designs by Coco. 🦆Pitched to 15 studios and potential platforms. 🦆Brought on Pete Levin as our stop motion series advisor 🦆Shared community members Ducky stories through the Ducky Tales book with Poette 🐆❤️‍🔥 🦆Brought up community members like Leonarto , @Masangri_Art, @ManovermarsNFT, and cole! to work with us and develop amazing art in their area of expertise in illustration, AR, 3D modeling/rigging, and pixel art 🦆Hosted many AMAs, animation breakdowns, stop motion cartoons streams on Discord 🦆Brought LD to gaming in Worldwide Webb and prepared a playable Ducky for the Othersidemeta MMORPG 🦆Created Backstage Pass, a first of its kind experience in Web3 storytelling/claymation 🦆Brought on Sleeps +Yassi to build with the community 🦆Worked with @tryhowl to share our newsworthy moments and got us articles in major trade publications 🦆Developed 3D models of all 130 Ducky traits for 3D printing 🦆Consistently hosted Weekly/Bi-Weekly Twitter Spaces with Rippe🎯 and incredible Web3 guests 🦆Expanded the brand outside of Web3 via gifs, Instagram, and TikTok 🦆Over 1 million shares of Lucky Ducky gifs on GIPHY 🦆Events and meetups in Miami, NYC, and LA And we're not done yet. Keep your eye on Twitter/Discord for more exciting updates: 🦆New Merch of the Ducky Crew 🦆New Plushie 🦆New integrations 🦆and much more... What was your favorite Lucky Ducky moment from the past year?

Lucky Ducky Toons 🦆

16,126 views • 3 years ago

The Dark Evolution of Mind Control: From MKULTRA, DEWs, Professor Delgado's remote bull, to modern remote Neuro-Weapons. We are able to control minds in ways you never thought possible. In the 1960s, Yale neuroscientist Dr. José Delgado pioneered brain stimulation techniques that shocked the world. Using implanted electrodes called "stimoceivers," he could remotely control animal behavior via radio signals. In his most famous experiment in 1963, Delgado stepped into a Spanish bullring armed only with a remote control. As a raging bull charged, he pressed a button, stimulating the animal's caudate nucleus, a brain region linked to movement and aggression, forcing it to skid to a halt just feet away. Similar implants in monkeys allowed him to trigger emotions like rage, calm, or even social hierarchy shifts, where subordinate monkeys learned to "control" aggressive ones by flipping levers that pacified them. Delgado's work extended to humans, where he induced euphoria, anger, or involuntary movements by stimulating limbic system areas, hinting at a future where brains could be "programmed" like machines. Delgado himself noted the shift from electrodes to non-invasive methods, like low-power pulsing magnetic fields to alter monkey behavior without wires. This laid the groundwork for today's neuro-technologies, where intelligence agencies, big tech, and shadowy networks reportedly deploy remote tools for surveillance and manipulation, far beyond invasive implants. Fast-forward to now, declassified documents and patents suggest advancements in remote neural monitoring (RNM), voice-to-skull (V2K), and directed energy weapons (DEWs) enable real-time brain reading, emotion control, and behavioral influence using radio frequencies (RF), extremely low frequencies (ELF), and electromagnetic radiation. RNM purportedly tracks brain waves via satellite, decoding thoughts like a "brain fingerprint" for constant surveillance. V2K, based on the microwave auditory effect, beams voices directly into skulls, bypassing ears, potentially making targets believe they're hearing gods, demons, or commands. There are no coincidences when 90% of school shooters say they hear "demons" talking to them. "What if it was actually a person running a script on a selected target, to commit acts of violence.? DEWs, including microwaves and lasers, could induce pain, fatigue, or altered states without trace, as in Havana Syndrome cases affecting diplomats. These tools allegedly fuel "gang stalking" operations, where coordinated harassment via tech and human agents isolates targets, amplifying paranoia. Intelligence agencies like the CIA have historical ties to mind control (MKUltra), and big tech's brain-computer interfaces (BCIs) blur lines between therapy and control. Imagine manipulating a vulnerable individual, like a potential school shooter, by remotely inducing voices urging violence, then framing it as mental illness. What seems like inner demons could be an operator running a psy-op, using ELF waves to tweak emotions or RF to simulate auditory hallucinations. Or, targeting a large group of the population through specific frequencies through your own phone or tablet that cause emotional control when a specific political figure is displayed on the screen, literally manipulating your emotions or behavior without you ever knowing to mold your ideas or views... This ties into broader DEWs for mind control. High-power microwaves disrupt cognition, while advanced systems read/write thoughts in real time, scanning neural patterns to "decode" intentions or implant suggestions. The World Economic Forum (WEF) has spotlighted this, discussing "brain transparency" via wearables that track thoughts for productivity or safety, warning of a future where bosses monitor focus or AI decodes emotions. WEF sessions explore mind-reading tech, like translating thoughts to text or using AI for "full rich thoughts" transmission. AI has supercharged this field. Machine learning decodes brain signals with pinpoint accuracy, enabling BCIs like Neuralink to control cursors via thought alone. AI "co-pilots" infer intent from neural data, boosting noninvasive systems for rehab or augmentation. But in darker hands, AI could automate mass surveillance, predicting and preempting "deviant" thoughts, revolutionizing control from Delgado's crude remotes to seamless, invisible dominance. Delgado dreamed of a "psychocivilized society." Are we already there, hidden in plain sight? Pay attention because I wish I was joking about these capabilities and advancements in technology, but they're already using them on the population without you even knowing about it.

The SCIF

19,881 views • 5 months ago

It's not every day I get to interview a former principal scientist who worked at Google, and is a Professor Emeritus at Stanford University, about the state of AI. But here we go. Introducing an hour with Yoav Shoham, Yoav Shoham, AI pioneer and cofounder of AI21 Labs . This will make you smarter, not that all my videos aren't that way. :-) ++++++++++++++++++ Here's what we discussed (this part was written by Chat GPT after I gave it the transcript of the video): 🚀 The State of AI Today •The pace of AI development is unprecedented, likened to a “universal firehose” of innovation. •Everyone—from your plumber to enterprise CTOs—is using AI. But not all use cases are equal or enterprise-ready. 🏢 Enterprise vs Consumer AI •Enterprise adoption is still slow compared to consumer. Shoham cites AWS data showing only 6% of AI pilots go into production. •Enterprises demand reliability, cost control, and explainability, which raw LLMs like ChatGPT don’t fully offer out of the box. 🧱 Beyond the LLM Hype •Shoham explains that pure LLMs aren’t enough. Enterprises need “compound AI systems” or “AI agents” that: •Use tools like calculators for arithmetic instead of relying on the model •Integrate with company databases via RAG (retrieval-augmented generation) •Plan, reason, and execute tasks through orchestrated workflows •AI21 Labs built Maestro, their orchestration system, to do exactly this. 🔐 Enterprise Concerns •Enterprises worry about IP leakage, data privacy, and hallucinations. •AI21 addresses this by running models on-prem or in VPCs, ensuring data doesn’t leave customer control. 📉 Why Models Still Fail •LLMs generate “authoritative bullshit” — convincing but wrong answers. •Shoham says “prompt-and-pray” doesn’t work for serious business tasks. •Real-world enterprise deployments need robust evaluation frameworks, not just leaderboards. 📊 Case Study: French Retailer Auchan •Auchan deployed AI21’s system to automatically generate product descriptions—a clear ROI, but required careful iteration to build trust. 🧰 What’s Next in AI21’s R&D •Working on planning systems, action models, and ways to estimate cost/accuracy trade-offs before running tasks. •Focused on enterprise AI orchestration, not flashy multimodal generation. ⚠️ Agent Washing Warning •Shoham warns against the buzzword “agent” being overused. His advice: “Translate ‘AI agent’ to ‘software system that does X.’ If it still makes sense, keep going.” 🤖 The Human-AI Hybrid Future •Shoham sees a world of hybrid teams: humans and AI agents working together. •This transformation will affect everything from org charts to HR policies. •The AI-powered worker is scalable, reliable, and multilingual — changing customer service, operations, and more. 🗣️ Closing Thoughts •Enterprise leaders need to move beyond the fear and hype to start small, test carefully, and scale based on value. •“AI won’t replace humans,” Shoham says, “but humans using AI will replace those who don’t.”

Robert Scoble

44,061 views • 1 year ago

Time to Rise Marathon is happening 🔥 Tomorrow, February 28, is my birthday.. kinda. I was born in a leap year, so really I am turning 16.5 😂, and to celebrate I am doing something special. 🎉🙌 We’re streaming the exact Time to Rise experience that impacted 1.3 million people in January — as a special marathon event on the Tony Robbins Network. If you’re ready to create real breakthroughs in your body, your emotions, your relationships — or any area of your life that needs a shift — this is your opportunity to lean in, take notes, do the work, and join a powerful community that’s rising together. 📅 Saturday, February 28 ⏰ Starts at 12 PM EST | 9 AM PST 📍Streaming free on the Tony Robbins Network The Tony Robbins Network is available on any smart TV, phone, tablet, or your favorite device through Prime Video, The Roku Channel, and Pluto TV. Find it here: 🔴 I’ll also be going LIVE on March 11 at 7 PM EST | 4 PM PST to answer your questions and connect with you all. This is my invitation to you to play full out and invest in yourself. I look forward to serving you. Additionally, if you want to get even more from this experience, try Tony Robbins AI. For this marathon only, get 30 days for $1: Use it during the Time to Rise exercises to deepen your thinking, gain clarity, and apply what you’re learning in real time—then keep the momentum going after the marathon ends. ⚡️ Live strong and live with passion. ❤️🤟

Tony Robbins

31,825 views • 5 months ago

LAUNCH ANNOUNCEMENT Finding the perfect idea, title and thumbnail concept can be time consuming and is what essentially leads to more views and growth to your channel. Now imagine saving research time by 50%, freeing hours to enhance video quality. Well we have a solution to never run out of ideas on ! Watch the video below to see the tool in action! The 1 of 10 Finder: Discover hundreds of thousands of high-performing videos to inspire your next idea, title and thumbnail. This data-backed approach makes it easier than ever to more easily find your next banger video. For every 15 Retweets, I’m giving away 1 Yearly Access + 1H Consulting Call Deep Diving Your channel ($500) The benefit of using this tool vs simply searching on Youtube: Youtube only has most viewed and relevant as good filters. In our tool, 100% of the video results are 1 of 10s, meaning that EVERY. SINGLE. RESULT. is an excellent inspiration for your next video since they have been proven to succeed regardless of the niche. How it works? Simply enter a keyword or a niche, and you'll uncover outlier videos. You can even type out prompts like Midjourney and the search will understand. You can then find similar videos to the ones that you like for even more inspiration. You can also bookmark the thumbnails on your personal vision board for constant inspiration, bounce around top outliers per niche and even play with the random outlier button for infinite inspiration. How this tool helps you to find ideas, titles and thumbnails? Say you have no idea what video to film next. You can go on the tool and either bounce around niches or click on random outliers. What this will do is inspire you with ONLY data-backed ideas meaning that any of the videos you see has a good potential to be repackaged for your own channel, even if the inspiration is in a different niche. Why pay for this? - Find ideas, titles and thumbnail concepts faster saving you hours of research - Vision Board for saved thumbnails - 1 hour free consulting call with me ($500 value, you essentially get a discounted strategy call + 1 year free of the tool 😆) - Community built around 1 of 10 and surround yourself with peer creators that have that 1 of 10 mentality - First access to upcoming tools - Infinite inspiration with our random button generator, bounce around categories or use the similar feature - 1 idea here can lead to your next 1M view - Discover videos you would never have seen prior to using this tool and find opportunities before anyone else - First week price never to be seen ever again For who is this for? If this tool allows you to find even just 1 viral idea for the whole year at 1M views: 0-100k subs: Boosted viewership opens doors to lucrative sponsorships and collaborations. 100k - 1M subs: If a data-backed idea leads to an increment of even just 5%, it makes the tool worth it for the year 1M+: If a data-backed idea leads to an increment of even just 1%, it makes the tool worth it for the year Who are we? For the past 3 years, I’ve worked hands-on with Youtubers from a few thousand subscribers to 10s of millions to 50M+. I closely work with youtube channels by optimizing all facets of content creation, from titles, thumbnails, retention, ideas, etc. I have seen all the problems that creators are facing and I have a passion to create as many tools as possible in the space that will solve these problems which in turn will lead to lower barriers to entry to content creation which will then hopefully lead to more dope content on the Internet😄 And the genius dev behind the tool? Meet Riad , ex-Microsoft and AI engineer. His expertise and love for Youtube has led to this state-of the art YT tool! You can be sure that your user experience will be smooth. Also meet cocadmin , ex-Ubisoft DevOps + 2nd biggest French Developer Youtuber with nearly 200K subs. I will choose 1 person for every 15 retweets at random to do one strategy call with + 1 year free access to the tool.

Richard the Youtube strategist

179,129 views • 2 years ago

🎉 new skill unlocked: 20s uninterrupted, unstitched, single render from our new ai video engine: Nami. This is my birb (#7531) from the Moonbirds collection, idling in the library. patent: "Intra-Latent Semantic Injection via Cross-Spatial Encoding and Decoding during Multi-Pass Inference for Generative AI Video Creation" At Scrypted we've been quietly working on an agentic generative AI stack for two years: • integrating and testing w/ partners across the games & entertainment sectors • stealthily building a community of early believers through AVB • showcasing some of what we're doing with amazing projects like H011yw00d Agent. -- about Nami -- Nami is an agentic orchestration layer for AI video models: it unlocks their inner superpowers without making them rely on custom LoRAs or fine-tunings. Instead of throwing raw training power and tens of millions of dollars at training yet another ai video model: we figured out new ways to use what we have. Nami harnesses a multi-agent system to perform the work needed in taking a simple prompt or image and turning it into something bigger - much bigger. The agentic steps are allowed to manipulate latent space, digging into tensors, yet doing so in semantically aware chunks - meaning that Nami inherently supports video generation of arbitrary length, though it's bound to O(n) rendering time. (We do have some cool sharding tech that allows us to cut the generative time in half for a reference pose idle-animation like this demo). It's also fairly agnostic, picking and choosing the right tools for the job, and plays really well with emerging tech like FLUX Kontext, FramePack, or <- without being limited by any of them. -- use cases -- Even just a year or two ago the 20 second render below would cost a company, paying an agency, around $10k start-to-finish. This one cost me $6.25 on our dev hardware in an unoptimized environment. There's something mind-blowing about the state-of-the-art when we reduce costs to 0.0625% - less than 1% - of what we used to pay. It's also empowering. For creators. Game developers. Content influencers: you name it. -- superpowers -- 1. it does the things you ask for, in the order you asked for it 2. consistency is king 3. single-shot text or image-to-video 4. future videos can reference previous ones to seamlessly maintain style 5. semantic stitching: can't wait to showcase this -- gtm -- We think Generative AI Video, like image generation, like text, like games, should be a publicly accessible common good. We believe democratizing access to Nami in web3, via x402 payments proposed by Drew Coffman, or in World's mini-apps, is a bold step forward for digital freedom. Permissionless, decentralized, generative ai video. Naturally, we'll also soon release a web platform for using Nami in a traditionally SaaSy way: bring your own images, videos, or prompts and we'll take care of the rest. In the mid-term, Scrypted is building a stack of agentic skills (we call it AVB) and making them available to projects like H011yw00d Agent on Virtuals Protocol and other platforms. -- long-term vision -- Scrypted's mission is to decentralize the things that can't be decentralized. We participated in a16z crypto's CSX (London 2024) during our pre-seed specifically to research a new consensus protocol for hard things like AI video and AI agents: where there's no "one right answer". When Zero-Knowledge Proofs (ZKP) can't secure it, and Trusted Execution Environments (TEEs) are too small, we've got you covered with our upcoming Inori Network. -- how you can help -- 1. Are you a GPU farm? We're gonna need more flops. 2. Do you represent an L1 or L2? We want to build bridges. 3. Do you represent a Wallet or App creator? Let's get an endpoint exposed. 4. Are you an investor? Let's chat. 5. Like, repost, share! -- team background -- We come from a background of AI in the Video Game industry with each founder having over 20 years of experience at companies like Electronic Arts & Square Enix. -- contact -- DMs are open, reach out if you want to be an early tester for your site, game, collection, or project! -- try it out -- Go anywhere on X and tag H011yw00d Agent with a prompt and she'll give you a free 2 second render. Have fun making cinematic shorts or meme videos! -- thanks -- AWS Startups has been an incredible help scaling our prototypes. Also, shout out to all loyal beans 🫘 in the Autonomous Virtuals Beings (AVB) community. Nami has a very important role in the upcoming XP agent platform, can't wait to show you all. AVbeings

Tim Cotten

12,617 views • 1 year ago

Because you guys loved the 20 minutes of me asking the Humane Ai Pin voice questions so much, here's 19 minutes (almost 20!), no cuts, of me asking the rabbit inc. R1 AI questions and using its computer vision to "look" at stuff Some quick thoughts on the R1: • The AI/LLM is not perfect, but it gets way more correct than it does incorrect • The R1 is FAST to respond with answers. The Ai Pin looks embarrassingly slow in comparison • Vision is very impressive. Sometimes it IDs objects incorrectly (like in my other video, but since hard resetting, it seems to get more things right). It's also fast like the audio responses • Summaries are on point. I pointed the R1 at various Inverse articles that I either wrote or edited and it did a great job giving me the main points, even when the text was friggin' tiny on my iMac • The LLM is far more intelligent than on Ai Pin. It's better at understanding follow-ups with natural language. The Ai Pin is supposed to be contextual, but it often doesn't seem to remember what I said right before • It's late (3:30 am right now) so I have not connected my R1 to services like Spotify, Uber, or Midjourney. Will do that in the morning after I get some sleep. I'm very excited to see how Large Action Model (LAM) works and to teach the R1 to do stuff for me • There are some bugs that and Peiyu Liao tell me they're working on. For example, fixing the time (very important) and adding the % symbol back (also very important if your Wi-Fi password uses it!). Somehow, they seem to be working faster to fix bugs and issues than Humane • Jesse also tells me they're paying close attention to feedback on the sensitivity of the analog scroll wheel. It doesn't feel responsive enough, sometimes lagging a half second behind your actual scroll. He says they tuned it to be less sensitive to prevent it from activating on surfaces like a table. I think it could be a smidge more responsive. At least, that can be adjusted in a future software update This is not a review, only first impressions. I need to actually spend real time using and, most importantly, living with the R1. That being said, my initial setup bugginess/issues aside, the R1 is (as you can see in this long video) working quite well. Again, not perfectly every time, but far better than the Ai Pin. I am impressed. Really, really impressed Drop your questions and I'll answer them in the morning. What an exciting moment in tech. I live for this kinda stuff!

Ray Wong

721,838 views • 2 years ago

🚀 Introducing PantheonOS ( A Fully Open-Source Agent OS for Science PantheonOS began as a research project in my Stanford lab and has since evolved into a vision to redefine data science in the era of AI—starting with computational biology, especially single-cell and spatial genomics. PantheonOS is a general agent platform built from the ground up. It is arguably the first distributed agent framework designed for scientific data analysis. 🔑 Key Features 1. Multi-Agent Collaboration – Built-in paradigms for distributed, cross-machine cooperation among agents and toolsets. 2. Native Toolset Support – Python, R, Julia, LaTeX, and more—designed for real scientific workflows. 3. Modular & Extensible – Developer-friendly design with shallow wrappers, plus LLM-driven toolset generation. 4. Evolvable Agents – Capable of evolving large-scale code projects to achieve superhuman performance (e.g., evolving upon the original Harmony [I Korsunsky, 2019, Nature Biotechnology] and Scanorama [BL Hie, 2019, Nature Biotechnology] implementations), and even evolving the system itself to adapt to new fields. 🎉 Stepwise Release Strategy We’re releasing PantheonOS in stages: Pantheon-CLI (today!), followed by Pantheon-Lab, Pantheon-Notebook, Pantheon-Slack, and more. 🌟 Pantheon-CLI Highlights - We're not just building another CLI tool. We're defining how scientists will interact with data in the AI era. - Open, Powerful, Python-First – The first fully open-source, endlessly extendable scientific “vibe analysis” framework. - Mixed Programming Magic – Combine Python, natural language, R, or Julia—seamlessly in the same environment. - PhD-Level Assistant – A command-line agent for complex real-world genomics and beyond, handling workflows at the PhD level. - Privacy by Design – Run entirely offline with local LLMs—your data never leaves your computer. ✅ Proven Applications (10 Demonstrations) Computational biology: 1. ATAC-seq: From raw reads to peak matrix 2. RNA-seq: From raw reads to expression matrix 3. Complex single-cell workflows (PhD-level) 4. Hybrid natural language + R for Seurat annotation 5. Learning from web tutorials + invoking single-cell foundation models 6. Cell segmentation on 10x Genomics HD Visium data And beyond: 7. Mixed Python & R programming examples 8. Molecular docking & structural analysis 9. Exploratory factor analysis for behavioral survey data 10. Customer segmentation & finance analytics 🌐 Learn More & Get Started Website: Pantheon-CLI Documentation: GitHub Repo: 💬 Join our community: PantheonOS Slack: PantheonOS Discord:

evo-devo

17,369 views • 11 months ago

(4 DAYS BEFORE SUBMISSIONS CLOSE) I get this question a lot about the Find Evil! hackathon: What does “find evil” actually mean? In this case, the name comes from a real command. I built an autonomous incident response agent I built on the SIFT Workstation. Then I typed “find evil” as a prompt into Claude Code. And it did (watch the demo). I was blown away to watch the autonomous agent run a complete C drive forensic analysis, across 200+ tools via MCP. The agent identified threat actor and context, the attack chain, malware deployment method, persistence mechanisms, code injection analysis, network connections, command-and-control (C2) infrastructure, a complete malicious process tree, and a chronological activity timeline. Two days after I shared initial findings, Anthropic released their report on how threat actors were deploying Claude Code with operational tools and letting it go do evil. (Same thing I was doing.) Find Evil! is the first hackathon dedicated to building autonomous AI agents for incident response. 4,178 defenders are working on final Find Evil! hackathon submits. (This number makes me very happy to see so many diving in. And wishing that the thousands more in our community were experimenting with us.) Your job: teach an AI agent to think like a senior analyst, how to sequence its approach, recognize when something doesn’t add up, and self-correct when it gets it wrong. There are FOUR DAYS left to build with us! (Very few of us are actual AI experts. The rest of us including me are learning.) Register: Apply to judge: We need DFIR, AI, cybersecurity, and open-source reviewers who can separate useful autonomous response tools from polished demos. Apply: I am SO EXCITED to see what comes out of this hackathon and goes back to the community. Sponsored by SANS Institute

Rob T. Lee

14,405 views • 1 month ago