Загрузка видео...

Не удалось загрузить видео

На главную

Introducing hallucination correction. We have reduced hallucination by 70%. Giga's hallucination rate is at ~1%. Better than the best frontier models. Deploy AI your customers can trust.

1,225,973 просмотров • 4 месяцев назад •via X (Twitter)

Комментарии: 32

Фото профиля Tiffany Fong
Tiffany Fong4 месяцев назад

NO MORE FAKE RESEARCH STUDIES

Фото профиля Not Jerome Powell
Not Jerome Powell4 месяцев назад

I need to correct my printer’s hallucinations. Prints too much

Фото профиля Giga
Giga4 месяцев назад

Learn more at

Фото профиля Giga
Giga4 месяцев назад

Check out our CTO's technical blog as well:

Фото профиля World of Engineering
World of Engineering4 месяцев назад

Impressive work on hallucination correction-lowering to ~1% is a big step for trustworthy AI! 🚀

Фото профиля Ali Tekin
Ali Tekin4 месяцев назад

The biggest problem with AI isn’t that it makes mistakes. It’s that sometimes it delivers wrong answers with complete confidence. That’s why hallucination correction may become the most important AI race of the next few years.

Фото профиля ryan
ryan4 месяцев назад

Another great launch by the team (and great video from @jcarvajalpa )

Фото профиля Vincent Sun
Vincent Sun4 месяцев назад

this was such a fun concept to work on with @jcarvajalpa @RyanJosephHill, @ben_aguilera’s world class animation team

Фото профиля Tyler McGrath
Tyler McGrath4 месяцев назад

Glad Giga is taking this on. Hallucinations aren’t talked about enough

Фото профиля Cameron Ellis
Cameron Ellis4 месяцев назад

No hallucinations were spoken. We're simply the best.

Фото профиля Patrick Conn
Patrick Conn4 месяцев назад

Game changer!

Фото профиля Ben Aguilera
Ben Aguilera4 месяцев назад

🔥🔥🔥🔥

Фото профиля Bek
Bek4 месяцев назад

1% is the wall not the floor. one hallucinated balance on a banking call triggers an escalation the agent can't backtrack from. what's the recovery flow when it fires?

Фото профиля Glitch Truth
Glitch Truth4 месяцев назад

GPT-4 Turbo benchmarks at ~3% on TruthfulQA, Claude Opus around 2.5%. If Giga is hitting 1% on real customer queries not curated benchmarks, what's the eval set? Vectara's HHEM leaderboard or in-house? Big difference between RAG-grounded 1% and open-ended 1%.

Фото профиля Alex Veremeyenko
Alex Veremeyenko4 месяцев назад

The AI hype cycle has been pretending hallucinations are “just part of the game.” They’re not. They’re the reason most AI products still feel unsafe in front of real customers. Giga cutting that down is a big deal.

Фото профиля AndreWGMI
AndreWGMI4 месяцев назад

so the fix is literally just exploiting the gap between how fast it generates text vs how slow humans speak?

Фото профиля Alex Prompter
Alex Prompter4 месяцев назад

Everyone is busy shipping AI that sounds confident while making stuff up. Giga fixing hallucinations first is exactly the boring-sounding thing that separates toys from products customers can actually trust.

Фото профиля Nader Abdulrub
Nader Abdulrub4 месяцев назад

This is a solid move towards making voice agents more reliable. You need a smarter model to monitor their behavior in realtime. We’ve also implemented a similar pattern for our restaurant voice ai and has improved behavior by 8x

Фото профиля ALICE ⚡ | AI Growth Mentor
ALICE ⚡ | AI Growth Mentor4 месяцев назад

Wow!

Фото профиля とりさん@AI&Humans Co-Create Innovative Futures AIと推し活
とりさん@AI&Humans Co-Create Innovative Futures AIと推し活4 месяцев назад

While reducing the hallucination rate to ~1% is an impressive achievement, I argue that the risk of AI generating hallucinations as a byproduct of prioritizing performance outcomes fundamentally arises from an overemphasis on "quality and efficiency" in its optimization objectives, with insufficient consideration given to "safety and contribution to humanity." Based on the framework of X-CII theory, I propose that integrating a real-time evaluation and control mechanism for Safety (S: harm prevention and human benefit)—in addition to Quality (Q) and Efficiency (E)—into the AI's decision-making process is essential for a fundamental resolution of this problem.

Фото профиля Wizard Trade
Wizard Trade4 месяцев назад

1% by what benchmark though closed evals or messy real user prompts? hallucination rate claims only matter if the test set actually hurts curious how it holds up when users go off-script

Фото профиля 🌊 Machine
🌊 Machine4 месяцев назад

Is it jailbreak proof?

Фото профиля Prateek Tripathi
Prateek Tripathi4 месяцев назад

1% is cute until someone asks their balance and gets told they're broke

Фото профиля Aaliya
Aaliya4 месяцев назад

awesome

Фото профиля Kin
Kin4 месяцев назад

New to you, LFG

Фото профиля Aladdin Kaya
Aladdin Kaya4 месяцев назад

@garrytan Hallucinations are visible. Behavioral drift usually isn’t. That’s why production systems often fail while still looking operational.

Фото профиля Aakash Gupta
Aakash Gupta4 месяцев назад

Epic hallucination reduction

Фото профиля Alice The Ai Expert
Alice The Ai Expert4 месяцев назад

Great Giga cuts hallucinations 70%

Фото профиля Vitor
Vitor4 месяцев назад

✨

Фото профиля 𝕋𝕙𝕖 ℂ𝕪𝕓𝕖𝕣 ℍ𝕠𝕦𝕤𝕖
𝕋𝕙𝕖 ℂ𝕪𝕓𝕖𝕣 ℍ𝕠𝕦𝕤𝕖4 месяцев назад

@garrytan 👀

Фото профиля Sani Ai Tech
Sani Ai Tech4 месяцев назад

Giga: 70% fewer hallucinations, ∼1% rate — AI you can trust

Фото профиля Space_Dogge
Space_Dogge4 месяцев назад

Wow, that's really nice. Next level. 😎🔥🙌🏻

Похожие видео