Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

An Imperial College eng professor gave four LLMs a problem set that graduate students had two months to solve. He had TAs grade the results blind alongside real submissions. Meta AI and Claude failed. ChatGPT ranked 27 of 36 students...while Gemini 2.5 Pro ranked 4 of 36 🤯

238,338 görüntüleme • 1 yıl önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

BREAKING: FIRE and College Pulse have just released the 2025 College Free Speech Rankings. Surveying 58K+ students from 257 institutions, the results find that free speech has been threatened in historic ways since the Israel-Hamas war began. Expand to learn more ⬇️ 🟡 UVA was ranked as this year’s best school for free speech! Michigan Technological University, Florida State University, Eastern Kentucky University, and Georgia Tech round out the top five. 🔴 Harvard University is ranked as the worst school for free speech for the second year in a row. Columbia University and New York University round out the bottom three — each earning an “abysmal” rating. 🔴 Campus deplatforming attempts spiked by 400% since Hamas’ October 7th attack on Israel. 🔴 7 out of 10 students are uncomfortable publicly disagreeing with a professor about a controversial political topic. 🔴 54% of students report it’s difficult to talk about the Israeli-Palestinian conflict. On 17 campuses where this issue was especially contentious, 75% or more marked the Israeli-Palestinian conflict as difficult to discuss. 🔴 1 in 4 students said it was unclear whether their college administration protects free speech. 🔴 1/3 of students approve of the use of violence at least “rarely” to stop campus speakers. 🔴 7 out of 10 students say it's at least “rarely okay” to shout down a speaker. Given all of this, it is unsurprising that American confidence in higher education is at a record low.

FIRE

166,545 görüntüleme • 1 yıl önce

GROK SURGES TO THE FRONT OF THE GENAI RACE AS GROWTH SKYROCKETS Grok is absolutely amazing, continuing to stun with incredible results. Web traffic jumped nearly 15% month-over-month in November 2025, the fastest growth in the generative AI industry, proving Elon’s xAI project isn’t just competing, it’s taking real market share from ChatGPT. Grok hit around 234.4 million visits in November, up from 204 million in October. That’s a staggering 1,300% year-over-year surge, pushing it past Perplexity and Claude in user growth and cementing it as the world’s #2 chatbot by market share. The breakout came with Grok 4.1’s release in mid-November. The update debuted at number 1 on LMSYS Arena, that’s the global benchmark where AI models are ranked through blind human evaluations. Grok’s new “Thinking Mode” scored 1483 Elo, a 31-point lead over every open model, beating Gemini 2.5 Pro and Claude 3. The upgrade also cut hallucinations (false answers) by two-thirds, expanded its context window to 2 million tokens, about 1.5 million words of memory, and dominated top reasoning and coding tests like, graduate-level logic, and the emotional intelligence aspect. U.S. traffic surged to 51.5 million visits, boosted by X integration and Grok’s unfiltered style. At just $0.20 per million input tokens, versus GPT-5.1’s $1.25—it’s winning with both speed and affordability. The momentum isn’t hype, it’s lift-off. If growth holds, Grok could reach 500 million users by mid-2026, forcing every rival to redefine what “intelligent” really means. To truthful AI winning! Source: X Freeze, NextBigFuture, CometApi, Langcopilot

Mario Nawfal

33,823 görüntüleme • 8 ay önce