Loading video...

Video Failed to Load

Go Home

Gemma 4 is here! Our most intelligent open models to date, are built on the same world-class research and tech as Gemini 3, and are sized to run and fine-tune efficiently on local hardware. Check out what Google Gemma 4 brings to devs: 💎 Advanced Reasoning: Deep logic tasks,...

269,657 views • 3 months ago •via X (Twitter)

0 Comments

No comments available

Comments from the original post will appear here

Related Videos

Before the week ends, let's acknowledge one of the most INSANE week ever for open AI, with 25+ notable open-weight drops across every modality: 🧠 LLMs → NVIDIA Nemotron 3 Ultra: 550B hybrid Mamba-MoE, only 55B active, 1M context, MMLU 89.1. NVFP4 variant claims ~5x throughput on Blackwell. First openly-weighted 550B hybrid Mamba-Transformer, closing the gap with frontier closed models. → Google Gemma 4 12B: fully open dense any-to-any (text/image/audio/video), 256k context, encoder-free, 140+ languages, AIME 2026 at 77.5. Shipped with a 23-checkpoint QAT wave (mobile ONNX + MLX). Most deployable model of the week. → StepFun Step-3.7-Flash: 198B sparse MoE VLM, ~11B active, SWE-Bench PRO 56.3. Apache 2.0. → Liquid AI LFM2.5-8B-A1B: edge MoE, just 1.5B active, 128k ctx, MATH500 88.8, MLX-ready. Best on-device option this week. → JetBrains Mellum2-12B-A2.5B-Thinking: their first open MoE, near-Qwen3-14B coding at 2.5B active. Apache 2.0. 🎨 Image gen (the surprise of the week) → Ideogram 4: their FIRST-EVER open weights. 9.3B flow-matching DiT trained from scratch. #2 overall behind GPT Image 2, top open-weight model on Design Arena + LMArena. Strongest open checkpoint for text-rich images, full stop. It has taste. Still can't believe this is open weights. 🔊 Audio & Speech (a breakout week for open TTS, 4 labs shipped) → Boson Higgs Audio v3 4B: 102 languages, 21 emotions, singing/whispering/shouting, sub-second TTFA. → RedNote dots.tts: the only fully continuous (no codec) open TTS pipeline, Apache 2.0. → Google Magenta RealTime 2: real-time music gen, <200ms latency, text+audio+MIDI. multimodalart ported it to PyTorch within hours with live ZeroGPU demos. → NVIDIA Nemotron-3.5 ASR: 600M streaming, 17x more concurrent streams vs Parakeet RNNT 1.1B. 👁️ Vision & VLMs → PaddleOCR-VL-1.6: SOTA document parsing at 1B params, Apache 2.0. → Baidu NAVA: 6.3B joint audio-video gen, best-in-class A/V sync, Apache 2.0. 🎬 Video, 3D & World Models → NVIDIA Cosmos3-Super: 64B omnimodal world model coupling action trajectories with video+audio gen, for Physical AI. → JD JoyAI-Echo: up to 5-min multi-shot text-to-video on LTX-2.3. → ByteDance Bernini-R + VAST TripoSplat (single-image-to-3D Gaussian splats, MIT).

Victor M

539,522 views • 1 month ago

Run Gemma 4 26b MTP on 8 GB VRAM GPUs at 25+ tokens/second. Flags included! local llm space is moving at terminal velocity. only 3 days ago google released gemma 4 26b a4b qat quants. more efficient than before, ran on 8gb vram at 20 tok/sec. and now just a few hours ago, mainline llama.cpp merged a massive update and we just shattered our own record. decode throughput went 25-40% up on the same 8 GB VRAM setup! Before MTP: 20 tps -> After MTP: 28 tps! llama.cpp just officially merged PR #23398 ("add Gemma4 MTP"), bringing native Multi-Token Prediction (MTP) support to Gemma 4 models. By running speculative drafting on the same 8GB VRAM RTX 4060 setup, my decode throughput on a 64k context instantly leaped to a blistering 25–27 tokens/sec thats 25-30% increase with the same hardware. Here is the architectural catch you need to know: Unlike the Qwen 3.5 and 3.6 series, which bake the MTP heads directly into the base GGUF, the Gemma 4 MTP head is not built in. You must download a separate, specialized MTP drafter GGUF (the assistant model) to act as the speculator. (I've dropped the download link in the replies). copy and try the exact flags: -m gemma-4-26B-A4B-it-qat-UD-Q4_K_XL.gguf --spec-type draft-mtp --spec-draft-n-max 6 --spec-draft-p-min 0.7 --spec-draft-model gemma-4-26b-A4B-it-assistant-Q4_0.gguf -c 64000 -v n-max 4 and p-min 0.7 is also worth checking out. benchmark on your setup and workflow. if you have a single 8 gb vram nvidia rtx 4060, 3060, 3070, 2080, 2070, grab the MTP drafter GGUF link in the comments and try it yourself. Check it out even if you have asmaller or a larger gpu, such as a single rtx 3090, 4090, 3060, 2060. MTP works for all gemma 4 sizes such as gemma 4 12b, gemma 4 31b etc. but remember to grab the correct mtp draft assistant models respectively. what are you benchmarking today

Alok

200,913 views • 1 month ago

Microsoft presents Windows Agent Arena Evaluating Multi-Modal OS Agents at Scale discuss: Large language models (LLMs) show remarkable potential to act as computer agents, enhancing human productivity and software accessibility in multi-modal tasks that require planning and reasoning. However, measuring agent performance in realistic environments remains a challenge since: (i) most benchmarks are limited to specific modalities or domains (e.g. text-only, web navigation, Q&A, coding) and (ii) full benchmark evaluations are slow (on order of magnitude of days) given the multi-step sequential nature of tasks. To address these challenges, we introduce the Windows Agent Arena: a reproducible, general environment focusing exclusively on the Windows operating system (OS) where agents can operate freely within a real Windows OS and use the same wide range of applications, tools, and web browsers available to human users when solving tasks. We adapt the OSWorld framework (Xie et al., 2024) to create 150+ diverse Windows tasks across representative domains that require agent abilities in planning, screen understanding, and tool usage. Our benchmark is scalable and can be seamlessly parallelized in Azure for a full benchmark evaluation in as little as 20 minutes. To demonstrate Windows Agent Arena's capabilities, we also introduce a new multi-modal agent, Navi. Our agent achieves a success rate of 19.5% in the Windows domain, compared to 74.5% performance of an unassisted human. Navi also demonstrates strong performance on another popular web-based benchmark, Mind2Web. We offer extensive quantitative and qualitative analysis of Navi's performance, and provide insights into the opportunities for future research in agent development and data generation using Windows Agent Arena.

AK

19,684 views • 1 year ago

Our 2025 Achievements at Propbase It has been a rockstar start to the year and we have so much more to come after so many achievements already! H1 - So far Launched Propbase 2.0 Technical Roadmap We unveiled our Propbase 2.0 roadmap, focusing on DeFi-inspired lending, AI-powered automation, and institutional-grade tokenization infrastructure. Why It Matters: This roadmap is our blueprint for transforming real estate investment. By integrating DeFi mechanics and AI, we’re scaling our platform to attract both retail and institutional investors, boosting our ecosystem’s total value locked (TVL). Tokenized and Sold Out Ramada Plaza by Wyndham We tokenized a Ramada Plaza by Wyndham property, and it sold out in just 11 hours. Why It Matters: This rapid sell-out shows the trust our investors place in us. It proves we can deliver high-quality assets, making real estate investment accessible with fractional ownership starting at $100. Revamped Our Website and Updated PROPS Vesting We launched a sleek new website and introduced a performance-driven PROPS vesting schedule to reward long-term holders. Why It Matters: Our new website makes it easier for users to explore our platform, while the vesting schedule—unique in crypto—builds confidence among token holders, aligning incentives for long-term growth. Switched to Native USDC Across Products We transitioned from USDT to native USDC for all our products, including crowdfunding, trading, and staking. Why It Matters: USDC’s reliability enhances transaction security and aligns us with global standards, making our platform more appealing to investors worldwide. Passed Certik Audit for APEX Smart Contracts Our Apex platform, a decentralized marketplace for property-backed tokens, passed a Certik audit with flying colors. Why It Matters: Security is our priority. This audit reassures our users that Apex’s smart contracts are rock-solid, enabling safe and scalable trading. Fully Launched Our Secondary Marketplace We rolled out our secondary marketplace, allowing users to buy and sell tokenized property shares with improved liquidity. Why It Matters: This marketplace solves the liquidity challenge in real estate, letting investors trade stakes anytime, anywhere, which drives engagement and trading volume. Deepened Integration with Aptos Ecosystem We focused on integrating with the Aptos blockchain, prioritizing TVL growth through Propbase Nexus and trading volume on our secondary marketplace. Why It Matters: Aptos’s scalability and low fees make our platform more efficient. This integration fuels our growth and strengthens our role in the Aptos ecosystem. Started Mobile App Development 💎💎💎Coming H2 2025💎💎💎 Propbase App Full-stack development for our iOS and Android apps, with a beta rollout planned for early users, featuring trading, staking, and governance. Why It Matters: A mobile app makes our platform accessible on the go, meeting user expectations and driving adoption among a broader audience. P2P Lending and Borrowing P2P lending and borrowing system, inspired by Aave and Compound, letting users borrow USDC using asset-backed tokens as collateral. Why It Matters: This DeFi feature unlocks capital for investors without forcing asset sales, boosting token utility and attracting DeFi enthusiasts. AI-Powered Liquidity Automation AI models to automate liquidity stability across our asset pools, using DEFAI for real-time risk management and market optimization. Why It Matters: Our AI ensures price stability and efficient liquidity, setting us apart as a tech-forward platform and building investor trust. Make sure you're here with us in 2025 H2 for Propbase to give you more than what they already have from H1

Propbase

46,508 views • 1 year ago

Security is and always will be one of the four pillars Propbase is built on. It is critical in building out a platform built for mass adoption. We ensure the security of our products and by extension our community that uses them. Propbase achieves a best in class security by implementing the follow at a minimum: Security Features Aptos Blockchain Foundation: Propbase operates on Aptos, a Layer-1 Proof-of-Stake blockchain renowned for its high-speed transactions, low fees, and Byzantine fault-tolerant consensus mechanism. This ensures unparalleled reliability and protection against network vulnerabilities. Military-Grade Wallet Encryption: Our Wallet Module uses advanced elliptical signature technology and military-grade encryption to secure your digital assets, giving you full custody and peace of mind. Certik & Hacken Audits: We’ve undergone rigorous smart contract audits by Certik and Hacken, earning top security rankings (Top 14# on Certik). These audits validate our platform’s integrity and compliance with industry best practices. CERTIK: HACKEN: Two-Factor Authentication (2FA): Admin and user accounts are protected with 2FA, including Google Authenticator integration, ensuring secure access and transaction approvals. Transparent Blockchain Registry: Every property transaction is recorded on our public title registry, verifiable on the Aptos blockchain. This eliminates fraud risks and ensures full transparency. Legal Compliance: Each property is held in a U.S.-registered LLC, providing true ownership with legal protections. Our rigorous due diligence process ensures all listings meet strict compliance standards. KYC Verification: A robust Know Your Customer (KYC) verification process to enhance user trust and regulatory compliance, further securing the platform against unauthorized access. Multi-Signature Wallets: Multi-sig wallets for both Web3 and non-crypto users, adding an extra layer of security for high-value transactions. Security measures are always assessed, updated and reimagined as needed. Our base level security is at an incredibly high level which brings peace of mind to property token and $PROPS holders 💎💎💎

Propbase

42,519 views • 1 year ago