Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Announcing the open-source release of Qwen3-VL! A powerful vision-language model that can operate GUIs, code charts from mockups, and recognize "everything" from daily life to specialized fields. Highlights: 🔹 Precise event location in videos up to 2 hours long. 🔹 OCR language support boosted from 19 to 32, with...

59,944 Aufrufe • vor 10 Monaten •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

We've officially released and open-sourced HunyuanImage 2.1, our latest text-to-image model. The new model delivers on our commitment to balancing performance and quality. With native 2K image generation, HunyuanImage 2.1 is an advanced open-source text-to-image model.🎨 ✨ New in 2.1: 🔹Advanced Semantics: Supports ultra-long and complex prompts of up to 1000 tokens, and precisely controls the generation of multiple subjects in a single image. 🔹Precise Chinese and English Text Rendering with seamless image–text integration: The model naturally integrates text into images, making it suitable for a wide range of applications such as product covers, illustrations, and poster design to meet the needs of various fields. 🔹Rich Styles and High Aesthetic: Capable of generating images in various styles—including photorealistic portraits, comics, and vinyl figures—it delivers outstanding visual appeal and artistic quality. 🔹High-Quality Generation: Efficiently produces ultra-high-definition (2K) images in the same time other models take to generate a 1K image. HunyuanImage 2.1 uses two text encoders: a multimodal large language model (MLLM) to improve the model's image and text alignment capabilities, and a multi-language character-aware encoder to improve text rendering capabilities. The model is a single- and double-stream diffusion transformer with 17B parameters. We've also open-sourced the weights of the the accelerated version with meanflow which reduces inference steps from 100 to just 8, and PromptEnhancer, the first industrial-grade rewriting model that enhances your prompts for more nuanced and expressive image generation. Now, creators turn complex ideas—like posters with slogans or multi-panel comics—into visuals faster than ever. We’re just getting started. Stay tuned for our native multimodal image generation model coming soon. 🌐Website: 🔗Github: 🤗Hugging Face: ✨Hugging Face Demo:

Tencent Hy

89,257 Aufrufe • vor 10 Monaten

🌐 BenFen | Next-Gen Multi-Currency Stablecoin Blockchain BenFen is a Layer1 blockchain purpose-built for stablecoin issuance, adoption, and payments, providing trusted and accessible on-chain payment infrastructure for global users and developers. 🟥 Technical Architecture:Built on the Move language, ensuring security, high performance, and accessibility. 🔹 Modular contract design ensures asset security and tamper-proof rules for reliable system operation. 🔹 Sub-second transaction confirmation with stable 10,000+ TPS, optimized for payment and interaction scenarios. 🔹 Supports zkLogin, enabling one-click wallet creation with Google/Apple ID. 🟥 BUSD Stablecoin Mechanism :Pegged to USDT/USDC, the core on-chain asset 🔹 Cross-chain 1:1 pegged minting with USDT/USDC, wit h native support for stablecoin GAS payments. 🔹 Supports mainstream G20 fiat-backed stablecoin conversions (eg, BUSD/BJPY, BUSD/BEUR). 🔹 Covers key scenarios like asset trading, RWA mapping, on-chain payroll, and daily consumption. 🟥 Native Features: Full-stack capabilities centered around stablecoins 🔹 Supports one-click issuance of stablecoins/RWAs, lowering development barriers. 🔹 Any issued stablecoin can be registered as a Gas token, with sponsored transactions available for Gas subsidies. 🔹 Zero-cost transfers for specific scenarios, enabling real "zero-fee payments". 🔹 Privacy accounts and payments: Supports hidden addresses and balances for privacy-focused use cases. 🔹 Multi-chain asset payments + G20 fiat settlement: Real-time settlement of USDT, SOL, ETH, and other major assets into BJPY, BEUR, BAUD, and other G20 stablecoins. 🟥 Native Cross-Chain Bridge 🔹 Proprietary native cross-chain protocol. 🔹 Supports BTC, ETH, BSC, Polygon, Optimism, Solana, TRON, Base, Avalanche, and more. 🔹 No third-party bridge is needed; assets can be cross-chain with just one click, offering fast, secure, on-chain verification and second-level fund arrival. 📱 BenFen Ecosystem 🔹 BenPay: An open, secure, and efficient comprehensive payment ecosystem, offering users a convenient and secure cryptocurrency payment channel. 🔹 BenPay Card: On-chain self-custodial payment card, with keys in hand, enabling global spending (Supports Apple Pay, Alipay, Amazon, Netflix, X, ChatGPT, etc.). 🔹 BenPay C2C: secure peer-to-peer decentralized trading marketplace. 🔹 BenPay DeFi Earn:A cross-chain yield farming protocol that allows stablecoins from major blockchains to be seamlessly transferred and deposited into high-APY farming pools. 🔹 BenPay DEX: One-click swaps and aggregated matching for stablecoins/crypto assets. 🔹 BenPay Lending: Decentralized finance protocol supporting BTC collateral and USDT lending. 🔹 BenPay Stake: Stake BFC to gain governance rights, ecosystem rewards, and airdrops. 🟥 BenFen DAO 🔹 Utilizes formal verification for security, supporting low-cost proposals and voting to promote community autonomous governance and rights protection. 🟥 Developer-Friendly Platform:Building a low-barrier development ecosystem 🔹 Provides standard SDKs/APIs/contract templates. 🔹 Supports full-cycle development scenarios like stablecoin issuance, RWA issuance, wallet integration, and DeFi application building. 🔹 Lowers development barriers to accelerate innovation. 🧭 Vision: To become the stablecoin value foundation for global on-chain payments BenFen is dedicated to building a next-generation high-performance stablecoin blockchain, with stablecoins as the core on-chain asset. Through native mechanisms, it connects cross-chain assets, stablecoin systems, and on-chain protocols to serve real-world financial interaction scenarios, creating a sustainable and high-frequency accessible stablecoin economic system. Learn more: #BenFen #RWA #Web3Payments #Stablecoin #MoveVM #zkLogin #CryptoInfra #Layer1 #PayFi #CrossChain #BUSD #DeFi #BenPay #OnChainSettlement #Web3Infrastructure

BenFen

37,286 Aufrufe • vor 10 Monaten

Excited to announce GR00T N1, the world’s first open foundation model for humanoid robots! We are on a mission to democratize Physical AI. The power of general robot brain, in the palm of your hand - with only 2B parameters, N1 learns from the most diverse physical action dataset ever compiled and punches above its weight: - Real humanoid teleoperation data. - Large-scale simulation data: we are open-sourcing 300K+ trajectories! - Neural trajectories: we apply SOTA video generation models to “hallucinate” new synthetic data that features accurate physics in pixels. Using Jensen’s words, “systematically infinite data”! - Latent actions: we develop novel algorithms to extract action tokens from in-the-wild human videos and neural generated videos. GR00T N1 is a single end-to-end neural net, from photons to actions: - Vision-Language Model (System 2) that interprets the physical world through vision and language instructions, enabling robots to reason about their environment and instructions, and plan the right actions. - Diffusion Transformer (System 1) that “renders” smooth and precise motor actions at 120 Hz, executing the latent plan made by System 2. We deploy N1 on GR1 robot, 1X Neo robot, and a large collection of simulation benchmarks. N1 achieves up to +30% boost in diverse manipulation tasks for household and industrial settings. While humanoid robots are the main focus of N1, our model also supports cross-embodiment. We finetune it to work on the $110 HuggingFace LeRobot SO100 robot arm! Open robot brain runs on open hardware. Sounds just right. Let’s solve robotics, together, one token at a time. Links to our Whitepaper, Github repo, HuggingFace model, and open dataset page in the thread: 🧵

Jim Fan

466,261 Aufrufe • vor 1 Jahr

Since the dawn of time, economies were powered by labor. Now, we’re entering a new age where AI becomes the workforce, and data + compute are the new raw materials. Enter $KAI in the lifeblood of new global autonomous capital markets. 💮 Ecosystem Plan - 💮 Token Ecosystem - 👇 ⸻ 2/ $KAI isn’t just a token. It’s a complete economic engine designed to: 🔹 Power products 🔹 Fuel agent networks 🔹 Incentivize data 🔹 Govern EQ+ infrastructure 🔹 Scale emotionally intelligent AI Here’s how it works 🧵 ⸻ 3/ Tokenomics (1B $KAI total supply): 🔹 Team – 10% (vested) 🔹 Ecosystem Fund – 30% (growth + integrations) 🔹 Incentives – 15% (UGC, quests, data) 🔹 Partnerships & Advisors – 15% 🔹Liquidity – 30% (initial trading on DEX/CEX) Vested. Scarce. Sustainable. ⸻ 4/ Market Dynamics: ~300M tokens in early circulation 🔹Liquidity on DEX + CEX 🔹Governance via staking 🔹5–15% deflation from fees + burns 🔹Modeled to drive long-term value via ecosystem velocity ⸻ 5/ Utility: 🔹 Arc Terminal – Stake $KAI to analyze sentiment across Web3, politics, culture 🔹 EQ+ API – Pay + stake to embed emotional intelligence in AI agents 🔹 █████ – Future-ready staking for sovereign AI workflows (TBA) Every product reinforces token lockups + value flow. ⸻ 6/ Circular Incentive Model: Fees → 🔹 70% to treasury 🔹 25% to rewards 🔹 5% burned 🔹Contribute data → earn $KAI 🔹Proof-of-Empathy → Incentivize unique, emotionally rich inputs Deflation + utility = sustainable ecosystem flywheel. ⸻ 7/ Roadmap: 🔹 Q3-Q4 2025 – Mint, Arc Terminal beta, staking 🔹 Q1-Q2 2026 – API v2, █████ beta, Kaiko Hub: leading AI ecosystem in UAE 🔹 Q3/Q4 2026+ – Scale to 1M+ agents, cross-industry integrations (health, edu, media) Emotion becomes infrastructure. And it scales. ⸻ 8/ TL;DR: $KAI powers the next evolution of AI economies. From empathy loops to agent networks. From data feedback to decentralized governance. This is not just about tokens. It’s about rebuilding value itself. 💮

KAIKO

37,495 Aufrufe • vor 1 Jahr

We’re excited to announce the release and open-source of HunyuanImage 3.0 — the largest and most powerful open-source text-to-image model to date, with over 80 billion total parameters, of which 13 billion are activated per token during inference.The effect is completely comparable to the industry’s flagship closed-source model.🚀🚀🚀 HunyuanImage 3.0 originates from our internally developed native multimodal large language model, with fine-tuning and post-training focused on text-to-image generation. This unique foundation gives the model a powerful set of capabilities: ✅Reason with world knowledge ✅Understand complex, thousand-word prompts ✅Generate precise text within images Different from traditional DiT architecture image generation models, HunyuanImage 3.0’s MoE architecture uses a Transfusion-based approach to deeply couple Diffusion and LLM training for a single, powerful system. Built on Hunyuan-A13B, HunyuanImage 3.0 was trained on a massive dataset: 5 billion image-text pairs, video frames, interleaved image-text data, and 6 trillion tokens of text corpora. This hybrid training across multimodal generation, understanding, and LLM capabilities allows the model to seamlessly integrate multiple tasks. Whether you're an illustrator, designer, or creator, this is built to slash your workflow from hours to minutes. HunyuanImage 3.0 can generate intricate text, detailed comics, expressive emojis, and lively, engaging illustrations for educational content. The current release focuses solely on text-to-image generation and future updates will include image-to-image, image editing, multi-turn interaction, and more. 👉🏻Try it now: 🔗GitHub: 🤗Hugging Face:

Tencent Hy

412,658 Aufrufe • vor 10 Monaten

🚨Update! Our new demo is LIVE 🚨 In this demo, we walk through the core features of Intelligence Cubed, a next-generation AI model platform built for research, experimentation, and ownership. 🔹 500+ Research Models Intelligence Cubed has grown from 200+ to 506 models, contributed by our expanding Research Fellow Cohort, including researchers, PhDs, and post-docs from Stanford, CMU, Harvard, MIT, and other top U.S. institutions. 🔹 Model Cards & Research Transparency Each model is linked to its original research paper and includes a detailed model card outlining its purpose, use cases, category, pricing, market traction, reviews, and public ownership percentage. 🔹 1.2M Public-Owned Models We’ve introduced Public-Owned Models, with over 1.2 million models available — all fully documented with research papers and comprehensive model cards. 🔹 Auto Router Not sure which model to use? Our Auto Router analyzes your question and automatically routes it to the most suitable model. In this demo, it selects an LLM Detection Survey model to answer the query. 🔹 Modelverse, Canvas & Workflows Users can explore models in Modelverse, try them instantly, add favorites to cart, and deploy purchased models in Canvas using drag-and-drop to build custom workflows. We also provide professionally curated workflows for immediate hands-on experience. 👉Try Now: #AI #Web3 #AIModel #DeFi #blockchain #LLM #OpenSourceAI #AIxWeb3 #DeAI #IntelligenceCubed

i³ (Intelligence Cubed)

116,576 Aufrufe • vor 6 Monaten