Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

We’re open-sourcing PorTAL, our framework for shared task representations and cross model LoRA adaptation. It now spans from hybrid attention models to multimodal systems including Gemma 4 E2B, Mistral 7B & Thinking Machines' Inkling. Code: ramp-public/portallib Models: Hugging Face /RampPublic

163,711 görüntüleme • 8 gün önce •via X (Twitter)

0 Yorum

Yorum bulunmuyor

Orijinal gönderinin yorumları burada görünecek

Benzer Videolar

Molmo by Ai2 - Open source SoTA Multimodal (Vision) Language model, beating Claude 3.5 Sonnet, GPT4V and comparable to GPT4o 🔥 They release four model checkpoints: 1. MolmoE-1B, a mixture of experts model with 1B (active) 7B (total) 2. Molmo-7B-O, most open 7B model 3. Molmo-7B-D, demo model 4. Molmo-72B, best model System Architecture > Input: Multi-scale, multi-crop images generated from the original image. > Vision Encoder: OpenAI's ViT-L/14 336px CLIP model, a powerful ViT, encodes images into vision tokens. > Connector: MLP projects tokens to LLM input space, followed by pooling for dimensionality reduction. > LLM: Decoder-only Transformer, various options (OLMo, OLMoE, Qwen2, Mistral, Gemma2, Phi) with diverse scales and openness. Model Variants > Vision Encoder: Consistent ViT-L/14 CLIP model across variants. > LLM: OLMo-7B-1024, OLMoE-1B-7B-0924, Qwen2 (7B, 72B), Mistral 7B, Gemma2 9B, Phi 3 Medium, offering different capacities and openness levels. Training Strategy > Stage 1: Multimodal pre-training for caption generation with new captioning data. > Stage 2: Supervised fine-tuning on a dataset mixture, updating all parameters. > No RLHF involved, Learning rates adjusted based on component types and pre-training status. > All the weights are available on Hugging Face Hub 🤗 > Compatible with Transformers (Remote Code) Kudos Ai2 for such a brilliant and open work! 🐐 Video credits: Allen AI YT Channel

Vaibhav (VB) Srivastav

80,474 görüntüleme • 1 yıl önce

🚨Update! Our new demo is LIVE 🚨 In this demo, we walk through the core features of Intelligence Cubed, a next-generation AI model platform built for research, experimentation, and ownership. 🔹 500+ Research Models Intelligence Cubed has grown from 200+ to 506 models, contributed by our expanding Research Fellow Cohort, including researchers, PhDs, and post-docs from Stanford, CMU, Harvard, MIT, and other top U.S. institutions. 🔹 Model Cards & Research Transparency Each model is linked to its original research paper and includes a detailed model card outlining its purpose, use cases, category, pricing, market traction, reviews, and public ownership percentage. 🔹 1.2M Public-Owned Models We’ve introduced Public-Owned Models, with over 1.2 million models available — all fully documented with research papers and comprehensive model cards. 🔹 Auto Router Not sure which model to use? Our Auto Router analyzes your question and automatically routes it to the most suitable model. In this demo, it selects an LLM Detection Survey model to answer the query. 🔹 Modelverse, Canvas & Workflows Users can explore models in Modelverse, try them instantly, add favorites to cart, and deploy purchased models in Canvas using drag-and-drop to build custom workflows. We also provide professionally curated workflows for immediate hands-on experience. 👉Try Now: #AI #Web3 #AIModel #DeFi #blockchain #LLM #OpenSourceAI #AIxWeb3 #DeAI #IntelligenceCubed

i³ (Intelligence Cubed)

116,576 görüntüleme • 7 ay önce

Today is a good day for open science. As part of our continued commitment to the growth and development of an open ecosystem, today at Meta FAIR we’re announcing four new publicly available AI models and additional research artifacts to inspire innovation in the community and help advance AI in a responsible way. More in the video from Joelle Pineau. What we’re releasing: 🦎 Meta Chameleon 7B & 34B language models that support mixed-modal input and text-only outputs. 🪙 Meta Multi-Token Prediction Pretrained Language Models for code completion using Multi-Token Prediction. 🎼 Meta JASCO Generative text-to-music models capable of accepting various conditioning inputs for greater controllability. Paper available today with a pretrained model coming soon. 🗣️ Meta AudioSeal An audio watermarking model that we believe is the first designed specifically for the localized detection of AI-generated speech, available under a commercial license. 📝 Additional RAI artifacts Including research, data and code to measure and improve the representation of geographical and cultural preferences and diversity in AI systems. We believe that access to state-of-the-art AI creates opportunities for everyone – not just a small handful of Big Tech companies. We’re excited to share this work and to see how the community learns, iterates and builds using this technology. Details and access to everything released by FAIR today ➡️

AI at Meta

380,822 görüntüleme • 2 yıl önce