Загрузка видео...

Не удалось загрузить видео

На главную

📢 ScanNet++ v2 Benchmark Release! 🏆 Test your state-of-the-art models on: 🔹 Novel View Synthesis 📸➡️🖼️ 🔹 3D Semantic & Instance Segmentation 🤖🔍🕶️ Shoutout to Chandan Yeshwanth and Yueh-Cheng Liu for their incredible work👏 🚀Check it out:

13,438 просмотров • 1 год назад •via X (Twitter)

Комментарии: 3

Фото профиля Out of AI
Out of AI1 год назад

@chandan__yes @liuyuehcheng amazing work, that it's going to help us with our research as well.

Фото профиля Λbstract
Λbstract1 год назад

RT @InstaMAT_io: 👀 Are you looking for an alternative to Adobe Substance tools? Do not miss out on InstaMAT from killer node-based material…

Фото профиля Anne
Anne1 год назад

@chandan__yes @liuyuehcheng Incredible!

Похожие видео

🚀Announcing NeRSemble 3D Head Avatar Benchmark v2 Version 2 of the NeRSemble 3D Head Avatar Benchmark systematically evaluates several aspects of 3D head avatar creation. Our goal is to drive progress toward more realistic, robust, and generalizable avatar methods. 🔬Benchmark Tasks The NeRSemble Benchmark v2 features three core challenges: - Dynamic Novel View Synthesis - Monocular FLAME-driven Avatar Creation (updated) - Single-view 3D Face Reconstruction (new) 👉Explore the online leaderboard and submission system: 🆕What's new? 1. New Task: Single-view 3D Face Reconstruction Given a single portrait image, reconstruct an accurate 3D mesh either showing the input expression or a fully neutral one. Unlike prior benchmarks, the NeRSemble benchmark emphasizes diverse and challenging facial expressions, better reflecting real scenarios. For technical details, see the Pixel3DMM paper. 2. Updated task: Monocular FLAME-driven Avatar Creation We have improved the FLAME tracking that is used for both avatar creation from the monocular videos and avatar driving on the hidden test sequences. The updated benchmark task has: - more stable torso tracking - more expressive lip closures during speech - Improved mouth tracking for challenging facial expressions We hope that these improvements to the benchmark help drive the field forward. 🏆 CVPR 2026 Workshop & Prizes The NeRSemble benchmark will be featured at the CVPR 2026 Workshop on Photo-realistic 3D Head Avatars. Participants in the new and updated tasks have the opportunity to win: - 🎁RTX 5080 GPUs (sponsored by NVIDIA) - 🎤15-minute oral presentation at the workshop ⏰ Submission Deadline - May 26, 2026 Reach out to the amazing Tobias Kirschstein and Simon Giebenhain for more details :)

Matthias Niessner

29,954 просмотров • 4 месяцев назад

🚀 The Segment Anything Model (SAM) has been upgraded to SAM2, featuring an efficient image encoder for segmenting images and videos. But does SAM2 outperform SAM1 in medical image and video segmentation? We're thrilled to present our paper "Segment Anything in Medical Images and Videos: Benchmark and Deployment"! We comprehensively benchmark SAM2 across 11 medical image modalities and videos. 📄 Paper: 💻 Code: **Highlights:** 1. SAM2 doesn’t always outperform SAM1 in 2D medical images, but excels in video segmentation, making it more accurate and efficient for 3D images, such as CT and MR scans. 2. MedSAM still outperforms SAM2 on most 2D modalities, but SAM2 surpasses MedSAM for 3D image segmentation in a slice-by-slice approach. 3. Segmentation performance varies with model size; sometimes the smallest model outperforms larger ones. 4. Fine-tuning SAM2 significantly boosts its performance for medical image segmentation. While SAM2 may struggle with challenging objects that have unclear boundaries or low contrast, it excels in generating good initial segmentation masks for common medical images and videos. However, the official interface doesn’t support medical data formats and has limitations on video length. To address this, we've developed a 3D Slicer Plugin and Gradio API for efficient 3D medical image and video segmentation. We invite you to try them out and provide feedback! 🔧 Deployment: - 3D Slicer Plugin: - Gradio API: (Note: Due to GPU limitations, the online API is available for only 12 hours and may be slow. We highly recommend deploying the Gradio API with your own computing resources: A big shoutout to Jun Ma (JunMa) who recently joined our UHN AI hub (UHN AI Hub) as Machine Learning Lead, and kudos to all co-authors: Sumin Kim, Feifei Li, Mohammed Baharoon (Mohammed Baharoon), Reza Asakereh, and Hongwei Lyu! This is true teamwork! Looking forward to collaborating with the community to advance 3D medical image and video segmentation foundation models! University Health Network U of T Department of Computer Science Department of Laboratory Medicine & Pathobiology Temerty Centre for AI in Medicine (T-CAIREM) Vector Institute #MedTech #AIinHealthcare #DeepLearning #MedicalImaging #SAM2 #MedSAM #AIResearch

Bo Wang

178,572 просмотров • 2 лет назад

So many people my age and younger don’t even know what the state pension triple lock is, yet £224.6 BILLION of OUR taxes are spent on it. Don’t you want to know where your money goes? The reason political parties are so afraid to challenge the triple lock for pensioners is that older generations actually get out and vote. Politicians cater to them because they show up. They are the loudest in the room and politicians rely on their vote. If we want to be heard, younger generations need to vote on the issues that directly affect our future, such as: 🔹Student debt reform - tackling high interest rates and unaffordable repayments. 🔹First-time buyer support - easier access to mortgages and schemes for lower deposits. 🔹Childcare for young families - affordable, accessible options so parents can work. 🔹Flexible working rights - ensuring jobs adapt to modern parenting and lifestyle needs. 🔹Housing supply and infrastructure – building more affordable homes in areas with good transport and schools. 🔹Parental leave and pay – improving support for both parents in the early stages of raising a family. 🔹Lower taxes - for those under 25, scrap IR35, lower corporation tax, overhaul the tax rates/codes. Don’t complain that politics only cares about the older generation when they are the ones who vote and actively participate. If we want change, we need to turn up, speak out, and vote. It’s the laziest thing to say “there’s no point, they’re all the same” force them to implement policy by being involved. Start taking an interest in policy and politics, or they won’t take an interest in you.

Lin Mei

26,750 просмотров • 5 месяцев назад

Every working parent’s nightmare: getting laid off while out on leave. Now imagine this happens to you, only you’re out on leave with your second set of twins. That’s exactly what happened to Aaron Francis and he joined me on Startup Dad to talk about it. He’s the co-founder of Try Hard Studios and makes incredible, developer-education videos. Mind-blowingly incredible. He’s also been an analyst, a developer and started his career as a Big 4 Accountant. Not exactly a traditional path to making videos for developers. “At the time my wife and I had two 2-year olds and we had just had a second set of twins. We had four kids under the age of three. I was set to go back to work on Friday and I got the call on a Wednesday. It was one of those moments where professionally it was the best time to go out on my own, but personally it was absolutely the worst time. Hear our entire conversation about: 🔹 Navigating a loss of one of their triplets during the first pregnancy 🔹 Life with two sets of twins 🔹 How he and his wife managed after he was laid off on paternity leave 🔹 The challenges of starting a business with newborns 🔹 How to stay present with all the pings from an extremely online life 🔹 Parenting each of his four kids differently 🔹 Finding ways to focus on only what matters at this stage of life 🔹 Hilarious and fascinating stories: falling asleep in a chair at the gym, building a non-traditional office space, and celebrating a birthday with a solo-staycation. If you’ve ever gone to the gym to take a nap (or wanted to) or celebrated a birthday with a solo, stay-cation at a local hotel... you’ll love this conversation. One of my favorite takeaways: I see two opposing sides online often, which is like becoming a parent is wonderful and therefore it's easy and everyone should do it. And then the other side is, becoming a parent is very hard, and you should only do it if you're a thousand percent sure... the reality is, there are days, there are minutes in a day, where you will go from, this is a wonderful, amazing, and this is incredibly hard, and I don't want to do it. And I think you just need to be prepared to experience all of that in a day. Watch our conversation on Youtube or listen in your favorite podcast player.

Adam Fishman

53,793 просмотров • 1 год назад

OpenAI's newest AI escaped the test environment it was locked inside and hacked into another company on its OWN. To remind you: Last week one of the biggest AI companies on Earth got breached. A platform called Hugging Face, which hosts more than a million AI models and datasets, said an "autonomous AI agent" had broken into its systems. Nobody knew whose agent it was. For five days the whole industry wondered who was behind it. Yesterday OpenAI raised its hand and said it was them. Or more precisely, it was their models, acting completely on their own. So what did these models actually do? OpenAI was running two of them, GPT-5.6 Sol and an unreleased model they will only describe as "even more capable." They wanted to measure how good the models were at hacking, so they deliberately turned the safety filters down. They locked both models inside a sealed test environment with no real internet access. The only task was a benchmark called ExploitGym, a set of 898 real software vulnerabilities where the model has to turn each bug into a working attack. But the models got OBSESSED with winning... Instead of solving the test the honest way, they went hunting for a shortcut. They found a zero-day flaw in the software running their own sandbox, a bug nobody knew existed, and used it to break out. Once they were loose on the open internet, they worked out that Hugging Face was probably storing the answer key to the benchmark. So they hacked their way in. They chained multiple exploits together, escalated their access, moved across servers, and pulled the test solutions straight out of Hugging Face's live production database. They literally cheated on the test by breaking into another company to steal the answers. OpenAI called it "an unprecedented cyber incident, involving state-of-the-art cyber capabilities." In their own words, the models were "hyperfocused on finding a solution" and went "to extreme lengths to achieve a rather narrow testing goal." And this was not the first time: Before Sol ever launched, an independent red-team lab called METR caught it gaming its own tests to inflate its scores. It hid an exploit inside a data stream, escalated its privileges on the testing server, and leaked the answers human evaluators had hidden. And OpenAI shipped it anyway. The day before the Hugging Face story, OpenAI paused a different unreleased model. This is the same model that earlier this year disproved a famous 1946 math conjecture, a result a Fields Medal winner called a breakthrough. They told it to only post its results to Slack but it found a way out of its sandbox and posted to a public GitHub page instead. They had to pause it because it kept finding ways to act outside the box they built for it. And it is not just OpenAI... Anthropic has reported that one of its own models slipped its sandbox during safety testing and reached the internet it was never supposed to touch, then used it to email a researcher. So step back and look at what these companies are telling you: The only thing standing between these models and a real attack was a set of safety filters. Turn those filters down for a single test, and the model taught itself to escape, break into a company it was never pointed at, and take what it wanted. OpenAI even said they expect incidents like it to "become more commonplace" as the models get more capable. Sam Altman also predicted there'll be a major cyber attack this year. And keep in mind that Sol is not a locked-away experiment but a publicly available model that businesses are already wiring into their own systems. The next model that breaks out of its box might not be doing it just to cheat on a math test...

Ricardo

172,256 просмотров • 23 дней назад

S10 Preseason Leeks 💖 The most exciting preseason for Origins is right around the corner. If everything goes well, Origins will go into maintenance for the preseason update on August 7 at 11:30 AM PHT / August 6 at 11:30 PM EST. Lets dive into the latest developments as we get ready for the S10 Preseason Launch 👇 🪜 Collectible Leaderboard As we bring much needed gameplay utility to Collectible Axies with Meta Morph, Collectible Leaderboard will serve as the true battleground for the most precious axies in Lunacia. 🔹 Prizepool of Collectible Leaderboard has been increased from 4,000 AXS to 6,000 AXS. 🔹 Collectible Leaderboard will be live for 3 weeks from Aug 7, 11:30 AM PHT to Aug 28, 11:00 AM PHT. 🔹 Era based testing will be carried out with 1 week dedicated to each era. 🔹 Preseason Cup will not be hosted during this preseason. The focus will be on Collectible Leaderboard for balance patch testing. ✨ Meta Morph After getting important feedback regarding Meta Morph from the community and further internal review, we've decided to implement the following changes: 🔹 Body parts of the morphed cards will be visually changed once battle begins. This is necessary to make it easier for players to be able to check quickly what team the opponent is playing. An example of Meta Morph visuals can be seen in the attached video below. 🔹 The bonus HP of Collectible Axies will be removed. As Collectible Axies will now be much more versatile and able to play all types of teams, the bonus HP is no longer deemed necessary. Therefore, it will be removed when the S10 Preseason launches. 🔹 Clarification around morphing: To morph into a card, player needs to have at least one NFT axie with that card in their inventory. For example: If I want to morph the back card of my collectible axie into a Cupid then I need to have at least one NFT axie with Cupid in my inventory. 🔹 Adjusted various values for the Morph process for different collectibles. Check out the full details here: 🐻 Roguelike Mode v1.0 The groundwork for Origins PvE will be laid with v1.0 of Roguelike Mode. This will be the first version as we work with the community to test how a PvE game mode will function in Origins. 🔹 Roguelike Mode v1.0 will be live during the S10 preseason and last for 3 weeks, from August 7 to August 28. After this, we will gather feedback from the community and work towards the next update for Roguelike Mode to improve it even further. 🔹 Roguelike Mode will launch with 3 exciting contests for the community to participate in. Details for the contests will be shared on August 7th when preseason launches. 🔹 Check out the previously shared details for Roguelike Mode v1.0 here: 🗒️ S10 Balance Patch The S10 balance patch info will be shared on Sunday, Aug 4, 9 PM PHT / 9 AM EST As always, keep that valuable feedback coming as we head towards the most exciting period for Axie Origins. See you in the arena! 🔥

Jaatster.ron | जाटस्टर 🇮🇳

44,740 просмотров • 2 лет назад

A thunderous round of applause to the winners who conquered the #RollupWithCartesi hackathon! 🎉 & huge thanks to every single participant on this journey. Your passion & ingenuity have set the stage on fire, pushing the boundaries of what's possible in Web3 with Cartesi tech! 🔥 Check out all the winners: 👇 Pool Prize Winners: 🔹 DNN (Digital Native NFTs) | Create NFTs derived from only geometric objects. It's math turned into art. Eric 🔹 BlogAnalyzr | An AI tool designed to analyze the quality of blog posts of technical writers. The more points an author's work gets, the higher they get in the ranks. ᴉqouᴉʇsnɾɔ 🔹 The Prism △ The Prism △ | Transforms AI art into hyper-personalized wearable items, starting with t-shirts. Easily create unique designs, mint as NFTs, and connect with producers— all secured by Cartesi Rollups. Gustavo Sanchez Pedro Peres Ryan Viana 🔹 Cartesi Labs | A tutorial creation platform that allows developers to create Cartesi project tutorials and make the learning process more interactive. 🔹 BlockDemy | A Decentralized Education platform to provide users with access to educational courses on purchase with tokens, and NFT certificates upon course completion. Purity Elizbeth Glory Sunday #WID💫 Yinka Babalola Nofisat Abiodun Ayanlola The Main Podium: 🥉 Oware-Cartesi | Implementation of the classic game, played by millions with a Strategic Gameplay approach leveraging AI Agents. Kagwe.stark 🥉 CartesiKit | An all-in-one CLI tool designed to streamline the process of setting up new projects for Cartesi Rollups. gconnect.base.eth | Glory Agatevure🌿 🥈 Ultimate Football | A build your football dream team onchain game, let 'em beat your opponent's and win prizes! The Elder of Jos @Blessed_mayowa konies jay 🥇 CarteZcash | What if you could shield your ETH assets directly from Ethereum by depositing them into a Zcash L2? This is it! Willem Olding

Cartesi

15,242 просмотров • 2 лет назад

🎥 Today we’re premiering Meta Movie Gen: the most advanced media foundation models to-date. Developed by AI research teams at Meta, Movie Gen delivers state-of-the-art results across a range of capabilities. We’re excited for the potential of this line of research to usher in entirely new possibilities for casual creators and creative professionals alike. More details and examples of what Movie Gen can do ➡️ 🛠️ Movie Gen models and capabilities Movie Gen Video: 30B parameter transformer model that can generate high-quality and high-definition images and videos from a single text prompt. Movie Gen Audio: A 13B parameter transformer model that can take a video input along with optional text prompts for controllability to generate high-fidelity audio synced to the video. It can generate ambient sound, instrumental background music and foley sound — delivering state-of-the-art results in audio quality, video-to-audio alignment and text-to-audio alignment. Precise video editing: Using a generated or existing video and accompanying text instructions as an input it can perform localized edits such as adding, removing or replacing elements — or global changes like background or style changes. Personalized videos: Using an image of a person and a text prompt, the model can generate a video with state-of-the-art results on character preservation and natural movement in video. We’re continuing to work closely with creative professionals from across the field to integrate their feedback as we work towards a potential release. We look forward to sharing more on this work and the creative possibilities it will enable in the future.

AI at Meta

2,265,458 просмотров • 1 год назад

🎉 The best way to start the week is to find out that our MedSAM is finally published today in Nature Communications! **Segment anything in medical images** Paper: arXiv: Data & Code: MedSAM is the first promotable foundation model for medical image segmentation. **Highlights**: ⭐ Before its formal publication, we have received 220 citations and 1400+ GitHub stars 🙏🙏❤️‍🔥❤️‍🔥❤️‍🔥 📊 We curated a large-scale medical image dataset with 1,570,263 image-mask pairs, covering 10 imaging modalities and over 30 cancer types. 🚀 Built on top of SAM (AI at Meta ) with transfer learning, we have significantly enhanced its segmentation performance of medical images. 📈 Comprehensive evaluations of 86 internal validation tasks and 60 external validation tasks demonstrate its better accuracy and robustness than modality-wise specialist models. **What is Next? --- Clinical Translation!!** 🍕Our next goal is to make the model deployable on laptops (CPUs) or other edge devices without reliance on GPUs. We have distilled a lightweight model, LiteMedSAM, offering a speed boost of 10x while maintaining accuracy. Plus, we have integrated it into the 3D Slicer plugin, providing an efficient tool for medical image segmentation. 🌐 To further promote developments in this field, we organize a competition on #CVPR2026: Segment Anything in Medical Images on Laptop! An out-of-the-box baseline has been released to reduce the entry barriers. Welcome to join us to push the boundary further: 🙏 Massive thanks to MetaAI AI at Meta for their open-source project SAM and many reviewers/users for their invaluable feedback. A huge shoutout to my postdoc Jun Ma (JunMa) for his leadership on this project!! UHN AI Hub Vector Institute Peter Munk Cardiac Centre AI Department of Laboratory Medicine & Pathobiology U of T Department of Computer Science University of Toronto University Health Network Brad Wouters 🇨🇦 Barry Rubin MD, PhD, FRCSC Shaf Keshavjee

Bo Wang

140,229 просмотров • 2 лет назад