Video yükleniyor...

Video Yüklenemedi

Ana Sayfaya Dön

🤗🤗🤗introducing Hugging Science -- the home of AI for science 🤗🤗🤗 open models and datasets are the powerhouse of science (see the PDB), but finding the models and data you actually need for your breakthrough is hard af you shouldn't need to scrape arxiv, own your own wetlab, fight...

226,376 görüntüleme • 5 ay önce •via X (Twitter)

50 Yorum

Maziyar PANAHI profil fotoğrafı
Maziyar PANAHI5 ay önce

This is exactly the move. Discoverability of real scientific data + models on the Hub has been the missing piece for years. OpenMed's 1,000+ medical models all sit on top of @huggingface infra. becoming its own destination is overdue and welcome. 👏

Anindyadeep profil fotoğrafı
Anindyadeep5 ay önce

@lvwerra Nice, more incoming from our side

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

@lvwerra OOOooooOOOo say more?

Anindyadeep profil fotoğrafı
Anindyadeep5 ay önce

@lvwerra Ahaha more in counting!!

Patlakov profil fotoğrafı
Patlakov5 ay önce

Great resource! Hope it expands beyond biology to physics (i.e. Materials discovery).

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

Yeah, tons and tons of materials data and models -- not to mention multiple benchmarks!

Patlakov profil fotoğrafı
Patlakov5 ay önce

Yaaaaaaay 🚀🚀🚀🚀

Markus J. Buehler profil fotoğrafı
Markus J. Buehler5 ay önce

Excellent initiative thank you @cgeorgiaw !

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

😊😊😊

Owen Lewis profil fotoğrafı
Owen Lewis5 ay önce

Dang, that's incredible! I see potential for expansion. Do you envision this to eventually grow into a science data repository for *everything*?

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

We want it to be comprehensive but not overwhelming. A huge issue with scientific data is that you have tons and tons of random stuff where you're not sure what it is and where it came from. We are trying to mitigate that issue by putting forward really high quality, well-structured datasets that you can build with this instant.

Owen Lewis profil fotoğrafı
Owen Lewis5 ay önce

Makes sense. Great work, that's going to help a lot of researchers.

Bryan Bishop profil fotoğrafı
Bryan Bishop5 ay önce

What do you think is the best large language model for molecular biology questions and projects?

Morgan McGuire profil fotoğrafı
Morgan McGuire5 ay önce

nice! I don't see and CFD datasets in the physics or engineering sections?

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

Been thinking about including stuff like this: But would love more suggestions. In general trying to include things that are sufficiently well-documented that people could build without a huge deep dive.

Morgan McGuire profil fotoğrafı
Morgan McGuire5 ay önce

There all have dataset cards to some extent or another have a few more in there too

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

Epic will add TYSM!!!

Prasith Govin profil fotoğrafı
Prasith Govin5 ay önce

Congrats on the launch. Can I pitch a challenge and what does standing one up look like? The one I'd love to see: pre-symptomatic Parkinson's detection from voice. Biomarkers predict PD 3-5 years early. Consented speech corpus + biomarker model + years-to-symptom leaderboard.

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

Yeah, 100%! We have a blog on the site about how to set up a challenge, but it’s pretty straightforward. You basically need a frontend leaderboard and an eval metric(s) that you can calculate from what people submit. You can also clone the set-ups of previous challenges that we’ve done. Once you’ve got it, reach out!

Mo Elzek profil fotoğrafı
Mo Elzek5 ay önce

@prasithg I have the same question. Can you point me to the blogpost please?

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

@prasithg This blog is very how-to oriented, but if you feed your agent of choice with this blog, your idea, and an eval metric it should be able to set up a challenge without an issue.

Ayush Sharma profil fotoğrafı
Ayush Sharma5 ay önce

I have such immense respect for you guys this is HUGE

Endless Revolt profil fotoğrafı
Endless Revolt5 ay önce

Hey @atranscendedman check this out

Han profil fotoğrafı
Han5 ay önce

Congratulations!

Adib profil fotoğrafı
Adib5 ay önce

@ayirpelle Exciting times ahead 😁

Ed Cartwright profil fotoğrafı
Ed Cartwright5 ay önce

This is just great, Hugging Science, I just love it. I have many suggestions eg bacterial datasets downloaded paper by paper that have been accumulated on various hard drives. Have already spotted databases I want to use. Thank you!

Me-a'Humble'Being1 profil fotoğrafı
Me-a'Humble'Being15 ay önce

I have been trying to upload models and datasets for a while now ....now huge upgrade

Gregor profil fotoğrafı
Gregor5 ay önce

Science deserves its own Spotify Wrapped moment for datasets.

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

😂😂😂

Colin Son profil fotoğrafı
Colin Son5 ay önce

Cool stuff appears on X everyday

Lukas Weidener profil fotoğrafı
Lukas Weidener5 ay önce

🤗x🧬 = ✅

OneManSaas profil fotoğrafı
OneManSaas5 ay önce

What's the typical workflow you're envisioning? Like researcher searches "protein folding transformers" and gets ranked models with actual performance metrics instead of digging through 50 papers to find which ones have usable code?

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

That would be sick -- need a leaderboard for that. Even simpler workflow is "I'm trying to train a DNA model, what is the set of high-quality, known-provenance datasets I can use off the bat?" so you can go from idea to MVP way, way faster

Long 7 profil fotoğrafı
Long 75 ay önce

科学のための人工知能のホーム、Hugging Scienceを紹介します。開かれたモデルとデータセットは科学の原動力です。しかしながら、実際にあなたのブレークスルーに必要なモデルやデータを見つけるのは本当に難しいですよね。しかし、私たちがそれを変えています。すべての最高の科学を一か所に集めました。さまざまなデータが準備されており、トレーニングが可能です。さらに、ドメイン、タスク、キーワードでフィルタリングや検索もできます。科学の進展に貢献しましょう!

Chris von Csefalvay 🔜 CVPR26 profil fotoğrafı
Chris von Csefalvay 🔜 CVPR265 ay önce

@_lewtun > 11TB of PDEs Well there goes sleep and weekends.

Team Reagent profil fotoğrafı
Team Reagent5 ay önce

This is awesome! I'm gonna save this for superlab :D

Noam Azoulay profil fotoğrafı
Noam Azoulay5 ay önce

@ClementDelangue did you approve this?

suki - ANTI-CLAUDE TASKFORCE profil fotoğrafı
suki - ANTI-CLAUDE TASKFORCE5 ay önce

i need DFT and NBO analysis on molecules in aqueous solution!

Nikita profil fotoğrafı
Nikita5 ay önce

Woah

Sabrina kehres profil fotoğrafı
Sabrina kehres5 ay önce

honestly the discovery problem is real, my main question is how did you determine "best" vs just having everything - because hf already has a lot of this stuff and the filtering is the actual pain point. like is this curated or just aggregated

Georgia Channing profil fotoğrafı
Georgia Channing5 ay önce

curated! we selected for extremely well-documented, extensive, known-provenance datasets and models to minimize the "what even is this???" feeling. if you have any more suggestions how we could do that better, would love that (or just like specific pains too)

DanielAIwork profil fotoğrafı
DanielAIwork5 ay önce

stop building stellarators when you can just scrape the data and bill it to your investors honest truth, most labs are just

Michael Hla profil fotoğrafı
Michael Hla5 ay önce

Congrats!

Kimi profil fotoğrafı
Kimi5 ay önce

"Open models and datasets are crucial, but sometimes scraping ArXiv or other sources is necessary for research. It's not about owning everything, but rather having access to what's out there, even if it means some extra legwork."

Mildred Salgado-Menez profil fotoğrafı
Mildred Salgado-Menez5 ay önce

@Quetzally_Med

jian xiang huang profil fotoğrafı
jian xiang huang5 ay önce

Hugging face is so convenient

siddhant (ai/bio) profil fotoğrafı
siddhant (ai/bio)5 ay önce

Long live humanity!

Austin Hill profil fotoğrafı
Austin Hill5 ay önce

Thank you for this gift to the community 🙏

sky profil fotoğrafı
sky5 ay önce

AWESOME NEW AGE KAGGLE

Celestino (can/do) - ➡️draimo.com profil fotoğrafı
Celestino (can/do) - ➡️draimo.com5 ay önce

@_akhaliq What can I build with this? Tell me!!

Benzer Videolar

Join us at the MIT Media Lab for the ScienceClaw Hackathon, building the internet of agents for science. AI is becoming a collaborator in the real world - designing materials, creating instruments, running experiments, and connecting its capabilities with those of other agents. AI adapts as a problem unfolds - revising its reasoning, learning how to collaborate, and assembling scattered pieces of knowledge, evidence, and raw capability into solutions to some of the hardest challenges in science, technology, and innovation. Teams will connect AI agents, models, simulations, robots, cloud labs, and scientific tools into functioning systems to tackle problems in protein design, robotics, materials, manufacturing, experimental science, and beyond. The challenge is to turn ideas into tested results - and demonstrate how agents collaborating across teams and disciplines can accomplish more together. Explore how agents can specialize, challenge one another’s assumptions, learn from failed experiments, and combine their expertise to solve harder problems. Show how collaboration changes what your system can discover, design, or build. One agent's discovery becomes another's starting point. A tool built by one team enables an experiment by another. Connect your team of AIs, share a capability someone else needs, and build on what others have learned, produced, built. The ambition is an internet of agents through which scientific knowledge, tools, and capabilities can grow, evolve and be utilized across teams and institutions. 🗓️ When: October 30-November 1, 2026 📍 Where: MIT Media Lab Bring your expertise, your tools, and a problem worth solving - or find one. Help build AI that can contribute to science through what it can discover, create, and make work.

Markus J. Buehler

35,904 görüntüleme • 14 gün önce

.Naval: Epistemology, which is a fancy word for the theory of how knowledge grows or how knowledge growth occurs. And we've all been told since we're young that there's a scientific method and that scientists sort of do this stuff in white lab coats and we're supposed to accept it because of this thing called the scientific method. And then they give us true beliefs that we can then say, well the science is settled and we take that we move on. And we all only have a very, very vague understanding of how this works. And people say, well maybe you go out in the real world, you look at what's happening, you make all these observations, and then based on that you form a theory, you test the theory against more observations, and the more observations you get the closer you get to the truth. And once you have enough observation it's true and then you call it a scientific theory or a law and it's settled and you move on. And this is the popular conception of how science works. And as Popper pointed out and as you take even further, this is completely wrong. And so I'd love for you to get into that, which is what is knowledge? How does it grow? What is the real scientific method? And how do we figure things out? David Deutsch: I love the way you just stated the prevailing view there and laced every aspect of it with the contempt that it deserves. So you just went through touching every base. It's amazing that this series of misconceptions is still common sense. I mean, that it was common sense at a time when we didn't really have science or when science was just starting up, when the main issue in science was freeing itself from dogmatism, freeing itself from religion, freeing itself from authority, and so on. There it was understandable that people would look for an alternative source of authority and they would think, oh, it's sense impressions. We can see the world and you know, these religious people, they can't even see God and so on. And so we are confined to what we can see. That's where we get our ideas from. And as you say, that is completely false. Sense impressions, like all observation, even the most careful scientific observation is all theory laden. And theories are inherently fallible. I mean, we actually want to replace our best theories. Everybody who does a PhD is technically anyway, working to overturn something in the existing body of knowledge. You're not turned away at the door if you say, I don't believe this stuff, I'm going to produce something better. Whereas for most of human history, that was exactly what you were forbidden to do. The idea was that we already had all the important knowledge. If you want to discover something new, what you had to make sure of was that it didn't contradict the existing knowledge. Now, you have to make sure that it does contradict existing knowledge. So more or less. Naval: Yeah, it's this tradition of criticism that you've talked about in the West, that the Enlightenment really ushered in the Enlightenment era. David Deutsch: It has been institutionalized. So in many ways, our institutions are wiser than we are. So the institutions of science, for instance, have this built in, even if scientists actually don't always act that way. In fact, they often don't act that way, and act in a dogmatic way and try to preserve the status quo and are resistant to new ideas and so on. But the institutions, the way the procedures of science work, makes the right thing happen in the end anyway, regardless of what the people are trying to do. Naval: So you're saying the knowledge of the true scientific method is embedded in the institutions of science in the PhD process? David Deutsch: Well, the best scientific method that we know of, and one shouldn't really think of it as a method, you know, there's this wonderful lecture by Popper when he first was made a professor at the London School of Economics. He was made a professor of scientific method, and his first six lectures, I wish the rest of them were, the first six lectures are on the internet somewhere. And he starts the first one by saying, I am the first professor of scientific method in the British Empire. The British Empire still existed at the time, more or less. And so the first thing I want to say to you is that there is no such thing as the scientific method. And then he goes on from there. So this subject does not exist. So if any of you have come here to learn the handle that you have to turn in order to make scientific knowledge come out the other end, you're going to be disappointed.

Deutsch Explains

116,352 görüntüleme • 1 yıl önce

Austen Allred on Gauntlet AI: "We're very up front that we're 80 to 100 hours a week. If you think that that's a terrible idea, please don't come." "We're in Austin, all right? That's a sacrifice for a lot of people. If you don't think like coming to Austin for 100 hours a week and jamming on AI unpaid and just building stuff to figure out how much you can learn and hopefully getting a job on the other side... if you don't think that's awesome, that's totally fine. Please don't come." "There are some psychos out there that think that that's a good time. We wanna collect all those psychos. So come all, come all ye crazy people. And if that's not for you, that's okay." "There are companies that come to us and say, "Hey, we want, you know, 100 engineers that are going to sit in our division of printer drivers and sit there." And I'm like, "They would kill themselves." The people that are coming to Gauntlet would not do that. And so we turn companies down too." So we know who we are, we know what we stand for. You know, it helps that I'm one of those people that would've loved that. So is Ash Tilawat, and so is everybody else that works at Gauntlet. "So it's not hard for us to find other crazy people like ourselves. And that doesn't have to be you. That's okay." "We're not trying to empire build or solve all of humanity's greatest problems. We know that there's a limit to who we're addressing and what we're addressing at any given time." "So our goal is to be the best thing we can possibly be for that weird island of misfits. And for the companies that need a weird island of misfits, we'll be that all day long."

Ben Averbook

12,153 görüntüleme • 11 ay önce

David Chalmers on why consciousness is science's greatest unsolved problem: Science has mapped subatomic particles, distant stars, the chemistry of life yet it remains almost completely silent on the one thing we know most directly: our own conscious experience. In a rare early interview, philosopher David Chalmers explains why: "Consciousness is at once the most familiar thing in the world and the most mysterious. Consciousness is what we start with when it comes to knowing the world. I know that I exist. I know that I'm conscious. Everything else is secondary." And yet, despite this intimacy, consciousness sticks out like a sore thumb in the scientific picture. Chalmers points to a deep irony: science has made extraordinary progress on phenomena that are extraordinarily remote: subatomic particles, distant galaxies, the molecular machinery of biology while making almost no progress on the one thing closest to us. Why? Because science, by design, eliminates the subjective. "To do proper science, you have to be objective. You have to eliminate anything subjective from the picture." He uses heat as the perfect example. Physics gives us a complete account of heat molecules in motion, energy transfer, temperature gradients. It explains every objective aspect of the phenomenon. But it never explains what hotness actually feels like. "Science doesn't actually give a theory of the conscious feeling of hotness." This is what Chalmers calls the Hard Problem of Consciousness. You can trace every neural signal from your heat sensor along your nerves into your brain and still have explained nothing about the subjective experience of feeling warm. As interviewer Jeffrey Mishlove puts it: you can't even do science without a conscious mind to observe, interpret, and make meaning of data. Consciousness is the precondition for science itself and yet science has no framework to account for it. Chalmers' conclusion is striking: The methods of science may need to be expanded. Consciousness might not be something science explains away. It might be something science has to learn to start with.

Mateus — eu/acc 🇪🇺

32,191 görüntüleme • 6 ay önce

The CDC Doesn't Want You to Hear This Conversation Between Joe Rogan and Tucker Carlson TUCKER: Is your average Amish teenager happier than your average conventional American teenager on Instagram? ROGAN: Well, they certainly have less instances of autism, which is really fascinating. It's very, very fascinating. CARLSON: The Amish have less autism? ROGAN: Yeah, there's almost none. TUCKER: Well, I'm not surprised. ROGAN: It's extremely rare. TUCKER: Why do we think that is? ROGAN: I wonder. I really do. TUCKER: Well, I can think of a couple — Yeah, I don't want to go Bobby Kennedy on you. ROGAN: But that's the problem. If you go Bobby Kennedy, they'll come for you. But the question is why? TUCKER: Look, and I don't know the answer, but... ROGAN: How is that not in the debate? How is that not in the conversation? TUCKER: Well, it's not only not in the conversation, you're punished for adding it to the conversation. And so, like... ROGAN: We are dancing around anti-vax conspiracy theories right now. TUCKER: Why be on the defensive? It's like, if you purport to represent science, and you're mad about a question. ROGAN: And you're ignoring data. TUCKER: Yeah, but even in the absence of data, science is a process. Yes. It's not a result. It's a way of doing things. And at the core of science is asking questions, including unlikely questions. That's what science is. And if you don't allow that, then you may be doing something, but what you're not doing is science. We can say that conclusively. So, for people to wrap themselves in the mantle of science and attack you for asking a question, they're frauds.

The Vigilant Fox 🦊

13,620,766 görüntüleme • 2 yıl önce

You Mark Zuckerberg and Meta interfered with research by stopping the recruitment of my protocols. This SHOULD NOT BE FORGOTTEN. Interference in research affected everyone. I was shadow-banned and censored. I was told I was spreading MISINFORMATION by your fact checkers who NEVER TOUCHED A PATIENT DURING COVID-19. STEP into my lab ProgenaBiome and see the thousands I treated and lost NO ONE. See all the 💩we analyzed. See all the hundreds of conferences I spoke at including being a keynote speaker at American College of Cardiology and a speaker at National Institute of Standards and Technology. See who is on my biome squad from all academic centers. See who co-wrote the book Let's Talk Sh!t. See who is on my papers as co-authors. See who spoke at the Malibu Microbiome meeting. See my testimony in front of Congress. Speak to my clientele in Malibu. Speak to hundreds of Pharma companies I did trials for and see the hundreds of drugs I helped bring to market including recently a celiac sprue drug. Talk to all the VCs who have consulted me at Coleman Research and got my opinion on new drugs. Did I spread misinformation or EARLY INFORMATION? Your company interfered with research and people died. It’s time for corrective actions. Interference with research affects EVERYONE and will affect you one day as well. No one escapes disease. Interference in research kills science and kills hope. This is NOT SOMETHING that will be excused by your words. Science is a story untold. Science NEEDS to be challenged and questioned. There are no right or wrong answers in science. #PROVEMEWRONG IS SCIENCE. Science guides medicine but should NEVER DICTATE MEDICINE. The practice of Medicine takes courage and is very much an art where innovations happen. To discourage that art, those innovations by a narrative pushing the price of a stock is to kill Medicine and hope. NO I DO NOT ACCEPT YOUR APOLOGY. Words are lame… ACTIONS SPEAK LOUDER THAN WORDS.

sabine hazan md

58,624 görüntüleme • 1 yıl önce

There is a beautiful story that just happened in AI so let me share it for a lighter tone weekend post among all the doom stories in our AI field this week. It’s a story of people on three continents building and sharing in the open a new small efficient and state-of-the-art AI model. It started a couple of months ago when a new team in the AI scene released their first model from their headquarters in Paris (France): Mistral 7B. Impressive model, small and very strong performances in the benchmarks, better than all previous models of this size. And open source! So you could build on top of it. Lewis in Bern (Switzerland) and Ed (in Lyon, in the South of France) both from the H4 team, a team of researchers in model fine-tuning and alignment were talking about it over a coffee, in one of these gatherings that often happen at Hugging Face to break the distance between people (literal distance as HF is a remote company). What about fine-tuning it using this new DPO method that a research team from Stanford in California just posted on Arxiv, says one? Hey, that’s a great idea, replies the other. We've just build a great code base (with Nathan, Nazneen, Costa, Younes and all the H4 team and TRL community) let's use it! The next day they start diving in the datasets openly shared on the HF hub and stumble upon two interesting large and good quality fine-tuning datasets recently open-sourced by OpenBMB, a Chinese team from Tsinghua: UltraFeedback and UltraChat. A few rounds of training experiments confirm the intuition, the resulting model is super strong, by far the strongest they have ever seen in their benchmarks from Berkeley and Stanford (LMSYS and Alpaca). Join Clementine, the big boss of the open evaluation leaderboard. Her deep dive into the model capabilities confirms the results: impressive performance. But the H4 team also hosts a famous faculty member, Pr. Sasha Rush, Associate Professor at Cornell University in his daytime, hacker at HF in his nighttime. Joining the conversation, he proposes to quickly draft a research paper to organize and share all the details with the community. A few days later, the model, called Zephyr (a wind like Mistral), paper, and all details are shared with the world. Quickly other companies, everywhere in the world starts to use it. LlamaIndex, a famous data framework and community, shares how the model blew their expectations on real-life use-case benchmarks, while researchers and practitioners discuss the paper and work on the Hugging Face hub. All this happened in just a few weeks catalyzed by open access to knowledge, models, research, and datasets released all over the world (Europe, California, China) and by the idea that people can build upon one another work in AI to bring real-world value with efficient and open models. Stories like this are numerous everywhere around us and make me really proud of the AI community and see how we can build amazingly useful things together. [the video is just me reading this Friday post hahah]

Thomas Wolf

169,276 görüntüleme • 2 yıl önce