Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

this is the new benchmark for models. 10 iterations on one generation is where Flux 3 Image starts to add a bunch of noise. I didn't understand at first why the image on the left was drifting so much, whereas on the right it wasn't. After doing a bit...

23,165 Aufrufe • vor 3 Tagen •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

✨ Grok Imagine Video is now live on Photo AI It's hard to explain how impressive this is because of the speed that xAI got itself from literally nothing to the top of the leaderboards Six months ago Grok's video model was a joke, it wasn't even close to any of the video models out there, it looked cartoony and wasn't there and nobody took it seriously Now it's here and it's instantly the #1 video model out there now, it shot above Kling (which I used before on Photo AI and usually my favorite) and above Runway Gen 4.5 which was just launched 6 days ago! Mmore importantly it's now above xAI's biggest competitors' models: Google's Veo 3 and OpenAI's Sora 2 Being the best video model doesn't mean it's flawless: video is incredibly hard and actually because it looks so realistic now when it does make a mistakes it's even funnier One thing I noticed is that it still has a hard time with is voice, it does it well for a majority of the video but then slips up and produces unintelligible blabbering (which is really funny to hear) in both English (video 1: "it's where I find my naim", what's a "naim"?), and tested it in Portuguese too (video 3 at the end is unintelligible Portuguese I believe) In many ways Grok Imagine Video also reminds me of Sora, it has that weird but funny Sora conversation style But guys it's REALLY really really really close to getting perfect, we're so close to having full video productions being to be able done in AI, actually you already can if you just cut out the bad parts already Very exciting and I'm grateful I can experience this

@levelsio

195,023 Aufrufe • vor 8 Monaten

The Sabotaging Practice of Over Supply and Sameness in the NFT Space. The current zeitgeist of the NFT space is that the same artists are doing the same kind of work five times a year, with project after project leaving a trail of disappointment and discontent among collectors and all of us watching in disbelief as huge resources are extracted from the space over work that feels like it could be left as an "artist study." I understand that you can do what you want with your money as collectors, but we are killing the whole space with this incestuous practice. No artist is that prolific to be able to do 5 collections of 100+ pieces each every year and actually deliver innovation and some kind of creative evolution. Of course, they can pretend play that the work has something new, but there is no precedent nor proof that that has ever happened in the speed that it happens in the NFT space. Again, people are free to through away their resources on whatever they want but with this way of doing things, we more and more are going to start seeing the consequences. Oh! There are consequences? Yes. Maybe unintended, but there are. Let's see. Let's start with the loss of belief in the NFT space as somewhere where emerging artists can come and find support for their experiments. Why even bother to bring experiments, innovation, and new ways to think of art on the blockchain if the same people have all the collectors hypnotized with their magical flutes? Why even try to come to a space where taking risks and challenging the status quo (the mission of art!!!) is overlooked? This makes the NFT space a social club and not a space for art. I guess it is fine, but IMO it is a recipe for disaster. New collectors stay away because the art will slowly but surely become stale and un-challenging. Why even bother to come and see what is happening here if you can't, as a collector, see new weird and up-and-coming artists? The amount of noise emitted by the same artists doing the same art over and over, drowns out any new voices. Again. A recipe for disaster. The NFT space is becoming a space of disappointment and doubt. We think that collections going to zero one after the other, over and over, is not damaging? I feel we are kidding ourselves. Disappointment piles up, and again, the people who will hurt are the emerging artists, the new blood, the ones who are willing to risk the most and, in return, put fire in this cold space of sameness. I love this space—don't get me wrong—it has changed my life, and I believe it has a ton of potential, but things need to change for it to become a beacon of light in art. But we need to support new voices. We need to support new ideas. The challenge is huge. I hope to contribute all I can to this change. I hope more and more see how exciting it is to go out and try to discover what else is out there and move this space forward. But again, I understand the leaps of faith needed, but if there is a space that is based on that, it's the NFT space...so there is hope. We will see. 📺by Boldtron

alejandro cartagena

98,261 Aufrufe • vor 2 Jahren

It's been several months since I curbed my phone addiction and it is so profoundly noticeable that my happiness has dramatically increased. I completely dropped doom scrolling, the news, politics, brain rot, and limit social media to mine and my friends' posts only. All the general negativity on the internet is like living in a home where everyone argues and is upset, even if it doesn't involve you and you don't engage with it, just being exposed to it every day is mentally draining. It's subtle, it creeps up slowly, and you don't even realize how mentally taxing it truly is. Now imagine the feeling of living in a home where everyone is peaceful and laughs, the feeling is equally infectious. You can even test this out by watching Friends or The Big Bang Theory with and without the laugh track on YouTube; spoiler, it's painfully cringe without it. The effect surrounding vibes have on your emotions is so strong, and it's hard to notice most of the time. I now find myself having uncontrollable fits of laughter at least once a week, which used to be closer to once a year. It's important to note that the things triggering these fits of laughing were always there, the difference is that my perception has changed. It's figuratively like a laugh track is playing in the background and all of a sudden, everything is now funny. I'm not trying to preach and tell people how to live their lives, but it's something I felt compelled to share since it's helped me so much. The habit was hard to break because the algorithm is designed to adapt to every person and lock you in, but I did it by substituting the phone with other things I enjoy. Right now it's learning the guitar and swimming. I'm developing skills I've never had, and it only compounds the feeling of happiness.

Macie Jay

74,257 Aufrufe • vor 1 Jahr

After a few more hours, I think I've figured out Opus 5. Opus 5 is trained to be more agentic than anything I've used. All Claude 5 models are like that. So what changes? The way to interact with Opus 5 or contextualize it won't work the same way as with other models. It loves exploring, so it doesn't need much guidance for it. Unique preferences, artifacts, and references compliment it well and enable cleaner and more effective exploration and execution. Now that it can explore more effectively on its own and understand intent better, the best thing to do is to get out of its way (e.g., it doesn't need examples of your preferences; a clear high-level description of it works best). It's truly agentic in that sense. A good first step to provide better context for Opus 5 is to distinguish between what's situational and what needs persistence. Regardless, persistent system prompts and CLAUDE.MD needs to stay lightweight. Remove memories and tool descriptions from these. CLAUDE.MD is also a great place to tap into progressive disclosure by linking command/skills to it. On the situational side, agent skills and auto-memory can leverage progressive disclosure and the improved ability of the model to use its external context/knowledge. Conflicting and unnecessary instructions, which are common at this layer (mainly to ensure reliability), are going to throw off this model easily. That's the biggest change I had to make. Simple, clean, and clear prompts and skills work best. I had to clean a lot of my skills and system prompts. The way I prompt remains the same (usually clear and well-scoped). MCP tool descriptions are also more descriptive and have been deduped from the system prompt. Anthropic released a guide on the new rules for context engineering, which was helpful here. I started to test the recommendations and created a little artifact with the things that worked along the way. This might feel like a lot of work. Believe me, it has been frustrating. But I think we can expect future frontier models to become more agentic and smarter at figuring out the right context/gaps. The best thing to do is to prepare for that now. Boris Cherny mentioned that Opus 5 is their least prompt-injectable model yet. I am not sure if that was something they intentionally trained for or if it emerged based on how it was trained, which is to be extremely agentic in nature and more direct in execution.

elvis

37,824 Aufrufe • vor 2 Monaten

This one was made with Seedance 2.0 Fast via Dreamina. This is pure Omni-Reference. The only character sheet I used was for these girls, Sari and Ploy. The dude with sarung here and the location were 100% prompted. I didn’t use a character sheet or reference for either of them. Even in Fast mode, Seedance 2.0 is bloody good and it still nails the hyper-vernacular vibe that I always aim for in my work. Seedance 2.0 is both exciting and scary for me 😆 It’s exciting because it is undoubtedly the best model currently available on the market. Trust me, you’ve seen the videos I’ve made so far right? The performance of the model It’s simply the best, period. It has helped me tremendously in creating a shit ton of stories about the region where I live, Southeast Asia. It has been the most exciting thing ever. The scary part is whenever a platform or company comes to me saying, “Hey, we have this new video model. Blah blah blah. We’ll let you know more soon.” It scares the shit out of me because the big question is whether it will be better than Seedance 2.0??? 😆😆 If not, I don’t even want to bother using it. I’ve come this far and achieved this level of quality with Seedance 2.0. That’s why I skipped Happy Horse, which I already tested. It’s also why I’m not bothering with Wan or anything else for now. Their current models are still far inferior to what we already get with Seedance 2.0. I don’t want to downgrade the visual quality. This is also why I need to be really honest. There are certain platforms that host their own in-house models and i'm still part of their CPP. However, because those models are still far behind the quality of Seedance 2.0, I haven’t used them that much. Seedance 2.0 has simply become the benchmark for me. The type of output I’m looking for is also extremely specific, so I can immediately feel it when a model cannot deliver what I need. Seedance 2.0 is definitely not cheap, but it gives me so much creative satisfaction and allows me to make whatever I want. I even have a team that low-key makes softcore erotic videos in the style of Vivamax 😆 I think I’ve trimmed down so many things in my AI workflow because my main goal is to focus on the content itself. If Seedance 2.0 Mini is released soon, I’m dead curious to test it. I think I want to create more stories that revolve around drama rather than highly technical cinematic shots. Seedance 2.0 Fast has been incredibly helpful, but I’m definitely curious to check out the Mini version. But the truth is that I’m completely tool-agnostic. I don’t care which company makes the model. I only care about the quality. You might remember when I praised Grok Imagine Video so damn hard because it was genuinely amazing back then. Then the quality kept getting worse and worse, so I stopped using it. But if it gets better again, I’ll definitely want to use it again. At the end of the day, quality is the only thing that matters.

MXVDXN // DAN

18,949 Aufrufe • vor 3 Monaten

introducing a new, very fun, LLM benchmark- the Game-of-Life Bench! the rules are simple: given an 8x8 grid following Conway's game of life rules, the goal is to create an initial pattern with at most 32 cells that can last the longest number of turns before dying/repeating. some results to highlight (with caveats detailed below): - gpt 5.1 lasts the longest with a 106 step run - claude models are really bad at this! they refuse to reason about this task and score < 25 points - deepseek r1 is the best open model with 102 steps. why? because i wanted to create a benchmark that has (i think) no practicality, but is still fun to look at, cheap, and still measures something interesting. i also am a big fan of the game of life. its absurdly simple rules leading to intractability is extremely cool to me. also, i saw a lot of work with LLMs trying to "predict" the next state in Conway's game of life, I think game-of-life bench is more fun because it's pretty open ended and only asks the LLM for the initial state. I also think this could be an RL env? but idk why you would ever train on this task haha i don't think this is a "serious" benchmark because it doesnt measure anything practical, but i still think it's a hard benchmark exactly because you can't predict what happens with your initial state many turns into the future; this is why i was initially expecting all LLMs to be bad at it, but turns out, some are clearly better than the others (the ordering may surprise you!) reminder: this is still a work-in-progress; (1) i am gpu-poor so could only do 10 runs for each model, even though total running cost is relatively low. maybe with some more credits i can run more seeds for each model. (2) i handpicked models which i think are at the frontier right now, plus some others that were on my mind. so, if you'd like to see a model on here, let me know. (3) i currently only do an 8x8 grid because i thought that by itself would be pretty hard for current LLMs, but of course we can increase grid sizes! (4) the coolest thing is, i dont think we can calculate the max possible number of states (yay undecidability!) you can go without repeating, so this is essentially a no-ceiling task, which is pretty cool! again, i did this mostly out of a desire to make LLMs do something fun. if this keeps me entertained for a few more days, i'd likely release a blog post on it. if it keeps me entertained for a week (and someone sponsors me), i'll put more work into it :P lastly, this is fully open sourced, so feel free to run this on your own!

Akshit

13,775 Aufrufe • vor 7 Monaten

Brooks Koepka leaving LIV Golf is an obvious blow to the league. But, it would be a considerably bigger blow if Bryson DeChambeau were to do the same. Bryson is contracted through the end of 2026, but he’s currently in discussions to extend early. Speaking exclusively to Tom Hobbs from Flushing It Golf, he spoke about the current situation: “I mean, look, it's confidential. I’m not going to share too much, but the conversations are in process. We have to get to a place where both parties have a good understanding of one another. It is getting to a place that makes sense for both sides. “And, I think that can happen, but you never know. Life throws curve balls and, obviously, we saw what happened today (Brooks Koepka leaving LIV Golf) and that was quite a shock to a lot of people and something that, you know, it is what it is. “People make decisions for whatever their needs and wants are and, ultimately, you have to respect it and move on and it feels like it was a mutual understanding and that's great. “I think that as a league now we have more opportunity to make some movements and I think that team has an opportunity to do some things differently than the past few seasons. So, we'll see where it all goes and where it all leads. Ultimately, it's quite interesting.” It certainly is interesting. Throughout the season, LIV Golf officials were trying desperately to convince media members that Brooks was happy and committed to the league and his franchise, when perhaps they were actually just trying to convince themselves. So, were the players aware he felt this way and did they expect the news to come so soon, before the end of his current contract? “There was always rumblings, but ultimately, it was a shock when I saw it today. I was like, whoa, all right, well, didn't know it was going to happen today. I didn't have that on my bingo card for the 23rd of December. “There's also, people can look at it as an opportunity. I always look at it as when one door closes another opens, right? And, I'm not going to speak for LIV, but I think that's what we're thinking of, at least from my perspective. If I was running the league, I'd look at it as an opportunity, not in a negative or positive way. It's just, it is what it is, right?” Bryson makes an interesting point. Brooks was never really committed to his captaincy of Smash GC or building a franchise, so perhaps a change of leadership and the opportunity that offers is actually a positive. There’s no point in having a big name player on the league if he’s not even prepared to wear his own team branding. So, now that Brooks has officially left, does Bryson think Brooks should be allowed back onto the PGA Tour? “I don't know, man. I don't know what they should allow or not… If they're going to be doing it by the book, they should do it by the book and not give any special exemption. But if there's a special exemption, it definitely opens the doors for others to do the same, which, you know, it's a slippery slope for sure.” Currently, no official pathway back to the PGA Tour for Brooks Koepka has been announced. So, any decision that is made will definitely be something all players on the LIV Golf League will be paying very close attention to. Follows on in a quote of this post. Bryson DeChambeau Crushers GC LIV Golf

Flushing It

1,206,807 Aufrufe • vor 9 Monaten