Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Global illumination with radiance cascades, now optimized, running at ca. 0.3 ms per frame (GTX 970 GPU). No denoising or temporal accumulation. Calculated from scratch each frame. Loving the technique. Link to the WIP paper: #gamedev #indiedev

168,530 Aufrufe • vor 2 Jahren •via X (Twitter)

9 Kommentare

Profilbild von Mytino (Asbjørn L)
Mytino (Asbjørn L)vor 2 Jahren

YouTube version of video:

Profilbild von Jan Orszulik
Jan Orszulikvor 2 Jahren

Ok, I like this a lot! I am adding this to the list I want to investigate further, took a quick glance at PoE 2 related paper, limitations are perfectly acceptable. Great work, both this implementation and mainly original technique from Alexander.

Profilbild von Bino🔜 VFM Tales of Neverland
Bino🔜 VFM Tales of Neverlandvor 2 Jahren

how well this translates to 3D?

Profilbild von Mytino (Asbjørn L)
Mytino (Asbjørn L)vor 2 Jahren

There are multiple ways of implementing it for 3D, such as the one used for Path of Exile 2 like Xor mentioned. The paper I linked has more information.

Profilbild von Mike Dailly™ 🏴󠁧󠁢󠁳󠁣󠁴󠁿🇺🇦🇵🇸💙
Mike Dailly™ 🏴󠁧󠁢󠁳󠁣󠁴󠁿🇺🇦🇵🇸💙vor 2 Jahren

That's damn fine work. ♥️

Profilbild von Zach
Zachvor 2 Jahren

I sort of skimmed the article "Radiant 2d is capable of calculating per-pixel global illumination in 2d in a constant time of about 30ms on a GTX3060." Am I missing something? How are you accomplishing this in 0.3ms?

Profilbild von Jan Orszulik
Jan Orszulikvor 2 Jahren

Hmmmmmm

Profilbild von DiTieM Games - Wishlist Dungeon of Astaroth
DiTieM Games - Wishlist Dungeon of Astarothvor 2 Jahren

what is the purpose of such fantastic visuals? game? engine? paper?

Profilbild von Mytino (Asbjørn L)
Mytino (Asbjørn L)vor 2 Jahren

Thanks, mainly game :)

Ähnliche Videos

-- What holds it together -- ✂️PAPER CUTS.- When the image is the end ⤵️ What you bring to life is the journey towards the image: how the pieces come to be as they are. The movement is the idea, and your image in Seedance is the final frame. ---------------------------------------------------- What happens to the object? 1⃣ By the time the object TRANFORMS into something else. It takes place in a continuous shot, a single camera that zooms in slowly and never cuts away. A cut would break the transformation; the hypnotic effect lies precisely in the fact that it doesn’t flicker. Template.- [@.RE IMAGE] is the LAST frame. [GLOBAL] [Technique + light + backdrop + focus]. [Real materials]. One continuous hypnotic transformation — a single take, no cuts. The same [OBJECT A] does not appear in pieces; it [VERB: melts / folds / blooms / dissolves] into [OBJECT B]. Slow surreal dreamlike drift, one unbroken slow push, slight stop-motion shimmer, no snapping. 0-1s: [state A, intact and recognizable]. 1-2.5s: [the change begins — "the same material begins to..."]. 2.5-3.5s: [the change at full — "...becomes..."]. 3.5-5s: [it settles]. The push eases to rest. Locked, exact match to the last frame. [LOGIC RULE] one continuous same-lens push, never cutting. [A] morphs into [B], [details that must not deform] stay legible, no warping. hypnotic drift. SFX: [a sound that also transforms]. no music, normal speed. 2⃣ When one thing LEADS to another, we need each step to be a link in the chain of causality. Template.- [@.RE IMAGE] is the LAST frame. [GLOBAL] [Technique + light + backdrop + focus]. [Real materials]. Hypnotic stop-motion paper cadence, slight frame-step, brisk causal montage, ~1.5s per shot, no naturalistic motion, no slow-mo. [cut] [shot + camera] / [carries on from previous cut] / [ACTION: what it DOES, not how it looks]. SFX: [beat-anchored hit]. [cut] ... (link 2 — chains from link 1) [cut] ... (link 3) [cut] ... (link 4) [cut] ... ease back to reveal / [link 5]. Locked, exact match to the last frame. SFX: ... [LOGIC RULE].- [materials], [what must NOT happen], no warping, [text legible if any]. ~1.5s per shot, no slow-mo. no music.

AlexandrIA

41,804 Aufrufe • vor 3 Monaten

CoDeF: Content Deformation Fields for Temporally Consistent Video Processing abs: paper page: present the content deformation field CoDeF as a new type of video representation, which consists of a canonical content field aggregating the static contents in the entire video and a temporal deformation field recording the transformations from the canonical image (i.e., rendered from the canonical content field) to each individual frame along the time axis.Given a target video, these two fields are jointly optimized to reconstruct it through a carefully tailored rendering pipeline.We advisedly introduce some regularizations into the optimization process, urging the canonical content field to inherit semantics (e.g., the object shape) from the video.With such a design, CoDeF naturally supports lifting image algorithms for video processing, in the sense that one can apply an image algorithm to the canonical image and effortlessly propagate the outcomes to the entire video with the aid of the temporal deformation field.We experimentally show that CoDeF is able to lift image-to-image translation to video-to-video translation and lift keypoint detection to keypoint tracking without any training.More importantly, thanks to our lifting strategy that deploys the algorithms on only one image, we achieve superior cross-frame consistency in processed videos compared to existing video-to-video translation approaches, and even manage to track non-rigid objects like water and smog.

AK

153,305 Aufrufe • vor 3 Jahren

Introducing KausaCompute. Running AI models privately is expensive and complicated. Traditional cloud providers require credit cards, KYC verification, and complex setup processes. For many developers, especially in emerging markets, accessing GPU compute remains out of reach. KausaCompute changes that. Deploy any Docker container with NVIDIA GPU, pay with USDC from Maze Pocket, no KYC, no credit card, no cloud provider account needed. GPU pricing starts at $0.47/hr with competitive rates across all tiers. KausaCompute lives inside KausaLayer Pocket. Open a pocket, swap SOL to USDC using the built-in swap feature, and start deploying GPU containers right away. Everything stays within one ecosystem. What it actually does: Private LLM endpoints. Deploy Llama 3, Mistral, CodeLlama, or any open-source model as a personal API. No one logs the prompts. No one reads the data. Full control. AI coding assistants that never see external servers. Speech-to-text processing for thousands of audio files in minutes instead of hours. Sentiment analysis across millions of data points. Document summarization at scale. Multi-modal AI that processes images and text together. All running on dedicated NVIDIA GPUs. A6000, T4, L4, L40, A100, H100, and more. Pick the hardware, pick the duration, deploy in one click. The billing is straightforward. USDC is deducted from Maze Pocket before deployment. Pro tip: Maze Pocket supports multiple pockets per wallet. Create a dedicated pocket just for KausaCompute to keep GPU spending separate from other activities like trading or transfers. Clean separation, easy tracking. KausaCompute is not another cloud provider. It is the fastest path from USDC to a running GPU, with zero identity requirements. Live now at

KausaLayer

10,945 Aufrufe • vor 4 Monaten

AI TENNIS ANALYSIS. A FULL COMPUTER VISION SYSTEM. BUILT ON YOLO, PYTORCH, AND KEYPOINT EXTRACTION. Take any tennis match broadcast, any camera angle, any resolution. Feed it into the pipeline. YOLO detects both players and the tennis ball frame by frame. No manual labeling, no pre-annotated dataset. A fine-tuned YOLOv5 model trained on a Roboflow tennis ball dataset handles the ball - the hardest object to track in any sport. Tiny, fast, constantly occluded. The model finds it anyway. Trackers maintain identity across frames so Player 1 stays Player 1 from the first serve to match point. But detection is just the start. A ResNet50 CNN trained in PyTorch predicts court keypoints from every frame - the corners, service lines, baselines, net posts. Fourteen points that define the entire playing surface geometry. From those keypoints the system builds a homography matrix and warps the broadcast perspective into a top-down mini court with real coordinates. Now every player has a position in real space, not pixel space. Every frame becomes a measurement. Every rally becomes a dataset. Player movement speed - calculated from position deltas between frames, converted to meters per second through the homography. Ball shot speed - measured from the ball trajectory across consecutive detections. Number of shots per rally - counted automatically through ball direction changes. All of this rendered live on the video as an overlay. A mini court in the corner showing both players as dots moving in real time. Stats updating after every point. OpenCV handles the rendering. Pandas handles the math. PyTorch handles the intelligence. YOLO handles the eyes. No Hawkeye subscription, no court-embedded sensors, no tracking chips in the ball. A Python script, a trained model, and a GPU. The full code is on GitHub. The tutorial walks through every module - from ball detector training to court keypoint extraction to the final statistical overlay. Professional teams used to need broadcast deals and proprietary hardware for this kind of analysis. Now you build it in an afternoon with open-source tools. Trading here: Computer vision didn't just enter tennis. It made the expensive stuff free.

zostaff

121,478 Aufrufe • vor 5 Monaten

Multi-Track Timeline Control for Text-Driven 3D Human Motion Generation paper page: Recent advances in generative modeling have led to promising progress on synthesizing 3D human motion from text, with methods that can generate character animations from short prompts and specified durations. However, using a single text prompt as input lacks the fine-grained control needed by animators, such as composing multiple actions and defining precise durations for parts of the motion. To address this, we introduce the new problem of timeline control for text-driven motion synthesis, which provides an intuitive, yet fine-grained, input interface for users. Instead of a single prompt, users can specify a multi-track timeline of multiple prompts organized in temporal intervals that may overlap. This enables specifying the exact timings of each action and composing multiple actions in sequence or at overlapping intervals. To generate composite animations from a multi-track timeline, we propose a new test-time denoising method. This method can be integrated with any pre-trained motion diffusion model to synthesize realistic motions that accurately reflect the timeline. At every step of denoising, our method processes each timeline interval (text prompt) individually, subsequently aggregating the predictions with consideration for the specific body parts engaged in each action. Experimental comparisons and ablations validate that our method produces realistic motions that respect the semantics and timing of given text prompts.

AK

126,635 Aufrufe • vor 2 Jahren

If you take a movement to unpack this visualization... You'll see how it simply breaks down how reality works. At frame 0 you have a static image. Everything is one, this is the monad. As soon as you hit frame 1 there is movement, there is change. Now you have two states, moving, or static. When Nikola Tesla says you can explain everything in frequency and vibration. The difference between frame 0 and 1, is vibration. The difference between movement and no movement. This is like binary logic we use in code which is made up of 0's and 1's. After frame 1, is when frequency emerges. Because the difference between frame 1 and all frames after is about how fast is the vibration/movement happening. If we skip forward to frame 50... You have a shape that begins to emerge, this is the 8 dots, then the 6 dots. Notice how unstable it is, it's 8 dots, then 6, then a moment with 4 in a rectangle These shapes are emergent properties. The first two emergent properties after the monad was vibration and frequency. Next comes shape (i'm skipping over rotation and direction). These shapes of dots can only exist when you have frequency and rotation. This frequency and rotation creates vortex energy. It's the same energy that things like your chakras use. Or the same energy we harness in devices like engines, airplanes, fans, blenders, hard drives, etc. It's also the same vortex energy you'll see in a tornado or hurricane. They are powered because they harness rotation and frequency(change/movement). Going back to the video, notice that it is inside the entire shape, the internal structure is manifesting before the external structure does. Then around frame 60 the hexagon of circles begins to rotate. First it was the two dots that moved and now it's a complex shape that is coming to life. This is a higher dimension (or lower depending on how you look at it) manifesting into existence. The internal state is "awakening" and experiencing it's own change like what happened to the whole shape in the first frames. But it is unstable. That's why it doesn't persist for long. If you think of the 8 dots being the octahedron, they map to the element of air. Air is in the material world, but it is not something you can see. The brief moments the 8 dots are visible is similar to that effect. They are only experienceable between a small frequency band of frames. Now here's where stability begins to appear in the internal structure. This is when the 4 dots appear. You'll see that the four dots, the square, is stable and persists the most visibly for the most amount of frames. The square represents earth in the platonic solids to elements mapping. Earth, is material, it's stable. We build our buildings in squares and with earth because it is a solid shape to build on. This visualization shows you why. Across different vibrations (frame rates) it can self sustain. Between this point and frame 180, you'll see a new emergent property. Which is depth. A new dimension is introduced at around frame 90 but really becomes visible at around frame 110. You can see a foreground and background. There is the shape of the dots, but also the triskellion wave happening in the background. Let's jump to frame 180. Notice how it is the same as frame 0 except... It's flashing. If you were paying attention, you'll notice you could see flashing at frame 90 and frame 120, but they didn't persist for long. At around 150 it started to reach stability and 180 it was solidified. Between frames 150 and 180 there is flashing, but the image is still moving. Only for a brief moment at frame 180 is the movement frozen and the flashing persists. Think of that like your computer screen. It's what your screen is doing right now as you read this. Even tho the text isn't moving, the screen is flashing at 60 or 120hz. The images appear on your device because this flashing brings things to life. The entire material realm and your physical body right now, is doing the same thing. While you look solid... You're flashing in and out of existence at very high frequencies. You can look at frame 180 and frame 0 as the same essence but it is the mid point between an octave change. In the video, the ying and yang was vertical, now it is horizontal. This is a phase shift. If you notice at exactly frame 180, the rotation freezes and then the direction of rotation changes. The process then repeats all the way to frame 360 but in the opposite sequence. Once it reaches frame 360, that is an octave change and the process repeats. Each time you repeat the process is a layering of the same patterns into higher octaves. This is the same as your chakras or how other things work. They are like russian nesting dolls where every octave is layering onto the next. The complexity of your body is a layering of basic principles that emerged in earlier stages. Your organs are built of systems that are built with cells that are built with proteins that are built with atoms and so on. The atoms, work just like your body at a basic level. Your body works just like the galaxies. At each level you'll have the same pattern. This is where the idea "As Above, So Below" from. The monad, splits in two, and so on and so on. One cell, splits into two through mitosis in the same logic. We could spend all day going through examples of how biology, physics, spirituality, etc. aren't really different. They are just categories that we use to dissect these frequencies and octaves of energy but they only start paying attention within the confines of materialism. The problem is, none of the sciences start at the root patterns. Because that is reserved for religion or spirituality. It's too woo-woo to take seriously so it's dismissed. And because of that... We're left ignorant on the simple explanations for how things work. Now you need some expert with tools you don't have access to in order to explain things. When you could be understanding them without the tools. The Yin and Yang symbol in this video is 3,000 years old. It's simple. Yet I just showed you how it explains deeper layers of reality.

Jamal ☯︎ 🔆🧘🏽🧠

13,149 Aufrufe • vor 5 Monaten

How to build chemistry in seconds: This is a great example of challenging, What challenging does is maintain the tension, it’s “fighting for position” Which really, what you’re avoiding is collapsing into her frame. She is trying to prize frame herself. “You wish you had this number” A lot of guys either collapse into her frame “Yeah I do, what’s your number” Or they just outright deny “No I don’t” Both suboptimal. You want to deny without qualifying yourself. Tease her back - “You’re mumbling I can’t even understand you” “Numbah?” “What’s a numbah?” Misinterpret “What you said you want my number? You can’t have it yet” Reframing the interaction knocks her off the pedestal without triggering her, or coming across try hard. What he did was perfect. Next, “Why are you getting so close to me” He’s challenging her prize frame, starting shit for more tension. “I didn’t give you permission. Stop touching my titties” Guys are gonna be like wtf is wrong with his voice But the intent and frame of your words matter more than the base. (Though a good voice still helps) “You like me already” Prize framing himself. But it’s also just him expressing his observation. “Yeah you wish I liked you” The girl now challenging the prize frame. And instead of going back and forth, another thing you can do is just call her bluff. He moved in closer to her and she broke instantly, wanting to kiss him. “I can’t kiss you right now” The only reason she didn’t was most likely a logistics issue. (Notice her looking over, checking who’s still around) He continues the challenging “Get out of here” “No you get out of here” Again - it’s more bluffing and calling out the frame He wants to stamp the fact that she likes him although it’s bit of an overextension. She’s starting to get tired of it and needs the interaction to move forward. “You want this or not because if not we’re gonna go home” The thing is he keeps playing around the “do you like me” frame Which is good to keep up the tension (which this girl likes) but it also keeps the interaction at a standstill. And eventually it will stall out and break. At this point - the opposition frame hurts your outcome more than it benefits. A better path forward is either to lead - pull her, help her with her logistics, move her away from friends, etc Or at the very least, instead of fighting for position, provoke escalation. “You can’t handle me ;)” “No you can’t handle me” Get her to challenge you and you can call her bluff to escalate like what happened earlier when he walked towards her. “Stop looking at me like I won’t do something to you right now ;)” Provoke, challenge her to set up the escalation window and it breaks the dancing around of who likes who while still maintaining tension. But bottom line, this interaction is a great example. Challenging = chemistry.

The Rizz Report

49,001 Aufrufe • vor 3 Monaten

🇺🇸 Orb UFO Video Released By Pentagon - Moving 500mph 😱 Incredible UFO Video from the latest drop 🛸👽 “Looks like a white round orb moving as fast as 500 miles per hour.” Newly released official infrared footage from a 2025 US Central Command mission in the Middle East shows a military sensor tracking an unidentified object. The Department of War has listed the case as unresolved. AARO states the available data is insufficient to determine whether the thermal contrast is a physical object or a sensor artifact. DOW-UAP-PR133, Unresolved UAP Report, Middle East, 2025 The sensor pans left to right to keep the area of contrast centered and changes magnification several times during the recording. No corroborating radar, range, or multi-sensor telemetry was released with the video. The report is part of the latest PURSUE tranche of previously classified UAP records. Hundreds of similar military reports remain in the active archive because they lack the complete sensor data needed for a conclusive identification. The United States Central Command submitted a report of an unidentified anomalous phenomenon to the All-domain Anomaly Resolution Office (AARO) consisting of 4 minutes and 58 seconds of video footage from an infrared sensor aboard a U.S. military platform in 2024. An accompanying mission report, DOW-UAP-D107, described the UAP as “look[ing] like a white round orbit… moving as fast as 500 miles per hour.” Video Description: 00:00-00:04: The sensor pans to track an area of contrast, keeping it generally centered within the frame. 00:05-00:50: The sensor decreases its level of magnification and continues to track the area of contrast, keeping it generally within the center of the frame. 00:51-01:21: The sensor decreases its level of magnification and adjusts its panning speed to track the area of contrast, keeping it generally within the center of the frame. 01:22-01:41: The sensor increases its level of magnification and adjusts its panning speed to track the area of contrast, keeping it generally within the center of the frame. 01:43-2:05: A rectangular digital overlay, or “focus box,” appears near the center of the frame. This display element is a graphical component of the sensor system and does indicate a physical change in the environment. The sensor pans to track the area of contrast, keeping it generally within the center of the frame before exiting the sensor field-of-view from the bottom edge. 02:06-02:21: The sensor pans across the scene and decreases its level of magnification. The area of contrast intermittently re-appears from the left edge of the frame. 02:22-02:28: The sensor pans across the scene, reacquiring the area of contrast. After the area of contrast re-enters the scene, the sensor continues to pan to track it, keeping it generally within the center of the frame. The sensor engages briefly stops panning and the area of contrast exits on the right side once again. 02:29-02:36: The sensor activates an auto-tracking reticle, which does not establish a lock on the area of contrast. After the auto-track reticle disengages, the sensor stops panning, causing the area of contrast to exit the frame from the right edge. 02:37-3:22: The sensor pans and requires the area of contrast, keeping it within its field-of-view before it exits the bottom edge of the frame. 3:23-4:29: The sensor continues to pan then zooms out and reacquires the area of contrast, now significantly smaller in the sensor’s field of view. The sensor continuously pans to keep the area of contrast in the field of view. 4:30-4:58: The sensor zooms in, and the area of contrast transits the field-of-view, exiting the top of the frame. The sensor cycles magnification levels multiple times.

Interstellar

22,784 Aufrufe • vor 11 Tagen