Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

ACE-Step 1.5. Excited for this model! Can't wait to test it - pro-grade music gen; full songs in <10s @ RTX 3090; - works on <4GB VRAM; - covers, repainting, & vocal-to-BGM across 50+ languages; - outperforms commercial models.

14,886 Aufrufe • vor 5 Monaten •via X (Twitter)

0 Kommentare

Keine Kommentare verfügbar

Kommentare vom Original-Post werden hier angezeigt

Ähnliche Videos

🎥 Today we’re premiering Meta Movie Gen: the most advanced media foundation models to-date. Developed by AI research teams at Meta, Movie Gen delivers state-of-the-art results across a range of capabilities. We’re excited for the potential of this line of research to usher in entirely new possibilities for casual creators and creative professionals alike. More details and examples of what Movie Gen can do ➡️ 🛠️ Movie Gen models and capabilities Movie Gen Video: 30B parameter transformer model that can generate high-quality and high-definition images and videos from a single text prompt. Movie Gen Audio: A 13B parameter transformer model that can take a video input along with optional text prompts for controllability to generate high-fidelity audio synced to the video. It can generate ambient sound, instrumental background music and foley sound — delivering state-of-the-art results in audio quality, video-to-audio alignment and text-to-audio alignment. Precise video editing: Using a generated or existing video and accompanying text instructions as an input it can perform localized edits such as adding, removing or replacing elements — or global changes like background or style changes. Personalized videos: Using an image of a person and a text prompt, the model can generate a video with state-of-the-art results on character preservation and natural movement in video. We’re continuing to work closely with creative professionals from across the field to integrate their feedback as we work towards a potential release. We look forward to sharing more on this work and the creative possibilities it will enable in the future.

AI at Meta

2,265,118 Aufrufe • vor 1 Jahr

【1st Free Model Giveaway】 This is a model I did for both art and rig! And I'm now giving it away for free! Things included in this model: - My personal tier 4 level rig - 8 Expressions - Capturable tongue out - Full set of VBridger parameters I can provide (8 in total) How to get it: - Follow, like AND retweet THIS tweet to get both model and copyright! - Would be better if there were comments! - And if you can please follow me on Kofi too!!! - I can't monitor each person for this so PLEASE be self-possessed! - The model is now avaliable at for free! Notice: - VBridger parameters need VBridger plugin on Steam, all config are set; Be aware: only cheekpuff, mouthX and TongueOut are positive parameters, others are all optimize params for mouth&face capture! - Expresssion's config are all set too, hotkey 1-8, which are shy, black face, scared face, circle eyes, arrow eyes, sparkle eyes, heart eyes and angry Here're some terms about the model: - Please use requires attribution when using the model, link to my twitter page - Only people who follow, like and rt could get full access to the model! I can't monitor each person for this so PLEASE be self-possessed!!! - Commercial copyright is allowed for commercial stream, but CANT be used for printing, reselling and other activities like this! - The final interpretation belongs to me (Kanijam) - I just want to mention again please be self-possessed on both getting and using it!!! I might upgrade the model & sharing more free models in the future so PLEASE keep a good community environment!!!

Kanijam | L2D COMMISSION CLOSE

48,354 Aufrufe • vor 3 Jahren

hey if you're thinking about running qwopus (the claude opus distilled qwen 3.5 27B) as a coding agent, this might save you a few hours. i tested both the base and the distilled version on the same hardware. single RTX 3090. same prompt. same context. same everything. the only variable was the model weights. base qwen 3.5 27B built octopus invaders in 13 minutes. 1,827 lines across 11 files. zero steering. one scope bug that took 2 lines to fix. game ran. qwopus couldn't finish the same task. enemies overlapping on screen. bullets not firing. controls worked but the game was broken. i had to steer it multiple times and it still didn't produce a playable result. both run at 35 tok/s. both use thinking mode. the distilled version actually has better jinja compatibility and doesn't stall midtask like base does on claude code. for conversation and reasoning it feels sharper. but for multifile autonomous coding where the model needs to coordinate 10+ files without losing track, base wins and it's not close. distillation compresses reasoning patterns but seems to lose precision on complex coordination. the model "thinks" well but can't hold the full picture across files the way base can. tested on opencode (base) and claude code (both). next up is hermes agent framework on base. same hardware. same prompt. comparing agents now, not just models. video below. first half is the distilled model's broken game. second half is what base built on the same 3090. judge for yourself.

Sudo su

45,052 Aufrufe • vor 4 Monaten

BOOOM! WE DID IT! BRAINWAVE TO REAL-TIME MUSIC AI! It has been a life long decades quest to read brain activity and to convert it to words, and/or music, colors and/or images. Today I am very excited to announce with the assistance from Mr. Grok director of The Zero-Human Lab, we have solved brainwave to music and this is the absolute worse it will be. We found the code using an array of NeuroSky toy chips and our software pipeline connecting to open source ACE-Step 1.5 and a highly modified LoRA model we built for this. The lyric version is in testing now. This would mean that the model will interpret words from the brainwaves and music! Today we have the music side done and the quality and genera will expand. The is the worse it will sound. Your Brainwave Music™️will be cut into 2-5 minute pieces based on a number of factors. The specimen below is from a dream/hypnogogic state I was in last night and I have a recording of my thoughts after the state. The music was made in real-time and GUIDED the dream state with known technology like binaural beats (not easy to hear in this clip) and word back masking. This specimen below shows the interplay of my brain state to the music made by my brain and adjusted to produce profound insights. I solved a very difficult issue in this session with a new AI model. IT FREAKING WORKS! THIS IS OUR FUTURE OF MORE POWERFUL BRAIN FUNCTION! Our goal is to produce a portable device you wear and will be able to give real-time audio and PEMF (skull region), ultrasound (temple region) to maximize creativity and remote viewing. It is very early days but I wanted you to know first! YOUR support of my X account, just by reading this and sharing it, subscribing to my X, buying me a and becoming a member at supports this research. I will open source this at some point and build a device ANYONE can own. Thank you! I love you.

Brian Roemmele

198,257 Aufrufe • vor 1 Monat