🚀 The #GeoAI Python package now supports feature segmentation... from high-resolution satellite and aerial imagery using text prompts, such as trees, buildings, etc. It efficiently processes large datasets with automatic tiling and can save results as a single image. Stay tuned for more features coming soon! 📓 Access the notebook: 🛠️ Explore the GitHub repository: 📚 Dive into the documentation: 📺 Check out the entire YouTube playlist: #GeoAI #geospatial #AI #Python #DeepLearningshow more

Qiusheng Wu
13,923 次观看 • 1 年前
🚀 A sneak peek of a new feature in... the #GeoAI Python package! Now you can train an image segmentation model for extracting features (e.g., buildings) from satellite or aerial imagery—all with just a few lines of code. 🛠️ Check out the GitHub repository: 📚 Dive into the documentation: 📺 Check out the entire YouTube playlist: #GeoAI #geospatial #AI #Python #DeepLearningshow more

Qiusheng Wu
11,284 次观看 • 1 年前
🚀 I spent the entire day training image segmentation... models from scratch! I've created several pretrained models to detect features like buildings, cars, ships, solar panels, and wetlands. Video tutorials are coming soon! 🤗 Check out the pretrained models on Hugging Face: 🛠️ Check out the GitHub repository: 📚 Dive into the documentation: 📺 Check out the entire YouTube playlist: #GeoAI #geospatial #AI #Python #DeepLearningshow more

Qiusheng Wu
15,084 次观看 • 1 年前
A sneak peek of a new feature in the... #GeoAI Python package! 🎉 Now you can detect cars from georeferenced aerial imagery using deep learning—all with just a few lines of code. Stay tuned for an in-depth video tutorial coming soon! 🛠️ Explore the GitHub repository: 📚 Dive into the documentation: 📺 Check out the entire YouTube playlist:show more

Qiusheng Wu
45,675 次观看 • 1 年前
The GeoAI Python package now supports object detection using... pre-trained models from the GeoDeep libarary ( The supported object types include cars, trees, birds, planes, aerovision, utilities, buildings, and roads. Try it out: GitHub: Notebook example: #geospatial #geoai #opensource #pythonshow more

Qiusheng Wu
87,167 次观看 • 5 个月前
🚀 Exciting news! The #GeoAI Python package now lets... you train land cover classification models with just one line of code. Leverage any PyTorch segmentation model from — with hundreds of image encoders & pretrained weights available. 📍 GitHub: 📓 Notebook: #GeoAI #Geospatial #DeepLearningshow more

Qiusheng Wu
16,210 次观看 • 1 年前
Create Stunning Time-Series Satellite Images in Seconds! The GEE... Data Catalogs Plugin v0.5 for QGIS is now available and it's a powerful upgrade. You can now create time-series satellite imagery with just a few clicks using a simple interface. The new version also supports direct downloads to your computer, making the workflow faster and more efficient. Key Features: - Access over 80 petabytes of satellite and geospatial datasets from Google Earth Engine - Generate animated time-series imagery effortlessly - Export results directly from QGIS to your local machine Useful Links: QGIS Plugin Page: GitHub Repository: Video Tutorial: #QGIS #geospatial #EarthEngine #Python #datascience #satelliteshow more

Qiusheng Wu
14,166 次观看 • 7 个月前
MiniMax H3 Instead of sharing the prompts for each... of these videos, I thought it would be more useful to share how I created that prompts. All of the videos were generated with text-to-video. First, find an image with the kind of scene, composition and mood you want to recreate. I used a few YouTube playlist thumbnails as references but Pinterest is also a great place to find inspiration. You can even use your own old or nostalgic photographs. Then upload the image to ChatGPT and ask it to describe the scene. The description it gives you can essentially become your text-to-video prompt. From there, you can generate completely new scenes with a similar composition, atmosphere and cinematic language. You can of course use the reference image directly with image-to-video or as a first frame. But if the original image isn't yours, I prefer using it only as visual inspiration and recreating the scene through text-to-video. This is the prompt I use with ChatGPT: "Describe the scene in this image in English, focusing primarily on what is happening, the characters, their actions and body language, the setting and the overall atmosphere. Also briefly describe the composition, framing, camera angle, approximate lens choice, lighting, color palette and cinematic aesthetic. Keep it concise and scene-focused rather than overly technical."show more

Kōda
51,449 次观看 • 17 天前
A New Era with V3🪄 V3's new engine introduces... significant advancements in output generation. Unlike V2, where the multi-model system processed prompts to produce a single output, V3 is designed to generate multiple outputs and logically link them together. This enhancement effectively removes limitations on output size, enabling more complex and expansive results. Key Features Seamless Multi-Output Generation: V3 has been trained to generate separate outputs and connect them logically. This advancement ensures that there are no longer any limitations on output size. Intelligent Image Creation: V3 improves image generation with better tools, allowing the AI to create as many images as needed and place them within the project’s context. It supports various formats like PNG, JPEG, SVG, and GLB. Web-Integrated Intelligence: V3 can now search the web for documentation and data, providing real-time context and up-to-date references. For example, if you run a restaurant and want to update your website, simply ask Alchemist AI to “generate this website in a more modern style,” and it will update all content accordingly. Improved Creativity and Output Quality: V3's creative capacity has significantly increased. Simple prompts now generate more complete and refined results, with the system efficiently combining multiple elements into cohesive outputs.show more

ALCHEMIST AI 🔮
48,748 次观看 • 1 年前
I vibe coded a visual PDF search app with... ColQwen2. This is how it works: - Store PDF files as images in a Weaviate AI Database vector database - Embed images and text with a multimodal late-interaction model (ColQwen2) - Generate token-wise (and summed) similarity maps to highlight image patches with high similarity Now I need to refactor the messy vibe-coded project. In the meantime, you can check out the Notebook this demo is based on to try it out yourself:show more

Leonie
34,494 次观看 • 11 个月前
Hailuo / MiniMax H3 is here! 🚀 Just ran... my first experiment with the new H3 model using an idea I previously tested on the 2.3 version — and the results are surprisingly impressive. This was a true one-shot generation: first attempt, no retries, no selecting the best take. More experiments coming soon as I explore what H3 can really do. 🔥 Diving deeper into Hailuo AI (MiniMax) ✨ #MiniMaxH3show more

Anissa
14,609 次观看 • 1 个月前
Introducing Replit ModelFarm, the fastest and safest way to... build your next Generative AI app. Available for free on Hacker and Pro plans till October 15th. It requires zero setup, zero configuration, and zero API keys. With Replit ModelFarm, you can build a working Gen AI app in as little as 3 lines of code. Get started by installing the Replit AI library in any Python, JavaScript, or TypeScript Repl. The library implements an API for text completion, chat completion, and text embeddings. It supports streaming so your users can see model responses in real-time rather than waiting on a single output. All Hacker and Pro builders will have free access to a selection of Gen AI models offered by Google Cloud Vertex AI through Replit ModelFarm. All models are accessible from the development environment and any deployed app.show more

Replit ⠕
229,285 次观看 • 3 年前
🚨 THE ESX BLOCKCHAIN EXPLORER IS NOW LIVE! 🚨... Hey EstateX Family! 👋 Big news from the tech front! After some serious work by our incredible dev team (shoutout to our superstar CTO Graham and our design and dev team 💪), we’re thrilled to announce the launch of the initial version of the ESX Testnet Block Explorer! 🔍 What can you do with it? - Easily track transactions - Watch the ESX token price and movements in real time. - Monitor block production live - Dive into transaction data from the team and our testing partners - And much more This is just the beginning! Over the coming weeks, we’ll be rolling out even more features, making the explorer smarter, faster, and packed with insights. Oh, and stay tuned... we’ll soon be revealing the first wave of projects building on the ESX chain 👀 👉 Check it out here: Very exciting time, family!show more

EstateX
124,350 次观看 • 1 年前
Big moment for text-to-speech. Qwen just open-sourced a text-to-speech... model that lets you clone voices, design new ones, and control speech using natural language. Let me explain what I mean: You can literally tell it "speak in a cheerful tone with slight nervousness," and it actually does that. No complex audio engineering needed. What makes this special: - 3-second voice cloning - Covers 10 languages: English, German, French, and more - Latency as low as 97ms for real-time applications - Supports both streaming and non-streaming generation The model comes in two sizes (0.6B and 1.7B parameters), so you can pick based on your hardware and quality needs. Three modes to work with: 1. Custom Voice: Use pre-built premium voices with instruction-based style control 2. Voice Design: Describe the voice you want in plain English (or Chinese), and the model creates it 3. Voice Clone: Provide a 3-second reference audio and clone that voice The best part? It integrates with vLLM for production deployment and has a simple Python package you can pip install. I've shared a link to the GitHub repo in the next tweet.show more

Akshay 🚀
31,249 次观看 • 7 个月前
gemini omniflash is actually f*cking cracked. you can animate/edit... any video with a text prompt. character swaps, object transforms, full environment changes without regenerating/rotoscoping. everyone using AI to to animate and edit videos right now hits the same wall. the clip comes out 90% right and you regenerate from scratch hoping the 10% fixes itself. it never does. the fix is using your video as the input. omniflash edits what's already there instead of rolling the dice again. here's what's in the system: > the two-layer premiere trick: generate the same shot twice (one with background removed), stack them, cut at one frame, instant scene change > character swap with a single reference image (plus the one line you need or the model keeps the original's features) > object transforms that leave the rest of the frame untouched: stone into glowing sphere, candles into flowers > style transfer from an image reference instead of text, way more accurate > why stacking edits in one prompt breaks everything and the exact step order that doesn't > the audio limitation nobody mentions and how to work around it i packaged every prompt, the edit sequence, and the premiere layering setup. RT + reply "OMNI" and i'll send it over.show more

Sulfur
36,571 次观看 • 2 个月前
ANTHROPIC JUST TURNED AI AGENTS INTO GIT REPOS Anthropic... shipped "ant" - a CLI that runs every Claude API endpoint straight from your terminal. The headline isn't the terminal access. It's that you can now version-control an AI agent as YAML in Git and have CI sync it to the Claude Platform, the same way you ship code. - Every API resource is a subcommand: messages, models, files, agents, sessions - Define an agent in a YAML file, check it into your repo, and keep it in sync with one update command - Spin up a session, send it an event, then pull every event and tool call back from the same CLI - Claude Code knows how to drive ant out of the box - it shells out and reads the results with no glue code Agents just stopped being prompts you babysit and became infrastructure you deploy.show more

BuBBliK
200,456 次观看 • 3 个月前
GITHUB JUST KILLED THE WORST PART OF VIBE CODING... they shipped a free tool called Spec Kit and it already crossed 120,000 stars the fix is stupidly simple instead of tossing vague prompts at an agent and praying it doesn't wreck your project Spec Kit makes the AI write a full structured spec before it touches a single line of code it works through the problem first figures out what you want to build asks about the gaps lays out the project then it starts coding you get fewer insane bugs, cleaner output and results you can predict the flow looks like this: /constitution for your rules and standards /specify for what you want to build /clarify for the open questions before you start /plan for architecture and stack /tasks for the ordered work /implement to run it it plugs into Claude Code, Cursor, Copilot, Codex, Gemini CLI and 25+ other agents 120,000 stars, 10,000 forks, open source, shipped by GitHub itself learning to drive agents like this is most of what separates people getting hired as AI engineers from everyone still fighting their promptsshow more

Atlas
501,832 次观看 • 1 个月前
Today, Nesa is excited to announce cross-chain support with... Sei. As part of this new development, Nesa will be featuring a dedicated ramp for Sei’s rich ecosystem of applications to seamlessly integrate with AI. Now any Sei dapp can run AI inference on hundreds of models supported by Nesa, including the largest LLMs and most popular models across Vision, Language, and Generative AI. For developers, we’ll have information soon on how to access AI inference on Nesa via smart contracts from the Sei blockchain. Stay tuned for more details on this ground breaking development for decentralized AI.show more

Nesa
79,690 次观看 • 2 年前
🎬 Motion Control has arrived! Take full control of... your AI videos with 12 dynamic camera shots — from smooth Dolly moves to dramatic Cranes and VFX like Explosions and Disintegration. Perfect for adding cinematic flare, stock footage, product ads, or just experimenting with storytelling. ✨ How it works: 1️⃣ Head to the new Video creation tool 2️⃣ Enter your prompt 3️⃣ Pick your camera shot from the Motion Control panel 4️⃣ (Optional) Add a Style or inspiration image as a start frame 💡Want extra consistency? Train an Element, generate your key stills, and animate them for a fully guided scene. More features are coming soon — including End Frames 👀 So stay tuned, and let us know what you create! 🎥 Lights, camera... Motion! Try it now on the Video page 👉🏻show more

Leonardo.Ai
984,421 次观看 • 1 年前