Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

Ahmedabad Crime Branch is making use of technical measures to avoid any stampede kind of situation. Anti stampede visual analytics,using reference area and crowd movement, head count algorithm. Anti-stampede algorithms on CCTV cameras are a crucial advancement in crowd management, leveraging AI and image processing to prevent dangerous situations...

339,758 Aufrufe • vor 1 Jahr •via X (Twitter)

11 Kommentare

Profilbild von Pranav Sarswat
Pranav Sarswatvor 1 Jahr

Ahmedabad Crime Branch is ahead in crowd safety with AI-powered anti-stampede algorithms on CCTV cameras! Real-time monitoring, crowd density tracking, anomaly detection & instant alerts help prevent stampedes before they happen. Leveraging advanced tech like Mask R-CNN and predictive analytics, they ensure safer public gatherings by detecting surges and unusual crowd behavior early. Challenges remain occlusion, privacy, and integration but this is a huge leap in proactive crowd management. AI + human action = saving lives!

Profilbild von Standing for Freedom Center
Standing for Freedom Centervor 2 Jahren

Like the plot to a dystopian movie, New York will now monitor social media writings, collect data, and use law enforcement to crack down on any expression it deems to be hate speech.

Profilbild von SK Reddy 🇮🇳 - A Political Saint.
SK Reddy 🇮🇳 - A Political Saint.vor 1 Jahr

@MrSinha_ Much appreciated. Otherwise India is huge in population & stampede has become new normal.

Profilbild von Power Of Youth
Power Of Youthvor 1 Jahr

Ahmedabad Crime Branch is utilizing advanced anti-stampede algorithms on CCTV cameras to enhance crowd management and prevent dangerous situations through real-time monitoring and predictive analytics.

Profilbild von Vikas 🚩
Vikas 🚩vor 1 Jahr

Awesome use of AI enabled camera on drones.

Profilbild von 🚩Divy_23 🇮🇳
🚩Divy_23 🇮🇳vor 1 Jahr

@ShivAroor @republic Please cover this news

Profilbild von Vinashaya cha dushkritam 🚜 🇮🇳
Vinashaya cha dushkritam 🚜 🇮🇳vor 1 Jahr

@SJTA_Puri it will be great if you may ask for its implementation during #rathyatra 2025

Profilbild von Mach10
Mach10vor 1 Jahr

What ahead and advance than Bangalore.

Profilbild von Vamsi Venkat
Vamsi Venkatvor 1 Jahr

@vijaygajera @ncbn @naralokesh might be helpful in our state

Profilbild von Venkatramana Rao P
Venkatramana Rao Pvor 1 Jahr

But all these works when there is proper planning. If it happens like that of in blore nothing can be managed.

Profilbild von Pranav Patel 🇮🇳
Pranav Patel 🇮🇳vor 1 Jahr

Wow

Ähnliche Videos

Kudos to Gujarat Police for revolutionizing law enforcement with cutting-edge technology! Gujarat Police is leveraging innovative technologies like night vision thermal detection drones, AI, real-time mapping, and human intelligence to transform policing in the state's remotest tribal district, Dahod. This forward-thinking approach has led to the successful detection of various cases. Some of the technologies being used include: •⁠ ⁠Night Vision Thermal Detection Drones: Equipped with thermal imaging cameras, these drones can track suspects in dark conditions and detect heat emitted by humans and animals. •⁠ ⁠AI-Powered Systems: Artificial intelligence is being used to analyze data, predict crime patterns, and identify high-crime areas, enabling police to allocate resources more effectively. •⁠ ⁠Real-Time Mapping: Advanced mapping technologies provide real-time information, enabling police to respond quickly and efficiently to emergencies •⁠ ⁠Human Intelligence: Trained personnel are working alongside technology to gather intelligence, conduct surveillance, and solve cases. The integration of these technologies has transformed policing in Dahod, making it a model for other districts to follow. Gujarat Police's commitment to innovation is truly commendable! Watch the inspiring story on India TV to learn more about Gujarat Police's groundbreaking initiatives! #GujaratPolice #PolicingWithTechnology #Innovation #LawEnforcement #Dahod #TransformingPolicing

Harsh Sanghavi

40,928 Aufrufe • vor 2 Jahren

Elon wants Treasury Dept to run on Blockchain as we now find out career officials are breaking the law every hour of every day by approving payments that are fraudulent or match funding laws passed by Congress Great news Elon Musk because Treasury Department is already running on-onchain and here are some of the things they are testing ✅DLT for Transparent Ownership Records: Digital ledger technology provides a transparent and immutable record of ownership for tokenized assets. Each transaction involving tokenized assets is recorded on the blockchain, creating a tamper-proof audit trail of ownership transfers. DLT ensures transparency and trust in the ownership history of tokenized assets, mitigating the risk of fraud and disputes ✅AI-Powered Asset Valuation and Risk Assessment: AI algorithms can analyze vast amounts of data to assess the value and risk of tokenized assets. Machine learning models can incorporate financial data, market trends, and other relevant factors to provide accurate valuations and risk assessments in real-time. AI-powered analytics can help investors make informed decisions about buying, selling, or holding tokenized assets based on their risk appetite and investment objectives ✅Automated Compliance and Regulatory Reporting: AI can automate compliance processes and regulatory reporting requirements for tokenized financial products. Machine learning algorithms can monitor transactions for suspicious activities, detect potential compliance violations, and generate regulatory reports automatically. By integrating AI-powered compliance solutions with DLT-based platforms, financial institutions can ensure regulatory compliance while minimizing operational costs and risks. ✅Decentralized Governance and Decision-Making: DAOs enable decentralized governance structures where stakeholders collectively make decisions about the management and operation of tokenized financial products. Token holders within DAOs can vote on key governance issues, such as asset allocation, dividend distribution, and protocol upgrades. Decentralized governance ensures transparency, accountability, and community participation in the management of financial product ---------------------- All of this is running on the same DLT that is the number one data focused blockchain used by @DeptofDefense 🇺🇸IRON SPIDR Powered by Constellation² $DAG

Dagnum²

17,293 Aufrufe • vor 1 Jahr

Experiments in progress. The one on the right has been learning for ~3 hours, the one in the middle for ~1 hour, and the one on the left just started a few minutes ago. The initial motivation for making the physical Atari was just to commit ourselves to a subset of algorithms that can make progress in this setup. This commitment rules out algorithms that require billions of samples to learn (or worse, require multiple environments running in parallel). Atari games are simple enough that we should be able to show learning on them in a short amount of time with no prior knowledge. Since then, I've realized that this setup is also a good way to compare different paradigms in robotics in a principled way. These paradigms are sim2real, learning from tele-operated data, and learning directly on the robots. So far, I have observed that getting sim2real to work reliably is hard. It requires tweaks that don't scale. Policies that can play perfectly in simulation fall apart because of latencies and the messiness of the real world. These aspects could be modeled to improve the simulation, but not without sinking significant human engineering hours. I have higher hopes for learning from tele-operated data, but that requires a human to learn the task first. These experiments are on my to-do list. I have to learn to play some of the games well through the robot. I’m half-decent at playing Pong and Ms Pacman now. Learning directly on robots is looking like the most promising approach. This approach takes away pesky distribution shifts and makes it possible to have algorithms that continually improve with more data and time without any human intervention. It feels great to let experiments run overnight and wake up to find improved policies. With learning on robots, I should, in principle, be able to go on a long vacation and come back to find better policies for complex tasks beyond Atari games. Whether that is possible with current learning algorithms is a different question.

Khurram Javed

52,110 Aufrufe • vor 9 Monaten

Major program launch: Data Analytics Professional Certificate! This large, five-course sequence takes you all the way to being job-ready as a data analyst, and shows how to use Generative AI as a thought partner to enhance your work in this role. Offered by on Coursera, this is taught by Sean Barnes, Ph.D., a Data Science & Engineering Leader at Netflix. Analyzing data remains one of the most important skills in where the world is going with AI. This comprehensive certificate takes you all the way to being job-ready. Each course comes with practical projects demonstrated in real-world contexts, such as analyzing sales data for a Korean bakery, video game sales trends across different regions, or identifying factors impacting customer retention for a communications company. You'll also work on estimating fire distribution for forest fire prevention, analyzing how a diamond's properties affect its market value, and developing predictive models for retail sales analysis, carbon emissions, and coral reef conservation. Here's some of what you'll learn: - How to define data and categorize it into its many types such as discrete & continuous numerical, structured & unstructured, time series, categorical, and know what insights can be derived from the different types of data categories. - How to differentiate between data-related job roles and their responsibilities, and how data flows through an organization from the moment of capture to decision-making. - How to perform data processing functions and apply conditional formatting in spreadsheets to extract business value from your data using statistical calculations and best practices for visualizing and interpreting data. - How to use LLMs for stakeholder analysis, data exploration, and data visualization. - Best practices for using LLMs for as a thought partner to data analysis work By the end of this professional certificate program, you will have learned core statistical concepts, analysis techniques, and visualization methodologies that will serve as the foundation for working as a data analyst. The world needs more data analysts, especially ones who know how to use modern generative AI. With data science roles projected to grow 36% by 2033, the skills taught in this program create new professional opportunities in data. Sign up here!

Andrew Ng

85,182 Aufrufe • vor 1 Jahr

AI is changing sports. Here is how. I sit down with Max Sebti, , founder and CEO of Score, and he gives me the latest about how sports is changing due to AI. What will you learn from this interview? 1. How AI Is Transforming Sports Using computer vision to analyze every movement, event, and play in real-time. Moving beyond basic stats to understanding impact and intent on the field. 2. What Makes SCORE Different Built on decentralized AI (Bittensor) and collective intelligence. Designed to work even with low-quality video—enabling access for high schools and amateur clubs. 3. Real Use Cases Player tracking, formation analysis, injury prediction, and in-game decision support. Visual tools like heat maps and frame-by-frame breakdowns. 4. Applications Beyond Pro Teams Empowering grassroots teams and scouts with elite-level insights. Parents filming Sunday league games could unknowingly be training data sources. 5. Fantasy Sports Integration AI-powered projections and analysis for fantasy leagues. Build-your-own tools for fans who want a data edge. 6. Injury Risk Detection Early signals from movement patterns that correlate with higher injury potential. Long-term value for athlete health and coaching adjustments. 7. Preparing for the AR/3D Future Compatible with lightfield displays and AR glasses (think Vision Pro). Real-time stats layered over gameplay during broadcasts. 8. The Role of Betting in Driving Innovation How sportsbooks and gambling tech are quietly pushing AI in sports forward. Inside view on how that funding and data are transforming scouting and coaching. 9. Startup Insights Bootstrapping vs. raising capital in deep tech. Hiring elite AI talent without spending $10M+ like Meta—thanks to open systems like Bittensor. 10. The Bigger Vision Creating a universal scoring system for athletes—objective, data-rich, and fair. Challenging legacy scouting reports with measurable intelligence.

Robert Scoble

65,908 Aufrufe • vor 1 Jahr

For generative AI to become an interesting art tool, we need much more control over the output. The slot-machine-like nature of pure text-to-image leaves too much to chance. Using the "Real-time Latent Consistency Model" that I'm using in the example here, is the first time I truly got a glimpse of a future where we'll be able to use our artistic skills and sensibility, to get control over AI image gen. Systems like these will never be able to match the quality or originality of a skilled artist, it won't surprise us in the same way an artist can. Things are a mess in terms of the training data these models are based on, and the questions about copyright concerns and about a time when everything will look the same are very valid. At some point capabilities like these will be embedded in photoshop, and anyone will be able to generate a pretty picture. But to create interesting designs, to tell original stories and to surprise us, we need creatives and artists with something on their mind. We'll be able to create immersive worlds, by making brush-strokes and sculpt marks, without needing to worry about all the dials, plugins, wires of our 3d and 2d tools today. I love to sculpt, I love to draw, and I love to explore new mediums and new ways to create. The Gen AI tools we have today are far from perfect, and things need to be steered in a better direction. For that we need artists to help point the way. Gen AI isn't going away.. it's too powerful and has the potential to allow us to tell stories like never before. Like all other big technological shifts, tech like this will come at a cost, but it will also open up new opportunities and empower a new generation of storytellers. I might be naive, but I believe that human ingenuity and creativity will persevere in this new world ♥️ #art #ai

Martin Nebelong

1,660,623 Aufrufe • vor 2 Jahren

Reinforcement Learning from Human Feedback (RLHF) is gaining traction. This field aims to make AI more responsible by including human values and preferences. In this video, Nathan Lambert, a research scientist and RLHF team lead at Hugging Face explores its inner workings, applications and industry impact. RLHF has gained the spotlight in recent years. The growth of language models like Anthropic’s Claude and OpenAI's ChatGPT have increased interest in human-feedback integration. "There are some rumors that Open AI had two teams; one was doing RLHF and the other instruction fine-tuning. And the RLHF team kept getting more and more performance." Understanding RLHF The RLHF process has three main steps: Pre-training: Much like with GPT models, the journey starts with pre-training on a large corpus of data. This can range from text data, web scrapes, to specialized datasets. Reward Modeling: This is the RLHF counterpart of supervised fine-tuning in large language models. This stage involves creating a reward model that resonates with human values and preferences. RL Optimization: This stage parallels reward modeling and reinforcement learning in traditional AI models. The AI system fine-tunes itself based on the reward model, employing reinforcement learning algorithms for that extra layer of optimization. The Data Challenge Data collection and curation in RLHF closely resemble the challenges you'd encounter in large language model training. Datasets from organizations like OpenAI can serve as a useful foundation. However, the need for high-quality, task-specific data cannot be overstated. Implementing RLHF: A Practical Guide If you’re someone who loves getting hands-on with AI libraries like Hugging Face, implementing RLHF is right way to do. It’s essential to understand its limitations. Think about model stability, over-optimization, and exploration strategies, much like you would when prompt engineering. Ongoing Research and Next Steps While he suggests that some basics figured out, there are layers of complexity that still need to be unraveled: 1. New Benchmarks: How do we measure the effectiveness of RLHF? 2. Preference Modeling: How can the model be made to understand human preferences better? 3. Interpreting RLHF: Much like explainability in traditional models, how do we make RLHF more interpretable? 4. System-Wide Evaluation: Going beyond individual performance, how does RLHF affect an entire system? The Transformative Power of RLHF Whether you're an AI developer, a business analyst, or a marketer, RLHF promises to revolutionize your domain. Imagine customer service chatbots that understand human emotions better, or content generators that align more closely with human values. RLHF is an emerging field that focuses on enhancing machine learning models through human feedback. While it tackles important issues like bias and ethics, its broader goal is to improve system performance across various applications. Whether you're deeply invested in the ethics of AI or simply curious about advancements in machine learning, RLHF offers valuable insights. If you're interested in the next wave of AI development, this area is definitely one to watch.

Muratcan Koylan

27,168 Aufrufe • vor 3 Jahren

New Short Course: Building AI Browser Agents! Learn how to build AI agents that interact and take actions on websites in this course, created in partnership with and taught by and @namangarg0, Co-founders of AGI Inc. AI browser agents can log into websites, fill out forms, click through web pages, or even place orders online for you. They use both visual information, like screenshots, and structural data, like the HTML or Document Object Model (DOM) of a web page, to reason and take action. With the complexity of webpages and multiple possible actions at each step, it can be challenging for an AI browser agent to complete an assigned task. Because these agents run long action sequences, a single error—like clicking the wrong button or misreading a field—can lead to unexpected outcomes or errors that compound over time. In this course, you'll understand how autonomous web agents work, their current limitations, and how AgentQ enables them to improve through self-correction. In detail, you'll: - Learn what web agents are, how they automate tasks online, their architecture, key components, limitations, and an overview of their decision-making strategies. - Build a web agent that can scrape website and return course recommendations in a structured output format. - Build an autonomous web agent that can execute multiple tasks, such as finding and summarizing webpages, filling out a form, and signing up for a newsletter. - Explore AgentQ, a framework that enables agents to self-correct by combining Monte Carlo Tree Search (MCTS), a self-critique mechanism for continuous improvement, and Direct Preference Optimization (DPO). - Deep dive into MCTS, learn how it finds an effective path, illustrated by an example of Gridworld animation, and use AgentQ to complete web tasks. - Understand AI agents' current state and future directions—including key factors shaping their evolution, such as hardware, algorithm innovation, and data availability. By the end of this course, you will have hands-on experience building browser agents and a deeper understanding of how to make them more robust and reliable. Please sign up here:

Andrew Ng

186,182 Aufrufe • vor 1 Jahr

PHOTON COUNTING CT is NOT a better CT It is a NEW imaging modality Photon Counting CT (PCCT) represents a transformative leap in medical imaging, not only as a molecular imaging modality but also as a technology offering ultra-high resolution and functional imaging capabilities. It is fundamentally more than just an enhanced version of traditional CT—PCCT introduces new ways of seeing and understanding the human body, providing critical insights at the molecular, structural, and functional levels. This positions PCCT as a unique imaging modality that requires a fresh approach to technical implementation, operational workflows, and financial planning. Despite the larger upfront investment, PCCT’s ability to drastically reduce downstream healthcare costs makes it a highly valuable investment in the long run. 1. Technical Innovations • Molecular Imaging and Energy Discrimination: Unlike traditional CT, which simply measures the total absorbed energy, PCCT counts individual X-ray photons and differentiates their energy levels. This allows for precise molecular imaging, revealing the composition of tissues and materials at a biochemical level. By distinguishing between different tissue types and contrast agents, PCCT opens up new diagnostic possibilities, such as identifying molecular biomarkers in tumors or distinguishing between stable and unstable plaque in coronary arteries. This capability shifts the focus of imaging from purely anatomical to both anatomical and molecular, offering more comprehensive diagnostic information. • Ultra-High Spatial Resolution: PCCT features significantly smaller detector elements compared to conventional CT scanners, allowing for ultra-high resolution imaging. This means clinicians can visualize fine structures such as microcalcifications in arteries, small lesions in soft tissues, or the intricate architecture of bones. This level of detail was previously unattainable with traditional CT. When combined with molecular imaging, this ultra-high resolution allows for the precise localization and characterization of disease at very early stages, which is essential for early diagnosis and intervention. • Functional Imaging Capabilities: PCCT also excels as a functional imaging modality. By capturing energy-resolved information, PCCT can provide insights into tissue functionality and dynamic physiological processes. For instance, it can detect changes in blood flow, tissue perfusion, and oxygenation without the need for additional contrast agents or scans. This functionality allows for real-time assessment of physiological processes, making it particularly valuable in cardiology, oncology, and neurology for evaluating organ function and monitoring disease progression. • Reduced Noise and Artifact Reduction: Photon-counting technology dramatically reduces electronic noise and imaging artifacts, such as beam hardening, resulting in clearer and more accurate images. The ability to deliver ultra-high resolution images with minimal artifacts improves diagnostic accuracy, reducing the need for repeat scans and ensuring that even subtle abnormalities are detected. 2. Operational Considerations • New Workflow for Molecular, High-Resolution, and Functional Imaging: The integration of molecular, ultra-high resolution, and functional imaging into routine clinical workflows introduces complexity that requires adaptation. Radiologists and technicians need specialized training to interpret and analyze multi-energy datasets that include molecular and functional information. PCCT produces a vast amount of detailed data, requiring clinicians to adopt new imaging protocols and refine their diagnostic approaches to fully leverage its capabilities. • Post-Processing and Data Management: PCCT generates richer, more complex datasets, which necessitates advanced post-processing tools and data management systems. Existing PACS and imaging software may not be equipped to handle such large volumes of data or to process functional and molecular information effectively. This means healthcare institutions must invest in robust IT infrastructure, including upgraded software and storage solutions, as well as provide additional training for staff on new imaging analysis techniques. • Revised Clinical Protocols: The molecular, functional, and ultra-high resolution imaging capabilities of PCCT will likely prompt changes in clinical protocols. For instance, the need for contrast agents may be reduced, simplifying patient preparation and decreasing the risk of adverse reactions. Additionally, the ability to monitor physiological functions in real-time through functional imaging could lead to more dynamic diagnostic procedures, such as assessing the effectiveness of interventions or treatments in real-time. 3. Financial Impact • Higher Initial Investment: PCCT systems are more expensive than traditional CT scanners due to their advanced technology, which includes photon-counting detectors and the computational power required for high-resolution, molecular, and functional imaging. While this upfront cost is significant, it is crucial to view it in the broader context of the downstream benefits and cost reductions that PCCT offers. • Downstream Cost Reductions: Although the initial capital investment is higher, PCCT’s ability to combine molecular, functional, and ultra-high resolution imaging leads to substantial reductions in downstream healthcare costs. Its superior diagnostic accuracy minimizes the need for follow-up tests, repeat scans, or invasive diagnostic procedures, such as diagnostic coronary angiographies. For example, in cardiology, PCCT can precisely differentiate between types of coronary plaque, reducing the need for invasive procedures to assess risk. • Lower Overall Healthcare Expenditures: By enabling earlier, more accurate diagnoses, PCCT can reduce the overall cost of patient care. Early detection of disease, particularly through its molecular and functional imaging capabilities, allows for more targeted treatments, potentially preventing the need for more aggressive and expensive interventions down the line. For instance, early-stage tumor detection via molecular imaging could lead to less invasive treatments, reducing hospital stays and improving patient outcomes, ultimately driving down healthcare costs. • Increased ROI Through Enhanced Patient Outcomes: Over time, the combination of molecular, functional, and ultra-high resolution imaging enhances diagnostic precision, which translates into better patient outcomes. Improved diagnostic accuracy reduces the incidence of unnecessary procedures, minimizes treatment delays, and results in more personalized and effective care. This leads to increased patient satisfaction, better healthcare outcomes, and greater patient throughput—all factors that improve the institution’s return on investment (ROI). • Competitive Advantage and New Revenue Streams: By adopting PCCT, healthcare institutions position themselves at the forefront of advanced imaging technologies. The ability to offer molecular, functional, and ultra-high resolution imaging creates a competitive advantage, attracting more complex and high-value cases. This can boost the institution’s reputation for excellence in diagnostics, leading to increased referrals, new patient populations, and expanded revenue opportunities. Summary Photon Counting CT (PCCT) is not just an evolution of existing CT technology—it is a molecular, ultra-high resolution, and functional imaging modality that fundamentally transforms the diagnostic landscape. Its ability to capture detailed molecular data, visualize minute anatomical structures with ultra-high resolution, and provide real-time functional imaging opens new possibilities for earlier and more precise diagnoses. While the financial investment in PCCT is larger, the reduction in downstream healthcare costs through improved diagnostic accuracy, fewer unnecessary interventions, and earlier disease detection far outweighs the initial expense. For institutions committed to advancing patient care and improving long-term financial outcomes, PCCT is an essential investment in the future of medical imaging. The video attached shows a patient accessing the Hospital for ACS. PCCT can provide ALL the imaging information of the concurrent imaging modalities (CXR, CAG, Echo, CMR) that you see around it... that's a lot! #PhotonCountingCT #MolecularImaging #UltraHighResolution #FunctionalImaging #FutureOfImaging #AdvancedMedicalImaging #EarlyDiseaseDetection #InnovativeCT #CuttingEdgeHealthcare #PrecisionDiagnostics #HealthcareInnovation #MedicalTechnology #CostEffectiveImaging #NextGenCT #PatientCareRevolution

Dr. Filippo Cademartiri

11,849 Aufrufe • vor 1 Jahr

Studies have shown ChatGPT outperforms human annotators for Structured Data by about 25% and costs 30x less. 1 In just 2 months, miners on SN33 running ChatGPT without optimization can’t survive. Today we announce SN33 is now ReadyAI to fully align with our mission 👇 SN33 is building a more performant and significantly cheaper alternative to Scale AI Today structured data is performed primarily by human annotation services like Amazon’s Mechanical Turk and Scale AI It is now more important than ever for every business and individual to make their data AI Ready. However, taking unstructured data and making it Structured Data using today’s tools is extremely costly. SN33 revolutionizes this process, unlocking immense opportunities for commercialization. We lay out the vision for it in this detailed blog post: Validators TODAY can monetize access to this structured data pipeline independently, but we’re streamlining this process, launching a frontend soon that any validator can opt into to provide bandwidth. We've received great feedback from the community, recognizing that what we're building goes far beyond Conversational AI. Building the world's largest annotated conversational dataset (which we've already accomplished) is just one of countless real-world applications for SN33's Structured Data pipeline. We're building a decentralized Scale AI, offering a full suite of Structured Data commodities—from text metadata tagging (available today) to fully customizable queries for company-specific data annotation use cases and image metadata tagging coming soon 👀. Thanks for all the feedback! It has been invaluable so keep bringing it to us! 🙏$TAO Openτensor Foundaτion 1 “ChatGPT Outperforms Crowd-Workers for Text-Annotation Tasks” shows “The zero-shot accuracy of ChatGPT exceeds that of crowd-workers by about 25 percentage points on average [...] Moreover, the per-annotation cost of ChatGPT is less than $0.003—about thirty times cheaper than MTurk”

David Fields

13,648 Aufrufe • vor 2 Jahren

"What is an AI Agent and why do they matter?" An agent is a program that autonomously completes tasks or makes decisions based on data. What do I mean by autonomous? The agent understands task intent, can plan steps to solve the problem, decide and execute and actions and adapt to the environment. Consider how many of us use AI chat interfaces today. You might ask ChatGPT to write an article from start to finish and get a one-shot response. You probably need to do some work to iterate on it yourself. An agentic version is more nuanced - it might write an outline, decide if research is needed, write a draft, evaluate if it needs work and revise itself. Unlike traditional AI models that simply respond to queries, agents are designed to be autonomous and proactive. Think of them as assistants that can not only understand what you need but also take initiative to accomplish tasks by using various tools and making decisions along the way. For example, an AI agent might help a marketing team by not just analyzing campaign data, but actively monitoring performance, adjusting budget allocations, and even drafting social media posts based on real-time engagement metrics. The significance of AI agents lies in their potential to transform how we work. In customer service, agents can handle complex inquiries by accessing multiple databases, processing payments, and updating records - all while maintaining natural conversations with customers. In software development, they can assist programmers by not just suggesting code but actively debugging issues, writing test cases, and even refactoring entire codebases. This level of autonomy and capability represents a fundamental shift from AI as a tool to AI as a collaborative partner. While there remain many unknowns, I'm excited about the potential for agents and we're thinking about how they can help users and developers on the web over in Chrome. The key to success will likely be finding the right balance between human oversight and agent autonomy, ensuring that these powerful tools enhance rather than diminish the human element in business operations.

Addy Osmani

30,412 Aufrufe • vor 1 Jahr

Announcing a new Coursera course: Retrieval Augmented Generation (RAG) You'll learn to build high performance, production-ready RAG systems in this hands-on, in-depth course created by and taught by , experienced AI and ML engineer, researcher, and educator. RAG is a critical component today of many LLM-based applications in customer support, internal company Q&A systems, even many of the leading chatbots that use web search to answer your questions. This course teaches you in-depth how to make RAG work well. LLMs can produce generic or outdated responses, especially when asked specialized questions not covered in its training data. RAG is the most widely used technique for addressing this. It brings in data from new data sources, such as internal documents or recent news, to give the LLM the relevant context to private, recent, or specialized information. This lets it generate more grounded and accurate responses. In this course, you’ll learn to design and implement every part of a RAG system, from retrievers to vector databases to generation to evals. You’ll learn about the fundamental principles behind RAG and how to optimize it at both the component and whole-system levels. As AI evolves, RAG is evolving too. New models can handle longer context windows, reason more effectively, and can be parts of complex agentic workflows. One exciting growth area is Agentic RAG, in which an AI agent at runtime (rather than it being hardcoded at development time) autonomously decides what data to retrieve, and when/how to go deeper. Even with this evolution, access to high-quality data at runtime is essential, which is why RAG is a key part of so many applications. You'll learn via hands-on experiences to: - Build a RAG system with retrieval and prompt augmentation - Compare retrieval methods like BM25, semantic search, and Reciprocal Rank Fusion - Chunk, index, and retrieve documents using a Weaviate vector database and a news dataset - Develop a chatbot, using open-source LLMs hosted by Together AI, for a fictional store that answers product and FAQ questions - Use evals to drive improving reliability, and incorporate multi-modal data RAG is an important foundational technique. Become good at it through this course! Please sign up here:

Andrew Ng

124,656 Aufrufe • vor 1 Jahr

Today we're announcing #GAIA1: a 9B parameter world model, trained on 4,700 hours of driving data, able to simulate complex and diverse driving scenes from video, text and action inputs. This model is 480x larger than the preview we shared earlier this year and the results are incredible. These videos are entirely synthetically generated by Wayve's generative AI, GAIA-1. But there is more here than just generating videos, GAIA is an entire world model. A world model allows us to simulate the future, conditioned on video, text and action inputs, which can be leveraged for making informed decisions when driving. Why is this game-changing for autonomous driving? 1. Safety. One limitation with AI systems like today's Large Language Models is that they are autoregressive, next-word prediction algorithms, but aren't necessarily aware of the implications of their decisions. A world model allows us to give our AI the capability to be aware of its decisions, by simulating the future, which is important for self-driving safety. 2. Synthetic training data. I believe synthetic training data is the future for AI, because it is safer, cheaper, and infinitely scalable. GAIA-1 unlocks unprecedented realism and diversity of synthetic data for self-driving. 3. Long-tail robustness. One of the biggest challenges for self-driving is long-tail robustness: dealing with the enormous magnitude of edge cases we see on the road. An advantage of generative AI is its incredible ability to recombine experiences in new ways. This is exciting for self-driving as it means we can learn from two edge case scenarios, and combine them to become a corner case. For example, we can experience driving in fog, and experience of jay-walking pedestrians, and GAIA can learn from these experiences to understand how to generate a fog+jay walking scenario. Check out many more videos in our blog or further technical details in our paper: Or come chat with our team who are at the International Conference on Computer Vision (#ICCV2023) this week in Paris in Booth 32 Jamie Shotton

Alex Kendall

631,909 Aufrufe • vor 3 Jahren

Chamath Palihapitiya believes AGI may already exist inside leading AI labs and the bigger story is that advanced intelligence is becoming cheaper and more widely available (Save this). Chamath Palihapitiya argues that the public may be focused too much on benchmark rankings, while frontier labs are already developing models capable of complex reasoning, coding, research, and tool use. The main question is how quickly companies will release these systems and how much access they will provide. AGI has not been officially confirmed and strong benchmark results do not necessarily prove that a model can perform every intellectual task like a human. However, AI capabilities are improving quickly, while the cost of running advanced models continues to fall. That combination is important because cheaper AI can be used by more businesses for customer service, software development, research, marketing, financial analysis, and automation. Competition is also accelerating among OpenAI, Anthropic, Google, xAI, Meta, and open source developers because as more companies release capable models, users gain more choices and prices continue to decline. This creates a powerful cycle in which better models attract more users, more usage generates more revenue and data, and lower prices encourage companies to apply AI to additional tasks. The biggest challenge is moving from impressive demonstrations to measurable business results. Companies still need to redesign workflows, train employees, protect sensitive information, and prove that AI spending is producing a real return on investment. AI agents could create the next major increase in demand because they can plan tasks, use tools, check their work, retry failed actions and operate for long periods without constant human supervision. Even if each AI task becomes cheaper, total usage could grow much faster as businesses use models across more departments and this could increase demand for GPUs, high bandwidth memory, networking equipment, electricity, cooling systems, and data centers.

Milk Road AI

13,501 Aufrufe • vor 18 Tagen

#VPDNews: The Vancouver Police Department (VPD) is adding new cutting-edge technologies to help keep the city safe. The new tools enhance frontline officer awareness, strengthen accountability, and build on the Department’s mission to protect public safety while balancing the privacy of both the community and VPD officers. In the air, VPD is the first police agency in Canada to deploy Skydio X10 drones for a Drone as First Responder Program. After extensive testing, six of the remote-piloted drone systems will be deployed. The drones have already been in testing for several weeks and are fully approved by Transport Canada. “The potential of the Skydio drone systems in our work is impressive for many reasons, not least of which is they will be able to link with our body-worn cameras,” said Inspector Wade Rodrigue, with the VPD’s Force Options Training Section. “For example, if an officer is in trouble, perhaps being assaulted, they can tap their camera three times which will automatically deploy a Skydio drone to their exact location at the direction of the pilot in command. Pilots can also fly the drones to a crime in progress, arrive first, and send their video feeds to responding officers on the ground as well as the Operations Command Center (OCC). That gives us better intel on what’s happening and can help responding officers to pursue suspects who may try and evade them.” The VPD will continue to adhere to its posted policyand procedures as well as those prescribed by Transport Canada and Nav Canada with respect to the operation of Remotely Piloted Aircraft Systems (RPAS) assets. The weatherproof Skydio drone launch/landing pods are installed on rooftops at strategic locations throughout Vancouver, including the VPD’s Tactical Training Centre. The drones will only record video when that function is activated by a pilot, and only when appropriate as per policy. Body-worn cameras are also expanding in function. The Axon body cameras used by VPD now have the ability to translate language in real time with Axon Assistant, allowing officers to understand at the push of a button whatever is being said to them in over 50 languages, and to have their reply translated into the language recognized as being spoken. Real-Time Translation adds to VPD’s current translation offerings and will be used when a quick translation is needed. “When you consider how multicultural Vancouver is, this translation ability is a game-changer,” said Sergeant Dermot O’Boyle. “We want to be able to help everyone in our city, including those who may not be fluent in English. Being able to understand what they’re telling us is a critical first step to getting them the help they need.” Body-worn camera footage can also now be live-streamed to the VPD’s Operations Command Centre so personnel there can see what the officer is seeing and dispatch additional resources as needed. This can be done at the officer’s request or based on the priority of the call when an urgent situation is developing. In addition, Axon’s real-time operations platform, Fusus, improves visibility and coordination across responding units and partner agencies by enabling Operations Centre personnel to view RPAS and body-worn camera video according to pre-determined operating protocols. Other new tools entering service include: ➡️ 73 Fleet 3 in-car video systems with automated license plate recognition cameras (ALPR) across the VPD fleet, helping officers spot vehicles of interest faster which are already proving effective, with one ALPR-equipped cruiser flagging 22 uninsured vehicles in just three hours ➡️ Holsters for conducted energy weapons and service weapons that automatically activate body-worn cameras when drawn, capturing critical moments right away “Combined, these technologies create a system that helps improve decision-making, response times, and overall public safety in Vancouver,” said Kevin Bernardin, Superintendent of Innovation and Technology at the VPD. All are designed with responsible AI and data use at their core, with safeguards for secure data handling, controlled access, and auditability. The AI does not make decisions about an individual, rather it gives the VPD the ability to respond more appropriately to emerging situations. Data is managed in alignment with local governance requirements, helping ensure it remains protected and under appropriate jurisdictional control, while giving VPD confidence that sensitive information is handled in accordance with British Columbia privacy standards. #VPD Drone Policy:

Vancouver Police

11,219 Aufrufe • vor 3 Monaten

🚨 BREAKING: ABB Robotics + NVIDIA close the sim-to-real gap with 99% accuracy! 👾 ABB Robotics is integrating NVIDIA Omniverse libraries into RobotStudio to deliver physical AI for industry, closing the gap from virtual training to real-world deployment with up to 99% accuracy. RobotStudio HyperReality, available second half of 2026, will fundamentally change how quickly manufacturers can scale production: reducing costs by up to 40%, accelerating time-to-market by 50%, and cutting setup and commissioning times by up to 80%. For decades, the deficit between simulation accuracy and real-world lighting, materials, and environments has limited manufacturers' ability to design advanced manufacturing processes in the virtual world. The only robot manufacturer with a virtual controller running the same firmware as the hardware, ensuring near-perfect correlation between simulation and real-world performance. The system uses physically accurate simulations and foundation models endlessly optimized with real-world data feedback. These models can train any number of ABB robots anywhere in the world with industrial-grade reliability. Foxconn is using RobotStudio HyperReality for consumer electronics assembly. Assembly robots are trained virtually using synthetic data to perfect multiple production processes across various scenarios, then moved to production lines with 99% accuracy. This eliminates physical training and tests, reducing setup times and costs. Workr is demonstrating AI-powered robotic systems at NVIDIA GTC 2026. Built on ABB technology, trained with synthetic data using NVIDIA Omniverse, deployed without operators needing programming knowledge . 🚨 I’ll be onsite in San Jose during GTC 2026, and will be showing all the cool stuff that ABB Robotics prepared this year! Can’t wait! 🫡 ~~ ♻️ Join the weekly robotics newsletter, and never miss any news →

Lukas Ziegler

22,482 Aufrufe • vor 6 Monaten