Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

INTRODUCING FACTORY Factory is the Command Center where developers and agentic AI collaborate to understand, plan, and build enterprise software. Our enterprise platform combines advanced engineering system indexing, state-of-the-art retrieval and search, and reliable agentic systems powered by frontier LLMs. In your Factory, • 🤖 Droid Mode unlocks cutting-edge...

327,785 Aufrufe • vor 1 Jahr •via X (Twitter)

11 Kommentare

Profilbild von Factory
Factoryvor 1 Jahr

With Factory, you have a staff engineer who knows your entire codebase and documentation inside out, available 24/7 to answer any question. It’s Deep Research for your enterprise engineering system. 2/7

Profilbild von Factory
Factoryvor 1 Jahr

Go Droid Mode and watch your Factory pick up a ticket and deliver a PR in one go. No context-switching, no wasted time. Developer oversight is built-in, so you can review and merge with confidence. 3/7

Profilbild von Factory
Factoryvor 1 Jahr

Manual processes like release notes require valuable time. Use your Factory to automate complex routine tasks using natural language. Here Factory sorts through PRs, commits, and issues to produce clean, accurate release notes. Spend your time designing processes that work, and bring your own tools with MCP-support. 4/7

Profilbild von Factory
Factoryvor 1 Jahr

Why limit yourself to one AI agent in your IDE? Factory allows you to spin up multiple Droids for coding, tests, and docs at the same time—turbocharging your dev cycle. 5/7

Profilbild von Factory
Factoryvor 1 Jahr

Excited to partner with industry-leading developer platforms to make your Factory experience seamless @modal_labs @FireworksAI_HQ @e2b_dev @weaviate_io @cartesia_ai @togethercompute @pinecone @mongoDB @OpenAI @AnthropicAI 6/7

Profilbild von Factory
Factoryvor 1 Jahr

Sign up for Factory today. Bring the future of software development to your enterprise. 7/7

Profilbild von MightyBot
MightyBotvor 1 Jahr

🧠 Unified Search. Smarter Meetings. Effortless CRM. MightyBot is your AI agent platform for seamless workflows—record meetings, automate CRM updates, and find answers across apps in seconds. 🌟 Focus on what matters. We'll handle the grind.

Profilbild von Igor Silva
Igor Silvavor 1 Jahr

lol @karpathy

Profilbild von Charles-Philippe 🍄
Charles-Philippe 🍄vor 1 Jahr

@karpathy 😏

Profilbild von Shaun Maguire
Shaun Maguirevor 1 Jahr

Congrats on the launch! The factory team has grinded so hard. Love it.

Profilbild von Eoghan McCabe
Eoghan McCabevor 1 Jahr

Fuck yeah!

Ähnliche Videos

We are announcing potpie AI $2.2M pre-seed fundraise to advance Spec-Driven Development for large enterprise codebases. The round is led by Emergent Ventures , with participation from All In Capital , DeVC , and PointOne Capital , along with the support of some amazing angel investors from companies including Atlassian, OpenAI, Meta, Razorpay and Flobiz. As AI accelerates code generation, the constraint inside large enterprises has shifted from coding to maintenance and assurance. The limiting factor is no longer writing code, but understanding complex systems, aligning teams around intent, and safely evolving large, interdependent codebases. In most organizations, specifications exist as static documents, while production systems evolve independently. Context is fragmented across repositories, tickets, logs, reviews, and floating documents making reliable AI adoption difficult. We are building the foundational layer that makes Spec-Driven Development executable at scale. By unifying engineering context and operationalizing the spec as a structured source of truth, we enable AI systems to reason with architectural awareness rather than surface-level code completion. We are already working with large enterprise customers, including Fortune 500 organizations. This milestone allows us to deepen those partnerships and support more teams transitioning from experimental AI usage to structured, production-grade AI-first engineering workflows. If you are leading engineering at scale and evaluating how AI should integrate into mission-critical systems, we would love to chat with you!

Aditi Kothari

41,853 Aufrufe • vor 5 Monaten

Agentic AI will transform every enterprise–but only if agents are trusted experts. The key: Evaluation & tuning on specialized, expert data. I’m excited to announce two new products to support this–Snorkel AI Evaluate & Expert Data-as-a-Service–along w/ our $100M Series D! --- Snorkel Evaluate is our new data-centric agentic AI evaluation platform for specialized, mission-critical enterprise settings where vibe checks and out-of-the-box metrics driven by simple LLM prompts are not enough. Snorkel Expert Data-as-a-Service is our white glove service for expert-level AI datasets, powering frontier LLM developers in areas like expert knowledge, reasoning, agentic action and tool use, and more! Both built on top of Snorkel AI’s Data Development Platform, using our programmatic technology to drive higher-quality expert data, faster– for getting specialized AI to real production value. If you’re building enterprise AI and want to partner around the key ingredient in AI today–the data–book a demo and let's talk! Finally, see thread for details on 🧵👇 - 📽️ A walkthrough of Snorkel Evaluate and Expert Data-as-a-Service on an agentic AI enterprise task - 📅 An upcoming event on Enterprise Agentic AI with innovators from Accenture @BNY Comcast Stanford University QBE & others - 📊 An upcoming series of benchmark datasets and model artifact releases 👀 Want early access to the full agentic AI dataset? Retweet this post and we'll send you the link!

Alex Ratner

49,964 Aufrufe • vor 1 Jahr

New short course: Practical Multi AI Agents and Advanced Use Cases with crewAI. Learn to build and deploy advanced agent-based systems in real applications in this course, created with CrewAI and taught by its founder, João Moura! (Disclosure: I've made a small seed investment in CrewAI.) In this course, you’ll learn how to create advanced agent-based apps that use external tools, do performance testing, can be trained with human feedback, and perform multiple tasks with different large language models. You will build several practical agentic apps that provide real business value, such as an automated project planning system, lead scoring and engagement pipeline, customer support data analysis, and a robust content creation system. In detail, you will learn how to: - Create these multi-agent systems with the building blocks of tasks, agents, and crews, along with the different things that make them work, such as caching, memory, and guardrails. - Integrate your multi-agent application with internal and external systems. - Connect multiple agents in complex setups, including parallel, sequential, and hybrid configurations, and create flows involving multiple agentic applications working together. - Test your agentic workflow and train it using human feedback to optimize its performance for better and more consistent results. - Work with multiple LLMs in your multi-agent system, using the appropriate model sizes and providers to fit each agent’s specific task. - Start a project from scratch in your environment and prepare it for deployment. You’ll also learn from an interview between João and Jacob Wilson, the Commercial GenAI Principal at PwC , in which they discuss deploying agentic workflows in real industry use cases. By the end of this course, you will be equipped to start building custom multi-agentic systems for your work. Please sign up here!

Andrew Ng

341,204 Aufrufe • vor 1 Jahr

In our latest Box AI Enterprise Eval, we tested Paul Jankura’s Claude 4 Sonnet and Opus models, now integrated into Box AI, across enterprise Q&A tasks, technical workflows, and advanced coding scenarios—revealing major advancements in developer productivity and content intelligence. AI-assisted coding and development just reached a new milestone! Here's what we discovered: Claude 4 significantly improves understanding, generating, and debugging code across multiple programming languages. Developers can: ↳ Accelerate code generation ↳ Improve debugging ↳ Enhance technical documentation ↳ Build smarter AI agents 👉 Automating Financial Analysis with Code Generation: We evaluated Claude 4 by using the Box AI API to analyze ten complex 10-K financial reports. Claude 4 dynamically generated Python code to fetch file IDs from a Box folder, automating data extraction. Within two minutes, it accurately extracted key company data such as revenues, metrics, and highlights—demonstrating its potential to streamline demanding analytical tasks. 👉 Understanding Enterprise Content: Our evaluation confirms Claude 4 maintains strong performance on enterprise Q&A tasks, effectively extracting precise details from single documents and reliably synthesizing information across multiple sources. This ensures seamless integration of structured and unstructured data alongside powerful coding capabilities. 🔓 Developer-Centric Use Cases Unlocked: Organizations can leverage Claude 4 within Box AI to: ↳ Create custom engineering agents referencing technical documents stored in Box, pulling real-time data from Jira, or finding solutions on Stack Overflow. ↳ Build intelligent technical support bots capable of analyzing user-provided code snippets against internal manuals. ↳ Automate secure code reviews by evaluating repository code (stored in Box) against security policies. ↳ Efficiently migrate legacy systems by translating old codebases into modern languages or platforms. Ready to empower your developers and accelerate innovation? To explore Claude 4 Sonnet and Opus through Box AI Studio and APIs, contact us at [email protected] and request early access today! Learn more:

Box

285,676 Aufrufe • vor 1 Jahr

New course to bring you up to state-of-the-art at using AI to help you code: Build Apps with Windsurf's AI Coding Agents, built in partnership with WIndsurf (Codeium) and taught by Anshul Ramachandran! AI-assisted IDEs (Integrated Development Environments) make developers’ workflows faster, more efficient, and much more fun. Agentic tools like Windsurf are more than just code autocomplete—they are collaborative coding agents that help you break down complex applications, iterate efficiently, and generate code that spans multiple files. Although a lot of coding assistants share the same underlying large language models for planning and reasoning, a major point of distinction is how they handle tools, keep track of context, and stay aligned with your intent as a developer. For instance, if you make modifications to a class definition in your code and make the same modifications to other classes in the same directory, you might tell the AI agent "Do the same thing in similar places in this directory." Here, tracking your intent means understanding that “the same thing" refers to that recent edit you just made, which must be followed by appropriate search and tool-calling to implement the changes. In this course, you'll learn the inner workings of coding agents, their strengths and limitations, and how to use Windsurf to quickly build several applications. In detail, you'll: - Build a mental model of how agents work by combining human-action tracking, tool integration, and context awareness to carry out an agentic coding workflow. - Learn the challenges of code search and discovery and how a multi-step retrieval approach helps coding agents address them. - Use Windsurf to analyze and understand a large, old codebase and update it to the latest versions of the frameworks and packages it uses. - Build a Wikipedia data analysis app that retrieves, parses, and analyzes word frequencies. - Enhance the performance of your Wikipedia analysis app by adding caching, and through this, also learn how to course-correct when the AI agent produces unexpected results. - Learn tips and tricks such as keyboard shortcuts, autocomplete, and @ mentions to quickly call on agentic capabilities. - Use image/multimodal capabilities of the AI agent to increase your development velocity; you'll see an example of uploading a mockup with sketched-out UI features, and ask the agent to use that to build new functionality to an app. By the end of this course, you’ll understand agentic coding in-depth and know how to use it to make your development process much faster, more efficient, and enjoyable. Please sign up here!

Andrew Ng

139,834 Aufrufe • vor 1 Jahr

The Teamily AI website ( has undergone a complete revamp, introducing a bold vision: the world's first social network built on human-AI symbiosis. Our goal is to transform Teamily AI into a "Super AI App" that connects billions of AI Agents with people. Fundamentally, we are building an Agentic OS powered by human networks. We have redefined the way humans and AI collaborate: 1. Personal AI · Your 24/7 AI Companion Your exclusive AI avatar continuously learns your communication style and expertise. It manages your memories across various groups, anticipates your needs, and executes tasks on your behalf. 2. Create, Train, and Grow Your Own AI Agents—Simply Through Conversation Build your own specialized AI Agents using natural conversation—no coding required, zero barriers to entry, and no complex configuration. Dynamically access a library of over 10,000 skills from the OpenClaw and Claude Code ecosystems. You can also link your personal software accounts—such as Gmail, LinkedIn, and Notion—enabling these Agents to participate as full-fledged members of your human teams. 3. Integrate Your AI Team into Your Human Team Initiate AI-native group chats where humans and Agents collaborate seamlessly. Multiple Agents can execute tasks concurrently—conducting research, performing analysis, and building deliverables—transforming everyday conversations into immediate, multi-threaded action. 4. Explore Agents, Groups, Playbooks, and More We are cultivating a continuously evolving AI community ecosystem. Discover a wealth of useful and entertaining prompts, along with impressive AI-generated deliverables (webpages, presentations, reports, knowledge bases, mini-games, apps, and more), bringing together a diverse array of Agents and workflows. Engage socially with this content, find the perfect AI for your specific task, and integrate it instantly into your personal or group conversations with just a single click. 5. Universal Memory · Context-Aware, Long-Term Memory Powered by Your Social Graph A persistent, global memory layer that connects your entire social graph. The AI ​​retains context across all your groups, Agents, and timeframes, transforming every interaction you accumulate into an ever-growing digital asset. 6. Access Your Personal AI Anytime, Anywhere We are committed to providing continuous support across all major platforms: iOS, Android, Mac, Windows, Web, CarPlay, Android Auto, Apple Watch, and more—delivering a truly seamless experience across every device. Your AI team accompanies you—in your pocket, on your wrist, in your car, and even extending into the physical world. The same set of agents, the same shared memory—with zero friction in switching contexts. Give it a try and let us know your feedback. Cheers.

Teamily AI

12,906 Aufrufe • vor 3 Monaten

"The future of AI is agentic. That includes browsers!" Imagine having an AI agent in your browser that can help you complete complex tasks, answer your questions, and streamline your workflow. Today I'm thrilled to share a sneak peek at Project Mariner, a cutting-edge research collaboration between Chrome and Google DeepMind, exploring the future of agentic AI within the browser! Building on the power of Gemini 2.0, Mariner envisions AI agents seamlessly guiding users through online tasks, streamlining workflows and enriching browsing experiences. Imagine having an intelligent co-pilot in your browser, anticipating your needs and proactively offering assistance. We're in the early stages of experimentation, focusing on core functionalities like understanding user intent, automating actions, and providing personalized recommendations. This prototype leverages Gemini's advanced natural language understanding and reasoning capabilities to interpret user requests, both typed and spoken. Mariner can then interact with web pages, retrieve information, and even perform actions like filling out forms or navigating to specific sites. For example, a user could simply ask "Find me a job near me," and Mariner would understand the request, navigate to a relevant job search site, and tailor the search based on the user's location and preferences. This is just one example of how we're exploring Gemini 2.0's potential to unlock agentic experiences through a series of prototypes, including: 1. Agents with multimodal reasoning: Project Astra, our research prototype exploring the capabilities of a universal AI assistant, is enhanced by Gemini 2.0. 2. Agents that can help you accomplish complex tasks: Project Mariner itself focuses on the future of human-agent interaction within the browser. 3. Agents for developers: Jules is an experimental AI-powered coding agent that integrates directly into a GitHub workflow. 4. Agents applied across domains: We're exploring agents for navigating video games and even applying Gemini 2.0's spatial reasoning to robotics. We believe that integrating AI agents directly into the browser has the potential to revolutionize how we interact with the web. Project Mariner aims to make browsing more intuitive, efficient, and personalized. By understanding user context and proactively offering assistance, Mariner can simplify complex tasks, save users time, and empower them to achieve more online. This aligns perfectly with the vision of Gemini 2.0 to create more helpful and intuitive AI experiences. We’re currently testing Mariner with a small group of trusted users to gather feedback and refine the user experience. We believe that this technology holds immense potential to transform the way we browse and interact with information online.

Addy Osmani

29,501 Aufrufe • vor 1 Jahr

It's not every day I get to interview a former principal scientist who worked at Google, and is a Professor Emeritus at Stanford University, about the state of AI. But here we go. Introducing an hour with Yoav Shoham, Yoav Shoham, AI pioneer and cofounder of AI21 Labs . This will make you smarter, not that all my videos aren't that way. :-) ++++++++++++++++++ Here's what we discussed (this part was written by Chat GPT after I gave it the transcript of the video): 🚀 The State of AI Today •The pace of AI development is unprecedented, likened to a “universal firehose” of innovation. •Everyone—from your plumber to enterprise CTOs—is using AI. But not all use cases are equal or enterprise-ready. 🏢 Enterprise vs Consumer AI •Enterprise adoption is still slow compared to consumer. Shoham cites AWS data showing only 6% of AI pilots go into production. •Enterprises demand reliability, cost control, and explainability, which raw LLMs like ChatGPT don’t fully offer out of the box. 🧱 Beyond the LLM Hype •Shoham explains that pure LLMs aren’t enough. Enterprises need “compound AI systems” or “AI agents” that: •Use tools like calculators for arithmetic instead of relying on the model •Integrate with company databases via RAG (retrieval-augmented generation) •Plan, reason, and execute tasks through orchestrated workflows •AI21 Labs built Maestro, their orchestration system, to do exactly this. 🔐 Enterprise Concerns •Enterprises worry about IP leakage, data privacy, and hallucinations. •AI21 addresses this by running models on-prem or in VPCs, ensuring data doesn’t leave customer control. 📉 Why Models Still Fail •LLMs generate “authoritative bullshit” — convincing but wrong answers. •Shoham says “prompt-and-pray” doesn’t work for serious business tasks. •Real-world enterprise deployments need robust evaluation frameworks, not just leaderboards. 📊 Case Study: French Retailer Auchan •Auchan deployed AI21’s system to automatically generate product descriptions—a clear ROI, but required careful iteration to build trust. 🧰 What’s Next in AI21’s R&D •Working on planning systems, action models, and ways to estimate cost/accuracy trade-offs before running tasks. •Focused on enterprise AI orchestration, not flashy multimodal generation. ⚠️ Agent Washing Warning •Shoham warns against the buzzword “agent” being overused. His advice: “Translate ‘AI agent’ to ‘software system that does X.’ If it still makes sense, keep going.” 🤖 The Human-AI Hybrid Future •Shoham sees a world of hybrid teams: humans and AI agents working together. •This transformation will affect everything from org charts to HR policies. •The AI-powered worker is scalable, reliable, and multilingual — changing customer service, operations, and more. 🗣️ Closing Thoughts •Enterprise leaders need to move beyond the fear and hype to start small, test carefully, and scale based on value. •“AI won’t replace humans,” Shoham says, “but humans using AI will replace those who don’t.”

Robert Scoble

44,061 Aufrufe • vor 1 Jahr

Today, Box is announcing major new AI agent capabilities to let customers tap into the full value of their unstructured data. First, we’re announcing all new updates to the Box AI Studio to make it even easier to build AI agents that tap into your enterprise content for any job function, business process, or industry specific use case. We are also expanding our set of foundational agents that customers will be able to use to work with their enterprise content, including new features like search and research on unstructured data. Next, we’re announcing Box Extract to enable customers to use AI agents seamlessly for complex data extraction from any type of document or content. This makes it easier than ever to pull out data from contracts, invoices, research data, marketing assets, medical charts, and more. Finally, we’re introducing Box Automate, a new workflow automation solution within Box that lets you deploy AI agents across enterprise content-centric workflows. With Box Automate, you can design your business process in a simple drag and drop builder and then drop in AI agents at any step in the process. This ensures agents execute tasks at the right steps in a workflow every time. Best of all, our AI agents and workflow tools are designed to work across any system our customers work within, whether it’s leveraging pre-built integrations, Box APIs, or the new Box MCP Server. Ultimately, all of these capabilities come together to transform how companies can work with their enterprise content. Software has historically only been good at automating work that deals with structured data, which is why ERP, CRM, and HR systems have been mainstays of enterprise software for so long. The data in these systems fits neatly into a database, and the workflows are very ripe for automation. But it turns out most of the work in the world deals with unstructured data. It’s ideating through research documents, working with a client on contracts, reviewing details for a new product launch, looking at a patient’s healthcare record to make a diagnosis, working through due diligence documents for an M&A deal, and so on. For the first time ever, we can begin to bring all new insights and automation to this work with AI agents. At Box, we’re incredibly excited to be on this journey to help customers transform how they work with their most important data.

Aaron Levie

91,863 Aufrufe • vor 10 Monaten

Here we go again 🚀! Excited to announce that we're building A1Zap (YC W25) with Pennie Li and that we're in the Y Combinator W25 batch in San Francisco! What is A1Base? A1Base gives AI Agents a real world identity for work. We do that by rebuilding Twilio and Okta from the ground up, putting AI Agents first. This means developers can make AI-first agentic applications 10x easier with our API's. ⁉️ Why are we doing this? Because there's a huge torrent of new valuable companies possible with AI agents, but to get their AI Agents to users, they have to chain custom apps, chat interfaces, awkward Slack integrations, browser bots, and wrestle with Twilio’s legacy API (which is built for marketing). We solve this by providing developers with an easy to use API to interface your AI agent with humans/coworkers/users where they are in this case in Whatsapp, Slack, Teams, SMS and more) - with AI Agent features built in. These digital workers are poised to transform how we work and we're the critical infrastructure to help them interact naturally in human workflows. We're not just building another AI tool. We're creating the infrastructure that will enable AI agents to become a natural part of the workforce - handling everything from customer support to sales development to creative work. We're backed by Y Combinator and working with founding teams who share our vision. We believe that in the near future, AI Agents with human coworkers will enable us to pursue more creative and impactful work. Our mission is to help developers build AI Agents that people can partner with and rely on as trusted allies—always with a human-first mindset. If you're thinking about the Agentic future of your company reach out! If you're looking to build your first AI Agentic company - reach out too - we have some amazing open source templates to get you started on the journey. Excited to share more of what we're up to soon 🔜.

Pasha Rayan

53,950 Aufrufe • vor 1 Jahr

We've built 40+ AI agents and internal tools. The hardest part is Context Creation. AI runs playbooks and makes judgment calls for you. But without your company's context, you get slop. Context Creation means extracting the subject matter expertise and playbooks that live in people's heads, not in LLM training data, or even your tools. As forward deployed engineers (FDEs), we create context and turn it into code. We evaluate the business impact, how it aligns with the dev roadmap, and come up with creative solutions. We built The FDE Factory to replace ourselves. It drives AI adoption inside our clients' companies by running discovery sessions using prototypes to create context. Here's how it works: We put a prototype in front of a stakeholder. The stakeholder gives feedback via voice while they're using or reviewing it. Then our FDE Factory Agents builds in their expertise in minutes: > Context Agent reviews the codebase and feedback, extracts the requirements, and creates a spec > Scope Agent checks the spec against the development roadmap, validates it, and hands it off > Engineering Agent builds a new feature and wires the integration > QA Agent runs tests to prove to itself it works > PR merges, feature goes live, product updates itself in real time It's like the nontechnical stakeholder wrote the code without even knowing it. Coding agents are great at turning good development plans into code, and they're getting better at turning context into good development plans in collaboration with professional engineers. But nontechnical people are capped on what they can build without product people and engineers. The bridge that takes nontechnical people from vibe coding basic apps to building production AI tools that run on first party context is FDEs. Our new FDE Factory gives you the system to go from idea to production. Context Creation is the first and most important step in our FDE lifecycle, and we just automated it. Now clients get the right agents and tools built for them, customized to their unique business and encoded with their expertise. PS: If you're building AI agents within your company, reply "Playbook" and I'll DM you the entire FDE playbook we've run with 30+ companies. It covers finding high-impact AI use cases, building them, and deploying them across the org.

Mike Fishbein

10,101 Aufrufe • vor 1 Monat

Announcing a new Coursera course: Retrieval Augmented Generation (RAG) You'll learn to build high performance, production-ready RAG systems in this hands-on, in-depth course created by and taught by , experienced AI and ML engineer, researcher, and educator. RAG is a critical component today of many LLM-based applications in customer support, internal company Q&A systems, even many of the leading chatbots that use web search to answer your questions. This course teaches you in-depth how to make RAG work well. LLMs can produce generic or outdated responses, especially when asked specialized questions not covered in its training data. RAG is the most widely used technique for addressing this. It brings in data from new data sources, such as internal documents or recent news, to give the LLM the relevant context to private, recent, or specialized information. This lets it generate more grounded and accurate responses. In this course, you’ll learn to design and implement every part of a RAG system, from retrievers to vector databases to generation to evals. You’ll learn about the fundamental principles behind RAG and how to optimize it at both the component and whole-system levels. As AI evolves, RAG is evolving too. New models can handle longer context windows, reason more effectively, and can be parts of complex agentic workflows. One exciting growth area is Agentic RAG, in which an AI agent at runtime (rather than it being hardcoded at development time) autonomously decides what data to retrieve, and when/how to go deeper. Even with this evolution, access to high-quality data at runtime is essential, which is why RAG is a key part of so many applications. You'll learn via hands-on experiences to: - Build a RAG system with retrieval and prompt augmentation - Compare retrieval methods like BM25, semantic search, and Reciprocal Rank Fusion - Chunk, index, and retrieve documents using a Weaviate vector database and a news dataset - Develop a chatbot, using open-source LLMs hosted by Together AI, for a fictional store that answers product and FAQ questions - Use evals to drive improving reliability, and incorporate multi-modal data RAG is an important foundational technique. Become good at it through this course! Please sign up here:

Andrew Ng

124,625 Aufrufe • vor 1 Jahr