Video wird geladen...

Video konnte nicht geladen werden

Zur Startseite

your team put the deploy skill in the repo so everyone would run the same steps it works, and on your laptop it is not the one running the order is enterprise first, then personal, then project a folder of your own with the same name wins, quietly, on...

12,609 Aufrufe • vor 13 Tagen •via X (Twitter)

9 Kommentare

Profilbild von Mempool in Plain English
Mempool in Plain Englishvor 12 Tagen

config precedence bugs are the quiet ones. same-name dir in cwd wins over repo skill and nothing logs it. spent an afternoon on this once, thought the script was broken

Profilbild von ⊹ Sofia l,
⊹ Sofia l,vor 12 Tagen

Silent local overrides defeat the reproducibility argument. How do you prevent drift between machines when the same name folder can quietly change behavior?

Profilbild von 努力配送每一天
努力配送每一天vor 12 Tagen

morning street whole day makes can accommodate you win all

Profilbild von BG翻70小九
BG翻70小九vor 12 Tagen

本地配置优先级反噬真的是团队噩梦

Profilbild von Bitget泛70西洲
Bitget泛70西洲vor 12 Tagen

本地环境隐形陷阱太坑了

Profilbild von 凡70 BG阿宁
凡70 BG阿宁vor 12 Tagen

这种本地配置差异最坑了简直是开发者的噩梦

Profilbild von 范70|BG
范70|BGvor 12 Tagen

本地覆盖生产配置真是团队协作的噩梦

Profilbild von 小栖欧易50凡佣
小栖欧易50凡佣vor 12 Tagen

这简直是埋在环境里的隐形地雷

Profilbild von 南风高返85 Gate
南风高返85 Gatevor 12 Tagen

这种本地覆盖逻辑简直是团队协作的噩梦

Ähnliche Videos

The number one question I get in the Claude Code / Cowork Community: "how do I share my Cowork skills with my team?" Here's the problem. You build a great skill. You zip it up. You drop it in Slack. Your teammate downloads it, uploads it, and maybe it works. Maybe they upload it wrong. Maybe you update the skill next week and nobody gets the new version. You're now maintaining skills through chat messages and hoping for the best. That doesn't scale. I just put out a video breaking down the three methods I've tested for sharing skills and plugins across a team. From dead simple to fully synced. Method 1: Shared drive (Google Drive, SharePoint, etc). You put your skill files in a shared folder. Teammates download and upload them into Cowork. It works, but updates are manual and there's no version control. Method 2: Built-in sharing on Team and Enterprise plans. You can share any skill directly with a colleague or publish it to your org directory. When you update the skill, everyone gets the update automatically. This is the easiest path if you're on a paid plan. The catch: there's no approval workflow for org-wide sharing, so set a clear owner. Method 3: GitHub repo. This is what I use. Your entire Cowork workspace -- skills, plugins, claude.md, folder structure, project files -- lives in a private repo. Teammates clone it. When you push an update, they pull it. Everyone stays in sync. You get version history, access control, and a single source of truth. The GitHub method sounds technical, but it's really just two steps: clone the repo, point Cowork at the folder. I walk through the whole thing in the video, including how to use .gitignore to keep personal files (like your morning briefing) out of the shared repo. This works for Cowork, Claude Code, and Open Codex. The infrastructure is the same. Full video linked below. If you've found a different approach that works for your team, I want to hear about it. Comment or reply and let's figure out the best practices together.

JJ Englert

16,176 Aufrufe • vor 6 Monaten

your agent reviewing its own work is not a check. it is a second opinion from the same source. this is the most common gap in agent systems and it hides in plain sight, because the step exists. there is a review. it just cannot do the thing you think it does. here is the mechanism. the model produced an output from a context. you then ask the same model, holding the same context, whether that output is correct. it answers fluently, because that is what it does. and the answer is drawn from the same distribution that produced the thing being judged. same weights, same window, same blind spots. if the reason the output is wrong is something the model does not know, the review does not know it either. if the reason is something the context does not contain, the review has the same context. the failure mode and the detector share a cause. > why it feels like it works because most of the time the output is fine, and the review says fine. agreement is not evidence of detection. a reviewer that says pass on everything agrees with reality most of the time too. what you actually want to measure is what happens on the cases that are wrong. that is the only place a check earns its name, and it is exactly the place where a self-review is weakest. there is research on this. Huang and colleagues at DeepMind showed at ICLR 2024 that intrinsic self-correction, revising without external grounding, does not reliably help and often makes things worse. > what to actually do move the check outside the model. a test that runs, a schema that validates, a file that exists or does not, an exit code from something you did not write. these are not smarter than the model. they are just not correlated with it, and that is the entire value. when the judgement genuinely needs a model, at minimum use a different family. same family means shared blind spots, and frontier judges measurably inflate scores for outputs that look like their own. and split the work by kind. anything objectively checkable goes to code. only the genuinely semantic calls go to a judge, and those get a rubric written as one line. a review inside the loop tells you the model is confident. a check outside it tells you whether the work is done. save this - then read the eval setup below

Hanako

14,325 Aufrufe • vor 2 Monaten

Hermes agent just left the terminal. 𝗛𝗲𝗿𝗺𝗲𝘀 𝗗𝗲𝘀𝗸𝘁𝗼𝗽 dropped yesterday. native app for macOS, Windows, and Linux. for months Hermes was the agent that learned your projects, wrote its own skills, and built a model of who you are. all of it buried in terminal logs. now it has a window. the important part is that it's not a wrapper. it runs the same agent core, the same sessions, memory, and skills as the CLI. you can start a task in the terminal and finish it in the app without anything resetting. the state is shared across every interface, not copied between them. what the GUI actually adds: → streaming chat that shows live tool calls and inline reasoning instead of a spinner → a preview rail that renders pages, code, and images right beside the conversation → an artifacts panel that collects every file the agent has ever produced → remote gateway mode, so you can point the app at a VPS and run the heavy work elsewhere → skills, cron, profiles, and gateways managed point-and-click instead of through YAML → voice mode, drag-drop files, and inline image generation remote gateway mode is the one worth slowing down on. the agent runs 24/7 on a $5 server while you control it from your laptop like a local app. other agent UIs are chatboxes with a logo. this one shows the autonomy instead of hiding it, so you watch the skills load, the tools fire, and the artifacts pile up as it works. it was teased in Jensen's GTC keynote. MIT licensed, local-first, no telemetry. if you already run Hermes, download it and everything is already there. your chats, memory, and skills carry straight over. i wrote a full masterclass on Hermes Agent that walks through the SOUL. md identity layer, the three-tier memory system, the self-evolving skills loop, and how to run three specialized agents 24/7. desktop is the interface that finally does all of it justice. the article is quoted below.

Akshay 🚀

51,540 Aufrufe • vor 4 Monaten

BlackRock runs on 20,000 people. Elon's Grok Bot runs the same shape for $300 a month, and it hires its own staff. You do not get an assistant. You get a company that hires. It does not throw ten agents at your problem and hand you the pile. It makes one agent that makes 10, and those ten make a 100. > LAYER ONE is one agent, the chief of staff, and it never touches the market > LAYER TWO is six desk heads, one job each, every one on its own computer with its own logins > LAYER THREE is whatever those six decide they need, spun up on the spot and shut down when the work is done Nobody writes a task list. You hand out job titles and the org fills itself in underneath. The swarm is never the same twice. Agents get spun up for one job, finish it, and are gone before I ever read their names. Not one of them sees the whole picture. The answer only exists after they hand off to each other. Wall Street cannot copy that. You cannot hire a hundred people for eleven minutes. BlackRock holds that shape together with a risk system called Aladdin. Mine holds it together with one agent that is only allowed to say no. I gave it $1,000 and told it to grow the money or get deleted. 15 hours later it was holding $3,900, on an address anyone can open and read. I was asleep for most of it, and I have still not written a line of code. The whole thing runs with my laptop shut, because none of it lives on my laptop. Setup is one evening. Create the chief, hand out the titles, run one trade on your screen while they watch, connect Telegram. Ten years ago a machine this shape had its name on a tower. Mine has a name I typed into a box. Save this while the whole thing still fits on one screen.

cvxv666

45,488 Aufrufe • vor 1 Monat

Anthropic just got outplayed again. Devs built the multiplayer assistant Anthropic couldn't, and open-sourced it. Claude Cowork is a solo desktop agent. You point it at a folder, give it a task, and it works through your local files on your own machine. The moment a teammate enters the picture, it has nothing to offer. Most real work does not happen alone. A teammate asks for a status update on something you own. The context they need is scattered across your meetings, your notes, and decisions made last week. Typing all of that out takes time you do not have. This is the gap Claude Cowork was never designed to cross. Rowboat Spaces is built on a different model entirely. Each person brings their own assistant into a shared channel. Your assistant is your second brain. It knows your meetings, your notes, and your open decisions. That personal context stays yours. When a teammate asks a question in the channel, you ask your assistant to brief them. It pulls from everything you know and delivers the answer on your behalf, attributed to you. Your teammate's assistant does the same, from their own context. Teams can draft specs, track decisions, and update shared files from plain conversation. Each assistant reads the full channel history, cross references it against what exists, and flags what is missing. The whole thing is open-source, and each assistant acts as the person it belongs to, not as a shared bot pulling from a common pool. The video below shows this in action. I joined a shared space and asked my team member for a status update. My team member asked their second brain to answer. A spec got built from that conversation, versioned, with every change tracked back to the message that triggered it. Rowboat GitHub: (don't forget to star 🌟) My co-founder also wrote a great article on building your second brain with Rowboat, and I highly recommend reading it as well. The article is quoted below.

Akshay 🚀

117,616 Aufrufe • vor 21 Tagen

Y Combinator CEO, Garry Tan, took the stage for 42 minutes at Startup School 2026 and explained how to build your own personal AGI better than any paid AI course. This is what he told the room: 1. The leverage is in your context, not the model. Tan watches hundreds of founders use identical models every batch. "There are 2x people and there are 100x people who are using the same Claude. Same weights, same context window size, same API. But the leverage is not in the weights." The gap between users is now bigger than the gap between models. 2. One person's output went up 400x. In 2013 Tan shipped maybe 14 useful lines of code a day as a YC partner, dead on the median for programmer productivity. "I did the math on my output, and I'm at about 400x what I did in 2013." 3. Agents run on a different working memory. Humans hold 7 things in their head at once. Every org chart and checklist ever built is a patch for that limit. "An AI agent holds a million tokens. That's about a thousand pages. Three Harry Potter books sitting open on its head all at once." You're still running your week on tools built for the 7-digit brain. 4. Markdown is code now. Tan's stack is mostly skill files: pages of plain English an agent can execute. "If you can write clear instructions in English, you're a programmer. The compiler is a language model." At YC, finance and events staff who never opened a terminal are building automations. 5. Your history is your moat. Tan's agent runs on a personal wiki: about 220,000 markdown pages covering 25 years of email, meetings, notes and decisions. "When my agent does anything, it does knowing everything I know. And that's the difference between an assistant and a colleague." No frontier model has your context. That's the one asset nobody can replicate. 6. Never do one-off work. Most people run a task with an agent, close the window and throw the learning away. Tan ends every task by having the agent turn what it did into a reusable skill file. "If you have to ask for something twice, you failed." Captured skills compound daily. Amnesia resets you to zero every morning. 7. Own your skill files before your employer does. A skill file is your judgment, extracted and executable. The only question is who controls it. "Own your skills because if you don't, your job becomes a skill file." Files in your repo compound your career. Files in the company's repo run your judgment without you. Watch it, then read the step-by-step guide on becoming an AI engineer.

Alex Prompter

250,159 Aufrufe • vor 2 Monaten

I OPEN SOURCED NOVAMP TERMINAL THAT MADE $100K+ FOR ROBINHOOD MEMECOIN TRADERS last memecoin cycle i was buying wrong tokens a lot. six figures in profit and i still have no idea how much i left on the table buying vamped tokens instead of the original. so this cycle i built the thing that tells you, and put it on GitHub for free. i came back with literally a cheat code for robinhood memecoin traders, and i am giving the whole thing away for free. NOVAMP, now open sourced on my GitHub. repo: [ here's what happens to you after u start using it: news breaks. a name catches. sixty seconds later there are 20+ tokens carrying that ticker or that name. you read number one's timeline. you buy number seven. it goes to zero and you write "rugged" in the group chat. nobody rugged you. you bought the copy, and the copy was never going anywhere. that is not a small leak. that is where most of retail's money on this chain quietly goes, and no chart on earth will ever show it to you. so i built the thing that tells you which one is real. paste a ticker or a name. it pulls every launch fighting over it and labels each one: > ORIGINAL - first, and clean > TAINTED - first, but the operator loaded it himself > CONTESTED - not first, but smart money is here anyway > VAMP - a later copy with nothing of its own > DEAD - nothing is trading CONTESTED is the one that pays for the whole tool. first is not the same as real. when the first launch carries a fat dev buy, a stack of wallets exempt from the opening tax, and a deployer with a hundred launches and zero graduations, being first only means he got there first with the bait. the money is on number three. you would never have looked. WHAT YOU ACTUALLY GET: > before you buy one command gives you the whole cluster: who was first, how many seconds behind each copy landed, top 10 concentration with the curve and the pool excluded, dev share, how many wallets were let past the opening tax, and whether wallets with a real track record are already in. every point of the score prints its own reason next to it, so you argue with a line and not with a number. > after you buy point it at your own wallet. it tells you which of your positions are copies and puts the original right next to each one. most people find out they are holding number six. better from a terminal than from the chart at 3am. > the whole chain rank deployers by how many copies they shipped and how many other wallets share their funder. one operator running eleven addresses stops looking like eleven people. > while you sleep put a name on a watchlist, get a telegram the second somebody copies it or the cluster flips to CONTESTED. it reads the ticker AND the token name, so a copy that takes a fresh symbol and keeps the name does not slip past. it folds cyrillic lookalikes, zero width characters, leetspeak, plurals, filler words. no key. no signer. no buy button. the build literally fails if signing code ever enters the repo. free, MIT, runs on your machine and not on mine. robinhood liquidity is back and you don't have to be exit liquidity this time. leaving the full repo below. i am developing it daily, more coming.

Oracle Boar

18,570 Aufrufe • vor 28 Tagen