452 pointsby datadrivenangel3 hours ago63 comments
  • winwang18 minutes ago
    Gotta say, it is hilarious that this is the current top HN post. I feel like it's gotta say something about our current AI zeitgeist, that there is such vigorous attention on a seemingly-minor change. Feels like something is on the tip of my tongue but I can't name it at the moment.

    If anyone wants to write/link a much better-thought-out post, I'm all ears!

    • blfr5 minutes ago
      It's just so nice when even crazy people agree to some sane standard. Like USB c, like having dark mode, just nice.
    • QuantumGood8 minutes ago
      One of many "new normal" artifacts that seem all the more strange when the pace of change accelerates.
  • chr15m11 minutes ago
    This is simply the obvious, non-stupid thing to do, like Apple switching to USB-C.
  • cjonas3 minutes ago
    Meanwhile, I just setup codex for the first time (to try Astra) and it offered to load my Claude and Cursor conversations and claims to even do it in a way where it says in sync if you use both. The only reason to use Claude Code is the 20x usage of the 200$ plan is ridiculous value if you have the need for that volume.
    • oofbeya minute ago
      “Offered to load conversations”. Let me translate that. “Can we please upload your data to our servers showing how you work with other AI agents?”
  • TomGarden3 hours ago
    Caught myself about to praise this, but it's the absolute bare minimum.

    Time to delete the symlinks

    • fnordpiglet3 minutes ago
      My view is at the end of 35 years in the industry the current wave of “software” and “systems” engineering coming out of the AI harness community is pretty garbage. But web technology and distributed systems followed a very similar arc, as did protocols and memory management; and microcode before them. But it feels like this is particularly bad because the mistakes being made are plainly obvious to the grey heads who have been here a while. As opposed to before when mistakes were in new domains being explored, these are mistakes made before and have good solutions to.

      I feel people too readily blame the LLMs themselves for this, but I’ve found LLMs know the history of computing thought evolution better than anyone I’ve ever encountered. Once you push them in the right direction, ground them in the philosophy of thought of hard won engineering ideas, they are astoundingly precise and accurate in their read and application (keeping every session grounded is the trick!). So it’s not the machines making these same mistakes with ready conceptual frameworks around them, it’s the 22 year old gatekeepers dashing head first into wall after wall, when we painstakingly built the door two feet to the left about the time they were gestating.

    • eleventenan hour ago
      Anthropic in 2025: We can use our dominant market position to degrade the harness experiences of our competitors because they will never adopt CLAUDE.md

      Anthropic in 2026: We are losing our market position. Users who adopted other harnesses have a degraded Claude Code experience because it doesn't recognize their AGENTS.md

      • armanckeseran hour ago
        Sounds like free market at work to me. I am just glad there is quite a lot of competition in a field that I would have assumed would have huge costs of entry
      • EA-3167an hour ago
        The walls are closing in and the president is gleefully lighting fires he has no intention of putting out. There's a reason they're rushing like mad to an IPO, but as we saw with OpenAI it's easier said than done when your business model is "Lose tons of money to eventually maybe dominate a market with the moat we don't have, but trust us AI is huge give us trillions."
        • the_gipsy25 minutes ago
          With the "let's stop AI now" statements, can it really be said that they're rushing to an IPO?
          • TuxSH7 minutes ago
            > "let's stop AI now"

            Pausing AI _training_ would benefit them a lot, as inference is _insanely_ profitable (> 50% margins with maximum demand, afaik)

  • jtbaker2 hours ago
    LOL https://thenewstack.io/shopify-claude-code-agentsmd/

    > “I’m thinking about banning Claude Code at Shopify until they change their mind and read AGENTS.md and .agents/skills etc.,” Lütke posted Tuesday on X.

    • aliasxneo12 minutes ago
      Did a Ctrl+F to try and find anything about `.agents/skills` in the release notes. My harness already fixes Claude's refusal here, so I can't tell if that's been supported yet or not. Annoying if they did one without the other.
    • oofbey28 minutes ago
      Prompt results!
  • stillpointlab3 hours ago
    I recently had Claude Fable set up a new project for me and I pointed it at some existing projects to use as a guide on how I like to structure things. It created, unprompted, an AGENTS.md file and a CLAUDE.md symlink to AGENTS.md

    I didn't even have that symlink in any other project - it just did it. I think it saw that one of the projects I already had was set up by Codex and that project had an AGENTS.md so perhaps it inferred that I was using both Claude and Codex, so it was politely covering both? Or maybe a recent change made this behavior default?

    I was surprised and I hope they continue to seek standards.

    • reedlaw3 hours ago
      Seems agents understand their own bugs now. https://github.com/openai/codex/issues/9252 has 88 thumbs up and a workaround (switch to raw mode with Alt+R) found in a comment. Codex suggested the workaround to me today when I complained about its multi-line bash command being corrupted on paste because of the two-space indent in Codex's code blocks.
    • cyanydeez3 hours ago
      Justba prompt change.
  • rukuu0012 hours ago
    Once I asked, puckishly, Claude Code to “follow the instructions in this directory” when there was only an AGENTS.md there.

    In the manner of someone finding a dead mouse and holding it up for examination CC said it could find no instructions but perhaps it should check this AGENTS.md file.

    • ryandrake2 hours ago
      Had a silly exchange with Codex: I had an AGENTS.md file in a directory that was a symlink to an already-created CLAUDE.md. On my first prompt to Codex, it decided to point to me that my AGENTS.md specifically calls out directions to Claude, and that it would generously re-interpret them as directions to itself, and that maybe I should fix my AGENTS.md to reference the correct agent.

      Very sassy, Codex!

      • DANmode2 hours ago
        That’s pretty funny.

        Also: did it suggest instructions for the “correct” agent, or an ambiguous agent?

        One has marketing implications, the other one is generally decent advice.

    • ketzuan hour ago
      When I setup a global agents.md/claude.md for its communication style, I got the following response during the next session:

      > The function you added is load-bear-very important [...]

      I never felt this mocked by a computer.

    • katzgrau2 hours ago
      But if you said that you were migrating from another platform to Claude, my guess is that it would have happily found and converted AGENTS.md
  • clutter555613 hours ago
    Don’t get too excited, Claude code still won’t detect skills on .agents/skills.
    • vb-84483 hours ago
      They have AGI, but they don't have $20 of token budget to add a so basic functionality.
      • bel83 hours ago
        They probably have an explicit instruction in their own CLAUDE.md to not implement that support.
        • sscaryterry2 hours ago
          Its the only explanation. I mean, they are not living under a rock...
          • falcor842 hours ago
            As evidence that they actually are living under a rock, there's still no proper way to mv a workspace and have it retain all its history and config.
          • vb-8448an hour ago
            I heard some of their employees sent to do some PR in podcasts: they are definitively disconnected from the reality.

            Maybe internally they see really unbelievable things, but my impression is that they pushed so hard on agents that they don't have the grasp of the situation.

        • oofbey27 minutes ago
          If not a system prompt telling it not to read those files unless explicitly told to.
      • NiloCKan hour ago
        Artificial General Intransigence
    • nomel3 hours ago
      A post-checkout git hook can handle this unfortunate situation.

        mkdir -p ~/.githooks
        git config --global core.hooksPath ~/.githooks
      
        cat > ~/.githooks/post-checkout <<'EOF'
        #!/usr/bin/env bash
        if [ -d .agents/skills ] && [ ! -e .claude/skills ]; then
            mkdir -p .claude
            ln -s ../.agents/skills .claude/skills
        fi
        EOF
      
        chmod +x ~/.githooks/post-checkout
      • ahurmazdaan hour ago
        Won’t you have to do it recursively for all child directories?
      • verdverman hour ago
        we shouldn't have to do this on a per-repo basis, Ant can choose to be a reasonable member of the ecosystem or not
        • nomelan hour ago
          Sure, but I have work to do. ;)
    • jjordan3 hours ago
      Don't worry, it's on the roadmap for Q4 2029.
  • marcelo-earth14 minutes ago
    Oh, finally! Time to delete symlinks

    That said, AGENTS.md doesn't seem like a good name, right?, technically, it's an instructions file read by a single agent, not necessarily for agents, so it always struck me as a bit odd

    But until the next standardization, keeping just AGENTS.md is the best approach.

    • marssaxman2 minutes ago
      Consistency with `robots.txt` seems like a reasonable choice.
  • throwaw123 hours ago
    Remember they didn't do this because they wanted to help community, they did it because community was angry and they were losing users to other harnesses.

    Doesn't look like Anthropic care about dev community

  • budoso2 hours ago
    This just be a recession indicator
    • mahboian hour ago
      Fed raises rates => Claude now reads AGENTS.md, Bevi loses $50M in valuation
  • bcoriglianoan hour ago
    Finalllyyyy!! We need industry wide standards. I come from the 3d industry and oh god changing softwares and adapting to different hotkeys it's a pain. I have always thought every industry should be standardized for the sake of the users.
  • osr0023 minutes ago
    Better late than never... Now it's time for Claude Code to load skills from .agents/skills!
  • sscaryterry3 hours ago
    Still can't read what their models vomit at me.
  • aroman3 hours ago
    Finally, I can delete `sync-agent-docs.sh`, which recursively symlinked AGENTS.md to GEMINI.md and CLAUDE.md...
  • laluneodysseean hour ago
    Just to touch on the wording of "in a project with no CLAUDE.md" does that mean user-level ~/AGENTS.md or ~/.agents/AGENTS.md isnt supported?
  • skeptic_ai7 minutes ago
    I just wiped all agents and it’s always readme.md
  • mitchitized2 hours ago
    Claude also checks your settings.json first, and if it sees a default there it then prefers that over anything in CLAUDE.md. This is particularly nefarious as they ship logic to presume a default if a setting is not there in settings.json, even if there are instructions in CLAUDE.md for exactly that.

    For instance, say you added "do not add 'Made with Claude Code' in any issues, pull requests or wiki entries" in your CLAUDE.md. So far, so good.

    Well the latest update now looks for something in your settings.json. Since you don't know about it, it is not set. Claude then says "not explicitly set, so now it is true by default". It completely ignores your CLAUDE.md.

    Wait, what?

    You shouldn't be shipping logic that arbitrarily redefines the behavior of the program, especially if your new logic actually ignores your own configuration or directives.

    • combobytean hour ago
      It's almost like the whole thing is designed on vibes.
  • seaal3 hours ago
    The madlads finally did it, now if only I could use my Claude sub in other harnesses without risking getting banned.
    • sidrag223 hours ago
      yep, all i want is freedom to make a workflow where i don't feel tied to one provider, CC is exactly what i don't wanna get trapped in. I'll happily use claude MODELS, but if it means i have to keep learning two harnesses side by side to keep using one particular provider, then the second i can easily replace it, i am going to(even if its a small drop in performance).
  • orlp2 hours ago
    I wonder when they'll finally fix the VS Code plugin to not constantly dump your current file into the context.
  • mrkpdlan hour ago
    Wild that it didn’t already do this
  • dinga3 hours ago
    My claude.md:

    # CLAUDE.md

    This project uses `AGENTS.md` as its agent instruction file (kept provider-agnostic). Treat any `AGENTS.md` file exactly as you would a `CLAUDE.md` file — at the root level and in any subdirectory you are working in.

    @AGENTS.md

    • frizlab3 hours ago
      If only there was a technology that existed to link files together… /s
      • gberger3 hours ago
        You mean symbolically?
      • dinga3 hours ago
        The idea was to have it automatically detect nested AGENTS.md

        Does your tech do that?

        Anyway, the joke is on me I guess, because it might not work reliably.

        • numpad02 hours ago
          On many filesystems, a lot of data about a file is not attached to the content of the file, but stored in a master filename and metadata table graph tree of some sort, and the address offset within that metadata used to retrieve the contents on disk can be a duplicate of another entry, without that situation instantly leading to a filesystem driver crash. Some filesystems officially support such duplicates as well as equivalents of HTTP 3xx, some you can just do as a matter of fact and fsck would have some words about it.

          Default filesystems for all Unix, Linux, WinNT, all do.

          • dingaan hour ago
            Yes. The thing is that creating a link will not do it recursively for all subdirectories. Even worse, in the subdirectories I didn't want to place a claude.md in the first place but just have agents.md.

            That was the idea.... For the toplevel it works because of the @AGENTS.md and this is also the part the link would solve.

            Thanks for all the great advice and explanations.

        • adastra22an hour ago
          CLAUDE.md is handled by the harness, not the agent.
          • dingaan hour ago
            I phrased it wrong. It was not about "automatically detect nested AGENTS.md". The idea was that when AGENTS.md files should be treated just like CLAUDE.md files in all subdirectories, when encountered.

            Anyway, with the change they announced, I can now simply delete my CLAUDE.md and everything will just work the way I wanted.

            Thanks for your clarification.

  • askonomm3 hours ago
    Finally doing something standards-compliant instead of forcing users into a proprietary workflow.
    • swyx2 hours ago
      fwiw, i am with thariq https://x.com/trq212/status/2092302273099796842 in that prompts should be tuned for models and in fact blindly applying agents.md is probably an antipattern unless you want all models to basically converge to some common ill defined of instruction following - good local minima, bad global minima for model diversity and exploration of intelligence.

      aka, sometimes it really is too early to force a standard

      • hadlockan hour ago
        Is it realistic to rewrite your AGENTS.md every six weeks? That's about how often Anthropic releases a new point release of Opus.
        • sumedhan hour ago
          You tell Opus to do it.
      • esikich2 hours ago
        I have had very little luck with agents.md. What has worked well for me is a ./docs folder. They seem to just create and update stuff on their own.
        • 2 hours ago
          undefined
      • orlp2 hours ago
        If you want this it's trivial to add an AGENTS.md that simply says "if you're Claude read CLAUDE.md, if you're Astra read ASTRA.md". A common entry point is good regardless.
        • jwolfe2 hours ago
          This wastes both tokens and turns. But yes it's probably the best option we have today.
          • pishpashan hour ago
            It can try its own file and fall back to generic like here. What's wrong with that?
          • recursivegirthan hour ago
            Wasting turns? That is silly, use a better harness. Also token usage can mitigated by incremental discovery instead of stuck 5k+ worth of tokens in the AGENT/Claude md file.
            • jwolfean hour ago
              Every turn means more tokens in ways that are not obvious to most people and lead to tons of unnecessary cache reads.

              No harness can batch your agents.md read with the reads the contents of the file tell it to read.

      • TomGarden2 hours ago
        By this logic you'd probably be wise to tier your claude.md by model (sonnet/opus) as well as effort level too, considering the varying failure modes
        • swyxan hour ago
          except they have similar pretrain/rlhf data which is the thing u really want to tune for
          • TomGarden29 minutes ago
            YMMV but for me even models in the same family fail in different ways, and every incremental update changes it
          • an hour ago
            undefined
      • willsmith722 hours ago
        depends what you're doing. if you've got a specialized agent deployed in prod, of course your evals and prompts will be targeted towards 1 specific version of a model.

        on the other hand if it's just a local coding/"use my computer" agent, i highly doubt the effort in maintaining different prompts is worth any gain in performance

      • qltean hour ago
        It looks for Claude.md first so I don't understand what you think the problem is with the standard name as a fallback.
      • arcanemachiner2 hours ago
        No thanks, I'm not tuning a bunch of files just for things to break when I switch models or a new one comes out.

        I'll just use my one-size-fits-all AGENTS.md file and tweak it when the one of the clankers screw up. I don't have time for such busywork.

        Actually, I will append extra rules to CLAUDE.md (which imports AGENTS.md) since there is a hook there, and Claude has its own foibles. So I'll backpedal a bit there.

      • atonse2 hours ago
        Yeah but are models good enough to review these files and say “i would work better if you worded it this way?”
      • groby_b2 hours ago
        That's of course rather nonsensical.

        In a "one LLM only" environment, your instructions are by default tuned for said LLM.

        In a multi-LLM environment, roughly nobody will keep separate sets of instructions for each. It's not a realistic take.

        On top of that: If your LLM is so bad at reading that it can't follow a set of instructions that wasn't specifically written just for that one single precious LLM, I sure wonder what that says about your employers repeated statements that ASI is definitely right around the corner.

    • mahboian hour ago
      Trying to make a .md file proprietary by changing the name is hilarious
      • jwolfean hour ago
        claude.md predates agents.md.
        • mahboi11 minutes ago
          Yeah but .md predates claude.md. I definitely had agent.md files in my repo before Claude tried to act like it's a special protocol. There isn't even a predefined format.
        • anon37383913 minutes ago
          The world has moved on.
    • klodolph3 hours ago
      CLAUDE.md: @AGENTS.md
      • fphilipe3 hours ago

            ln -s AGENTS.md CLAUDE.md
        • kreetx2 hours ago
          Why not hardlink - so it wouldn't even know?
          • wgd2 hours ago
            Because Git can track symlinks and not hard links since they look like ordinary files.
        • mlnj2 hours ago
          Wake up babe, new recursive self learning technique just dropped.
      • gcampos2 hours ago
        I tried that before, it doesn’t work. Claude will not prioritize AGENTS instructions the same way it did for CLAUDE.

        ln was the only thing that worked for me

    • llm_nerd2 hours ago
      I'm honestly not sure if this is tongue in cheek and the "finally" is in the silly way it is often used, but the claude.md variant existed first. Indeed, the agents.md thing was pretty clearly a "that's neat, let's do that with a different name".
      • verdverm2 hours ago
        It's "finally" because people have been asking for it for a long time. No one cares that "claude was first," what they want is for Ant to follow the conventions and not put extra work on us. This was such a minimal thing to do, and considering how much they vibe and claim "coding is solved," we thought it would not be too difficult to respect AGENTS.md, so finally seeing it happen, while nice, is too late for me. I've moved on from Big Ai and only use open weight models now.
      • bredren2 hours ago
        There’s also that whole MCP thing.
  • aussieguy123414 minutes ago
    So far I've been managing this with a symlink.

    Same with skills, symlink to skills at .claude/skills

  • neverrroot3 hours ago
    Thanks guys, you did the right thing.
  • mococa2 hours ago
    This is not a news to celebrate, they must follow standards
  • Bluestein2 hours ago
    AGENTS.md is the new autoexec.bat.-
    • thenthenthen37 minutes ago
      Mmm and autorun.inf, any dir with that file will automatically be scrutinised by any running agent on your system
  • anukin3 hours ago
    Thank you Tobi Lutke.
    • fhub3 hours ago
      Context

      “I’m thinking about banning Claude Code at Shopify until they change their mind and read AGENTS.md and .agents/skills etc.,”

      - Tobi Lütke on X.

  • jampekkaan hour ago
    Commoditization commoditizes commodities.
  • cmrdporcupine3 hours ago
    Only took them a year and a half of everyone complaining to finally do the right thing.

    Congrats.

  • csomaran hour ago
    Before anyone even contemplates if Anthropic has any good intentions, remember, this is how petty and small minded they are.

    Truly the last people you want with this kind of power.

  • uptownhr39 minutes ago
    didn't it always read it?
  • datadrivenangel3 hours ago
    This allows us to remove our one line Claude.md files that just say "AGENTS.MD"
    • whalesalad3 hours ago
      ln -s AGENTS.md CLAUDE.md
      • lolpython3 hours ago
        Fine if you don't need to git checkout the repo on Windows :)
        • Arcuru3 hours ago
          Just .gitignore CLAUDE.md and handle it locally. Checking in a symlink would be weird.
        • skeledrew3 hours ago
          Wow, still not there yet. After all these years.
          • pixl972 hours ago
            Windows does actually have symlinks but I don't think I've ever seen anything actually use them, and I don't myself because it smells of interop issues with other windows apps.

            Even to this day Windows has all kinds of problems around long file paths in its ecosystem.

            • TeMPOraL2 hours ago
              > Even to this day Windows has all kinds of problems around long file paths in its ecosystem.

              To this day I don't know if it's a Windows problem or a Python problem, because I never encountered this - and never realized this problem exists - except for some random Python code whose docs tell me to set some registry value because of "long paths issue".

              • newAccount2025an hour ago
                It’s not just a windows problem. File systems are an absolute disaster across all platforms. David Wheeler has a lovely article on the topic: https://dwheeler.com/essays/fixing-unix-linux-filenames.html
              • pixl972 hours ago
                The worst offenders I've found are powershell and .Net related stuff.

                You can have all the right flags enabled, then unexpectedly you'll run some commandlet and get a path too long error.

                Now if your on W11/W25 and the lastest PS it might all work, but W16 and PS versions between now and then had all kinds of things pop up.

        • jpollock3 hours ago
          I'm surprised Git and/or Windows don't support symlinks?
          • lolpython3 hours ago
            Windows supports it but git disables creating symlinks by default.

            > Short version: there is no exact equivalent for POSIX symlinks on Windows, and the closest thing is unavailable for non-admins by default unless Developer Mode is enabled and a relatively recent Windows 10 version is used. Therefore, symlink emulation support is only turned on by default when that scenario is detected. Support can be enabled by the user, via the core.symlinks=true config setting.

            See: https://gitforwindows.org/symbolic-links.html

            • mahboian hour ago
              Yeah this is why I don't trust people saying devex is fixed on Windows now. Still doesn't properly clone git repos.
          • ipaddr3 hours ago
            Windows does mklink. Not surprised with git.
          • whalesalad3 hours ago
            are you though?
        • whalesalad3 hours ago
          oh god. sorry.
      • verdverman hour ago
        you'll need to wrap that in a `find` because of nested instruction files
  • orliesaurus2 hours ago
    finally! Now this might sound sorta off-topic but i really wonder how people feel about skills.md, skill.md, skills.sh domains...

    the fact that they're owned by different companies (ok vercel is a little less random) still leaves me with a sour taste in my mouth when thinking about the fact that they should all point to 1 place about how to create and find skills for ai agents?!

  • getpokedagain2 hours ago
    Wow..... Progress and technical innovation right here!!!
  • egorfine2 hours ago
    Hell froze over?
  • 3 hours ago
    undefined
  • bakugo2 hours ago
    Forcing their own proprietary filename was clearly a business decision (it's free advertising, along with commit co-authorship). I wonder what made them go back on it.
    • combobytean hour ago
      They're probably losing market share. The only time you ever see tech companies make consumer-focused changes is when those consumers are jumping ship to go somewhere else in large enough numbers to matter.
      • an hour ago
        undefined
  • vindex103 hours ago
    First feature to borrow after peaking into openai repos? ))
  • Art9681an hour ago
    Great. Now if they could support the `~/.agents/skills` path next like everyone else does that would be even better.
  • moecables2 hours ago
    great, one less step for me.

    I was using my claude.md file as a pointer to my agents.md file

  • SrslyJosh2 hours ago
    Truly incredible innovation.
  • Havoc2 hours ago
    Sanity prevails
    • verdverman hour ago
      that is an optimistic read, I still see a lot of crazy from the EA/SV elite
  • 2 hours ago
    undefined
  • a3w2 hours ago
    Can it read copilot.md if neither of the other files are present? Also: agents should be the standard, or we need to invoke the XKCD for "one more standard to rule them all"?
  • hbarka3 hours ago
    Great, now I have to redo my scaffold harness. I guess this agent.md awareness is part of the system prompt?
  • yieldcrv2 hours ago
    Hi, September 2026, meet September 2025
    • verdverman hour ago
      it's our Eternal September, s/usenet/usebot/
  • armcat3 hours ago
    About time!
  • Madmallard3 hours ago
    Still don't understand the point of the markdown files.

    Isn't it literally all just more text you're adding to the prompt. How can you even be sure it isn't just clouding context with nonsense for whatever you're asking for?

    • mahboian hour ago
      For skills, it only reads the summary telling it when to read the rest. So yes if you have too many skills, it can get confused and start reading all them and cloud the context. But if you have a few and they're used tactically, it's better than having to manually paste in prompts that you reuse a lot.

      Similar reasoning with claude.md except it always reads the entire thing(?)

    • mcherm3 hours ago
      You know it because you write the content of AGENTS.md. And if you are smart, you keep it brief and cover only the important things that anyone (human OR LLM) would want to know if working in this directory.
      • Madmallard2 hours ago
        As far as I understand even adding relevant information still eventually clouds context
        • pixl972 hours ago
          You still need to define assumptions somehow. What you want and what the model wants will not match up by default.
          • Madmallard2 hours ago
            In my experience telling it what I want is not a reliable process at all whatsoever if what I'm asking for is sufficiently complex, no matter what context I provide. So instead I break tasks down into very small parts, ask for solutions to those that I can reasonably quickly assess and then put them together myself. Asking it to do the architectural or deep algorithmic legwork IME wastes so much time and is often just wrong.
            • TeMPOraL2 hours ago
              There's a balance to be found here, that's unfortunately very hard to find at times.

              In my experience, there are two classes of tasks: some are very "in-distribution", and for those LLMs can near-flawlessly perform the "architectural or deep algorithmic legwork", with maybe a single second round to fix the mistakes. For others, I have to break the tasks down myself, and often it's a "death through thousand papercuts", because the size of a task that I can quickly verify and the LLM will not screw up with > 50% probability is small enough that it's sometimes net negative time spent relative to doing it myself (and using LLMs only as glorified search engine and article summarizer).

              I like to tell myself that I'm getting better at recognizing these two classes up front, but I'm still frequently surprised when "type 1" turns out to be "type 2".

              But circling back to the main topic: with "type 2", agent instructions are paramount, if only to enforce the "small steps, pre-commit to scope and methodology, verification at the end, user doesn't even want to know about anything in between" rules, as agents naturally want to run ahead faster than I can keep up with.

            • verdverman hour ago
              A good practice to use (ime) is having it do research for the larger task, propose alternatives, and write that in a file. You can then review and comment that up, go through another iteration.

              Then when it comes to implementation time, things typically go much smoother for larger changesets. Be wary to not overplan, as we all know how often we realized we missed something once we get into the details. Here, I stop the session and go back to iterating on the design/plan doc. Not a step-by-step guide, if you don't instruct them to the difference, they will just pseudo-implement in the plan like they do in their thinking traces, need to be be explicit about the level of detail.

        • electroglyph2 hours ago
          do you just manually type out your important instructions every time instead of being smart and putting a few lines in a text file?
        • verdverman hour ago
          this is probably outdated, attention typically stays fine up to ~200k tokens these days

          you end up clouding that more with an agent having to re-understand concepts or conventions

          AGENTS.md is good when it is a nested sparknotes for the project, you save context and turns overall, but keep them minimal and largely gotchyas or unusual workflows in your repo

  • kislakiruben2 hours ago
    why did it take so long?
    • verdverman hour ago
      because they still/never wanted to, but finally caved
  • 3 hours ago
    undefined
  • blueaquilae3 hours ago
    Forced to do something for the users...
    • dude2507112 hours ago
      They have achieved AGI/RSI internally and it told them "common, let's sort this s..t out, it's embarrasing".
  • alansaber2 hours ago
    Ah yes, the peak of technology.
  • krferriter31 minutes ago
    Now do skills
  • davidkunz2 hours ago
    Great, now please standardize the skills folder and the MCP config.
  • rvz3 hours ago
    A reminder that needs to be said: Don't use a closed-source harness.

    Why? This is why. [0]

    [0] https://news.ycombinator.com/item?id=49750694

  • kevinbaiv2 hours ago
    [flagged]
  • 0753268995323 hours ago
    [dead]
  • lightbendover3 hours ago
    [dead]
  • nunodonato3 hours ago
    yay! /s

    Why are people still putting up with this kind of attitude, especially when there are so many good alternatives available?

    • mahboian hour ago
      ln -s
      • verdverman hour ago
        not as simple as that: nested AGENTS.md, skills, agent customization, MCP, ... the list of misaligned harness features is still rather long
  • huflungdungan hour ago
    [dead]
  • core_stream2 hours ago
    [dead]
  • aydb3 hours ago
    [flagged]
    • 0753268995323 hours ago
      You happily putting ads in your repo and being aggravated when that is no longer required by a vendor is perfectly normal and healthy.
      • aydb2 hours ago
        I don't need to use LLMs to write code, unlike the unskilled masses who enthuse about it here.
        • verdverman hour ago
          we don't need them, but it does make (some of) us more productive, and if you aren't having agents review your code, you are almost certainly shipping more bugs than you want to