you are viewing a single comment's thread
view the rest of the comments
[–] 6 points 1 day ago (4 children)

Agents md file doesn't actually do anything except using more tokens https://arxiv.org/pdf/2602.11988

  • source
  • hideshow 4 child comments
  • [–] 6 points 1 day ago (3 children)

    I feel like you either didn't read your article or didn't read the post article

  • source
  • parent
  • hideshow 3 child comments
  • [–] 4 points 13 hours ago (2 children)

    To quote the research paper conclusion...

    Conclusion We evaluate the impact of context files on coding agent performance for four common coding agents on SWE-BENCH and the novel CTXBENCH, built from recent GitHub issues and less popular repositories containing developer-written context files. We find that all context files consistently increase the cost and number of steps required to complete tasks. LLM-generated context files have a marginal negative effect on task success rates, while developer-written ones provide a marginal performance gain, neither statistically significant. Our trace analyses show that instructions in context files are generally followed and lead to more test- ing and broader exploration; however, they do not function as effective repository overviews. Over- all, our results suggest that context files don’t improve coding agent performance, and should only contain specific additional instructions beyond what is already available in the codebase. This high- lights a concrete gap between current agent-developer recommendations and observed outcomes, and motivates future work on principled ways to automatically generate concise, task-relevant guid- ance for coding agents.

    That sounds like "Agents.MD doesn't work" to me.

    Am I missing something?

  • source
  • parent
  • hideshow 2 child comments
  • [–] 1 point 12 hours ago (1 child)

    Agents.md doesn't exist to make sure generated code functions, it's there to make sure it's written well.

  • source
  • parent
  • hideshow 1 child comment
  • [–] 2 points 8 hours ago

    What is the difference between the models writing code "well" and their performance in this context? Are we referring to readability?

    Genuine question. If we use agents to read, edit, and review code, why do we care about readability? That's a human constraint. Unless attempting to do those three is not effective and thus requires human attention to correct issues which would justify readable code. If that's the case; why use the agent to edit the code in the first place?

  • source
  • parent