225 points pszypowicz 1 hour ago 89 comments

nfRfqX5n 1 hour ago | parent

Crazy part is: can’t tell if this intended or a bug

vaylian 1 hour ago | parent

Can you think of a bug that makes sense in this case?

hgoel 1 hour ago | parent

Vibe coding

thejazzman 1 hour ago | parent

It is most certainly the kind of subtle bug the LLMs love to slip in and burn you in production

Not sure what y’all are thinking with these unwritten conspiracy theories that begin and end with “it’s intentional”

jpitz 1 hour ago | parent

Assuming that it's intentional, what's the motivation?

kriro 52 minutes ago | parent

Get more people to turn on telemetry.

kennethops 1 hour ago | parent

I like to give my graces to people and the companies who are typically not trillions of dollars. Have an incredible amount of resources that many countries would like to have, with fewer of the obligations. I'm going to chalk this up to its intended

serial_dev 1 hour ago | parent

Someone on the Claude Code team is probably wondering the same...

Razengan 1 hour ago | parent

Claude/Anthropic has been sus from the start:

https://www.thatprivacyguy.com/blog/anthropic-spyware/

+ not letting users change their email, or remove their payment methods, etc.

jaapz 1 hour ago | parent

many of these can also be chalked up to the fact that all of their products are extremely vibed

nibbleyou 1 hour ago | parent

I cannot set a password for login, on logging in it says we've sent a login code but it's a link instead...

chrisjj 1 hour ago | parent

Vibe-coding at its best.

tjoff 1 hour ago | parent

Nice find, though I'd rather read the prompt that was used to write this article. It is about ten times longer than it needs to and is quite painful to read.

sandrello 1 hour ago | parent

Based on my experience with these tools so far, this seems exactly the kind of subtle but extremely severe bug that sneaks in when you start piling up layers of AI generated patches to a codebase without caring too much about the code.

hgoel 1 hour ago | parent

Yep, especially with long contexts (and moreso if the last thing you were working on in the same context involved telemetry too). AI sneaks in weird conditions like this and then does the entire "You're absolutely right" thing if you're paying enough attention to catch it.

crazygringo 1 hour ago | parent

It seems to be exactly the opposite, a "fully human error":

https://news.ycombinator.com/item?id=49815363

Nothing to do with AI patches at all, nor was it a bug. It was intentional human behavior, a temporary rollout setting, that seems to have made sense.

But I guess that doesn't fit the "narrative".

q3k 51 minutes ago | parent

> It seems to be exactly the opposite, a "fully human error"

"LLMize the succeses, humanize the failures." is the PR strategy at play here. Anything goes well it's because AI did it, anything goes bad it's because a human didn't catch it.

crazygringo 24 minutes ago | parent

Wow people are cynical here.

No matter what happens to be the truth, HNers have a cynical narrative to fit it.

fg137 30 minutes ago | parent

One thing I do know is that an Anthropic employee is definitely NOT going to blame this on the model they are using.

tpurves 59 minutes ago | parent

Except that, and this I find slightly amusing, is a situation where they definitely don't want to publicly blame the ai model when there are mistakes.

Agentlien 58 minutes ago | parent

I used to actively use Msty for local models because it just worked and had a lot of nice advanced features. A few months ago they released their beta version of a Claw-like UI and mentioned using a version of it to develop Msty itself. Well, that was around the time I stopped using Msty because every update started breaking things and two updates in a row included bugs which wiped all my configs, chats, and history.

HotHotLava 35 minutes ago | parent

"extremely severe" - aren't we laying it on a bit thick here? The whole impact seems to be that users who have telemetry turned off got this feature ~2 days later, when the bug was noticed.

code_runner 12 minutes ago | parent

a "feature" that every user has asked for - which is the equivalent of changing claude.md to agents.md - and the release isn't even smooth because the telemetry isn't wired up quite right.

for an organization that is being used as a model for new agentic software development practices.... and every software exec on earth is trying to reshape their organizations after - its a pretty stupid bug for a feature that should've been straightforward in the first place + took forever for them to get around to.

its just kind of emblamatic of the rough edges that exist EVEN FOR SIMPLE THINGS whenever human judgement is totally removed the equation.

code_runner 10 minutes ago | parent

ps: apparently thats not how you spell emblematic - but I'm actually glad to leave a little humanity around given the topic :D

Traubenfuchs 1 hour ago | parent

Issue 95690, opened 3 days ago, 500k+ engineers, a simple CLI...

AGI was reached like 2 weeks ago, latest claude 5.x models rule supreme and oftware engineering is solved?

lucfranken 1 hour ago | parent

Isn't that how they just release all features? So they can do progressive roll outs and telemetry on issues with it?

Not sure if they later move the code from inside the flag check to the main code or that they keep the flag check.

But if they would keep all features behind a flag that would not make most sense as you then would have not many features without telemetry.

fg137 53 minutes ago | parent

This is what's written in their release notes for 2.1.277:

> Added AGENTS.md support: in a project with no CLAUDE.md, Claude Code reads AGENTS.md instead

My dumb brain tells me none of this is rolled out progressively (as of that version). You either have it or not.

nijave 36 minutes ago | parent

It's been like this at least months. I turned telemetry off a few months ago and it basically prevented all stages rollouts of new features from working.

Anthropic commonly gates behind feature flags that require telemetry until they're "promoted" and defaulted on.

A little bit annoying you can't manually control the flags without telemetry but I think the title is a bit click bait.

vorticalbox 1 hour ago | parent

I use cursor and claude, its kinda annoying having to have the same skills in both so I made a ~/.agents/skills folder then ln both cursor and claude skills to point to that folder.

which works except that claude uses .skills/synced which is uses to sync changes to skills from claude servers into the skills folder.

every other agent I have used just directly syncs into .skills so it ends up duplicating skills

0m13 1 hour ago | parent

can i ask claude to change these settings for me, and will it enable telemetry as implied checkpoint when i ask it to enable AGENTS.md?

rvz 1 hour ago | parent

I am once again (for the third time) [0] asking you to stop using a closed-source harness.

[0] https://news.ycombinator.com/item?id=49760449

pbasista 52 minutes ago | parent

I understand the motivation for such a concern. But Claude Code in particular has its source code available on GitHub [0].

So I am unsure if it is fitting to call it a "closed source" harness.

[0] https://github.com/anthropics/claude-code

wccrawford 47 minutes ago | parent

It's not open source. It's more "source available", since it's published, but you legally can't do anything with it. Other than maybe build it yourself, for yourself, I guess.

teekert 1 hour ago | parent

Uhm I turned off everything there is to turn off in my Pro plan, and Claude just read my agents.md with no issue? So... What telemetry am I missing? Or did they JUST update? (I updated my docker image 40 minutes ago to get Opus 5.5, am on version 2.1.280, so not the mentioned 2.1.277, so it's fixed?)

mpoteat 1 hour ago | parent

Sorry folks, this is a rollout artifact, we needed a way to turn this off remotely via feature flags if it broke something, and with telemetry off you don't get those. It's already been fixed as part of v2.1.281 releasing today.

The mod is source available here: https://github.com/anthropics/claude-code/tree/main/mods/age...

Apologies again folks, this was a fully human error on my part - I should've found a better way to launch with a kill-switch.

piltdownman 1 hour ago | parent

Much respect for the prompt response, humility, and frank disclosure.

ndbe 1 hour ago | parent

I don't understand, isn't coding solved already?

mpoteat 1 hour ago | parent

The AGENTS.md support was implemented via our new extensibility system for CC, called Mods, which is launching soon-ish. A mod is a plugin with a new type of hook, which we call a function hook.

If folks play around with it, I would love feedback on the relevant issue: https://github.com/anthropics/claude-code/issues/91870

Mods allow quite a bit more customizability and control. I really believe in the idea.

rickette 1 hour ago | parent

The only difference between CLAUDE.md and AGENTS.md is the filename. Do you really need a whole plugin system to support that use case? Feels like this could have been a one line change.

oblio 57 minutes ago | parent

One line change? Pffft. That means you're still looking at the code, you're behind the times.

locknitpicker 54 minutes ago | parent

> The only difference between CLAUDE.md and AGENTS.md is the filename. Do you really need a whole plugin system to support that use case? Feels like this could have been a one line change.

I don't think this is a reasonable assumption. The document format in CLAUDE.md is whatever Anthropic specifies, where AGENTS.md is a common ground format that is expected to be supported by any agent, be it from Anthropic or not.

https://agents.md/

You might argue that differences are small or negligible, but that is just an expectation.

BowBun 51 minutes ago | parent

I think you're giving these files too much credit. There are no specs, they are freeform text. In this sense they are the same. The expectation of what could be in it by each vendor means nothing unless it's enforced.

ljm 51 minutes ago | parent

It's not a format though is it? it's literally just an extension to the system prompt in plain markdown.

There is no rhyme or reason to the structure of this file, just like with most things in AI. It's best effort human language.

logifail 48 minutes ago | parent

> The document format in CLAUDE.md is whatever Anthropic specifies

Q: Do Anthropic actually specify a document format?

taormina 37 minutes ago | parent

No, they do not.

stingraycharles 51 minutes ago | parent

Yes, but at the same time, it’s also a good, simple use case to test a new plugin system. I can totally understand that.

xandrius 47 minutes ago | parent

Don't you run bizantine ralph loops on remote environments with codex security checks, coupled with jev, grok, open router and a fully independent openclaw (on a maxed out mac mini inside a caveau in an undisclosed location, with open telegram) to change constants? You're going to be left behind.

doublerabbit 40 minutes ago | parent

You sound like my type. hey, wanna come over to myspace so I could twitter your yahoo till you google all over my facebook?

Anonyneko 7 minutes ago | parent

pdpi 32 minutes ago | parent

The plugin system itself was probably already in the making, and they just chose to implement this tiny feature as a plugin to try it out.

As for the only difference being the file name, that's an untested assumption. Up until now, Claude hadn't supported AGENTS.md, and it's a simple application of Hyrum's Law that somebody, somewhere, was taking advantage of that to give one set of instructions to Claude, and a different set of instructions to some other provider. The correct behaviour in the presence of both files is not obvious, either.

Having a kill switch for the change is a perfectly reasonable safety measure in case something goes horribly wrong.

redox99 25 minutes ago | parent

That line of thought is the reason why everything gets so overengineered.

Read CLAUDE.md if it doesn't exist read AGENTS.md you don't need to overthink it so much.

arcfour 5 minutes ago | parent

This line of thinking is something you will quickly be disabused of once you try supporting software that hundreds of millions of people use.

chrisweekly 18 minutes ago | parent

With apologies for not just testing this myself (currently AFK), doesn't it still work to have a CLAUDE.md file containing just `@AGENTS.md`?

chrisjj 4 minutes ago | parent

[delayed]

d5lt5 49 minutes ago | parent

Sounds like you've read the deepseek harness paper.

dybber 9 minutes ago | parent

Will you extend your plugin to read skills and rules from `.agents`? Or should we write our own plugin/mod for that?

michaellee8 1 hour ago | parent

Really loved this Tibo-level responsiveness, if Anthropic can keep it up with this level of service, I am pretty sure a lot of people will just ditch their ChatGPT subscription and just move to Claude.

d5lt5 53 minutes ago | parent

On the other hand, if Anthropic is to follow the industry standards, this would never have happened in the first place. It's not like the feature gates are the frontier of software development.

bpodgursky 12 minutes ago | parent

This is the most symbolic and unimportant change in history (you can literally just symlink), I think people will be fine.

d5lt5 7 minutes ago | parent

Symlink requires admin, and most people do not run CC as admin with bypass permissions though.

bpodgursky 5 minutes ago | parent

Right... you can just do it yourself.

arcfour 3 minutes ago | parent

? You need root to run ln? Since when?

BowBun 51 minutes ago | parent

Because they respond to HN threads about their products? Which are likely Claude hooks monitoring for activity in the first place? Come on...

At least make an argument for switching vendors based on the quality or price of their service.

user43928 25 minutes ago | parent

After I was mildly disappointed with GPT-6 Sol and Luna not improving intelligence and only cutting the price, I'm running Opus 5.5 today.

After the last month or so in the Codex app, I was pleased with the Claude app.

It might be a case of the grass always being greener on the other side, but this is what stands out:

After 3-4 hours of usage, the weekly usage limit moved by only 1%.

Compared to Astra where I can watch the limit draining live, this is a great improvement.

I'd estimate it 3x cheaper, and that's with a 450k context limit instead of the 258k in Codex.

So far Opus 5.5 appears less prone to stopping for no apparent reason at checkpoints in the middle of a longer task.

It doesn't open an internal browser with a useless comparison page, where it then proceeds to add notes despite no one having asked for it.

It is a breath of fresh air: I get the response in the chat, while the Codex app recently loves randomly opening artifacts instead.

Opus 5.5 xhigh made great progress on the task, more so than Astra High, but that could be random chance.

Oh, and the 'Auto' mode actually works and does not force me to instead run 'Full access' like in the Codex app, lest it blocks even 'git push'.

m3kw9 24 minutes ago | parent

Sure a fast response on HN would make people switch. Try better rates, infra, limits etc.

cowboylowrez 9 minutes ago | parent

What we need is a low level but constant drumbeat against openai in general. In general the AI situation is overleveraged and underpoliced, with the occasional hints of AI gone wild. If openai were to just be left to die, we could let that financial mess unroll and bail out the leftovers, I don't like bailouts anymore than the next guy but with this administration its almost a guarantee if things go south because this adminstration can charge administrative fees of maybe $20-30 billion (which goes to trump), get Sam Altman to serve one or two years in a cushy resort type fed place for the hugging face hacking and put openai's processes on github as a premium feature, say $10000 a month to access (which again goes to trump).

I know I know, why are we giving money to trump? Its because he's going to take it anyways so can't we at least apply some window dressing?

grim_io 1 hour ago | parent

The same guy writing readme's for my vibeslopped toy projects is also the readme writer at Anthropic, what a coincidence ;)

Maxion 1 hour ago | parent

Well now, this is how you do community outreach

aviperl 1 hour ago | parent

Ouch.

I've had to send such messages, but internally at work, not on HN!

Have a great day, human.

OtherShrezzing 1 hour ago | parent

>Apologies again folks, this was a fully human error on my part

Blink twice if you need help

fg137 56 minutes ago | parent

> a fully human error

Would be interesting to know how much time you/your team spent on that design decision

rachr 44 minutes ago | parent

The correct design was in the AGENTS.md but they didn't have telemetry on

senko 39 minutes ago | parent

I hope disabling /r if telemetry is disabled is also unintentional...

mort96 38 minutes ago | parent

"Rollout artifact"? This is Claude-speak isn't it? I have never ever heard anyone call a bug like this a "rollout artifact" before.

ako 35 minutes ago | parent

I was probably an agent that made the change, and the same agent that commented here on HN.

criley2 9 minutes ago | parent

I don't think "bug" is the correct term. They put a feature behind a feature flag, and feature flags don't work if you turn them off (via telemetry). That's "Working As Designed™".

davidmurdoch 12 minutes ago | parent

How are you planning to turn this off remotely when telemetry off?

chrisjj 6 minutes ago | parent

[delayed]

mgaldys4 1 hour ago | parent

Even if this was an honest rollout mistake, the design is indefensible. Reading a local file should never depend on a remote feature flag, and silently skipping it with no warning is worse. I've tried to give Claude Code the benefit of the doubt, but this crosses a line.

tehlike 1 hour ago | parent

Not everything is malicious. The author of the feature already responded on why this happened

chrisweekly 11 minutes ago | parent

I'm not the person you replied to, but your response misses their point: rollout issue aside, the approach is flawed by design. The feature author didn't address that at all.

arrowsmith 1 hour ago | parent

Claude Code also doesn't read AGENTS.md by default if there's a CLAUDE.md it can read instead. This isn't limited to your repo, e.g. if you have a ~/CLAUDE.md then no AGENTS.md will be read.

To always read both, you have to switch the 'Project instructions' setting to the non-default `claude-md-and-agents-md`.

Just in case anyone is wondering why their AGENTS.md still isn't being read.

gmponyo 1 hour ago | parent

This is not the only case where Anthropic has done stuff silently without giving users any information about changes that would hurt them.

cowpig 1 hour ago | parent

The number of people raw-dogging software that executes arbitrary instructions on their machine coming from a 3rd party server just absolutely baffles me.

The same people who've spent years of their career making sure that never happens.

pmlnr 59 minutes ago | parent

Urm... no tests caught this? How?

msp26 57 minutes ago | parent

Claude Code Remote control only works with telemetry enabled too.

fg137 31 minutes ago | parent

This in some way sounds like VSCode's bug of always adding Copilot as a co-author of git commit regardless of user settings.

If people only glance over the code agents generate for them and don't bother to spend even half a minute thinking through what's actually happening, this is inevitable.

Certainly this kind of things happened before LLMs existed. But I'm not optimistic about the direction of how things are going.

BiteCode_dev 16 minutes ago | parent

Anyway, my claude.md contains only this:

@agents.md