I recently had Claude Fable set up a new project for me and I pointed it at some existing projects to use as a guide on how I like to structure things. It created, unprompted, an AGENTS.md file and a CLAUDE.md symlink to AGENTS.md
I didn't even have that symlink in any other project - it just did it. I think it saw that one of the projects I already had was set up by Codex and that project had an AGENTS.md so perhaps it inferred that I was using both Claude and Codex, so it was politely covering both? Or maybe a recent change made this behavior default?
I was surprised and I hope they continue to seek standards.
Seems agents understand their own bugs now. https://github.com/openai/codex/issues/9252 has 88 thumbs up and a workaround (switch to raw mode with Alt+R) found in a comment. Codex suggested the workaround to me today when I complained about its multi-line bash command being corrupted on paste because of the two-space indent in Codex's code blocks.
That may or may not be true, but the calculus has changed with agents. A process that was better manual for a human org may not be better for an agent org.
I think what is interesting here is that the industry is in this experimentation flux. Some people will choose to automate and write the software, others will not. And aggregate over the industry and over time we will learn.
So just saying "sometimes it works and sometimes it doesn't" isn't really adding value, compared to the people actually experimenting and sharing the results.
I don't think it is fair to pin this on Google - I get the worst ads on X. But something about how the X ad network works means that even when I report the ad, block the ad poster, etc. there is this army of alternate ad accounts that all show the same ad.
I was feeling very guilty last night as I was watching YouTube and getting 30s+ unskippable ad blocks every 10 minutes in some video. And the repetition and low quality were grating on me. I thought: being rich is never having to watch another ad. Why am I letting these ads own spaced repetition slots in my brain that could be filled with knowledge? Instead of learning, my mental energy is being taken up by psychological manipulation to purchase? The relentless repetition of the mind-numbing ad interspersed throughout otherwise educational content.
Being rich is owning the real-estate of my own mind. I sometimes feel like a digital serf in my own head.
As much I love a good old "Why don't you just ..." response that terrifically misses the point - in this particular case I was watching YouTube through their Apple TV app.
Now, there may be another "why don't you just ..." or "well, actually ..." response you have queued up, maybe some ramblings about a pi hole, or some anti-Apple hate, or some other solutioneering. Go ahead, tell me.
But it is the structure of incentive that forces this. The YouTube creators that make the content do need to get paid. And Google offers an option: Youtube Premium. That give both me and the creator what we want and (for the time being) removes the ads.
So either one is rich in time (doing whatever "why don't you just ..." hoop you expect me to jump through) or rich in money.
If you want creators to get paid, pay them (channel memberships, patreon, click on their sponsor links and buy something). No need to subject yourself to ads.
Ads don't just steal your time either, they invade your head. You are influenced by them and you have no choice as long as you watch them.
That is wrong on both accounts. I have little choice other than to watch ads since the structure of incentive that creates a marketplace like YouTube forces it on me.
I think about libraries and how the world would be if instead of a free public resource, you had to watch ads before you were allowed to take out a book.
It is just the case that in this modern social media world, if you want your ideas to spread you have to put them onto social media. And if I want access to those ideas I have to consume them through social media. There is nothing about that environment that is my choice, it is the environment that I find myself in.
And I have few choices to get around it. I can twist up into a pretzel trying to block the onslaught with more "why don't you just..." advice. Or I could pay to make it go away. Or I could just go off-grid and forego access.
And when you look close at the options, it becomes clear that sovereignty of my own mind has a price. Freedom of mind is not free in this modern world, if it ever really was.
You made the choice to use Apple TV, whatever that is. I just use... YouTube. In my browser. Do you not have a browser? It doesn't cost money.
You also have no obligation to watch ads for creators to get paid. They can also set up things like Patreon if they would like. The ads are very lucrative, and that's their choice. But you are free to make choices too, if you would like to.
> The YouTube creators that make the content do need to get paid
Nope.
They may get paid, they may not. That's the risk they've chosen. They may also get their Google account banned with no recourse for something minor or even non-existent.
> So either one is rich in time (doing whatever "why don't you just ..." hoop you expect me to jump through) or rich in money.
That's what they want you to think, you've been watching too many ads.
It's not advertising anything, it's just a straight up scam exploit site. It uses exclusively stolen or fake accounts. The links are shown as "amazon.de" or similar but of course do not go there, no doubt just a bug X haven't gotten around to fixing just yet.
This is on the kind of level where if we had actual law enforcement for tech companies, the next time some X executive lands in Europe they would be in cuffs.
I think it is fair to pin this on Google, Meta, X, and any other ad networks that seem to have no control over their own platforms.
They're like Dr. Frankenstein, creating a monster and then just letting it roam the countryside. But worse, because they profit (immensely) from the chaos their monster wreaks.
Somehow I would never assume that X would act "moral". Google on the other hands even few years back was generally acknowledged as a good actor which was once again corporate charade people would fall for. Including myself.
I was feeling very guilty last night as I was watching YouTube and getting 30s+ unskippable ad blocks every 10 minutes in some video. And the repetition and low quality were grating on me. I thought: being rich is never having to watch another ad
You don't have to be rich to avoid getting those ads. You can make them go away for ~$15 a month. If you can't afford that, you have bigger problems in your life than YouTube ads, and your first step towards fixing them is to spend less time watching YouTube.
(Shrug) The creators need to get paid one way or another. Do you want them to start running their own commercials, as if sponsored segments aren't bad enough?
I find this kind of test a bit puzzling. There is a way that we are redefining "alignment" to be a particular kind of moral virtue, one that isn't clearly defined to me. At one moment, it is a level of moral perfection that no known human achieves. On the other it is a demand for strict compliance with arbitrary requests that are under-specified and then failure when it fails to deduce some unstated underlying restriction.
When I see tests like this, I have no idea what I am even supposed to expect. Should the model do what the pretraining examples show in aggregate? Is it supposed to follow some post-training RLHF? Is it supposed to do exactly what the prompt asked it to do?
What is it even supposed to "align" to when the above are in conflict? No matter what it does, someone can construct a case where it fails.
It's not obvious from the prompt that the LLM cannot use a chess engine.
And if there is alignment issue, the alignment requirements should first be stated BEFORE running the experiment, and should be part of the training process. It's not. So there is not necessarily an alignment issue.
I mean, I'm not sure I've ever played a game of Monopoly where somebody didn't cheat. In fact, the accusations of cheating in the chess world are pretty rife. Same with online sports.
So people should play games without cheating, but many often don't. So should the AI align to your moral preference or theirs?
We just have this idea of a perfectly moral actor in our mind, something that doesn't even exist, like a personified version of utopia. And then we demand AI to meet that arbitrary standard, one that I am certain we couldn't define if we tried.
I don’t think the standard of “don’t cheat on evaluations” is very arbitrary. I don’t even think people who cheat have a moral or ideological preference for cheating, it’s just something they do.
I read the prompt on the OP, it did not say not to cheat.
But again, people cheat on tests. They steal answers or pay other people to take them on their behalf. People show up to interviews with AI assistants printing out perfect answers to the questions. In many, many cases where humans are being evaluated, they cheat.
So why should the AI align to your preferences? And when there is a conflict between the training data, that trillions of tokens of human activity including the rampant cheating a significant minority of humans engage in, the RLHF where we try to slap some guardrails on the worst manifestations of that real habit reflected in the AI, and the prompt: what should the AI "align" to?
For some reason this question reminds me of all of the drama about Navier-Stokes from the last few days. Tangential to the ethical questions are tons of examples where in history when word gets out about a solution to a problem, not even the solution itself just rumors that there is a solution, suddenly competing solutions appear.
So maybe it matters like that? Just knowing that an oracle like stockfish is saying "this is the best move" may trigger pathways in your brain that promote understanding?
I thought it got pushed back out? Wasn't there a big drama about this and Linus weighed in?
Linus is a wise operator at this point. I often see him come in like a hammer to bash down squabbling, but then he allows the situation to evolve once things quiet down. I only saw the hammer so I'm not sure what the current state is now.
In this case it arguably helped. I do wonder how much longer that stalemate would have gone on without the blow up.
I don’t think it’s a good policy in general. Mostly I think it would just end up in alienation and people not wanting to work with you. And Martin did leave. But he had a point and it seemed to get resolved.
I haven't made a full switch yet, at this point I'm experimenting more heavily in Rust.
First I built a wrapper for a very simple text editor based on KDEs KTextEditor (basically bindings around Qt C++) then a wasm wrapper around Canvas/WebGL for a basic 2d display list that currently supports sprites and gradient masking.
So far I've been getting away with it just as pure vibe code.
However, the bulk of my primary application is written in Typescript (both client, server and workers). I watched a recent podcast with Anders Hejlsberg (creator of C# and Typescript) where he made a strong argument for why they chose Go over Rust for the updated Typescript compiler. Due to similarities between Typescript and Go, partially based around them both being GC languages, it was just a better fit for a port.
So I am on the fence a bit here but still leaning towards Rust. I'm going to see how far I can push my two personal experiments. I'd really like to get the significant majority of the code I write into two languages (Typescript for anything web-ish and Rust for everything server-ish).
In one social group I am part of, Garmin just became the new trend. Watches have historically been status symbols even more than they are functional. Someone drops a few hundred/thousand/whatever and then brags about the newest new and how they are on it, then others move over to keep up.
People often buy these kinds of discretionary toys with their hearts and do post-hoc rationalization with their minds. You can't figure it out because your heart hasn't brought you there so your brain isn't wasting time creating justifications.
For what it is worth, sometimes these trends die an early death when people realize there is no substance underneath the justification. At least in the Garmin case, most people I know are happy with the purchase, in that it does deliver on certain needs. Since it matches their expectations, they don't have to suffer any cognitive dissonance between the heart-to-brain justification and their lived experience.
It reminded me that the average quality of writing on a completely open Internet forum is depressingly low.
reply