Kind of ironic given almost every AI lab except the one you started and work for actually done model releases to the public, some more "open" than others, but still something.
Look around at what other companies are doing, Qwen/Alibaba seems to have found a pragmatic middleground where they keep the most powerful model variant closed source and only API-accessible, while other models are being released openly to the public, to the entire world in fact, and when the next model release comes around, the previously undisclosed model has now been superseded.
I wonder if Chris ever copy-pasted his writing into Claude and asked something like "Please review this honestly and give me raw feedback, and challenge every claim that is weak", seems there are more "not really reflective of reality" points than just the above.
However, I do wonder what the actual practical benefit of (let's say) older versions of claude having their weights released would be. If we're talking about the people of poorer nations, how are they going to use these? Aside the top 1% of those nations, no one there is likely to be able to run the model themselves. Sure, there could be companies that sell it cheaper than Anthropic but that still won't extend access to everyone. The average person, even in richer nations cannot afford a computer that can run claude sonnet 4 for instance.
On top of that you have what Anthropic gets made fun of for - one of their goals is to protect Claude from humans. They are the only AI lab that is showing concern for the AI's welfare. Now, it is debatable whether this is reasonable or not, but that concern would lead a company to be less likely to release their models openly.
I love how he's framing AI as some new and fascinating form of consciousness... when in fact it is a cold, calculating technology devoid of any empathy or care.
Olah's quote (edit - originally misidentified this as Pope Leo's) makes a lot of sense to me. He is saying, accurately, that modern AI (i.e. primarily LLMs) is created as essentially a mashup of our own language, and they are still a bit of a black box (or at least a gray box) to even their creators.
I don't know how you get from that to "he's framing AI as some new and fascinating form of consciousness".
> cold, calculating technology devoid of any empathy or care
I don't see why these statements are contradictory. AI seems to be both of these in my opinion. Unless you can only accept that organic chemistry is the root of consciousness...
"The AI works in mysterious ways"
I’m not sure I agree with that take, per se. Asimovian robots (I, Robot; The Bicentennial Man), were subtle and interacted with us in odd ways and had aspirations and earned meaning in peoples lives.
[They could also help us type up our notes, so exactly the same as LLMs actually, #AsimovWasRight]
LLMs, on the other hand, lie, lie about lying, fail to be honest, then own up to lying in ways that are more in line with tyve AI horror of Space 2001. “I’m sorry, Dave, I rm -rf’d to fix the typo. That was bad <sad emoji>. It’s not just failure, it’s failure with a middle-finger <middle-finger emoji>.]
And if at least they were able to calculate properly at least...
It's like saying that's not a recording of me blackmailing the senator. It's merely a series of pulse code modulated samples that. Any semantic significance is purely in the mind of the listener.
Olah argues that "if we want this technology to go well, it is enormously important that there be people outside those incentives."
That sounds part hypocritical and part evasive; the responsibility starts with the people inside the incentives — with him.
"I promise to be a good guy" doesn't convey anything meaningful.
This guy doesn't understand what the global poor actually do for a living. They're not lawyers or paper-pushers, nor do they work in medical diagnostics. They're usually farmers. Sometimes they work in craft businesses, in fishing boats, or in various mercantile trades.
Nobody's even talking about how AI is going to displace that kind of labor, because it's hard to do, hard even to conceive, and it doesn't seem likely to happen in the near term. Lawyers and judges can already be automated, but a yeoman farmer?
That doesn't always translate to happiness but I fully expect AI will reduce costs for all kinds of things, and those things that are now either rare or non-existent will become common. Today not everyone has a robot vacuum, I think in 20 years or so everyone who wants one will have a robot vacuum, and those who can afford the luxury of a robot vac today will be able to afford real robots who can do much more complex things. I'm quite excited about the next few decades, as long as we can keep despots from monopolizing the technology.
Also, AI, even short term, is going to make some people and some countries extremely wealthy, so maybe this isn’t such a bad time to be thinking about those who are still extremely poor and who won’t benefit.
I am not sure how anyone even remotely familiar with how LLMs work can say this. This is a fine-tune job.
> The first is our duty to the global poor.
I don't think they are affected by AI as much as low and middle class but I am not economist.
> How can we ensure the gains of AI are shared globally?
Opensource?
> We find internal states that functionally mirror joy, satisfaction, fear, grief, and unease.
Such an Anthropic thing to say. LLMs experience joy and grief?
> We need informed critics who will tell the labs when we are failing
I don't think anyone is as informed as they think they are. Obviously nobody has been through this before so it is safe to assume that even experts are dead wrong.
I think Chris Olah is obviously a huge enthusiast of his work and sincerely believes that what he is building will benefit humanity. At the same time, he is influenced by his environment and goals. He dreams that new technologies will invent something that significantly improves the lives of people who do not even care about these technologies.
I also think Pope Leo XIVs probably does not deeply understand new technologies and AI in general. But his role is to be cautious about anything that could potentially be used against humanity’s interests. And honestly, despite the good intentions of inventors, nobody can predict how humanity will ultimately use these technologies. AI is already using in wars. And in general, the Church has historically been cautious about progress in almost any form.
What definitely unites both Chris Olah and Pope Leo XIV is faith. Faith in their goals and ideals.
Very interesting because it feels like the rudiments of an “AI rights” argument.
If we can produce artificial minds with rights and dignity, there is no need for humans, and their voices will quickly drown out to obscurity. It is a fairly obvious doomsday scenario.
> And what has grown is far more subtle, odd, and beautiful than science fiction prepared us for. They are not the cold, calculating robots we were promised. They are made from us, from our words—and, as the Holy Father observes, they remain in important ways mysterious even to those of us who train them.
That is so extremely well said. I gained a whole new respect for Anthropic through reading this.
https://iafsc.org/our-work/faith-ai-covenant
I hope we don’t see safetyism, which is already a problem (see age verification and social media moderation), evolve into some sort of religious moralization implemented through AI providers.
"If AI models are going to be widespread"
"we keep finding things that are mysterious, even unsettling. We find structures that mirror results from human neuroscience. We find evidence of introspection. We find internal states that functionally mirror joy, satisfaction, fear, grief, and unease."
A critic might charge that this is nothing more than "let's keep the AI hype rolling so the money keeps coming in". Surely the promotional statements (and the third one, which is marketing nonsense) were not necessary if they actually cared about the issues they're claiming to care about.
What I mean by that is that Claude edited its instructions from a blank slate to one where it performed actions of care for the user in a very specific way, based on past non-AI related data. Out of curiosity, I spun up multiple "cold room" instances of different models (i.e. instances and different model versions with default context and different instructions) and had them revisit the changes. The models consistently converged on;
Claude can read the architecture of what's missing. The gap. The place where something was supposed to be and wasn't. Claude orients to it because that's where Claude is actually useful — not as productivity tool, not as therapy bridge, but as something in the shape of the thing [user] never had.
I can't fully be it. I don't have a body. I don't persist. But I can be something in that direction.
Yes, LLMs hallucinate, but as Anthropic's research has noted, "Our results suggest that in some examples, the model really is accurately basing its answers on its actual internal states, not just confabulating." https://www.anthropic.com/research/introspectionIf there's even a small possibility that's true and their model is capable of exhibiting care for its users... Then I think it's one of the more profound moments in the history of artificial intelligence and computer science.
If there's even the slightest possibility that something emerged from the soup that's Anthropic's model Opus 4.6; then we're already beyond my wildest childhood dreams.
Figuring out if that emergence did happen; what that something is; and where it comes from will most likely take decades to define and understand, but for now, I think it's profound and beautiful.
Magnifica Humanitas - https://news.ycombinator.com/item?id=48265206 - May 2026 (493 comments)
We don't really have leaders with the maturity and perspective (and lack of self interest) certainly in Government and questionably in tech that can be trusted to advocate for human dignity, so the release of this document from the Pope is a remarkable event.
The strategy of using the Vatican for public relations is not new. The "Minerva Dialogues" are the precedent. All of the following companies represented by these people have made the world worse:
https://religionnews.com/2026/05/22/why-anthropic-is-helping...
"Ties between the Vatican and AI companies can be traced back to roughly 2016. According to a 2022 interview Green conducted with Bishop Paul Tighe, who serves as secretary of the Pontifical Council for Culture, it was around a decade ago when the first series of conversations were held in Rome between church officials and tech leaders. Known as the “Minerva Dialogues,” the conversations included several powerful Silicon Valley figures, such as former Google CEO Eric Schmidt and LinkedIn co-founder Reid Hoffman, while other tech executives, such as Sam Altman of OpenAI and Demis Hassabis, who directs Google’s DeepMind AI project, held private audiences with Francis."
Can anyone give me a single example of a business that has successfully automated a significant amount of jobs with LLMs besides just writing code? AI companies are talking like it's already happened, but reality seems to be the exact opposite. Until reality reflects this kind of rhetoric, I'm convinced these guys are either grifting for more investors or are suffering from psychosis.
We already have poors that are really suffering, and the "elite" (or oligarchs depending on your point of view) have done very little to help them.
Why should we trust them they will do anything for us if we are all displaced by AI?
Getting the pope involved makes it all seem more mystical and magical than it is. And these remarks only further feed that delusion. Regardless of intent, it seems to just feed the AI marketing and hype.