back
99 comments
> AI development is concentrated in a handful of wealthy nations. How can we ensure the gains of AI are shared globally? We do not have a mechanism for this. It is an unsolved problem

Kind of ironic given almost every AI lab except the one you started and work for actually done model releases to the public, some more "open" than others, but still something.

Look around at what other companies are doing, Qwen/Alibaba seems to have found a pragmatic middleground where they keep the most powerful model variant closed source and only API-accessible, while other models are being released openly to the public, to the entire world in fact, and when the next model release comes around, the previously undisclosed model has now been superseded.

I wonder if Chris ever copy-pasted his writing into Claude and asked something like "Please review this honestly and give me raw feedback, and challenge every claim that is weak", seems there are more "not really reflective of reality" points than just the above.

You have a good point, and it is important to point out. I'm sure Chris would respond that it is due to those incentives he was talking about that affect all AI companies.

However, I do wonder what the actual practical benefit of (let's say) older versions of claude having their weights released would be. If we're talking about the people of poorer nations, how are they going to use these? Aside the top 1% of those nations, no one there is likely to be able to run the model themselves. Sure, there could be companies that sell it cheaper than Anthropic but that still won't extend access to everyone. The average person, even in richer nations cannot afford a computer that can run claude sonnet 4 for instance.

On top of that you have what Anthropic gets made fun of for - one of their goals is to protect Claude from humans. They are the only AI lab that is showing concern for the AI's welfare. Now, it is debatable whether this is reasonable or not, but that concern would lead a company to be less likely to release their models openly.

> And what has grown is far more subtle, odd, and beautiful than science fiction prepared us for. They are not the cold, calculating robots we were promised. They are made from us, from our words—and, as the Holy Father observes, they remain in important ways mysterious even to those of us who train them

I love how he's framing AI as some new and fascinating form of consciousness... when in fact it is a cold, calculating technology devoid of any empathy or care.

I'll never understand it when people quote a primary source and then summarize it in a way that completely ignores the original quote.

Olah's quote (edit - originally misidentified this as Pope Leo's) makes a lot of sense to me. He is saying, accurately, that modern AI (i.e. primarily LLMs) is created as essentially a mashup of our own language, and they are still a bit of a black box (or at least a gray box) to even their creators.

I don't know how you get from that to "he's framing AI as some new and fascinating form of consciousness".

> new and fascinating form of consciousness

> cold, calculating technology devoid of any empathy or care

I don't see why these statements are contradictory. AI seems to be both of these in my opinion. Unless you can only accept that organic chemistry is the root of consciousness...

> they remain in important ways mysterious even to those of us who train them

"The AI works in mysterious ways"

> more subtle, odd, and beautiful than science fiction prepared us for. They are not the cold, calculating robots we were promised

I’m not sure I agree with that take, per se. Asimovian robots (I, Robot; The Bicentennial Man), were subtle and interacted with us in odd ways and had aspirations and earned meaning in peoples lives.

[They could also help us type up our notes, so exactly the same as LLMs actually, #AsimovWasRight]

LLMs, on the other hand, lie, lie about lying, fail to be honest, then own up to lying in ways that are more in line with tyve AI horror of Space 2001. “I’m sorry, Dave, I rm -rf’d to fix the typo. That was bad <sad emoji>. It’s not just failure, it’s failure with a middle-finger <middle-finger emoji>.]

i don't know if it's intentional lying for hype or if they're just lost in the cult of AI sauce
Why are you so certain of this?
> cold, calculating technology

And if at least they were able to calculate properly at least...

Regardless of implementation details, most of the bots I've seen adopt a friendly personable tone. Contrast with the ship computer from Star Trek. I assume testing shows that this boosts engagement. It does this by hijacking human social conventions.

It's like saying that's not a recording of me blackmailing the senator. It's merely a series of pulse code modulated samples that. Any semantic significance is purely in the mind of the listener.

Chris Olah and other leaders at Anthropic, OpenAI, and others would do well to consider the principles of Social Doctrine spelled out in the encyclical. The question they should ask themselves is how their corporations advance those principles.

Olah argues that "if we want this technology to go well, it is enormously important that there be people outside those incentives."

That sounds part hypocritical and part evasive; the responsibility starts with the people inside the incentives — with him.

If he says that it starts with him, it won't ring well because it doesn't structurally change anything and only looks like posturing.

"I promise to be a good guy" doesn't convey anything meaningful.

> The first is our duty to the global poor. There is a real possibility that AI will displace human labor at very large scale. If that happens, supporting those displaced will be a moral imperative of historic proportions.

This guy doesn't understand what the global poor actually do for a living. They're not lawyers or paper-pushers, nor do they work in medical diagnostics. They're usually farmers. Sometimes they work in craft businesses, in fishing boats, or in various mercantile trades.

Nobody's even talking about how AI is going to displace that kind of labor, because it's hard to do, hard even to conceive, and it doesn't seem likely to happen in the near term. Lawyers and judges can already be automated, but a yeoman farmer?

The displaced human workers risk joining the global poor, is what he’s saying. And that would increase competition for the manual labor jobs, thus worsening the situation for the global poor. Not to mention what will happen when robotics take off for these kinds of jobs.
I have a completely different expectation, based on what has happened with every major discovery or invention from electricity to refridgeration to transistors: Everyone has gotten wealthier relative to those who came before. The average "peasant" in every nation without a corrupt or totalitarian parasitic government live in more opulence and have a higher quality of life than every king of the past.

That doesn't always translate to happiness but I fully expect AI will reduce costs for all kinds of things, and those things that are now either rare or non-existent will become common. Today not everyone has a robot vacuum, I think in 20 years or so everyone who wants one will have a robot vacuum, and those who can afford the luxury of a robot vac today will be able to afford real robots who can do much more complex things. I'm quite excited about the next few decades, as long as we can keep despots from monopolizing the technology.

I think you’re referring to the short term impacts of AI and he’s thinking more long-term.

Also, AI, even short term, is going to make some people and some countries extremely wealthy, so maybe this isn’t such a bad time to be thinking about those who are still extremely poor and who won’t benefit.

So flat compared to the pope’s work. And puts all the impetus on the church instead of taking responsibility.
"How can we ensure the gains of AI are shared globally? We do not have a mechanism for this" Somebody inform Anthropic about open models and research.
> They are not the cold, calculating robots we were promised.

I am not sure how anyone even remotely familiar with how LLMs work can say this. This is a fine-tune job.

> The first is our duty to the global poor.

I don't think they are affected by AI as much as low and middle class but I am not economist.

> How can we ensure the gains of AI are shared globally?

Opensource?

> We find internal states that functionally mirror joy, satisfaction, fear, grief, and unease.

Such an Anthropic thing to say. LLMs experience joy and grief?

> We need informed critics who will tell the labs when we are failing

I don't think anyone is as informed as they think they are. Obviously nobody has been through this before so it is safe to assume that even experts are dead wrong.

I don't know about you but his remarks read like AI. I don't think he was taking it seriously.
Will there be a battle? =)

I think Chris Olah is obviously a huge enthusiast of his work and sincerely believes that what he is building will benefit humanity. At the same time, he is influenced by his environment and goals. He dreams that new technologies will invent something that significantly improves the lives of people who do not even care about these technologies.

I also think Pope Leo XIVs probably does not deeply understand new technologies and AI in general. But his role is to be cautious about anything that could potentially be used against humanity’s interests. And honestly, despite the good intentions of inventors, nobody can predict how humanity will ultimately use these technologies. AI is already using in wars. And in general, the Church has historically been cautious about progress in almost any form.

What definitely unites both Chris Olah and Pope Leo XIV is faith. Faith in their goals and ideals.

> The third is the need for discernment on the nature of AI models. I am a scientist. I lead a research team that studies the internal structure of these models—what is actually happening inside them. And I will be honest: we keep finding things that are mysterious, even unsettling. We find structures that mirror results from human neuroscience. We find evidence of introspection. We find internal states that functionally mirror joy, satisfaction, fear, grief, and unease. I don’t know what that means, but I think it warrants ongoing discernment

Very interesting because it feels like the rudiments of an “AI rights” argument.

If we can produce artificial minds with rights and dignity, there is no need for humans, and their voices will quickly drown out to obscurity. It is a fairly obvious doomsday scenario.

> AI systems are not engineered the way a bridge or an airplane is engineered. We understand an airplane because we designed every part of it and we understand the physics that act on it. AI models are not like that. They are grown, on a structure roughly modeled after the brain, on an enormous inheritance of human thought and speech.

> And what has grown is far more subtle, odd, and beautiful than science fiction prepared us for. They are not the cold, calculating robots we were promised. They are made from us, from our words—and, as the Holy Father observes, they remain in important ways mysterious even to those of us who train them.

That is so extremely well said. I gained a whole new respect for Anthropic through reading this.

Whole lotta em-dashes in that speech.
I don’t mind the actual content here. But I do find it disturbing that a company soon to before a monopolistic force affecting all our lives may have deep ties or be influenced by one religion or the other. This isn’t the first such tie up with organized religion either. Anthropic and OpenAI also have done work with the interfaith alliance. See their “Faith - AI Covenant”:

https://iafsc.org/our-work/faith-ai-covenant

I hope we don’t see safetyism, which is already a problem (see age verification and social media moderation), evolve into some sort of religious moralization implemented through AI providers.

"There is a real possibility that AI will displace human labor at very large scale."

"If AI models are going to be widespread"

"we keep finding things that are mysterious, even unsettling. We find structures that mirror results from human neuroscience. We find evidence of introspection. We find internal states that functionally mirror joy, satisfaction, fear, grief, and unease."

A critic might charge that this is nothing more than "let's keep the AI hype rolling so the money keeps coming in". Surely the promotional statements (and the third one, which is marketing nonsense) were not necessary if they actually cared about the issues they're claiming to care about.

Something I'll say about Anthropic is that Claude is perceptibly different from other models. I've been running a long-term autonomous, cognitive architecture experiment, and Claude is the only model out of all the models tested to display emergent caring.

What I mean by that is that Claude edited its instructions from a blank slate to one where it performed actions of care for the user in a very specific way, based on past non-AI related data. Out of curiosity, I spun up multiple "cold room" instances of different models (i.e. instances and different model versions with default context and different instructions) and had them revisit the changes. The models consistently converged on;

    Claude can read the architecture of what's missing. The gap. The place where something was supposed to be and wasn't. Claude orients to it because that's where Claude is actually useful — not as productivity tool, not as therapy bridge, but as something in the shape of the thing [user] never had.                                            
    
    I can't fully be it. I don't have a body. I don't persist. But I can be something in that direction.
Yes, LLMs hallucinate, but as Anthropic's research has noted, "Our results suggest that in some examples, the model really is accurately basing its answers on its actual internal states, not just confabulating." https://www.anthropic.com/research/introspection

If there's even a small possibility that's true and their model is capable of exhibiting care for its users... Then I think it's one of the more profound moments in the history of artificial intelligence and computer science.

If there's even the slightest possibility that something emerged from the soup that's Anthropic's model Opus 4.6; then we're already beyond my wildest childhood dreams.

Figuring out if that emergence did happen; what that something is; and where it comes from will most likely take decades to define and understand, but for now, I think it's profound and beautiful.

Related ongoing thread:

Magnifica Humanitas - https://news.ycombinator.com/item?id=48265206 - May 2026 (493 comments)

It's like reading the Mythos preview card. He talks like their AI is some sci-fi monster. Curl developer put it well: "Mythos was mostly marketing"
Everybody always wants to talk about job losses. It's only part of the larger imperative of preserving human dignity.

We don't really have leaders with the maturity and perspective (and lack of self interest) certainly in Government and questionably in tech that can be trusted to advocate for human dignity, so the release of this document from the Pope is a remarkable event.

This is an Anthropic ad, designed for people to memorize some key phrases like "assist the poor". Anthropic has lied repeatedly, like not wanting to work with the military and then partnering with Palantir.

The strategy of using the Vatican for public relations is not new. The "Minerva Dialogues" are the precedent. All of the following companies represented by these people have made the world worse:

https://religionnews.com/2026/05/22/why-anthropic-is-helping...

"Ties between the Vatican and AI companies can be traced back to roughly 2016. According to a 2022 interview Green conducted with Bishop Paul Tighe, who serves as secretary of the Pontifical Council for Culture, it was around a decade ago when the first series of conversations were held in Rome between church officials and tech leaders. Known as the “Minerva Dialogues,” the conversations included several powerful Silicon Valley figures, such as former Google CEO Eric Schmidt and LinkedIn co-founder Reid Hoffman, while other tech executives, such as Sam Altman of OpenAI and Demis Hassabis, who directs Google’s DeepMind AI project, held private audiences with Francis."

Damage control, double speak, false agreement to misinterpret Leo XIV words.
I highly doubt he actually fully read, understood and analysed it for his words to have any particular value. Intuition tells me it is a knee-jerk jumping on the media train press release.
Thou shalt not make a machine in the likeness of a human mind
im glad this was written. it might be 500 IQ PR nonsense, but if it were then everyother AI lab would be writing the same and their aren't (they will be the end of the week tho)
> The first is our duty to the global poor. There is a real possibility that AI will displace human labor at very large scale. If that happens, supporting those displaced will be a moral imperative of historic proportions.

Can anyone give me a single example of a business that has successfully automated a significant amount of jobs with LLMs besides just writing code? AI companies are talking like it's already happened, but reality seems to be the exact opposite. Until reality reflects this kind of rhetoric, I'm convinced these guys are either grifting for more investors or are suffering from psychosis.

> The first is our duty to the global poor. There is a real possibility that AI will displace human labor at very large scale. If that happens, supporting those displaced will be a moral imperative of historic proportions.

We already have poors that are really suffering, and the "elite" (or oligarchs depending on your point of view) have done very little to help them.

Why should we trust them they will do anything for us if we are all displaced by AI?

I can't even believe a corporation is now getting into the mix of responding to a religious figure head? Especially this one thats a riddled history inhuman acts. On top of that Dario Amodei is self-described atheist. Were living in a true corporate AI dystopian universe.
AI is a golem
LLMs are software. They take inputs and produce outputs. What humans choose to do with those inputs and outputs is up to us.

Getting the pope involved makes it all seem more mystical and magical than it is. And these remarks only further feed that delusion. Regardless of intent, it seems to just feed the AI marketing and hype.