Tag Archives: Science

Treating Populations, Not People

I’ve been following Nick Norwitz, MD/PhD, for a while now. He’s quite the researcher, that’s for sure. And his latest video covering his cholesterol situation is wild: I Bet My Life Against Cholesterol Dogma (And Won)

In the past Nick had severe inflammatory bowel disease, which altered his life significantly. At one point he dropped below 100 pounds, ended up in intensive care, and drove him to question his life. I totally understand. I’ve been there a few times myself. So, out of desperation he tried a therapeutic ketogenic diet, which sent his disease into full remission. Now, that fact alone isn’t supposed to be possible, but it’s actually the side story at the moment.

The real issue here turned out to be his cholesterol. As a result of that new diet, Nick’s LDL (low-density lipoprotein) climbed from a so-called healthy 90 to near 600, while his total cholesterol tapped out at 700! Under traditional medical science, those numbers represent an emergency right now and over time a death sentence from a heart attack or a stroke. Treatment was needed right away. But Nick made a different decision. As he says, for seven years he has been “running an experiment that should be killing me.” But he kept at it because “the thing that is causing my cholesterol to go so high is also saving my life.” That’s quite a position to be in. But it’s also one in which many people find themselves when dealing with two intractable medical conditions that doctors can’t simultaneously treat without side effects.

Since Nick’s been dealing with massive LDL for so long while also declining standard treatment with statin drugs, I was curious why he hadn’t yet gotten a cardiac calcium scan. Those scans are fast and relatively inexpensive. Why wouldn’t he want to check for plaque in his coronary arteries? It just seemed odd to me given how comprehensively he researches his own conditions.

Well, recently, he finally went ahead and got his scan. The result? Zero! No plaque at all! Personally, I wasn’t surprised because I’m familiar with his research, and I know others who have had similar results from similar conditions. But I’m sure he was relieved given his own history and also living with being attacked constantly for his unique approach. The result of zero plaque “doesn’t just poke a cholesterol dogma, it shatters it,” he says. It does, indeed. But medicine changes slowly. Perhaps in a hundred years or so the protocols will change to embrace new research. For now, though, he’ll likely remain an aberration, just like all the other aberrations out there. The difference here, though, is that Nick publishes prolifically, so he’s leaving a detailed scientific paper trail that other researchers are noticing.

Anyway, it’s a good story on its own for Nick’s gut and his heart. But the part I want to point to comes later when he steps back and explains why he thinks this happened. His argument concerns context. A single blood marker like cholesterol, or even height, only means something within a particular person’s situation. For example, Nick uses basketball superstar Shaquille O’Neal to illustrate the point. Shaq is seven foot one because of lucky genes. But Nick says that other people could be the same height because of a tumor messing with their growth hormones while the tumor eventually quietly kills them. Same measurement, different reality. “Context is everything.”

Then Nick makes the point that stuck with me because I’ve experienced it many times myself dealing with the medical industrial complex. It’s a very real phenomenon. Modern medicine, he says, often misses context entirely when dealing with individual patients. The context of the person gets traded away for algorithms that let the system run efficiently at scale as it implements particular protocols or policies. And as much as individual doctors may want the best for their patients, and Nick believes they do, “modern medicine treats populations, not people.”

That’s the line right there. I got it right away. Experienced it many times. Painfully. Nick says that treatment protocols get built on averages drawn from large groups of mostly sick people. That may work well enough for most people most of the time to get them out of some acute condition or to enable them to live longer with a chronic condition. But it fails the outliers. These are the people whose numbers on one test look alarming for a reason the average never accounts for while their other markers are exceptionally healthy. In Nick’s case, that reason is high cholesterol driven by a therapeutic diet rather than any other known disease. He returns to the idea near the end and tells us this: “Don’t settle for being treated like the population average, because almost none of us are.” He’s right. But the “settle” bit is challenging because the medical community as a system isn’t so warm and fuzzy when confronted with people who question it.

There’s much more in the video. He briefly reviews his experiment where he added Oreo cookies to his diet that cut his cholesterol sharply, which he offers as an indicator that his fat-burning metabolism may be the real driver to his cholesterol markers breaking the established reference ranges. He also explains why he walked away from a traditional medical career to instead work as a medical communicator. On placing that bet against decades of established cardiology, he says, “The data judges the ideas, not the credentials backing them.” His YouTube channel has over a million subscribers now, so he’s touching more people on any given day than would be possible had he gone into clinical medicine.

Systems are built to scale. But they rarely consider individuals, especially individuals whose conditions don’t fit published protocols. That’s why it’s critical to always do your own research, question your doctor, and act in your own interest. The doctors may care about you to a certain degree, but the protocols they implement don’t.

Good luck!

Bill Joy’s Future

Was Bill Joy Correct Back in 2000? What His Warning Means for Science, Technology, and AI Today.

Since there’s a lot of distracting and unfortunate AI doom in the media these days, I figured I’d go back and revisit Bill Joy in April 2000 for some history on the pending catastrophes we keep hearing about today. I remember that Joy’s massive article “Why the Future Doesn’t Need Us” hit really hard twenty-six years ago. Joy was a cofounder of the iconic Sun Microsystems, after all. He helped build the Internet. And here he was now warning that three technologies might eventually lead to human extinction. Unlike today, though, that negative perspective was a pretty novel idea back then. The technologies he cited that could end us all included robotics, genetic engineering, and nanotechnology. His core argument was actually pretty simple. These were not like past technologies. These new things could potentially replicate themselves, and that’s the bit that would change everything if they were used as weapons or just got loose by accident.

I first read Joy’s article at a coffee shop in Cupertino, California right across the street from Sun where I worked in software systems marketing. The article was widely read at Sun and also across Silicon Valley and for months also generated wild discussions about Joy and his analysis. Many people just called him crazy. But that only demonstrates to me that those who made such flippant statements never read his article or thought deeply about his arguments. Others knew better, though. They knew full well that Joy was documenting in detail the very real risks of rapidly developing technology and the consequences of ignoring them.

That is, of course, the standard and pervasive culture of Silicon Valley. The valley may be big, but it’s remarkably insular as well. I didn’t know Joy at the time I read his article, but I went on to meet him several times at Sun and worked closely with his teams promoting projects like SPARC, Solaris, Java, Jini, and later on JXTA. I didn’t really know him well, of course, but he was always friendly and professional to me. He was quiet, too, and I always found him a serious thinker who obviously knew far more than he ever expressed. Sun was filled with such characters. They all fascinated me to no end.

The timing of the Joy article is also interesting. He said he had been working on the essay since his 1998 conversation with Ray Kurzweil and continued revising drafts through 1999. When Wired published the piece in April 2000, the tech world was at its peak. The NASDAQ hit a high of 5,048 in March of 2000 just a few weeks before the article dropped. And at that time Sun’s stock reached $250 a share, which gave the company a market cap around $200 billion. That was a significant achievement for 2000. Sun was one of the hottest companies in the valley back then, and it was quite an experience working there. The place was buzzing with activity. I loved it. Many of us did. So, it was into that environment of overt tech optimism running at manic levels that Joy published his thoughts about our potentially perilous future.

Andy Bechtolsheim, Vinod Khosla, Scott McNealy, and Bill Joy at the Sun Reunion in Silicon Valley in October 2019. Photo by Jim Grisanzio.

Boom!

Then everything blew up. The bubble burst. By mid-April 2000 the NASDAQ suffered its worst week in history and dropped more than 25 percent. Companies started dumping workers like I’ve never seen before. Joy’s dark warnings about unchecked technology landed precisely as that optimism crashed. He obviously couldn’t have seen the future, but his timing was remarkable.

Joy’s warnings took on an even darker tone the following year. After publishing the Wired article, he signed a book contract to expand on the concepts. He moved into a hotel room in New York City and surrounded himself with gloomy books covering plagues and nuclear bombs and other such material he was studying about the future risks of technology. Then on September 11, 2001 came the terrorist attacks we’ve come to know as 9/11. I knew a few people who worked at the Sun building in New York City, but I didn’t know Joy was also in the city at the time. He said he stood in the streets with everyone else and watched the impossible happen in real life. The next morning he went back outside and observed a long line of sanitation trucks parked on Houston Street ready to haul away the rubble. Everything below 14th Street was closed, he said. “It was quite a compelling experience, but not really, I suppose, a surprise to someone who had his room full of the books I was reading,” he said in a TED Talk. “I was not surprised that it happened at all.”

Joy eventually abandoned the book project. I point this out just as an aside since the event occurred shortly after he published the article I’m writing about here in this post. Still, it does reflect the feeling of the times. How much had changed in Silicon Valley and the United States in just one year.

Who Is Bill Joy?

Joy was born in 1954 in Michigan, and he was considered a child prodigy. He started school early, was reading by age three, and later excelled in math and science. He even graduated high school at 16. He loved books and thinking, and that became his escape from the world. He also loved science fiction and devoured Heinlein’s “Have Spacesuit, Will Travel” and Asimov’s “I, Robot” with its Three Laws of Robotics. He wanted to be a ham radio operator, which were the Internet hackers of their day, but he couldn’t afford the equipment. On TV, Star Trek inspired his imagination, and Gene Roddenberry’s “The Prime Directive” clearly resonated with him. You can actually see that ethic woven into his writing thereafter.

At Berkeley in the 1970s, he created the vi text editor, which to his surprise, was still widely used more than twenty years later Some hard core developers still use it even now. He also developed the Berkeley version of the Unix operating system and was a key contributor to the TCP/IP network stack. When the other founders of Sun Microsystems (Andy Bechtolsheim, Vinod Khosla, and Scott McNealy) invited him to join them, he participated in the creation of advanced microprocessor technologies and software technologies such as Java and Jini. As co-designer of three microprocessor architectures — SPARC, picoJava, and MAJC — he helped drive innovations that shaped modern computing.

By the time he wrote his famous Wired essay, Joy was only 45 years old and at the peak of his influence among developers in Silicon Valley. But Joy was far more than just a coder. He was well connected to the broader scientific community as well. That’s what made his article so jarring to so many people. He was not an uninformed critic chiming in from outside with yet another opinion in the media, which we’re all familiar with today as we read the news. He was a core architect of the digital age who was expressing deep doubts about where his own work was leading. His self-reflection was pervasive during this time in his writings and during his conference presentations.

The Kurzweil Meeting

Joy’s concern seemed to begin at George Gilder’s Telecosm conference in 1998 when he met Ray Kurzweil, who was an inventor and futurist. Kurzweil talked about how the rate of technological improvement was accelerating and also how humans may merge with robots or download their consciousnesses to achieve near immortality. I remember attending several talks on this topic of immortality when I moved to Silicon Valley in early 2000. It always sounded so silly to me. I wondered how such smart people could take that stuff so seriously. But even now some people in these circles talk about downloading themselves. It still sounds silly. The other bits, though, about intelligent robots and genetic engineering were much more reasonable given my own experience working in the biotech industry before I joined Sun. Joy had heard such talk before and always felt sentient robots were science fiction. But hearing it from someone he respected changed things. Kurzweil gave him a preprint of “The Age of Spiritual Machines,” which outlined a utopian future where humans gained near immortality by becoming one with robotic technology. I don’t know how far Joy goes with respect to robotic sentience, but it’s clearly more than I’m willing to accept.

Nevertheless, Joy’s unease intensified after reading the book. He felt sure Kurzweil was understating the dangers. Then he found a passage in the book describing a dystopian future where machines become so capable that humans depend on them completely. The passage argued that we wouldn’t consciously hand over control to the bots. Instead, “the human race might easily permit itself to drift into a position of such dependence on the machines that it would have no practical choice but to accept all of the machines’ decisions.” In other words, we would gradually grow dependent. I can surely see that as a potential reality, no question about it.

But that passage came from Ted Kaczynski, the Unabomber! Joy admits this realization was uncomfortable to say the very least since he was taking a point from a terrorist seriously. Many people have said the same thing after reading Kaczynski’s words, and he’s actually still cited even today. Kaczynski’s bombs had killed three people and wounded many others. One bomb gravely injured David Gelernter, one of Joy’s colleagues and friends. But Joy felt compelled to confront the argument because, however uncomfortable, he saw merit in that single passage about the unintended consequences of technology.

The Self-Replication Problem

Joy’s main concern centers around one key difference between powerful 21st-century technologies and those of the 20th-century. Nuclear weapons required huge facilities and rare materials. But genetic engineering, nanotechnology, and robotics, what Joy called GNR technologies, require less infrastructure and can potentially make copies of themselves. A bomb explodes once, but a self-replicating machine doesn’t stop. And that could be a serious problem if something goes wrong.

This matters because knowledge spreads freely. You can’t control ideas at all like you may be able to control uranium. Once people know how to genetically engineer bacteria or design tiny self-replicating machines, that knowledge exists in the world and will move rapidly. A small group, even one person, could potentially cause massive harm. Joy calls this “knowledge-enabled mass destruction.” As he wrote, “I think it is no exaggeration to say we are on the cusp of the further perfection of extreme evil, an evil whose possibility spreads well beyond that which weapons of mass destruction bequeathed to the nation-states, on to a surprising and terrible empowerment of extreme individuals.”

He made that point even sharper later. The real danger, he said, is no longer nation-states but individuals or small groups now empowered with “pandemic power.” These new digital, self-replicating technologies give extreme individuals the kind of destructive capability once reserved only for governments. That shift changes everything because the ramifications of mistakes or ill intent can’t be calculated or controlled.

The advancement of technology was clearly moving faster and Joy knew it. He learned about complex systems and non-linear systems from physicists Stephen Wolfram and Brosl Hasslacher in the early 1980s. These are systems where small changes move in unpredictable ways and where feedback loops create unexpected outcomes. Thus, they are extremely difficult to predict. Later, Joy deepened his understanding of these issues after conversations with Danny Hillis, a pioneer of parallel supercomputers and co-founder of the Long Now Foundation, biologist Stuart Kauffman, and Nobel laureate Murray Gell-Mann. Hasslacher and Mark Reed, a leading researcher in molecular electronics at Yale, also gave him insight into molecular electronics, which is the manipulation of matter at the atomic and molecular level where individual atoms replace transistors. When you get to this point in Joy’s article you can’t help but realize that he’s going well beyond just being a smart software developer who happened to strike it rich by helping found a successful tech company in the valley.

Joy knew that the merger of computers and physical sciences was creating enormous power, which up to that point others hadn’t expressed clearly in the popular tech media. By 2030, Joy calculated, we would likely build machines a million times as powerful as personal computers of 2000. That’s enough computing power to make the scenarios that worried him technically possible. As he wrote, “But now, with the prospect of human-level computing power in about 30 years, a new idea suggests itself: that I may be working to create tools which will enable the construction of the technology that may replace our species. How do I feel about this? Very uncomfortable.” He admitted he had once been too optimistic about nanotechnology. Having struggled his entire career to build reliable software systems, it seemed to him more than likely that this future would not work out as well as some people had imagined.

Three Scenarios

Joy walks through what could actually happen with these technologies. What’s interesting is that evolution itself would be one of the driving forces moving these technologies from a positive outcome to something very negative in the wrong hands.

One — Robots might simply out-compete humans for resources the way better-adapted species have always displaced others. We wouldn’t need a robot uprising. Hans Moravec, a roboticist and futurist at Carnegie Mellon, argued that “in a completely free marketplace, superior robots would surely affect humans as North American placentals affected South American marsupials.” So, economic forces alone could push us aside. Also, the dream of robotics includes downloading our consciousnesses into machines. But Joy questions whether a downloaded consciousness would be human in any meaningful sense. The robots would not be our children, and on that path our humanity might be lost entirely. This is one of the reasons why I take Joy seriously. He sees the obvious problem with the downloading issue in a way that others in this field simply do not.

Two — Genetic engineering gives us power to create devastating plagues, either by accident or intention. Joy calls this the “White Plague” scenario, which references a Frank Herbert novel where a molecular biologist weaponizes his knowledge. We now know these profound changes in biological sciences could be imminent and could challenge all of our notions of what life is, and Joy points out that the public remains skeptical even by the standards of 2000.

Three — Nanotechnology could produce “gray goo,” which are self-replicating nanobots that consume the biosphere. Eric Drexler, an engineer and pioneer of molecular nanotechnology, warned in 1986 that “tough omnivorous ‘bacteria’ could out-compete real bacteria. They could spread like blowing pollen, replicate swiftly, and reduce the biosphere to dust in a matter of days.” Joy notes grimly, “Gray goo would surely be a depressing ending to our human adventure on Earth, far worse than mere fire or ice, and one that could stem from a simple laboratory accident. Oops.”

The Manhattan Project Parallel

Joy uses the atomic bomb as his template. After Hiroshima and Nagasaki, physicists were shocked by what they created. Oppenheimer later said the physicists “have known sin.” But there was a real opportunity to prevent a nuclear arms race by internationalizing nuclear power as documented in 1946 via the Acheson-Lilienthal report and the Baruch Plan. It failed, though, because political distrust and competitive pressure got in the way. Within years, the Soviets had the bomb, and the arms race was on. This is a pattern that will repeat in the future.

Freeman Dyson, a theoretical physicist who worked on nuclear weapons and later advocated arms control, captured the moment: “The glitter of nuclear weapons. It is irresistible if you come to them as a scientist. To feel it is there in your hands, to release this energy that fuels the stars, to let it do your bidding. To perform these miracles, to lift a million tons of rock into the sky. It is something that gives people an illusion of illimitable power, and it is, in some ways, responsible for all our troubles, this, what you might call technical arrogance, that overcomes people when they see what they can do with their minds.”

Up to that point, everyone feared nuclear bombs as the ultimate expression of madness. But Joy feared that we were repeating the pattern with even more dangerous technologies, and the commercial incentives for their production would be enormous. Nations compete. Corporations compete. Individuals compete. Researchers want and need breakthrough innovations. The momentum builds, and pretty soon it’s almost impossible to stop. We are being propelled into the new century with no plan, no control, no brakes. However, the driver is not military necessity this time but instead private sector economic gain and competitive pressure. This is human nature, though. Our history is literally filled with these processes. It’s who we are. Joy was simply stating the obvious.

The Relinquishment Argument

This is the part that bothered many people the most. Joy argues for “relinquishment,” or the voluntary decision not to pursue certain lines of knowledge or technology because they are too dangerous. This goes against everything we believe about the value of knowledge and open inquiry, especially in Silicon Valley and across the scientific community generally.

That may be why Joy’s article struck such a nerve. The general population is used to various power centers attempting to curtail their freedoms for whatever reason of the day. But regular people are largely powerless to do much about it beyond voting, which is time consuming, or protesting, which brings its own personal risks. The scientific and technological elite, however, represents something different. They are one of the power centers themselves, and here was one of their own, a high profile one at that, advocating for relinquishment. You have to give it to Joy. He was brave to express such thoughts directly in the face of powerful people.

So Joy asks, what if the unlimited pursuit of knowledge puts us all in mortal danger? He points out that we have done this before. At a 1989 nanotechnology conference, Joy said, “We cannot simply do our science and not worry about these ethical issues.” The United States unilaterally abandoned biological weapons development because the logic was clear. These weapons were easy to replicate and could easily end up in malicious hands. We would be more secure if nobody developed them, and we embodied this concept in the 1972 Biological Weapons Convention and 1993 Chemical Weapons Convention.

Joy further quotes Thoreau: “We do not ride on the railroad. It rides upon us.” Then he asks directly, “The question is, indeed, Which is to be master? Will we survive our technologies?”

Joy isn’t advocating the stopping of all research. He just prefers being thoughtful and strategic about which lines of inquiry we pursue and which ones we intentionally avoid. It means international agreements, verification systems, and scientists adopting ethical codes like the Hippocratic oath. It requires transparency and cooperation. Within that context, he sounds reasonable, right?

Personal Responsibility

Joy writes honestly about his own sense of responsibility: “I feel, too, a deepened sense of personal responsibility, not for the work I have already done, but for the work that I might yet do, at the confluence of the sciences.” That statement carries some weight because Joy was someone who helped create the technologies that might enable the dangers he himself feared.

But he finds hope in the Dalai Lama’s “Ethics for the New Millennium,” which argues that the most important thing is to conduct our lives with love and compassion for others, and that our societies need a stronger notion of universal responsibility. This awareness of greater spiritual principles at play in a world of technology is rare among the valley elite. Joy was saying that neither material progress nor the pursuit of knowledge is the key to happiness. We need to find alternative outlets for our creative forces beyond the culture of perpetual technological and economic growth.

In the TED Talk he gave six years after publishing his article, Joy made the point even clearer. The solution, he said, can’t be technology alone. We need both better public policy and deeper moral progress. He spoke of the need for “the head and the heart,” which echoes Russell and Einstein. He also argued that scientists, technologists, and businessmen must be held personally accountable under the law for the consequences of their inventions. That must have been shocking to Joy’s peers! Yet, today they face no such responsibility. That, he believed, had to change. What’s striking about hearing him articulate these perfectly reasonable points is that they aren’t taken that seriously during the time he was writing or even now. Many thoughtful people say similar things at technical conferences or in political speeches. We all clap and support the concepts. Yet, very little actually changes. Or at least the changes take so long and occur in such small steps that we’re left unsatisfied in the present moment.

Was Joy Right?

Joy’s article was unique for its time in the popular press. When it appeared on Wired’s cover in April 2000, it created quite the rumble in tech circles. Joy even said it generated thousands of letters to the editor, many of them overtly hostile. Wired had been a cheerleader for the digital age for nearly a decade. Its shift from cheering to warning marked an important and surprising moment in the digital community. Also, Bill Joy wasn’t some alarmist outsider like many others. He was one of the core architects. His warning came from inside the cathedral and it certainly resonated. At the time, I was working in marketing at Sun, and one of my jobs was to promote Sun’s technologies to the media. In press interviews during this period reporters would always bring up Joy’s article even if the meeting was booked to cover another issue. We had to draft briefing and messaging documents to prepare executives, managers, and engineers that we brought into all interviews because we knew Joy’s article would always come up in the discussion. I remember many stressful meetings and awkward interviews from this period.

The article also presaged much of what we are experiencing now but not necessarily always in the ways Joy anticipated. His specific predictions about nanotechnology have not materialized yet. And the gray goo scenario is now considered flawed and implausible. Most scientists believe built-in limitations make runaway nanotechnology improbable. But the science is never settled, so we’ll just have to see how things evolve in the future.

But Joy’s underlying concerns were clearly prescient. And we have to deal with that. He worried about knowledge-enabled destruction, powerful technologies becoming widely available, and complex systems we don’t fully understand. All of these have become more relevant as artificial intelligence has advanced today far faster than most people expected in 2000. Interestingly, Joy only explicitly mentioned artificial intelligence once in his article, possibly because he was writing at the tail end of the second “AI winter.” Yet his concerns about self-replicating technologies and systems moving beyond human control resonate well with some of the modern AI risk debates these days.

In 2025, Joy himself reviewed his Wired article in a comprehensive session at Berkeley that covered scientific advancements over the last 50 years. He said he was called a doomer for his article in 2000, but to him it was just a matter of pragmatic realism. “So if we look back 25 years, the risk was real. It’s happened roughly like I said, and we haven’t done anything about it, which is really kind of frustrating and disturbing.” He also connected his original argument from 2000 to the technologies of today. Whereas 20th century technologies require rare raw materials, the thing about “21st century technologies is that they only require information. So, they’re fundamentally different and more dangerous and difficult-to-impossible to control.” One example he cites is that AI today is “verging on being able to do recursive self-improvement,” which is the very process he warned about decades ago.

He also warned about mirror life, which he described in the Berkeley session as a possible threat to the biosphere and cited recent reports from Stanford and Nature. He framed it as just one more example of powerful technologies that could cause catastrophic harm if misused. His concern remains fixed on crazy people using advanced technologies “to do bad things as their first act, and we wouldn’t have the chance to stop them.” His conclusions weren’t encouraging, which certainly fits the tone of his original piece in Wired back in 2000. Technology today has become more powerful, the financial incentives have become more intense, and society still hasn’t done enough to address the dangers he identified years ago. It seems like his predictions were pretty close to an observation of reality today.

In the same 2025 session, Joy said that the early calls for an AI pause and stronger rules for deployment were justified. “For the AI risk I think the people weren’t wrong that we need a pause, and we need rules, and we need to be aware that we’re giving very powerful tools to everybody. I mean, we have the bill of rights and freedom of speech, and we say it’s just information and everybody should have access to everything. That means if we spend a trillion dollars collectively in our society making something which has some good uses and also can be dangerous I have to give it to everybody independent of the fact that I’m then thereby giving it to the people for whom it can be dangerous. Or should I say, well, that’s some specialized knowledge and maybe it should be more closely held.” Fair point. But where to draw the line is the issue and that goes well beyond just talking about technology.

So, Joy recognizes the benefits of AI here, just as he recognizes the value of other technologies. But he also reserves much more of his focus to describe in detail his concerns about the risks. That fits well with the views he expressed in 2000.

He said that the field of advanced technology has quickly become another arms race driven by enormous financial incentives, which repeats the pattern of technical arrogance that he warned about in relation to nuclear weapons in his original Wired article. “People are so excited about what they can do with their minds, they just can’t help themselves. And there’s such a pot of money out there that they just can’t stop.” He said society was now “skating right past” the period when the technology is still flexible enough to control. Once everyone has money and status invested in the race, he warned, changing direction would become much harder. There’s a “social physics” he describes that makes things seriously complicated and extremely difficult for people to consider the consequences. The danger, though, isn’t necessarily the technology itself, which also has many benefits. Instead, the issue is that the technology in the wrong hands can cause problems that can’t be prevented or fixed.

That 2025 session at Berkeley is the first time I’ve heard Joy address his Wired article in detail in the context of today’s debates over technological risk. Hopefully, he’ll keep contributing to the debate because unlike many others in the media at present, Joy can speak articulately about the benefits of the technology without also losing sight of the risks. He has that credibility because he’s lived the history directly. Perhaps he will because now he’s working on an AI startup with his son, daughter-in-law, Claude, ChatGPT, and Gemini as the only employees. So, Bill Joy is very much in the game. It will be fascinating to see what he builds and how that differs from the current implementations of AI.

The Sun Sets

After leaving Sun in 2003, Joy moved into venture capital and focused on green energy and climate change investments at Kleiner Perkins. He also worked as a principal investigator and chief scientist at Water Street Capital. Now at 71, he’s focusing on his AI startup with his kids. And aside from his 2025 session at Berkeley, he’s mostly remained silent in public regarding the debate he actually started 26 years ago. It’s almost as if he said what he wanted to say in 2000, briefly repeated many of the same concepts in 2025, and now we’re left to digest his thoughts and hopefully implement his lessons all these years later. I doubt we will. I saw very little press coverage or social commentary of his 2025 talk at Berkeley.

Over the years, Joy didn’t stay stuck in doom or alarm. After the Wired article he actively tried to move toward better outcomes. He invested in solutions and backed innovations in education, new options to preserve the environment, and a major $200 million biodefense fund aimed at closing the gaps that could lead to a pandemic. He came to believe that we can’t solve the management of dangerous technology with just more technology alone. Instead, we need better policies, markets that price in the true cost of potential catastrophes, and a much deeper moral awareness. That combination of thoughts remains relatively rare today. Perhaps that explains the silence in the market that continues to just focus on doom rather than solutions.

Joy put it simply in that earlier talk at TED. “We can’t pick the future, but we can steer the future.” Over the years technologies have changed, but the fundamental challenge he identified in 2000 remains relevant now. Figuring out how to pursue knowledge and innovation while also maintaining enough wisdom and caution to survive the unintended consequences seems to be a question others should carefully consider today. Yet, few are doing so. In fact, the mania is running faster than ever and being driven by many people who have never heard of Bill Joy.

In his article, Joy compared coding to Michelangelo releasing statues from marble. He described his software engineering in a similar way to those ecstatic moments when the code emerged from his imagination as if it were already waiting in the machine to be freed. He ended his essay with that same image. After eighteen pages of text exploring multiple scientific disciplines and warning about existential dangers from the exploitation of technology, he wrote, “I am up late again, it is almost 6 am. I am trying to imagine some better answers, to break the spell and free them from the stone.”

Well, twenty-six years later, we’re all still up late as well. We’re still searching for those answers. But now we go forward in the wild world of AI, which probably represents the biggest paradigm shift enabling new opportunities in tech since the Internet itself. We should embrace those opportunities, but will we also consider and mitigate the risks?

The AI Doom Vibe Change

For a couple of years now, the single story pounding our heads about AI all day every day has been exclusively about looming disaster. AI takes the jobs, then it takes everything else, then few people get rich. A love story. It was an intentional positioning of the technology, obviously, but the question remains why. Well, Cal Newport has a good hypothesis that I wrote about the other day. But others are noticing as well. So, the phenomenon of the old pitch evolving into something new is probably real.

In a recent episode of The AI Daily Brief, host Nathaniel Whittemore says the previous extremist narrative may finally be cracking, and he cites a fair number sources to substantiate his claim. It’s a different take from Cal Newport’s but there is some overlap. The signals are faint, he says, but they’re showing up in two key places at once so that may mean that the shift will likely have some legs. I think Whittemore may be pulling his punches a bit by saying “the signals are faint” just to cover himself since this shift has been so recent, such as really only the last few weeks. I think the shift is clearly underway. Remember, the IPOs are coming soon, baby! The companies and their simps can’t continue with the doom rhetoric. The American public has rejected that strategy. And it’s interesting that some public opinion polls in China lack this doom positioning. Anyway, back to Whittemore’s daily brief.

The first place Whittemore notices the vibe shift is within the never-ending chattering class in the media. He points to Ezra Klein’s recent New York Times article, “Why the AI Job Apocalypse (Probably) Won’t Happen.” I like the “probably” bit. But coming from a big voice on the political left and also one that’s outside the AI bubble, Klein may carry some weight in Whittemore’s eyes that a similar post from others, say, Marc Andreessen, simply wouldn’t. Klein cites economist Alex Imas from the University of Chicago and also a wider body of economic research to make a case rooted in Jevons Paradox. When something gets cheaper, we tend to use more of it, not less. So although computers may have changed or even eliminated specific tasks, the cost savings created enough new demand that the occupations expanded overall. As Klein puts it, “Every enthusiastic AI adopter I know is working harder than ever because there is more they can do.”

Whittemore points to more data that’s emerging. Software engineering, which is the job category most exposed to AI, is the one where postings have actually increased recently. Citadel Securities cites the increase at 18 percent since May of last year. Federal Reserve numbers also show software engineering jobs at their highest level since November 2023, although the current number is still well under the previous mark three years ago. Also, Stripe Atlas just hit 100,000 incorporations, with Q1 up 130 percent year over year. As Derek Thompson says, “AI agents are better at creating firms than destroying jobs.” A new trend?

The second place the shift is showing up for Whittemore is in markets themselves. Anthropic’s revenue, according to SemiAnalysis, has gone from 9 billion to more than 44 billion this year, which is roughly doubling every six weeks. Atlassian’s stock jumped about 30 percent recently after strong earnings with customers using its new Rovo AI tool growing their own ARR at twice the rate of those who weren’t. The skeptics have been questioning how you justify trillions in infrastructure when seats only sell for 20 dollars a month. Well, that’s being answered by the move from seats to tokens taking place recently in the intelligent agent era. A single engineer with Claude Code might burn through hundreds or thousands of dollars in tokens each month, and the companies selling those tokens cannot keep up with demand.

There’s another piece of the vibe shift worth noting, one which I found most interesting since I’ve worked in both industries. The Associated Press recently reported on construction companies teaming up with big tech to push back on community opposition to data centers. Rob Bear of the Pennsylvania Building and Construction Trades Council told the AP that communities should figure out what they actually want from these projects rather than just saying no. “If you don’t ask, you’re never going to get,” he said, pointing to things like better project plans or money for local schools and infrastructure. Whittemore’s take is sharper. He calls it “an insane indictment of how poorly tech companies have run these projects that the issue has gotten this bad” given how many ways there are to make data centers genuinely valuable to nearby communities at a fraction of the total cost. He’s spot on. The AI companies deserve the public backlash. We’ll see how they adapt to the very real world they are now entering.

Even the AI labs are softening their messages. Sam Altman recently wrote that “jobs doomerism is likely long-term wrong” and that OpenAI wants “to build tools to augment and elevate people, not entities to replace them.” Whittemore says this is a meaningful pivot from a company whose stated goal used to look a lot more like replacement.

But Whittemore is careful not to declare victory too fast. The AI transition will still be painful for specific workers and communities, and history shows we generally don’t help them much at all when economies move through technological advancements. But he ends on a hopeful note.

“I find it extremely encouraging to feel the collective foot being taken off the gas of the AI doomerism for just a moment. If nothing else, it creates an opportunity to have a different type of conversation. One that’s neither doom nor utopia, but about how to adapt to and maximize the opportunity of the change that’s here and coming. I think the more time we spend on that conversation rather than in the extremes, the better off we’ll be.”

The AI Doom Fever Finally Fades

Is the AI Doom Fever Breaking? (It’s About Time!) — Cal Newport, AI Reality Check, Deep Questions Podcast

For many years now executives leading the big AI companies have been telling the public that their own products will destroy the economy and gut the white-collar workforce. As Cal Newport observes this is the rough equivalent of a Pfizer executive announcing a new pill that cures psoriasis but also turns half the population into zombies. But lately the extremist rhetoric on AI has started to soften significantly. Instead of totally replacing entire segments of the workforce, these new AI systems will now simply augment existing workers and also lead to massive new opportunities for employment. That’s quite a radical shift in attitude, especially coming from people whose breathless messaging has been so bold. 

Nevertheless, tech companies are still laying off tens of thousands of employees and citing AI as the reason. So, we’ll see. Newport thinks the previous over-the-top positioning on AI resulted more from culture, whereas the recent shift in tone is likely more tactical. His analysis is comprehensive and seems pretty accurate given that he’s been pushing back on this rhetorical issue for years now. 

A Strange Sales Pitch

In his podcast, Newport runs through many of the recent doom statements. Mustafa Suleyman of Microsoft AI has suggested that AI will be capable of automating most knowledge work within roughly a year. Dario Amodei of Anthropic has warned that the technology will soon replace up to half of all entry-level white-collar jobs in finance, consulting, and tech. Sam Altman of OpenAI speculated last summer about a future in which AI would “do everything,” leaving humans to find new ways to “participate” in the world.

“That’s basically what we’re getting from the AI CEOs,” Newport says. “And I think it’s just lunacy.” Newport has held this position for some time now. And it’s an opinion shared online by many advanced engineers who have been working with AI systems for a long time. However, there are many so-called social media influencers in the AI space who still push the end-of-the world theme. It’s been an odd experience for sure. Granted, software executives have a long history of bragging in their corporate earnings calls that their systems will enable customers to cut expensive employees. It’s hard to think, though, of another industry whose leaders so cheerfully predict that their products will wreck civilization while making their founders and a few insiders rich beyond their wildest dreams. Seems like a difficult sell, eh?

The New Vibe

In late April 2026, Altman posted on X that OpenAI wants to “build tools to augment and elevate people, not entities to replace them,” and added that “jobs doomerism is likely long-term wrong.” A few days later, Nvidia CEO Jensen Huang pushed back even harder. In a Fortune interview posted May 2, he called the half-of-jobs prediction “ridiculous” and warned that becoming a CEO can leave a person with what he described as “a God complex,” speaking as if their position alone gave them the authority to predict civilization-scale outcomes. Huang also estimated that AI has already created more than half a million jobs, because companies that adopt it grow faster and hire more, and noted that demand for software engineers is actually rising.

Newport reads Huang’s comments as a slightly concealed dig at Amodei, but perhaps it was a subtle signal to the industry that things will be shifting. Either way, the tone has clearly migrated from outright apocalypse to smooth augmentation. Who knows. But at least it’s a welcome change for those directly affected by the recent massive layoffs attributed to AI systems that haven’t even been fully built and deployed yet. 

Where the Doom Came From

To understand the old rhetoric, though, Newport argues that you have to look at the tech culture of San Francisco and Silicon Valley, especially among engineers, and especially content articulated in a few internet forums in recent decades. The most influential was LessWrong, founded by Eliezer Yudkowsky and devoted to refining the art of human rationality. Also, the blog Slate Star Codex, written by Scott Alexander, helped push the same themes into a wider readership. Out of this loose network grew the rationalist movement, which is the idea that if you trained yourself to think like a logical engineer, you could overcome cognitive bias and act more effectively in the world. Newport, who was trained in computer science at MIT, recognizes this culture. “I’m around engineers. I am an engineer. I know this way of thinking,” he says, adding that his wife once told him, “Don’t take me to the MIT Christmas parties because you guys are all so weird.”

Two important offshoots followed. One was effective altruism, which applies something called expected-value reasoning to charitable giving and was made famous — and then infamous — by Sam Bankman-Fried when he led FTX. The other was the existential risk community (X-risk), which argued that very rare disasters with very large costs deserve serious consideration right now. Newport summarizes the X-risk crowd as focusing especially on three threats: asteroid strikes, deadly pandemics, and superintelligent AI. Nick Bostrom’s 2002 paper Existential Risks is the foundational text, but the wider X-risk literature also covers nanotechnology, nuclear war, and what Bostrom calls “totalitarian lock-in.” Newport doesn’t mention Bill Joy’s shock article “Why the Future Doesn’t Need Us” in WIRED in 2000 but it certainly fits the paradigm.

The X-risk crew organized a closed-door conference that produced the open letter signed by Stephen Hawking, Elon Musk, Bill Gates, and many of the leading AI researchers of the day. Newport places it in Puerto Rico in 2017, but the actual event was the Future of Life Institute’s “Future of AI: Opportunities and Challenges” conference in San Juan in January 2015. The 2017 follow-up was the Beneficial AI conference at Asilomar in California. Robert McMillan’s WIRED piece from January 2015, “AI Has Arrived, and That Really Worries the World’s Brightest Minds,” captured the elite anxiety that emerged from the Puerto Rico meeting and may be the article Newport has in mind when he describes the era. There were other similar pieces in the elite media during this time period as well. The elites aren’t shy with the media. 

ChatGPT and the Hero Complex

Then ChatGPT arrived. For people who had spent ten years writing footnoted lists about the coming superintelligence, it felt like the moment they had been preparing for. Newport thinks this was both terrifying and intoxicating. “What if we were right about this risk,” Newport imagines them thinking, “and not only were we right, but it’s happening?” The rationalists sensed they were going to be Neo. They were going to be John Connor. They were the ones who saw it coming and would now lead everyone else through it. It may be hard for normal people to think this way, but we are talking about the tech elite, after all. They do actually live in a different world, one that’s in many ways disconnected from the normal reality of people who have to work for a living. Newport stresses that it’s important to realize that the current AI companies we see now all grew from that culture. 

OpenAI, Newport says, originally presented itself as a nonprofit AI safety organization heavily shaped by X-risk concerns. It was started as almost a hobby project for the rationalist crowd before commercial ambitions reshaped what it is now. Anthropic was founded by former OpenAI staff who, according to Newport, felt their old employer was not being rigorous enough about safety. Grok came out of the same orbit. The CEOs, Newport says, were not playing 4D chess with investors. They were just talking the way everyone they knew talked in Silicon Valley. The trouble started when their companies got too big to keep speaking only to their own closed subculture. As Newport puts it, “we’re not, you know, in the Mission District anymore.”

Why It’s Breaking Now

Newport sees three potential forces that may be accelerating the recent change in AI positioning:

First, there is real IPO pressure building. As OpenAI and Anthropic move toward public markets, more sober East Coast investors (who wear suites, Newport says) are quietly asking the founders to stop terrifying the customers they hope will pay for AI products. 

Second, public opinion is turning. A Quinnipiac poll from March 2026 found that 55 percent of Americans now believe AI may do more harm than good in daily life, which is up from 44 percent a year earlier, with about seven in ten people expecting fewer job opportunities in the future.

And third, journalists are running out of patience. Ezra Klein’s May 2026 New York Times column “Why the A.I. Job Apocalypse (Probably) Won’t Happen” reports that the economists he interviewed are skeptical of mass joblessness. Also, a recent Ronan Farrow piece in The New Yorker even raises the question of whether Altman is actually a strong chief executive for a trillion-dollar company.

Newport says there may be additional pressures in the market pushing AI executives to temper their rhetoric in recent months, but his analysis on the three issues above seems pretty comprehensive as a working hypothesis. 

A Welcome Maturation

The Silicon Valley monoculture has finally collided with the rest of the country, Newport argues. Wall Street realism, journalistic scrutiny, and ordinary public sentiment are forcing the language to evolve to the realities of the market. Newport sounds almost relieved. Somebody, he suggests, finally had to tell these founders to “stop talking like you’re Sarah Connor from Terminator 2.” His understated parting advice still applies. Take AI seriously, but not everything you hear about it.

AI’s Perpetual Present

I’ve been reading “Why We Need Continual Learning” by Malika Aubakirova and Matt Bornstein recently. I also listened to a podcast interview from Malika on a16z . Now, I’m no AI researcher or developer. But I do like exploring the scientific foundations on which advanced software tools are built, especially since I use these applications every day and hope to leverage them more in the future. So although I don’t fully understand what’s actually happening underneath, poking around a bit is an interesting exercise. What follows below is what I’ve learned from the article. Consider it a work in progress. If you want the expert version from Malika and Matt, go read their original piece for a deep dive. This text here is just me working through things as best as I can at my level. At the end of this post, I include a list of terms and definitions. I’ll make that a standard feature in similar upcoming posts for my own short-term recall practice and also for long term memory consolidation. Memory practice (the human kind) is a hobby of mine.

Anyway, here we go. The authors open their article on continual learning by referring back to Christopher Nolan’s “Memento,” which is a film about a man named Leonard Shelby who suffers from anterograde amnesia that prevents him from forming new memories. Every few minutes his world resets and he wakes up in the same perpetual present with no idea what just happened in the past. He tattoos notes on his body and carries Polaroids as memory aids just to function throughout the day. It turns out that he’s very resourceful because he uses whatever he can in his environment to get by. He even appears pretty capable within any given scene in the movie. But, as the authors put it, his tragedy is that “he can never compound. Every experience remains external.” So, I guess that means he can’t learn based on his present moment to prepare for the future like most of us who have normal memories.

That seems to be a good general description of where AI models are right now. Back before I knew about this issue, I actually inadvertently tripped over it when I first used ChatGPT and Grok a few years ago. It was clear from my chats at the time that the models were not “learning” from our conversations at all. I kept spinning around in circles explaining myself over and over again. And, in fact, some of those earlier models didn’t know even basic facts from current events, which was shocking since AI was sold to us as being so super smart. That’s when I realized that the “learning” for LLMs took place at some point in the past and then they were locked shut while life continued on. That experience of an AI not knowing simple bits in the news rarely happens now so the user experience has improved significantly. However, there’s a lot more to it that I didn’t realize from those first few frustrating conversations.

What’s Actually Happening When You Type Into That Text Box

Here’s what I didn’t fully understand before reading the article. When you type into a chat window and stuff happens before you get an answer, that process is not the model learning anything from your input. It’s reading what you gave it and generating a response. When the conversation ends, the model does not carry that conversation forward in its memory. The next conversation starts from exactly the same place as every other new conversation. Initially, that felt unnerving so I had to figure out ways to leverage the knowledge from the LLM without all that forgetting going on.

The text box we type into is just a door into the system. What matters is the context window behind the door, which is everything the model can see at once. So, your message, the whole conversation history, any documents you shared, and any background instructions — all of these things represent what the model is working with when it responds. And it has a size limit. When it fills up, older content gets dropped to make room for new content. So if you spend an hour explaining your company’s internal processes to an AI assistant and then start a fresh conversation the next day in a new text box, the AI has no memory of the previous conversation. You have to start over. Not because it forgot. Because it never learned in the first place.

There’s a name for this phenomenon. The article calls it in-context learning, which is really just the model making smart use of whatever sits in front of it right now. It’s temporary by design. The model reads, responds, and moves on. It’s similar to glancing at your notes before a meeting rather than actually deeply studying, internalizing, and using the material beforehand. When the meeting ends, those casual notes go back in the drawer and are forgotten.

The Frozen Model Problem

To understand why this matters, you need to know a little about what’s inside these models. During training, a model reads an insane amount of text and gradually adjusts billions of numerical values called parameters or weights. You can think of each weight as a dial on a pipe connecting two nodes in the network controlling how much signal flows through. The model trains by turning billions of those dials very slightly over and over again until it gets good at predicting language. That right there is really impressive to me given the scale of information these models are working with. But when the training process ends, all those dials get locked. That stage represents deployment. The model then goes out into the world with its knowledge frozen in place.

Training works because it’s a compression process. The model can’t store everything it reads verbatim. It has to find the underlying patterns, generalize the data, and build something compact that transfers to new situations it’s never seen before. The authors describe this as lossy compression, and that lossiness is actually what produces what seems like intelligence to us when we talk to an AI. When I first read that I thought of a camera compressing a RAW file to a JPEG file. The RAW image contains all the available data but it’s a massive size and requires editing in post production to produce a beautiful image. The JPEG, however, is much smaller because it’s been compressed by the camera to just what’s needed to display a good quality image at a certain size. I’ve always understood that process in photography, but I didn’t realize that LLMs are going through a similar process.

Here’s another way to think about it. Remember when you first learned how to ride a bike? You didn’t read the entire manual every time. You just got some guidance from a friend or a parent and you practiced. You fell down a few times and adjusted your technique, and then eventually your brain distilled your experience into something automatic and compact. That’s compression. You still remember falling down, but that falling down process is no longer helpful for riding once learning has taken place. What remains is the final skill of balancing to ride. An AI model that memorizes every training sentence perfectly would be less useful, not more, because it could retrieve but never generalize. It would behave more like a simple retrieval system than a sophisticated learner.

The painful irony the authors identify is this. The very mechanism that makes these models powerful during training is exactly what we stop them from doing once they’ve been deployed. We freeze the compression at the moment of release and replace it with what’s called external memory. That clarified the argument for me. The system is layered, and each layer is essentially a workaround for the fact that the compression stopped. Understanding that made the next part of the article click.

The Filing Cabinet

To compensate for frozen models, developers have built elaborate scaffolding systems, such as chat histories, retrieval databases, system prompts, external document stores, and more. All of these things make up what the article calls external memory. They are flexible and they live outside the model’s internal, frozen weights. When you need information, the system retrieves it and feeds it into the context window. Then the model reads it and responds.

This architecture works as is and the authors are honest about that. However, they make a point I hadn’t considered before. “A bigger filing cabinet is still a filing cabinet.” Retrieval is not learning. The model is looking things up, not actually knowing them. It just does it very quickly and uses natural language so you get the impression you are talking to someone who is intelligent.

Here’s another practical example. Say a hospital deploys an AI assistant to help with real world clinical decisions. That model was trained on medical literature through some cutoff date. A major new clinical trial or medical policy comes out afterward that changes how doctors treat a particular condition. The hospital can feed that paper into a retrieval database so the AI can surface it when it’s relevant. But the model doesn’t internalize that new research the way doctors would after reading it, applying it to patients, observing the outcomes, and revising their practice accordingly. The AI can retrieve the abstract. But it can’t reason from the new finding the way someone who has truly learned it can in practice. That’s the limitation these researchers are trying to fix.

The same problem exists in cybersecurity with treats evolving daily. A frozen model can be given descriptions of new attack patterns through retrieval, but it can’t compress and generalize from those patterns the way an analyst does who has spent months chasing a specific class of threat. The knowledge stays external. It never becomes part of what the model actually knows unless the model is updated with a new learning process, which is time consuming and very expensive.

What Real Learning Requires

So what’s the alternative? The article introduces a concept called continual learning, which is the field of research aimed at letting models actually update their weights based on new experience after deployment. Not just read notes. Actually learn live like humans do.

And here’s where the Memento metaphor really makes sense. The authors say that today’s AI is stuck in Leonard Shelby’s perpetual present. The scaffolding, the Polaroids and tattoos, and other memory aids work well enough within any given scene. But the model can never compound in real time. Every new thing it encounters stays external.

Think about the difference between a doctor who simply retrieves a recent study and a doctor who has spent years treating patients with that knowledge fully and personally internalized. Or consider the difference between someone who has your email history in front of them and someone who actually knows how you think over time. The article frames this cleanly. “The difference between ‘Here is what you responded to this email before’ versus ‘I understand how you think well enough to anticipate what you need’ is the difference between retrieval and learning.” Even in normal human memory, immediate retrieval is necessary to manage your present experience. However, it’s also required that your present experience be embedded into long term memory for continual learning.

The authors bring up Fermat’s Last Theorem as one powerful example of the kind of hard discovery problem they have in mind. Mathematicians worked on the issue for 350 years. Eventually the problem was solved by Andrew Wiles. But he didn’t crack it by retrieving the right papers. He solved it by working in near total isolation for seven years, and inventing entirely new mathematical techniques to bridge two previously disconnected fields. That kind of discovery required genuine compression, generalization, and creative combination. Not simply fast retrieval. And the article asks directly whether a model that can’t compound from experience could ever do anything like that. The honest answer is they don’t know yet.

Why Updating Weights Is So Hard

At this point I had to ask myself if real time continual learning is so important, why can’t the LLM models do it now? The short answer is that updating a model’s weights after deployment is genuinely dangerous and technically unsolved at scale.

The most obvious problem is called catastrophic forgetting. When you update a model’s weights to learn something new, it tends to overwrite what it already knew. New learning crowds out old learning. If you fine tune a general model specifically on medical records, it might get better at clinical language while getting noticeably worse at everything else because the new training has nudged weights that were also doing other jobs. The model gets better at one thing and potentially worse at everything it was already good at. When you understand this you can really appreciate how humans have benefited from millions of years of evolution. The AI machines seem rather clunky by comparison. When humans learn, new neural connections are made in the brain that stick for a long time as new learning is layered on top. But even in humans, old learning and memory does actually fade gradually over time if a specific neural pathway isn’t continually or at least occasionally reinforced. It just takes a very long period of time. With AI systems, however, new learning can wipe out old new learning immediately. The authors didn’t address this issue directly in humans, but the example seems similar if you study biology.

There’s also the problem of data poisoning. If a model’s weights can be updated through interactions after deployment, bad actors could gradually manipulate its behavior through carefully crafted inputs over time. Unlike a one-time attack, poisoned weights persist across every future conversation. The damage would live in the model itself so safety alignment would degrade unpredictably immediately or some time in the future. The article notes that “even narrow fine-tuning on benign data can produce broadly misaligned behavior,” which is a sobering thought to sit with. Yet we all know this would happen right away based on our own experience being online every day fighting bots and hackers.

These aren’t hypothetical concerns. They’re real problems without clean solutions yet.

Where Things Are Heading

The article maps out a spectrum of approaches to continual learning that are organized around a question I found clarifying: where does the compaction actually happen? It seems there is a stack of technologies managing the process.

On one end you have pure retrieval. No compaction. The model just reads notes. That’s most of what exists today. In the middle there are modules, which are attachable and specialized components that let a model develop some expertise in a specific domain without retraining the entire thing from scratch. A hospital might attach a medical module to a general model so it performs at a specialist level on clinical questions, while the same base model with a different module handles legal contracts. Each module is swappable independently. That’s a practical and reasonable middle ground for now.

On the far end you have full parametric learning, where the model’s weights actually update from new experience after deployment. This is the goal, but it remains largely unsolved at scale with the current technologies. But there are serious research efforts moving in this direction with things like test-time training where the model runs brief learning cycles before it generates a response. Also there are self-improvement approaches where models like AlphaEvolve have generated their own training data and genuinely improved from it, at least within constrained problem domains like mathematics.

The authors frame the path forward as layered. In-context learning stays as the first line of adaptation because it works now and keeps getting better. Modules offer some personalization and domain specialization. But for genuinely novel problems, adversarial scenarios, and knowledge too tacit to put into words, models may eventually need to compress new experience directly into their parameters after training. Otherwise, as the authors put it, we stay stuck in Memento’s perpetual present.

What I Took Away

I started reading this article as someone who uses AI tools every day without really thinking much about what’s happening underneath. What I came away with is a better sense of the gap between what these systems appear to do, what they’re actually doing, and what they’ll potentially do in the future. Right now they can respond to new information and adapt to what you give them. And most times they feel like they understand you. But the reality is that they don’t compound. They don’t learn. They don’t internalize new experience the way continual learning systems or humans would. Their dials are locked. And until engineers figure out how to update those dials safely and continuously after deployment, the models we’re using now are doing something more like reading notes than actually learning from the experience. That’s a distinction with a very big difference.

Check out the original article and Malika’s podcast for the technical details. Below is a list of related terms and definitions.


Continual Learning: Vocabulary List

This list of terms below is based on the a16z article “Why We Need Continual Learning” by Malika Aubakirova and Matt Bornstein and also the podcast with Malika discussing the article. Some definitions closely reflect the article itself, but others expand into broader concepts from the field for additional context. I error checked the terms and definitions with Grok, ChatGPT, Gemini, Perplexity, and DeepSeek.

Agentic Loops

A mode of operation where the model works autonomously step by step toward a goal without you typing each instruction. Each step produces output that feeds into the next. This process can go on for many cycles. The article identifies two related problems as steps accumulate: (1) the immediate symptom is coherence degradation, where the agent loses the thread and starts making poor decisions, and (2) the underlying cause is that maintaining a growing context becomes increasingly expensive and inefficient. Both concerns together represent why the article frames agentic loops as one of the pressure points on the current in-context learning paradigm. For example, an agent tasked with researching a topic, drafting a report, checking sources, and revising the draft might handle the first twenty steps cleanly. But by step eighty the accumulating context has grown so large and costly that the agent starts losing track of earlier decisions and repeating work it already did.

Attention Heads

A key mechanism inside transformers that allows the model to weigh how relevant each part of the context is to every other part when generating a response. Multiple attention heads run in parallel, each learning to focus on different kinds of relationships in the text. One head might learn to track grammatical agreement between subject and verb across a long sentence, while another tracks thematic connections between paragraphs. Together they allow transformers to handle complex, long range dependencies in language that earlier architectures struggled with. For example, in the sentence “The lawyer who argued the case, despite the objections raised by her colleagues, ultimately won,” an attention head helps the model correctly connect “won” back to “lawyer” across all the intervening words.

Catastrophic Forgetting

When a model updates its weights to learn something new, it tends to overwrite what it already knew. In other words, new learning crowds out old learning and sometimes dramatically. This is one of the central unsolved problems in continual learning, and one of the main reasons models are not updated continuously after deployment. Think of it somewhat like overwriting parts of a hard drive. The new files go in, but the old ones can be partially or fully lost. For example, if you fine-tune a general purpose model specifically on a medical records archive, the model will get better at clinical language but noticeably worse at writing poetry or explaining history because the new training has nudged weights that were doing other jobs.

Compression / Compaction

The process of taking a vast amount of raw information and distilling it into something compact and generalized. During training, a model compresses an enormous amount of human writing into its parameters and finds the underlying patterns rather than storing things verbatim. The article uses “compaction” as a broad organizing term for how deeply new information gets digested, which ranges from not at all (pure retrieval, where facts just sit in a database) to fully (weight-level learning, where the model actually internalizes new knowledge). For example, rather than memorizing every recipe ever written, a well-trained model compresses the underlying logic of cooking: how heat transforms food, how flavors balance, how techniques generalize across cuisines.

Continual Learning

The broader field of research aimed at letting models learn from new experience after deployment, ideally by updating their weights rather than relying on external scaffolding. It’s the opposite of the current norm, where training and deployment are completely separate and weights are frozen the moment a model is released. The goal is something closer to how humans learn continuously from experience without needing to be retrained from scratch every time the world changes. For example, a customer service model using continual learning could gradually internalize patterns from thousands of resolved support tickets over time and get genuinely better at its job rather than just retrieving past examples.

Context Window

The full body of text the model can see at once when generating a response. It includes your message, the full conversation history, any documents you shared, and any background instructions passed to the model. It has a size limit measured in tokens. When it fills up, older content must be dropped to make space for new content. For example, if you have a long conversation with an AI assistant and then ask it to recall something you mentioned earlier, it may not be able to answer because that part of the conversation has already been pushed out of the window.

Data Poisoning

One of several serious governance and security risks the article raises around continuous weight updates. If a model’s weights can be updated after deployment interactions, bad actors could gradually manipulate its behavior through carefully crafted inputs over time, which is a slow and hard-to-detect form of corruption that lives in the weights rather than just in the context. Unlike a one-time prompt injection attack, poisoned weights persist across every future conversation. The article groups this alongside other unsolved challenges: alignment degradation, the impossibility of unlearning toxic knowledge, auditability failures, and privacy risks from user interactions being compressed into parameters. For example, an adversary could repeatedly feed a customer-facing AI subtly misleading information about a competitor’s product until the model begins reproducing those inaccuracies on its own with no obvious sign of tampering.

Distillation

A process involving two models: (1) a large, capable, frozen teacher and (2) a smaller student. The student is trained to match the teacher’s outputs as closely as possible and absorb its knowledge in a more compact form. The result is a smaller, more efficient model that performs nearly as well as the larger model on the tasks it was trained for. It’s like an apprentice learning by closely watching and mimicking a master until the skill becomes their own. For example, a large hospital system might use a massive general-purpose model as the teacher and distill its medical reasoning capabilities into a smaller model that can run efficiently on local hospital hardware without requiring a cloud connection.

External Memory

Anything outside the model’s weights used to store and retrieve information. Chat history, databases, document stores, and agent notes are all examples of external memory. Information gets fed back into the context window when necessary. In current deployment architectures, the model typically does not update its weights from that information during inference. The key limitation is that external memory requires retrieval. The model has to be given the right information at the right moment, and if it isn’t, the knowledge might as well not exist. For example, a legal AI might have a database of ten thousand case summaries it can search, but if the retrieval system surfaces the wrong cases, the model has no way to compensate from its own knowledge.

Few-Shot Learning

The ability of a model to perform well on a new task after seeing only a handful of examples, rather than requiring thousands of training samples. Transformers are surprisingly good at this when examples are provided in the context window. Meta-learning approaches aim to make weight-level, few-shot learning just as effective, so the model can internalize new tasks from just a few examples even without them being available in the context. For example, if you show a model three examples of how you want your emails formatted and then ask it to format a fourth, it adapts immediately without any retraining. That’s few-shot learning in action.

Fine-Tuning

A more targeted form of additional training done after the initial training run. Instead of training from scratch on everything that’s known, you take an already-trained model and update it on a smaller or specific dataset. The new information shapes the model’s behavior for a particular use case without rebuilding it from the ground up, but the process still risks catastrophic forgetting if pushed too hard. For example, a company might take a general-purpose language model and fine-tune it on thousands of their internal support conversations, so the model learns the company’s terminology, tone, and common issue patterns without losing its broader language capabilities.

Gradient Descent

The mathematical process by which a model adjusts its weights during training. It measures how wrong the model’s predictions are on a given example and then calculates which direction to nudge each weight to reduce that error slightly. It’s called “descent” because the process is navigating downhill on a mathematical landscape, always moving toward lower error rates. Repeat this across billions of examples and the model gradually gets much better. For example, if the model predicts “cat” when the correct answer is “dog,” gradient descent works backward through the network to figure out which weights contributed to that wrong answer and adjusts them a tiny amount. Do that enough times and the model learns to tell cats from dogs reliably.

In-Context Learning (ICL)

Everything the model reads and uses during a single conversation without updating its underlying knowledge. You paste in a document, it reads it and responds. You describe a task, it follows your instructions. But when the conversation ends, none of that experience changes the model itself. The next conversation starts with the same frozen weights as always. This is a smart use of temporary information, but it’s not genuine learning. For example, if you spend an hour teaching an AI assistant about your company’s internal processes and then start a new conversation the next day, the model will have no memory of the previous conversation. You would need to paste in that information all over again.

Inference

The act of a model generating a response from input. It’s the opposite of training. Training occurs when the model learns by adjusting its weights. Inference occurs when the frozen model performs and takes what it knows and produces an output. Any time you send a message and get a reply, that’s inference. The term “inference-time compute” (below) builds on this and refers specifically to spending extra computational effort during inference to get a better result. But plain inference just means the model is running, not learning. For example, asking a model what the capital of France is and getting back “Paris” in a fraction of a second is inference in its simplest form. No learning took place. The model generated an output from its existing weights without updating them.

Inference-Time Compute

The current dominant paradigm for improving model performance by spending more computational effort at the moment of response rather than updating weights. This includes chain-of-thought reasoning, tool use, search, and iterative problem-solving, all of which cost more compute at response time but produce better results. The article positions this process as a workaround, a scaling of what already works rather than a true solution to the learning problem. Test-time training is the most aggressive form of this learning because it actually runs gradient updates on new information during inference, which begins to compress it into weights in real time. This process sits at the boundary between the current paradigm and genuine parametric learning. For example, when you ask a model a complex math problem and it works through each step before giving a final answer rather than just guessing immediately, that is inference-time compute. The model is using more processing in the moment to arrive at a better result.

Instruction Tuning

A form of fine-tuning where the model is trained specifically on examples of instructions paired with ideal responses. It’s one of the main reasons modern models are so much better at following directions than earlier versions, which tended to just complete text rather than actually do what you asked. The model learns not just facts but the shape of helpful behavior, including how to interpret requests, how to structure answers, and when to ask for clarification. For example, an early language model asked to “summarize this article” might just continue writing in the same style as the article. An instruction-tuned model understands that the request calls for a concise, distinct summary and produces one.

KV Cache

Short for key-value cache. A technical mechanism that stores intermediate computations during inference so the model does not have to redo them from scratch for every token it generates. The article discusses it specifically in the context of KV cache compaction where the cache functions as a form of non-parametric memory but grows substantially as conversations and agent loops get longer. The authors argue that learning to compress this cache more efficiently is one of the meaningful challenges in moving from pure retrieval toward more durable knowledge storage. For example, in a long agentic task, the KV cache holds the computed representations of everything the model has processed so far. Without it, each new token would require reprocessing the entire history from scratch, which would be prohibitively slow.

Lossy Compression

Compression where some information is permanently lost in the process, as opposed to lossless compression where everything can be recovered exactly. For LLMs, the inability to store everything verbatim during training forces the model to find patterns, generalize, and abstract. That forced abstraction is precisely what makes the model seem intelligent and useful in new situations it has never seen before. A JPEG image is the familiar everyday example. Save a photo as a JPEG and the file shrinks dramatically because fine detail is discarded. But if you zoom in close enough you can see the degradation. For most purposes, though, the image is perfectly usable. The tradeoff is the point. For a language model, the equivalent is that it cannot recite every sentence it ever trained on, but it can write a new sentence in any style on any topic because it extracted the underlying structure rather than memorizing the surface.

Meta-Learning

Teaching a model how to learn rather than what to learn. The model is pre-trained in a way that positions it to update quickly and effectively with just a few new examples, rather than requiring extensive retraining. It’s the difference between educating someone to be a quick study versus simply giving them a lot of facts to memorize. A quick study can walk into an unfamiliar subject and get up to speed fast, whereas someone who only memorized facts cannot. For example, a meta-learned model shown three examples of a new classification task, say sorting customer complaints into categories it has never seen before, should be able to generalize accurately to new complaints after just those three examples rather than needing hundreds.

Modules

The article uses this as a broad middle-ground category on the compaction spectrum that sits between pure retrieval and full weight-level learning. In practice, modules can take several forms: adapter layers, LoRA-style weight updates, memory components, or cached representations. What they share is the ability to specialize a general-purpose model for a specific domain without retraining the entire model from scratch. They offer more than retrieval in that some digestion of information happens, but less than full parametric learning in that the core model is typically left unchanged. For example, a hospital might attach a medical module to a general-purpose model so it performs at a specialist level on clinical questions, while the same base model with a legal module performs at a specialist level on contract review, with each module being swappable independently.

Multi-Agent Architectures

Systems where multiple AI models work in parallel with each one handling a slice of a larger task and communicating results to each other or to an orchestrating layer. If a single model is limited by its context window, a coordinated group of agents can collectively handle far more. But this shifts the problem rather than eliminating it. Each agent still faces its own context limit, and coordinating many smaller contexts introduces its own complexity for the system to manage. It’s a non-parametric workaround for scale, not a solution to the underlying constraint. For example, a research task that would overflow one model’s context window might be split across ten agents, each reading a different section of source material with a coordinating agent assembling their summaries into a final report.

Neural Network

The underlying computational structure of an LLM. It’s a network of interconnected nodes organized in layers, loosely inspired by neurons in the brain. But the analogy should not be pushed too far. Each connection between nodes has a weight that determines how strongly one node influences another. During inference, information flows forward through the layers, gets transformed at each step, and eventually produces an output. The network learns by adjusting those weights during training until it gets good at its task. For example, in an image recognition network, early layers might learn to detect simple edges and colors, middle layers might learn to recognize shapes, and later layers might learn to identify objects. Language models work on the same principle but applied to sequences of text.

Parameters / Weights

The billions of numerical values inside a model that encode everything it learned during training. Each value represents the strength of a connection between two nodes in the neural network. During training, these values get adjusted gradually until the model becomes good at predicting language. After training they are frozen, and the model’s knowledge and capabilities are entirely determined by those fixed numbers. “Parameters” and “weights” refer to the same thing and are used interchangeably throughout the article. For example, frontier models contain billions or even trillions of parameters. Each one is a small dial that was tuned during training and now stays locked in place, collectively encoding an enormous amount of compressed knowledge about language, facts, and reasoning patterns.

Parametric Learning

Learning that actually updates the model’s weights based on new experience, as opposed to in-context learning which uses information temporarily without changing anything permanent. It’s the deeper form of learning the article is ultimately arguing we need more of. When a model learns parametrically, new knowledge gets compressed into its weights the same way training data did and becomes a durable part of what it knows rather than a note it holds briefly and then discards. For example, a parametric update after a model encounters thousands of conversations about a new programming language would leave it genuinely better at that language going forward across all future conversations, not just within the session where it learned.

Regularization

A cautious approach to weight updates that penalizes changes to parameters deemed important to existing knowledge. Before updating a weight, the system estimates how critical that weight is to the model’s current capabilities. If it’s very important, the update is constrained or slowed down. This is one of the older approaches to continual learning and helps manage the stability-plasticity dilemma. But it tends to be brittle at scale. Think of it like a renovation rule that protects load-bearing walls. You can still remodel, but certain structures are off-limits because removing them would collapse the building. For example, EWC (Elastic Weight Consolidation), one of the most cited regularization methods, computes an importance score for each weight after training on a task and uses that score to resist changes when training on subsequent tasks.

Reinforcement Learning (RL)

A training approach where a model learns from feedback signals rather than from labeled examples. It tries things, receives a reward or penalty based on how well it did, and adjusts its behavior accordingly over many iterations. The article mentions RL-based feedback loops as one direction in continual learning research where models could improve from real-world deployment signals like user corrections or task outcomes. However, it’s not the central mechanism the authors emphasize. The core focus of the article is on compaction, weight updates, and memory structures. For example, the systems that learned to play chess and Go at superhuman levels used reinforcement learning by playing millions of games against themselves and adjusting strategies based on wins and losses rather than being taught explicit strategies.

Retrieval-Augmented Generation (RAG)

A common approach to giving models access to current or specialized information without retraining. Instead of baking knowledge into weights, you build a searchable database the model can query at response time. The retrieved content gets injected into the context window and the model uses it to generate its answer. It’s purely non-parametric. The model retrieves information but never internalizes it. The limitation is that retrieval only works if the right information gets surfaced at the right time, and no amount of retrieval can substitute for knowledge the model needs to reason with flexibly. For example, a financial AI might use RAG to pull in the latest earnings reports before answering questions about a company’s performance because that information changes constantly and cannot be baked into training data.

Safety Alignment

The work done during training to make a model helpful, honest, and safe to use. It involves carefully curated training data, human feedback on model outputs, and specific training objectives designed to shape the model’s values and behavior. One of the serious risks of continuous weight updates after deployment is that alignment can degrade unpredictably even from adding seemingly benign new data. It seems that fine-tuning on almost anything can shift the weights that govern behavior, not just the ones governing the specific knowledge update. For example, researchers have shown that even brief fine-tuning on ordinary instructional text can weaken safety guardrails in ways that are not obvious until the model is probed specifically for harmful outputs.

Self-Improvement

An approach where the model generates its own training data, filters out low-quality results, trains on the high-quality results, and repeats the cycle. It learns from its own work rather than from human-provided data and can improve capability over repeated iterations in constrained settings. The article cites AlphaEvolve and AlphaProof as examples of this kind of closed-loop improvement. But these systems operate in constrained domains like mathematics and algorithm optimization, not open-ended real-world learning. The article uses these examples to illustrate iterative self-training loops, and what qualifies as a genuinely new discovery in this context remains debated. For example, AlphaEvolve used self-generated solutions and automated evaluation to discover improvements to algorithms that human programmers could not find because it worked within a well-defined problem space where correctness could be verified automatically.

Stability-Plasticity Dilemma

The fundamental tension in any learning system between staying stable, meaning not forgetting what it already knows, and staying plastic, meaning remaining able to learn new things. Push too hard toward plasticity and you get catastrophic forgetting. Push too hard toward stability and the model cannot adapt to anything new. Solving this dilemma is one of the core engineering challenges in continual learning, and no approach has fully solved the problem at scale. The dilemma exists in biological brains too. From what I understand about biology, human memory consolidation is strongly associated with sleep and offline processing, which suggests the brain has its own version of this stability-plasticity problem built right in. For example, a model trained to be highly stable might refuse to update its belief that a particular drug is safe even after being shown new clinical evidence, while a model trained to be highly plastic might update so aggressively that it forgets basic grammar rules after a week of medical fine-tuning.

State Space Models (SSMs)

An alternative to traditional transformer architecture that the article highlights for offering a fundamentally better scaling profile for long contexts. The article describes them as using fixed memory layers interspersed with normal attention, which unlike transformers does not grow unboundedly with every token added to the context. Traditional transformers scale quadratically with context length, while SSMs aim for near-linear scaling. However, this remains an active area of research rather than a fully settled property. The article treats SSMs as a promising architectural direction for enabling much longer agentic loops rather than a definitive solution to the broader continual learning problem. For example, a transformer handling a 100,000-token conversation requires vastly more compute than handling a 10,000-token request. But an SSM handling the same expansion would ideally require only proportionally more, which could make very long agentic tasks far more practical.

Temporal Disentanglement

A core limitation of parametric memory since a model’s weights do not separate timeless facts from information that changes over time. Both get compressed into the same parameters and are tangled together with no internal label distinguishing what’s permanent from what’s mutable. This makes continual weight updates risky because changing a time-sensitive piece of knowledge can corrupt stable knowledge stored in nearby weights. The article frames this as one of the fundamental unsolved problems standing between today’s frozen models and genuinely adaptive ones. For example, the fact that two plus two equals four and the fact that a particular person holds a particular job title are both encoded somewhere in the weights. Updating the job title risks disturbing the arithmetic, because the model has no mechanism for knowing which facts are stable laws and which are contingent facts about the world.

Test-Time Training

An approach that blurs the line between training and responding by letting the model do a small amount of learning before it generates a final answer. Rather than relying entirely on what it learned during the original training run, the model runs brief gradient updates based on what it’s currently seeing and then responds. The article describes this as running gradient descent on test-time data, compressing new information into parameters at the moment it matters, and treats it as one of the more substantive moves toward genuine continual learning because it actually changes weights at inference time. For example, if a model is asked to analyze a long, unusual technical document, test-time training would let it briefly train on that document before responding, compressing its key patterns into weights rather than just reading it as context. This method potentially produces a much more accurate analysis as a result.

The Bitter Lesson

A well-known observation in AI research. It holds that given more compute and data, general methods that let models figure things out at scale consistently outperform clever human-engineered solutions over time. Every time researchers have tried to hardcode structure and shortcuts into AI systems, the simpler but more scalable approaches have eventually won. The article invokes this phenomenon to question why we still hand-engineer memory and compression pipelines rather than letting models learn to do it themselves. For example, early chess programs used elaborate human-crafted rules about piece values and board positions. They were eventually crushed by systems that simply learned from millions of games with minimal human guidance and relied on scale rather than cleverness. The same pattern has repeated across nearly every domain in AI.

Token

The basic unit of text that a large language model processes. A token is roughly a word, though it can also be a fragment of a word, a punctuation mark, or a short common sequence like “ing” or “un.” Models do not read text the way humans do, character by character or word by word. Instead, they break input into tokens first and then process the sequence. The size of a context window is measured in tokens, not words or characters. For example, the sentence “The cat sat on the mat” would be broken into something like seven tokens, roughly one per word. But a word like “unbelievable” might be broken into two or three tokens: “un,” “believ,” “able,” because it’s less common and gets split into recognizable subunits the model has seen frequently.

Training Run

The large-scale and expensive process of building a model’s knowledge by exposing it to massive amounts of data and adjusting its weights. Training involves feeding these huge datasets through the network repeatedly and using gradient descent to nudge weights toward better predictions. The process runs on clusters of specialized hardware for weeks at a time and consumes substantial amounts of electricity. It’s all carefully controlled, occurs before deployment, and produces a fixed set of weights that define everything the model knows. Once training ends, the weights are frozen and the model goes out into the world as-is. For example, training a frontier model like GPT-4 or Claude is estimated to cost tens or hundreds of millions of dollars and requires specialized data centers. This is precisely why continuous post-deployment learning is so appealing because rerunning a full training run every time the world changes isn’t practical.

Transformer

The dominant architecture underlying most major AI models today including Claude, GPT, and Gemini, and more. At its core, a transformer predicts the next token in a sequence of text based on everything that came before it. It generates outputs token by token at very high speed. That sounds simple but at scale it’s not. The architecture was trained on so much human-generated text that it models statistical relationships in language and attempts to produce behavior consistent with understanding context, logic, and meaning. For example, when you ask a transformer-based model to explain a complex idea, it makes predictions about what a good explanation would look like given your question based on patterns it absorbed from vast amounts of human writing on similar topics. That’s why it seems smart. It’s familiar. Whether the final output constitutes genuine understanding is a separate philosophical debate that the article doesn’t address.

The Real AI Boom from Marc Andreessen

Marc Andreessen has been right about a lot of things. He co-invented Mosaic, which inspired Netscape and helped popularize the web that was invented by Tim Berners-Lee at CERN. He identified the concept of “software eating the world” well before most people knew what that even meant. In 2011, he predicted that within a decade five billion people would own smartphones. The number turned out to be six billion. Close enough. So when he sits down and tells us that 2025 was the most interesting year of his life, and that he expects 2026 to exceed it, maybe we should give him a quick listen.

Now, although Andreessen has been right about many tech trends, he’s also a supreme cheerleader and entertainer so you have to listen with a skeptical ear. I don’t think he’s malicious at all, but we have to question and observe carefully by doing our own research. And we have to watch how capital flows through markets and vendors because at present AI represents probably the biggest spend in recent years. Anyway, I enjoy listening to Andreessen because he forces me to think about what’s possible within a world that has gone mad. He’s always building. Contrast that optimistic view to the doomers who only wreck things and offer nothing of actionable value in return. That’s the tell. Marc hedges himself at many points, as well, so that’s good enough for me. It’s a reasonable starting point, anyway.

So, let’s take a look at his latest take on AI: Lenny’s Podcast: Marc Andreessen | The real AI boom hasn’t even started yet

In this recent conversation on Lenny Rachitsky’s podcast, Andreessen laid out a view of the AI moment that cuts across almost everything we hear in the megaphone media. The panic about job loss, he says, is based on a fundamental misreading of the world we actually live in. The fear that AI will make young workers obsolete gets it almost exactly backwards. And the comparison people keep making to past technological disruptions uses the wrong baseline entirely. His argument isn’t simple. But it is coherent. And once you hear it and understand his framework, the conversation we’ve been hearing about AI starts to look a little different.

What he Got Wrong

Before going further, it’s worth saying that Andreessen is the first to flag his own record. “I’ve been wrong about tons of things,” he says, joking about having buried his failures somewhere behind the shed. One example includes his famous debate with Peter Thiel about whether technological progress had stalled. Andreessen originally argued the optimistic side, of course, and insisted that progress was still happening. He now gives Thiel’s argument much more credit than he once did, at least in part. Thiel’s core claim was that we had plenty of progress in bits, meaning software, the internet, and digital technology, but very little progress in atoms, which is represented by the physical or built world. And the evidence supports him. Look around. We see bridges built in the 1930s. Dams built in the 1910s. Cities founded in the 1880s. “What have we done recently?” Andreessen asks. “Where are the new cities? Where are new dams? Where is the California high speed rail?”

He doesn’t mention that Japan and China and some nations in the European Union have had high speed rail for decades, but he surely knows that because he travels constantly. I found that missed opportunity odd. Perhaps he’s just commenting on how spotty the physical world has evolved. The infrastructure build out in China in recent decades dwarfs anything in the West, yet he fails to mention that as well. The built world, he says, is simply not that different from fifty years ago. That seems true. If you compare 1870 to 1930 or 1930 to 1970 the physical transformation was dramatic during both of those periods. Then compare 1970 to today, and it’s far less impressive than it feels. After hearing that I have to think he’s focused much more on the United States than the rest of the world.

The reason for this lack of progress, he argues, is structural. Red tape. Rules. Restrictions. Politics. Regulations. Unions, cartels, and monopolies that have every incentive to prevent rapid change. Healthcare is his favorite example. “By rights AI is going to have a dramatic impact on the healthcare system in very positive ways,” he says, “but large parts of the medical system are cartels.” Doctors are a cartel. Nurses are a cartel. Hospitals are a cartel. And then on top of that, all of these systems are increasingly acting like a government monopoly. ChatGPT is almost certainly a better doctor than your doctor today, he says, but it can’t get a license to practice medicine. It can’t prescribe medications. It can’t perform procedures. The technology is ready. The institutions are not. That’s an interesting take. But I’d much prefer my doctor leveraging AI as a tool instead of letting AI take the lead. There’s no way I’d trust ChatGPT running a spinal surgery on my back. No way. It’s not smart enough.

But this is also why he doesn’t expect AI to transform everything overnight. “There are real structural impediments in the economy and in the political system that prevent rates of change anywhere near the rates people had in the past.” Maybe AI causes us to revisit those assumptions for the first time in decades. That would be the optimistic view. It surely is time to build, as he once famously said. It’s true that the deep state is the impediment in modern times, but it’s also true that much of the previous innovation he’s talking about took place with no guardrails whatsoever. Keep in mind that much of the building that took place previously during the industrial revolution was driven by the very concept of “cartels” that he criticizes as the gatekeepers of today. If he were challenged on this, he’d admit it, I’m sure. I don’t think he’s trying to hide here. I just think he’s passionate about progress and anything that holds that back needs to be overcome.

The World is Not What You Think

To understand why Andreessen is still broadly optimistic despite all of the above, you first have to accept something uncomfortable. His view is that despite everything we’ve felt over the past fifty years, technological progress in the actual economy has been extraordinarily slow.

“It’s felt like we’ve been in a time of great technological change,” he says, “but actually if you look for evidence of that, like statistical evidence, analytical evidence, you basically can’t find it.”

Economists measure technological progress through productivity growth, which is essentially a mathematical expression of how much technology is actually moving the needle in the economy. And by that measure, the United States has been running at roughly half the pace it sustained between 1940 and 1970, and about a third of the pace it ran between 1870 and 1940. What about the smartphones and social media and cloud computing that felt so revolutionary? In terms of measurable economic impact, they barely registered compared to the era of electrification, railroads, and mass manufacturing.

So when people worry that AI is going to blow up the economy the way past technologies did, they’re actually comparing it to a fifty-year stretch of relative stagnation. The real baseline, the comparison that should inform how we think about what’s coming, is the period from 1870 to 1930. And people who lived through that era didn’t experience it as disruption. They experienced it as abundance.

“If you go back and you read accounts of 1870 to 1930,” Andreessen says, “people just thought the world was awash with opportunity.”

That may be true for some people, or even many people, but if you are familiar with that period in the United States, it’s hard to miss how disruptive and painful those times were for many people. Do you really think the air in New York City was as clean back then as it is now? Do you really think workers had a better life working the mills in Boston in the late 1800s than the pampered kids working in air-conditioned offices in Los Angeles or Silicon Valley now? I get Andreessen’s point, and I largely agree from a macro perspective. But it would be nice if he would at least recognize some of these inconvenient facts during his interviews. Technological progress isn’t always just about some innovation measured in a clean record-breaking economic report. What did it take to get there? Go back and talk to the people who broke their bodies to build all of that infrastructure from 1870 to 1930.

The Demographic Time Bomb

Now layer on top of this a second fact that rarely enters the AI conversation — parts of the world are depopulating. That’s the message in the media. It’s also why many people have accepted mass migration as a solution. We’ll leave that insanity for another day.

Nevertheless, birth rates across the developed world, and increasingly across the developing world, have fallen below replacement level. The United States, China, Japan, Korea, and many countries in Europe are all heading toward population decline over the next century. Andreessen argues that this creates a context in which AI isn’t just a nice productivity enhancement. It’s a necessity.

Here’s the part that almost nobody is talking about. Without AI, the more pressing economic crisis wouldn’t be too many jobs disappearing. It would be too few workers to fill them, too little consumer demand to sustain growth, and an economy gradually hollowing itself out. “You’d be looking at these very dystopian scenarios of an economy self-euthanizing over time,” he says. That is the world we were heading into before the models arrived. Here Andreessen gets quite negative and sees no alternative but to embrace the robots. That’s one view, certainly. But it’s also not the only view.

“If we didn’t have AI, we’d be in a panic right now about what’s going to happen to the economy,” he says. And the only reason we’re not in that panic is because the technology showed up at exactly the right moment. “The timing has worked out miraculously well. We’re going to have AI and robots precisely when we actually need them.”

This reframes the job loss anxiety entirely. In a world where working-age populations are shrinking and immigration is increasingly politically constrained, the workers who do exist are not going to be displaced. They are going to be scarce. “The remaining human workers are going to be at a premium, not at a discount,” he says.

The Job Loss Panic is the Wrong Panic

Andreessen is direct about what he thinks of the mainstream narrative on AI and employment. “The job-substitution, job-loss thing is very reductive. It’s an overly simplistic model.”

His reasoning is straightforward. We’ve been in an era of such slow economic change that even a dramatic acceleration from AI would only bring us back to historical norms, not blow past them. “Even if AI triples productivity growth in the economy, which would be a massively big deal, it would take us back to the same level of job churn that was happening between 1870 and 1930.” And that era was one in which people felt surrounded by opportunity, not displacement. Again, he’s focusing on one part of the equation here, which is expected given his own personal net worth. Does he see the other side? It’s hard to tell.

Even in the more radical scenario, where AI genuinely transforms entire industries overnight, the economics don’t point toward mass misery. They point toward something much stranger and more interesting, which is a collapse in prices. Deflation.

Think through the mechanics. Massive productivity growth means more output for less input. You’re substituting AI for human workers, or for entire categories of effort. The result is gluts of goods and services across affected sectors. And from those gluts, prices fall. The thing that costs you a hundred dollars today costs ten dollars tomorrow. “That’s the equivalent of giving everybody a giant raise, right? Because now they have all this additional spending power.” That spending power fuels growth, creates new fields, and leaves everyone materially better off.

And even if some unemployment does emerge at the far end of that process, the social safety net becomes dramatically cheaper to run, because the costs of everything it covers — healthcare, housing, education — have collapsed along with everything else. “There’s no scenario in which everybody’s just poor,” he says. “In fact, it’s quite the opposite.”

He’s careful to note that none of this is bold or precise prediction. “Everything I’ve just described is just a very straightforward extrapolation on very basic economics.” More likely, he expects the process to be incremental rather than overnight. But even the incremental version, in his view, is a fundamentally good news story. This is a typical view for those who have experienced the many boom-bust cycles in Silicon Valley.

The Philosopher’s Stone

Before getting into what people should actually do about all this, Andreessen takes a detour through the history of alchemy that turns out to be the most memorable passage in the conversation. Take Isaac Newton. He spent decades obsessed with a problem he could never solve. The problem was the philosopher’s stone. It’s a hypothetical process for transmuting lead into gold to convert the most common thing in the world into the most rare and valuable thing in the world. Newton never cracked it. Nobody has ever cracked it. But people have surely tried and that’s led to a great deal of fraud over the years.

“Now we literally have a technology that transfers sand into thought,” Andreessen says. “The most common thing in the world, converted into the most rare thing in the world. AI is the philosopher’s stone.” That may be a stretch. But I appreciate the optimism. He’s not being metaphorical, though.. Silicon is literally made from sand. And what comes out the other end is something that seems to be reasoning, creating, diagnosing, writing, and codind. Newton would have approved. Andreessen believes we’re largely there now, but I have to think we have a ways to go yet.

Jobs and Tasks

Here is where Andreessen makes a distinction that most people miss entirely, and it reframes the whole anxiety around careers.

A job, he says, is not the atomic unit. It’s a bundle of tasks. “Everybody wants to talk about job loss, but really what you want to look at is task loss. The job persists longer than the individual tasks.”

This has always been true. In the 1970s, no vice president typed their own correspondence. They dictated memos to secretaries, who typed and mailed them. When email arrived, the secretary’s role shifted. Instead of typing letters, secretaries printed incoming emails and hand-delivered them to the executive’s office, then typed the executive’s handwritten replies and sent them back. Today, executives handle their own email, while their assistants manage travel, logistics, and scheduling. Both roles survived. Both roles changed completely.

The question most people are asking, will my job disappear, is the wrong question. The right question is which tasks in your job are about to rotate out, and whether you’re ready to absorb the new ones that replace them. For people who work for Silicon Valley companies, this is normal. Life changes, sometimes radically, like clockwork every six months or so. It’s just normal.

The Original Calculator was a Person

Andreessen traces a history of coding that puts the current moment into context. The word “calculator” originally referred not to a machine but to a person. Rooms full of human beings performing mathematical calculations by hand, sometimes thousands of them, for insurance companies calculating actuarial tables, military logistics, and government agencies. He’s right about this. I can remember as a kid when my father took me to his office at Grumman Aerospace. I saw massive rooms of dozens of designers hunched over boards sketching aircraft schematics by hand. Things changed when computer-aided design and manufacturing (CAD/CAM) came along, which enabled one engineer to handle what used to take a whole team to achieve. It’s real.

Back to Andreessen. After human calculators came machine computers and machine code. The first computers had no programming languages but instead were programmed in ones and zeros. Then came punch cards. Then assembly language, which was essentially machine code with a layer of readable English on top. Then higher-level languages like C, which compiled down to machine code. Then scripting languages.

That last transition is the relevant one for the masses of developers. When JavaScript and Python and Perl arrived, there was a loud argument in the technical community about whether scripting was real programming. “Real programmers,” the objection went, write code that compiles to machine code. They do their own memory management. They understand every layer. Scripting was cheating because the complexity of previous tools was now buried in the system. That’s automation.

The scripting skeptics were wrong. Those languages swept the world. Most coding today happens through scripting languages that have abstracted away multiple layers of detail that programmers once managed by hand.

AI coding is the next layer. The day job of the best programmers right now is not writing code. It’s arguing with AI bots. “They sit there and they shift from browser to browser, terminal to terminal,” Andreessen says, “and their day job now is kind of arguing with the AI bots trying to get them to write the right code, then debug it, fix the problems, change the spec.” He’s stretching here again because even a casual conversation with advanced software engineers will reveal that they are still writing most of the code. True automaton of code generation will take more time. It’s coming though. And fast.

But if you don’t know how to write the code yourself, you can’t evaluate what the bots are giving you. You can’t tell when something is wrong. Multiple engineers have told me this as well. Your understanding still has to go all the way down as much as possible. If the goal is to be one of the best software people in the world, he says, you want to understand every layer of the stack, including how the AI itself works. And AI, conveniently, is your best friend for learning all of that. Ask it to teach you. Have it quiz you. “There’s never been a technology before where you can ask it: teach me how to do this thing.”

The Mexican Standoff

For the product managers, engineers, and designers who make up a large share of the technology workforce, Andreessen has a characteristically vivid way of describing what’s happening right now.

“There’s a Mexican standoff happening between those three roles. Every coder now believes they can also be a product manager and a designer because they have AI. Every product manager thinks they can be a coder and a designer. And then every designer knows they can be a product manager and a coder.”

And here’s the interesting bit. They’re all more or less correct. AI is now good enough at all three functions that someone with deep expertise in one area really can use it to do credible work in the other two. The irony, he notes, is that all three will eventually realize they can also replace their manager with AI, aiming the guns, as he puts it, up the org chart. That’s the next phase of the standoff. The question is not whether the roles are converging. They clearly are. The question is what that means for how people should build their careers.

Build an E, not a T

The conventional career advice for the past decade has been to become T-shaped. Develop deep in one discipline with broad familiarity across adjacent ones. During the conversation, Lenny proposed an evolution of that model, something more like an E or an F laid on its side with multiple genuine verticals of competency rather than just one. Andreessen agreed. It’s a useful frame and it’s also commonly known among tech types.

Scott Adams, the creator of Dilbert, put the underlying principle well. “The additive effect of being good at two things is more than double,” Andreessen says. “The additive effect of being good at three things is more than triple. You become a super-relevant specialist in the combination of the domains.”

Adams himself was a pretty good cartoonist and a pretty good student of business. Neither skill alone would have produced Dilbert. Together, they produced one of the most successful cartoons in history. In Hollywood, the directors who can also write are not just doubly valuable. They’re categorically different from everyone else.

Andreessen’s friend Larry Summers, the former Harvard president, frames it as an economic principle. Don’t be fungible, Summers would tell people. Don’t be a cog. Don’t be replaceable. Andreessen extends the point further. If you’re just a designer, just a product manager, or just a coder, you can in theory be swapped out. But someone who combines those domains in a way that’s genuinely rare becomes “massively important because you’re one of the only people in the world who can do that combination.” This isn’t easy to do. And keep in mind that everyone’s doing it so the ability to compete still remains at the core of any career.

All you can do is keep leveraging all the tools you possibly can. But there’s also something deeper here about how AI functions as a learning tool, not just a productivity tool. You can watch it work and learn from what it’s doing. You can ask it where you went wrong. You can run one AI against another, have one write the code and a second one debug it, and let them argue. These are skills, Andreessen says, that are going to become incredibly valuable.

Three Layers of What AI Actually Changes

When Andreessen talks to the most forward-thinking founders right now, he describes three distinct layers of transformation, each more profound than the last.

The first is AI redefining the product itself. Take Adobe and Photoshop, a decades long franchise in image editing. Is AI a feature to be added to Photoshop for smarter editing? Or do users just stop editing images entirely because they’re generating new ones from scratch? The answer varies by domain, but the question is being asked everywhere.

The second layer is AI redefining jobs within companies. If you have budget for a hundred coders, do you still want a hundred but now each doing ten times more? Or do you now only need ten? The best founders are working through this right now.

The third layer, the one he says hasn’t quite emerged yet, is AI redefining what a company even is. “Can you have entire companies where the founder does everything,” he asks, “because what the founder is doing is overseeing an army of AI bots?” Bitcoin is probably the most spectacular example. Instagram and WhatsApp achieved enormous outcomes with tiny teams. The holy grail of the one-person, billion-dollar outcome has existed as an aspiration for years. AI may finally make it achievable at scale.

Determinate vs. Indeterminate Optimism

Andreessen’s investment philosophy at Andreessen Horowitz is built on a framework he borrows, and gently argues with, from Peter Thiel. Thiel distinguishes between determinate optimists, people who believe the future will be better because they are going to do a specific thing to make it so, and indeterminate optimists, people who believe things will improve without being able to say exactly how. Thiel has historically been skeptical of the latter, seeing it as a polite name for wishful thinking.

Andreessen pushes back. A16Z’s strategy is firmly indeterminate optimist, and he doesn’t think that’s a weakness. The founders they back need to be determinate optimists. Elon Musk is the archetype. He’s building the electric car, building the solar panels, getting to Mars. He’s very specific, very committed. Founders get to run their companies and put their hand on the steering wheel. VCs don’t. And history remembers Henry Ford, not the seed investor who funded him alongside nine other car companies that failed.

But the virtue of indeterminate optimism at the portfolio level is that it runs as many experiments as possible, across as many smart people trying as many interesting things as possible. “The great virtue of the capitalist system, of the American economy, of Silicon Valley, is we don’t just have one determinate optimist and we don’t just have ten. We have a thousand, and then ten thousand.” You don’t have to pick the winner in advance. You have to make sure the system that produces winners keeps running.

Beyond the Human Ceiling

Andreessen admits he has always struggled a little with the concept of Artificial General Intelligence (AGI). There’s the cosmic definition, which is essentially the singularity, a moment where self-improving machines race so far ahead of human judgment that human decisions become irrelevant. He doesn’t think we’re heading there, at least not in the near-term. And then there’s the more general industry definition, which the co-founder of Anthropic has described as AI that can perform a broad set of the most economically valuable tasks as well as a human. We’re getting close to that, if we’re not already there.

But Andreessen thinks even that definition undersells what’s actually coming. The reason is that human skill level is not a theoretical ceiling. It’s a biological one.

Human fluid intelligence caps out around an IQ of 160. That’s the Einstein level. At 140, you’re looking at the world’s best research scientists and best-selling authors. At 130, the sharpest lawyers. At 110, a strong line manager. At 105, a careful small-business accountant. The range of impressive human cognition, 110 to 160, is determined entirely by what fits inside the human skull. There’s no theoretical limit beyond that. We’re just capped by our own biology.

Current AI models are already testing in the 130 to 140 range by some measures, Andreessen says. They will reach 160. And then, unlike humans, they won’t stop. “I think we’re going to have AI models relatively quickly that are going to be 160, 180, 200, 250, 300,” he says. And the question that follows from that, he says, is not threatening. Would the world be better off with more Einsteins or fewer? More, obviously. The same logic applies to machines that think at that level or beyond it.

He’s candid about what that means on a personal level too. “I know a bunch of people who are smarter than I am,” he says. “And I know it because when I talk to them, at a certain point I’m just like, this person is outthinking me and they’re going to keep outthinking me.” He reads ten books on a topic and forgets almost everything a few days later. His memory isn’t perfect. His processing has limits. Living with those constraints is just the condition of being human. Having tools available that don’t share those constraints is something we genuinely haven’t experienced before, and he thinks we’re not fully appreciating what that means.

What Marc Reads

Andreessen is an active consumer of information. His media diet reflects the same disciplined thinking he applies to everything else. He reads X for what’s happening right now, and old books for what’s timeless. Everything in the middle — magazines, newspapers, newsletters — he approaches with heavy skepticism. Go back and read old newspapers, he says. None of the predictions played out. None of it was actually relevant. The problem with the middle, he says, is that by the time a magazine article hits publication it’s often already out of date.

There is a big exception, though, which is directly accessing specific domain practitioners and the content they produce. These are the people who are actually doing real work and the things they’re writing or talking about. Podcasts, long-form conversations, technical newsletters written by people with real skin in the game. That’s where the real signal lives. “The world is awash in that today in a way that it wasn’t as recently as ten years ago.” People love talking about what they do, and for the first time in history, the rest of us have direct access to them.

What this actually means for you

The philosopher’s stone, the thing Newton spent his life chasing, turned out to be real. It’s just that what it transmutes isn’t lead into gold. It’s curiosity and effort and deep knowledge into something that, for most of human history, only a very small number of people ever got to become. Andreessen is not predicting a frictionless utopia. The structural impediments are real — the cartels, the red tape, the politics, the regulations — that prevent meaningful progress in atoms for the last half a century. And they will continue to slow things down. So, if you’re building in the physical world, budget accordingly, I guess.

But at the level of the individual, the picture is different. The person who goes deep in at least one domain, uses AI to extend genuine competency into two or three others, and treats the technology as a teacher rather than just a tool, that person can move into the best labor market in fifty years. Not because the world is getting easier but because they will be genuinely hard to replace.

“People who really want to improve themselves and develop their careers should be spending every spare hour at this point talking to AI,” Andreessen says. “Train me up. Superpower me.”

The real AI boom, he seems to be saying, isn’t the one that’s been happening or that we hear on social media and the news. Instead, it’s the one that starts when people understand what they’re actually holding in their hands.

We’ll see. I’m certainly all over it. Every. Single. Day.

Aaron Siri at Kennedy

Imagine a world where American drug companies were required to prove that their vaccines were “safe and effective” by testing them against inert placebos? And then imagine that those companies were liable for their vaccines and patients were allowed to sue the companies if those vaccines turned out to be not “safe and effective” but instead caused harm? That’s not this world, that’s for sure. Not even close. To explore the details of these issues, see Aaron Siri’s presentation at the Kennedy Center — An Evening with Aaron Siri – Millennium Stage (March 23, 2026) where he reviews data in his new book Vaccines, Amen. Also, here’s his new YouTube channel. Fair warning up front, though: if you’re not familiar with this topic, you’ll be shocked. I’ve been following the issue for decades so nothing in this content surprises me one bit.

The ICU

Seems quiet and clean now. But 30 hours in the ICU was loud, bloody, messy, smelly, and painful. Can still hear the screams from the guy next to me, and the never ending moans from the guy two bays over. No one slept. Everything dragged on for what seemed like forever. I just focused on my breath. One in, one out. Next. One in, one out. That kept the panic at bay. One time in the middle of the night I tried to gently turn over but ripped out all the wires from my chest. Then a small army of nurses flew in with little flashlights shining them in my eyes. Nurses aren’t normal, btw. They are extraordinary. Without them, we all die. Anyway, every day is rehab day now. Progress is painfully slow but steady.

The ICU. Japan. Four days later. My bay on the right.
The ICU. Japan. Four days later. My bay on the right.

The Decline and Fall of Peter Attia

Over the years Dr. Peter Attia built a reputation as an influential voice in popular longevity medicine. He’s been pervasive online. And recently he was hired by CBS News based on his credibility. So, he’s big time. His podcast draws millions of listeners, and his book became a bestseller. People trust him with their health decisions. I was an active follower until about 2018 when his arrogance finally got the best of me and I just got tired of him. It wasn’t much of a loss for me to dump him because there are plenty of other excellent people in the field who question the established medical paradigm. But Attia kept going and getting rich in the process, which is common among the top tier people in the field. But apparently, he got himself in way over his head. Hopefully, he’ll come crashing down as fast as he rose. It’s unlikely. But I can still hope.

Check out this report about Attia from Joseph Everett — Hidden Data: How the Top Longevity Doctor tricked us all — that reveals a pattern of overconfidence, selective data presentation, and dismissiveness toward ideas that challenge his worldview. In the video, Everett analyzes Attia’s claims and public statements and does several on-camera interviews with Dave Feldman, Dr. Nick Norwitz, and Dr. Chris Masterjohn to provide some scientific context and to counter some of Attia’s most confident assertions. Here’s a review of Everett’s report. 

The Epstein Connection

I never thought Attia would fall based on his association with Jeffrey Epstein because until last week no one even knew he was hanging out with that crowd. But that’s where we are. The Epstein files revealed more than 1,700 PDFs mentioning Attia. The emails show a close relationship that continued even after Jeffrey Epstein’s crimes were publicly known. Attia met Epstein in 2015 through Epstein’s ex-girlfriend. He called Epstein “literally one of the most interesting people I’ve ever met.” He expressed interest in visiting Epstein’s island, and he stayed at Epstein’s apartments. He attended dinner parties with Epstein and his so-called famous friends. 

Attia’s timeline with Epstein is particularly striking. According to Attia’s own book, on July 11, 2017, his wife rushed their one-month-old infant to the hospital because the baby had stopped breathing and his heart had stopped beating. His wife stayed in the hospital for four days, pleading with Attia to come home. He said he couldn’t because he was in New York with “important work.” The next day, July 12, Attia emailed Epstein confirming he could meet him. That’s difficult for any parent to imagine. But it provides a view into the dark space in which Attia lives so freely.

When the Miami Herald published its piece on Epstein in November 2018, Attia claimed he was “repulsed and nauseated.” Yet he stayed in touch with Epstein for at least four more months. In December 2018, he asked Epstein about the “fallout from recent story.” In February 2019, he was still emailing with the subject line “Where are you these days?” There are many other disgusting exchanges between the two in those files. Read them if you want.

The connection to Epstein itself might not be damning. Time will tell. Plenty of people got caught up in Epstein’s orbit but haven’t been implicated with crimes. But the same overconfidence that led Attia to maintain this relationship shows up repeatedly in his health claims.

The VO2 Max Problem

Attia built much of his longevity framework around VO2 max. He called it “perhaps the single most powerful marker for longevity.” He was emphatic. He brought it up on podcasts constantly. The claim appeared on 60 Minutes. It’s in his book. He says everyone should know their VO2 max and track it.

But there’s a problem. The studies he cites to support this claim never actually measured VO2 max. In Everett’s video, Chris Masterjohn, PhD, points out that you can search some of these papers for the word “oxygen” and find nothing. What the studies measured was how long people lasted on a progressively harder treadmill test. That’s not the same thing. I remember people online questioning Attia on this issue and many other issues. Attia generally ignored them. My impression was that Attia grew to be connected and protected. Or at least he was acting that way. 

And when Attia made a table for his blog showing VO2 max levels, he simply relabeled data from a paper that measured something else entirely. The table made it into his book. He tells people they should spend $200 annually to have their VO2 max measured at an exercise science lab, when all you actually need is a treadmill, according to Masterjohn. 

Masterjohn explains that most people in the general population can’t even reach their VO2 max during these tests. They give up because of heart palpitations, leg pain, or feeling like they’ll throw up. They’re never hitting their actual maximum oxygen consumption. But Attia presents this metric as if it’s the single most important thing you can measure. It shows how Attia grew to become totally disconnected from regular people just trying to get healthy. It’s clear in retrospect that Attia was focused on other clients who could pay him millions. Everett points out in his video that Attia would show up on video podcasts wearing several different $300,000 watches. 

The real problem is that focusing so intensely on VO2 max could lead people to spend hours doing one repetitive motion instead of diversifying their training. As Masterjohn says, gymnasts and pole vaulters have eight years on the general population for lifespan. That’s a testament to the breadth of functional training, not endless cardio optimization.

Hiding Inconvenient Data

Now let’s move to cholesterol and statins. When a major study in Cell Metabolism showed that atorvastatin (the most profitable drug in history) slashed GLP-1 levels while worsening glycemic markers and insulin resistance, Dr. Nick Norwitz (MD, Ph.D.) started discussing the issue online. Two weeks later, Attia published a newsletter that appeared to be responding to Norwitz’s analysis. That seems like Attia’s back handed style. 

But Norwitz noticed something strange. Attia’s newsletter tried to discredit the study, but it completely ignored the most relevant graph. As Norwitz puts it in Everett’s report, “There’s one graph of relevance. An 8-year-old can tell you what’s going on. There’s a big red line. It goes down.” Attia showed the figure that contained this graph but clipped out the specific panel that showed the problem. It’s hard to think that wasn’t intentional, Norwitz says. 

Nick Norwitz and the New Generation

Norwitz himself represents exactly the kind of rigorous, open-minded research Attia claims to support. At age 30, Norwitz has already published 54 peer-reviewed papers. His scientific impact score is higher than Thomas Dayspring, the lipidologist Attia regularly cites as his go-to cholesterol expert. And Dayspring is 50 years older. Norwitz is prolific online, too, but he embraces the conversation with both technical and non-technical people — unlike Attia. 

Norwitz recently gave some credibility to another researcher, Dave Feldman, who is a popular software engineer who has been researching LDL cholesterol for years now. Feldman developed an alternative cholesterol model by conducting his own personal experiments and publishing the results online and in scientific papers. Norwitz tested some of Feldman’s views and lowered his LDL cholesterol from 384 to 111 by eating Oreo cookies for 16 days. When he tried statins for 6 weeks, he only managed to lower his LDL from 421 to 284. The Oreos worked better than the drugs. Why the Oreos worked at all is interesting.

The point isn’t that people should eat Oreos instead of taking statins. The point is that our understanding of cholesterol metabolism is more complex than Attia presents it. And when young researchers like Norwitz demonstrate this complexity, Attia either ignores them or dismisses their work outright.

Dismissing Dissent

Perhaps the most revealing episodes with Attia involve Dave Feldman. When Feldman appeared on Attia’s podcast in 2018, the exchange revealed something troubling about how Attia pushes his influence and bullies people. He interrupted Feldman a staggering 66 times, according to Everett’s count. He called Feldman’s investigation “brain damage.” His treatment was arrogant and dismissive to say the very least. He told people interested in cholesterol questions to “sit down, shut up for a minute, and pay attention.” I remember listening to that episode. It was infuriating. And it was clear that it was actually Attia who was out of his league — not Feldman. I mean, when a software engineer goes up against an MD you’d expect the engineer to lose in a fair fight. But Feldman more than held his own in the match. In fact, he was really impressive. Ultimately, though, when some so-called “expert” reacts like Attia did, it’s likely an indication that they have been exposed. I was mostly done with Attia at that point, at least on the cholesterol issue.

The irony is that Feldman’s questions were reasonable and backed up with some excellent, albeit early, data. According to Feldman and Everett, Most research showing that LDL cholesterol is dangerous comes from studies of metabolically unhealthy people. But what happens to people who have high LDL and also have high HDL and low triglycerides, which are both markers of good metabolic health? Do they still die early? What about the many other biomarkers that demonstrate good health in these people? These questions were totally dismissed by Attia.

Attia further said that this line of inquiry was pointless. He challenged Feldman to crowdfund research if he thought his ideas had merit. Feldman, always up for a good challenge, did exactly that. He raised $350,000 for the Lean Mass Hyper-Responder study. The results showed that metabolically healthy people on a keto diet with sky-high LDL didn’t develop more arterial plaque in their hearts. Imagine that. Attia’s response? Silence. And when invited to co-author a journal editorial about the findings by his own former head of research, Bob Kaplan, Attia declined and said there were people “far more reputable” to discuss it with. 

Meanwhile, Feldman went on to publish eight peer-reviewed papers on the subject and continues to work closely with Norwitz on cholesterol research.

The Real Issue

The pattern is clear. Attia presents himself as committed to open-minded scientific inquiry and he strongly lectures people about that. Yet his own behavior says otherwise. He once stated we should “go back to our original ideals: open minds, the courage to throw out yesterday’s ideas when they don’t appear to be working, and the understanding that scientific truth isn’t final.” But when faced with obvious data that challenges his positions, Attia either ignores it or attacks the messenger, many times in childish, unscientific ways. You can see him online repeatedly rolling his eyes, smirking, shaking his head, raising his voice, and using harsh language when he’s simply questioned by anyone. From his powerful platform, he still maintains extreme confidence in his claims, which seem more shaky than ever. That’s not science. That’s ego. And it’s the same attitude that kept him emailing Jeffrey Epstein months after the world knew what Epstein had done.

I wonder how many online health gurus who called him a friend will come out publicly and question Attia now. So far, too many are silent. I wonder, do these people think we’ll not notice?

Make Protein the Priority

Dr. Donald Layman on Building Muscle, Losing Fat, and Reducing the Metabolic Decline of Aging

Dr. Donald Layman has spent decades studying protein and amino acids. He’s a Professor Emeritus at the University of Illinois, and he’s published more than 120 peer-reviewed papers. His work has transformed how we think about dietary protein. And unlike many scientific experts promoting their views online, Layman’s research cuts cleanly through the confusion about how much protein we need, when to eat it, and why quality matters more than most people realize. 

This post is based on eight interviews with Layman in the last few years, and it’s focused exclusively on his insights and recommendations about protein metabolism, muscle health, and optimal nutrition. All quotes are Layman’s. See the resource list below.

The Muscle-Centric Philosophy

Layman developed what he calls a muscle-centric approach to nutrition. He explains that nutrition comes down to two critical tissues: the brain and skeletal muscle. Everything else in the body adapts and regulates, but these two tissues must be supported well because they determine our quality of life. Layman says, “If you keep muscle healthy, you’ve got a good shot at avoiding obesity, avoiding diabetes, avoiding cancer as you age.”

This focus makes sense when you understand what muscle does. It serves as our largest reservoir for glucose, holding about 75 to 80 percent of our total glucose storage capacity. When muscle health declines, however, we lose this metabolic buffer. Then fat droplets accumulate in muscle cells, and over time this creates insulin resistance and makes it harder for muscles to accept carbohydrates. This cascade leads to hyperglycemia in the blood and eventually diabetes.

But muscle does more than manage glucose. Layman says, “Whether you’re 16 or whether you’re 65, you have to build 250 to 300 grams of new protein [every day] just to replace and repair what you already have.” This constant daily turnover happens regardless of age. The difference is efficiency. Young people enjoy hormonal protection that makes protein synthesis easier. But older people must work harder on both diet and exercise to maintain the same results. The difference is significant. 

The body operates in a constant state of protein flux. Every protein in your body gets broken down and rebuilt in a continuous cycle. Layman says that we replace the equivalent of every protein in our bodies about four times per year. This metabolic demand never stops. The liver must produce proteins 24 hours a day to maintain blood protein levels, manufacture enzymes, and support immune function. When dietary protein falls short, the body simply pulls amino acids from muscle tissue to keep these essential processes running. If dietary protein continues to decline over many years, people loose a significant amount of muscle mass. Sometimes this loss is hidden when people gain weight, but when they lost that weight the muscle loss becomes obvious. 

Children require only about five grams of net new protein per day for growth. Adults need none for growth, but they require vastly more protein for maintenance and repair. The adult body must synthesize 250 to 300 grams of protein daily just to stay even. This remarkable fact challenges the common assumption that children need more protein than adults. The opposite is true. In reality, adults face a more demanding protein requirement because the efficiency of their protein metabolism declines as they age. Also, adults are not as active as children and so their muscles don’t receive the signal from exercise to grow. 

The RDA Problem

The Recommended Dietary Allowance for protein in the United State sits at 0.8 grams per kilogram of body weight. Layman has spent years explaining why this number falls way short. The history is important. The RDA came from nitrogen balance studies conducted decades ago using conscientious objectors during World War II. Researchers put these men in special suits to collect nitrogen leaving their bodies. They lowered protein intake to zero, then gradually increased it until nitrogen inputs matched outputs. This became the basis for protein requirements.

There are big problems with this early research, however. “The RDA is the average requirement. By definition, that’s only the average. Half of the people are above average.” The RDA represents a minimum for survival, not optimal health. Also, nitrogen balance studies carry inherent flaws. All amino acids contain different amounts of nitrogen and plant proteins have more non-essential amino acids than animal proteins. When researchers use nitrogen analysis, they consistently overestimate the actual protein content.

Layman now argues for a fundamental shift in how we think about dietary requirements. He says we should stop talking about protein as a requirement and start talking about essential amino acid requirements. Layman says, “We actually don’t need protein in the diet. We need nutrients and the nutrients are essential amino acids.”

Based on dietary surveys, about 45 percent of Americans consume protein below recommended amounts. That’s an incredible statistic given that RDA is the minimum requirement. Women face particular challenges, especially those over 60 and between 18 and 22. Older women prefer carbohydrates to protein in their diets during a time when their bodies are aging and require fewer total calories to function. Their resting metabolism literally drops about 100 calories per decade naturally. However, younger women often intentionally restrict protein for appearance or moral reasons when many of them adopt extreme diets like veganism. Only women in the middle age range are getting the minimum amount of protein.

How Much Protein Do We Actually Need?

Remember that the American RDA for protein is  0.8 grams per kilogram of body weight. However, Layman recommends that most adults should consume between 1.2 and 1.8 grams of protein per kilogram of body weight. For practical purposes, this translates to roughly 0.5 to 0.8 grams per pound. The exact amount depends on several factors including age, activity level, and metabolic health. Layman’s recommendation isn’t unreasonable, although it may sound shocking to some people since it’s substantially higher than what the U.S. government has been saying for decades. Many others in the protein field go even higher to 1.0 – 1.2 grams of protein per kilogram of body weight just to account for those days when it’s impossible to get the minimum. Life is variable so people have to account for things like travel, schedules, emergencies, etc. 

When determining your dietary protein target, Layman says that you should use your ideal body weight rather than your current weight if you carry excess fat, which is at this point most Americans. He explains that adipose (fat) tissue does not require protein maintenance the way muscle does. For someone who weighs 200 pounds but should weigh 170, calculate protein needs based on 170.

The upper limit matters less than most people think, and Layman is actually somewhat conservative in his recommendations. He says that his research supports a protein intake of up to 1.8 grams per kilogram without concerns. “I don’t think the data really supports going above 1.8 g per kg,” but he emphasizes this comes from lack of research rather than evidence of harm. Studies on protein intake from 0.8 grams per kilogram up to 3 grams per kilogram show that protein appears totally safe across this entire range. Note that most professional and amateur athletes who are knowledgable about protein exceed Layman’s recommendations. Yet Layman sticks to his levels because that’s what his data demonstrates. Actually, Layman rarely talks about professional athletes. Instead, his focus and expertise is on the biochemistry and health of the general population through their entire lifespan. 

Protein Quality Makes the Difference

Not all protein delivers equal benefits. Layman focuses on three key amino acids: leucine, lysine, and methionine. Leucine especially triggers muscle protein synthesis through the mTOR pathway. Animal protein foods deliver about 8 to 10 percent leucine, while plant proteins typically contain only 6 to 7 percent. This presents a problem for vegans as they age. Smart vegans supplement. But if people who have adopted vegan diets don’t supplement, they generally find that over time they lose muscle mass and overall lean tissue as they age. You can clearly see this with the naked eye in vegan populations that aren’t overweight. 

This difference compounds when you consider digestibility. Animal proteins offer about 95 percent digestibility. Plant proteins drop to 60 to 75 percent, and that varies also with different cooking techniques. When you eat beans, for example, your body cannot access nearly half the protein on the label. “The idea that beans are a substitute for beef is a really dumb idea,” Layman says. Strong language like that is rare from Layman, yet he has decades of research to back up his claims. 

The distinction matters most as we age. “Under 30 it doesn’t matter when you eat your protein. It doesn’t matter very much the quality of the protein. But once you cross 30, now the efficiency of how you put it all together makes a big difference in how you’re going to repair or remodel your protein for healthy aging.”

Americans who shift toward plant-based diets, with the extreme being pure veganism, typically increase their grain consumption substantially rather than eating more beans, chickpeas, and almonds. “Americans get 80% of their plant-based protein from wheat. And wheat is a very poor quality protein.” When people eat primarily grains at the RDA level, they become deficient in two or three essential amino acids.

The wheat problem is deeper than most people realize. About 60 percent of American protein comes from animal sources such as fish, eggs, milk, and meats. The remaining 40 percent, though, comes from plants, and 80 percent of that plant protein comes from wheat. That’s a problem because it makes it challenging to get all the required amino acids in their proper proportions. Wheat is deficient in lysine, tryptophan, threonine, and leucine. When people shift toward more plant-based eating by simply consuming more wheat products, they create multiple amino acid deficiencies. These shortfalls affect everything from metabolic signaling to fat burning to brain function through tryptophan and serotonin production. Remember that the body must have all the essential amino acids in specific ratios, so if you don’t consume them the body will simply take them from storage — the muscles. 

Layman emphasizes that although relatively few Americans are vegans the general population is already largely eating a plant-based diet. Over 70 percent of our calories come from plants, but more than 80 percent of those plant calories come from added sugars, oils, hydrogenated fats, and highly refined carbohydrates. Junk food, basically. The number one plant in the American diet is french fries, followed by tomato sauce on pizza and lettuce. “We don’t need a more plant-based diet. We need a better one.” Here Layman distinguishes himself from the current carnivore trend. He’s perfectly ok with an omnivore diet as long as the non-meat portions are based on healthy, properly prepared, food. 

The body requires 20 different amino acids to build protein tissue. Nine of these amino acids are essential, which means that the body cannot manufacture them and must obtain them from food. The remaining 11 are non-essential. But that term can be misleading because the body still needs them. The difference is production capability, not importance. When you lack essential amino acids, protein synthesis stops completely. The body cannot substitute one amino acid for another or skip amino acids in a protein chain. So the body has no choice but to tap the stocks and get those amino acids. Where does it go to find those amino acids? To the muscles. 

Leucine stands out among essential amino acids for its unique signaling role. Beyond serving as a building block, leucine activates the mTOR pathway that initiates muscle protein synthesis. You need approximately 2.5 to 3 grams of leucine per meal to trigger this response. Animal proteins deliver this amount in modest portions, while plant proteins require much larger servings to reach the same threshold. This is one reason why people on plant-based diets must consume more total calories because the body is always hungry and looking for those required amino acids. 

Lysine becomes particularly important when evaluating plant proteins. Grains contain very little lysine, which creates a major limitation in grain-based diets. This shortage explains the traditional practice of combining beans with rice or corn in cultures relying heavily on plant foods. The combination provides complementary amino acid profiles that neither food delivers alone. Experienced vegetarians and vegans know this issue well and monitor it carefully when they eat. It’s not a full proof method to acquire all the essential amino acids, but at least it’s an attempt to recognize the issue.

Methionine supports glutathione production, one of the body’s master antioxidants. Layman’s research shows that when protein intake drops toward the RDA level, glutathione levels decline. To maximize glutathione levels in older adults, protein intake needs to reach at least 1.2 grams per kilogram. This requirement is 50 percent higher than the RDA, revealing how the minimum standard fails to support optimal metabolic function. There is more than enough research on protein now. One wonders why the RDA remains so low. 

Timing and Distribution

The first meal after waking is most importance. Layman consistently emphasizes getting at least 30 to 35 grams of high-quality protein within an hour of rising. This first protein bolus stops the catabolic overnight fasting state that builds during sleep and triggers muscle protein synthesis. 

After that initial meal, distribution throughout the day becomes more flexible. In weight loss studies, Layman found success with 35 grams of protein at breakfast, 35 at lunch, and about 50 at dinner. But he allows variation based on individual preference. “The reality is we have really good data about the protein at breakfast. We have pretty good data about protein at dinner and we have zero data about protein at lunch.”

Layman’s own diet reflects this understanding. He typically consumes 40 to 45 grams at breakfast, 15 to 20 at lunch, and 50 to 60 at dinner. He finds that large midday meals make him sleepy and less alert. Everyone’s different. People just have to experiment to see what fits them best as long as they are getting what they need in any given day. But it’s important to hit that minimum effective dose per meal of around 30 grams of high-quality protein. This amount provides about 2.5 to 3 grams of leucine, enough to trigger muscle protein synthesis. Younger people need less. But older people need higher amounts and sometimes require 40 to 50 grams to achieve the same anabolic response.

The Fasting Dilemma

Time-restricted eating has grown in popularity recently, but Layman urges caution for anyone over 40. In fact, he’s emphatic on the issue. “I don’t think anybody over the age of 50 should ever fast.” The muscle mass lost during extended fasting becomes permanent without aggressive resistance training during the recovery period, and very few people realize this until they directly experience the process themselves. I certainly have. And it’s shocking.

But Layman distinguishes between proper time-restricted feeding and true fasting. Eating within an eight or ten hour window can work if you maintain adequate protein intake. But going 36 hours or longer without food creates a catabolic crisis that strips away lean tissue. Remember, the body must absolutely have all 9 essential amino acids at all times not only to build and maintain muscle but also for hundreds of regular biochemical processes just to maintain health. 

The problem intensifies with some popular one-meal-a-day approaches advocated by some people online. Protein synthesis can process only so much protein at once, typically maxing out around 40 to 50 grams. When you try to consume 100 or more grams in a single meal, your body cannot use all the amino acids for building tissue. You miss opportunities to stimulate synthesis multiple times throughout the day. You see these people online. Their bodies transform quickly. They tend to look shredded but also gaunt. They don’t look healthy at all. 

Research on protein synthesis shows effects lasting four to five hours after a protein-rich meal. This window suggests that we should space our protein intake every four to five hours to optimize results. Cramming all protein into one meal wastes the stimulatory potential of well-timed protein distribution throughout the day. 

Protein and Weight Loss

Layman conducted extensive weight loss research demonstrating protein’s protective effects on lean tissue. He found that typical weight loss results in about 50 percent muscle loss and 50 percent fat loss. This composition change proves devastating, especially for older people who clearly struggle to rebuild lost muscle. But a higher protein intake can change this situation dramatically. “We can make the weight loss 95% fat 5% muscle,” Layman says. The key lies in maintaining at least 100 grams of protein per day for women during caloric restriction combined with resistance exercise. Men need proportionally more protein based on their larger size.

Layman taught his research subjects a simple a visual method to make implementation much easier than counting grams. He says that protein and carbohydrates should look equal in size on the plate. Four ounces of meat appears roughly equivalent to half a cup of rice. This one-to-one visual balance creates roughly the right macronutrient ratio for a meal. 

Higher protein intake during weight loss provides additional advantages beyond muscle preservation. Protein carries a higher thermogenic effect than carbohydrates or fat, which means you burn more calories digesting and processing protein. Protein also provides superior satiety, reduces hunger, and makes caloric restriction more tolerable.

The metabolic advantage of protein extends beyond the thermic effect of food. When you maintain muscle mass during weight loss, you preserve your metabolic rate. Each pound of muscle burns more calories at rest than a pound of fat. But losing muscle during dieting creates a vicious cycle where your metabolism slows down and makes further weight loss harder and weight regain more likely.

Layman’s research reveals that people using skim milk during weight reduction lost the advantage of both protein quality and satiety. The fat naturally present in dairy products serves important functions. He recommends reduced fat products rather than fat-free versions, allowing people to control total fat intake while maintaining the benefits of naturally occurring dairy fat.

The body’s overall composition during weight loss matters far more than the number on the scale. A 60-year-old who loses 30 pounds but half of that comes from muscle ends up metabolically similar to an unhealthy 80-year-old. And visually, these people look obviously gaunt. The muscle loss accelerates the aging process and reduces functional capacity. In contrast, losing mostly fat while preserving muscle keeps you metabolically younger and maintains strength and mobility, both of which are critical to maintaining health as you age. If you think that strength isn’t a factor over time, just observe any elderly person after a few falls and hospital stays. Their bodies dramatically reduce in size as their health rapidly declines. 

GLP-1 medications also present particular challenges during this time. Sure, people lose weight rapidly on these drugs. But without adequate protein and resistance exercise, about 50 percent of the weight loss comes from lean tissue. Layman warns — strongly — that this muscle loss becomes especially problematic because these people often struggle to rebuild what they lost after they discontinue the medication. 

Exercise and Protein Work Together

Resistance training and protein operate synergistically. Layman estimates that building muscle comes down to about 75 percent resistance training and 25 percent protein intake. That may seem anti intuitive given his research on the biochemistry of protein synthesis. However, you cannot compensate for lack of exercise with more protein. The protein intake is the foundational requirement to enable people to maximizes the benefits of training. 

But as you age you don’t have to go crazy in the gym. The definition of resistance training is broader than most people assume. Layman says that gradual stretching and low impact workouts represent a major part of resistance exercise. His weight loss studies with middle-aged women used Nautilus machines without adding weights. The women simply moved through the full range of motion, which emphasizes the eccentric or stretching process. This protocol produced significant improvements in body composition and demonstrates that even a minimal level of moment is beneficial. 

For people intimidated by gyms, Layman recommends yoga, Pilates, or rubber band stretches and movements at home. He says, “The best exercise is the one you’ll do consistently!” The critical factor lies in creating mechanical stimulus that tells muscles they need to maintain or grow their mass. And while muscles are moving under even mild stress they signal other lean tissue to grow, such as bones, ligaments, and tendons. Movement is critical to health. 

Timing protein around exercise matters less than most people think. Layman’s research shows that resistance exercise in a fasted state can trigger protein synthesis mechanisms. But without adequate amino acids available, the signal cannot translate into actual tissue building. Layman says you can “trigger those processes” with exercise or leucine alone, “but you need the complete amino acid mix.” Here Layman emphasizes the interdependence of proper nutrition and proper exercise. 

The Kidney Myth

Concerns about protein harming kidney function persist despite evidence to the contrary. Layman explains that higher protein intake actually increases kidney size and improves the body’s glomerular filtration rate. The kidney adapts to increased protein load by becoming more efficient at clearing urea and creatinine from blood.

Research comparing 0.8 grams per kilogram of protein with 1.6 grams per kilogram shows accelerated clearance rates at higher intake levels. Layman says that “the rate of creatinine clearance, the rate of urea clearance, actually accelerates. GFR becomes more efficient.”

But protein restriction produces the opposite effect. When people consume low protein diets below or even at the RDA level, their kidneys actually shrink and clearance capacity decreases. For healthy people, protein intake up to one gram per pound appears completely safe for the kidneys. Some markers of kidney function improve simply with better hydration because protein requires more water for processing.

Practical Protein Sources

Eggs, dairy, fish, and meat provide the most complete amino acid profiles with the highest digestibility. Whey protein stands out as particularly effective due to its rapid digestion and high leucine content. Layman himself often mixes whey protein powder with Greek yogurt to combine the fast-acting whey with the slower-digesting casein found in yogurt.

For budget-constrained people, eggs and ground beef offer an excellent value. Layman says that when people shift from junk food and quick-service meals to protein-focused diets, they often actually spend less money despite buying better quality food. Various cuts of meat, fish, chicken, ham, cheese, and milk all qualify as functional protein sources since you don’t have to worry about getting all nine essential amino acids. Milk provides one of the easiest ways to fine-tune protein intake. Layman says, “Milk is one gram per ounce. I like milk. So if I need nine more grams at my meal, I just have nine ounces of milk.”

For plant-based eaters, though, the protein challenge intensifies. Beans have to be combined with grains to provide complete amino acid profiles, but even then, digestibility problems persist. Layman suggests that vegetarians and vegans should consider supplementing with essential amino acids to ensure adequate an intake of leucine, lysine, and methionine.

The dairy industry transformed itself based on Layman’s research. In 2003, he addressed 250 research and development professionals from dairy companies. He told them their yogurts contained nothing but sugar and urged them to develop Greek yogurt with higher protein content. Chobani launched within a year, and the entire yogurt market shifted toward protein-rich products. The proliferation of protein shakes and high-protein foods in stores today traces back to this research that shows the importance of protein quality and quantity. Or you could just eat a a traditional diet based on whatever culture you are from since they all generally contain enough protein without having to always think about reading labels and supplementing. 

Greek yogurt also offers advantages over regular yogurt because the straining process concentrates protein while removing excess liquid and sugar. However, Greek yogurt contains more casein than whey. Layman addresses this by adding whey protein powder to Greek yogurt, which creates an ideal blend of fast and slow-digesting proteins.

Whey digests rapidly and floods the bloodstream with amino acids within an hour. This quick availability makes whey excellent for stimulating muscle protein synthesis. Casein, however, digests more slowly and provides a steady release of amino acids over several hours. The combination delivers both immediate stimulus and sustained amino acid availability.

Cheaper protein sources work perfectly well for most people too. Ground beef provides complete protein at lower cost than premium steaks. Chicken thighs cost less than breasts and often taste better due to higher fat content over breasts. Whole eggs deliver superior nutrition compared to egg whites despite containing fat and cholesterol. The key lies in choosing real animal foods rather than processed protein products that promise convenience but deliver inferior amino acid profiles.

Layman conducted weight loss studies specifically examining low-income populations. These studies revealed that shifting to a protein-focused diet actually costs less than the typical American diet heavy in processed foods, chips, candy, and fast service restaurant meals. The perception that healthy eating costs more often reflects comparison of premium organic products with budget processed foods rather than realistic and simple alternatives. Learn to cook. And don’t forget to consider the cost of disease that results from decades of eating poor quality food.

The Carbohydrate Question

Layman takes a balanced approach to carbohydrates that focuses on individual needs and activity levels. He says that the brain and red blood cells require at least 100 grams of carbohydrates per day. If you fail to eat this minimum, your body converts protein into glucose through gluconeogenesis. This process wastes dietary protein that could otherwise build and repair tissues, which is critical as we age. 

Also, exercise intensity determines carbohydrate needs above this baseline. Activities below 65 percent of maximum heart rate primarily burn fat for fuel. Walking, cycling, and low-intensity exercise fall into this category and require minimal carbohydrates beyond the 100-gram minimum. Once intensity crosses the 65 percent threshold into moderate and high-intensity work, the body shifts toward carbohydrate metabolism. Resistance training, running, competitive sports, and high-intensity interval training all demand immediate carbohydrate fuel.

Layman practices what he preaches. He plays tennis and exercises intensely, so he includes carbohydrates in his diet. He says he needs adequate carbs to feel good and compete well during exercise. Without sufficient carbohydrates, he says his performance suffers because he operates above the threshold where fat alone can meet energy demands.

Layman’s one-to-one visual ratio of protein to carbohydrates on the plate creates an effective and visual framework for most people. This approach naturally limits carbohydrate intake while ensuring adequate protein. For someone eating 30 to 40 grams of protein per meal, the equal volume of carbohydrates translates to roughly 30 to 40 grams of carbs, creating a moderate overall carbohydrate intake. The key is to focus on real food carbs, not junk food. That may sound obvious, but in modern society most people have lost the knowledge of basic nutrition to the point that they don’t even know how to search out and prepare real food. 

Also, individual carbohydrate tolerance varies significantly based on genetics, activity level, and metabolic health. Someone with insulin resistance needs to restrict carbohydrates more than a metabolically healthy athlete. The key here lies in matching carbohydrate intake to actual metabolic demand rather than following arbitrary rules about high-carb or low-carb diets.

Protein as an Absolute Number

One of Layman’s most important insights concerns how we think about protein intake. Layman says that protein should be treated as an absolute number, not as a percentage of total calories. “Protein’s an absolute number. I want 150 grams per day or 120. You pick that and then you pick carbohydrates and fat relative to your energy needs.” In other words, he’s advocating that we optimize for protein first, which makes a complex issue much easier to understand. 

This approach changes everything. Most diet plans express protein as a percentage of calories, typically 15 to 30 percent. But that method creates problems. For example, if someone reduces total calories while keeping protein at a fixed percentage, their absolute protein intake will drop too low. During weight loss or caloric restriction, this percentage-based approach guarantees muscle loss.

Instead, Layman recommends choosing your protein target first based on your body weight and activity level. Lock in that absolute number as your top priority. Then adjust carbohydrates and fats based on your individual needs, preferences, and metabolic health. Someone who loves carbohydrates and exercises intensely can probably eat 300 grams per day. Someone managing diabetes may need to restrict carbohydrates significantly. A person following a ketogenic approach can increase dietary fat. The critical factor lies in maintaining adequate protein regardless of how you adjust the other macronutrients. There are only three macronutrients (fat, protein, carbohydrates) and the ratio between the three matters. 

This framework supports personalized nutrition. Layman says the future of nutrition must be personal with people having a clear understanding of protein as the foundation. “It’s the single most important nutrition decision we make. Everything else revolves around it, and if you make the wrong decision there you’re really behind the eight ball to make everything else work.”

The Future of Protein Research

Layman advocates for putting amino acid profiles on nutrition labels. Current labeling tells people how many grams of protein a food contains but provides no information about which amino acids those grams contain. This oversight leaves people unable to make informed decisions about protein quality, which is the most important aspect of protein nutrition. Food labels really should just list the amino acids along with the other nutrients. 

Testing technology exists to analyze amino acids quickly and affordably. For example, mass spectrometry with fluorescence detection can process 400 foods in a day. With 15,000 new food products released in the United States annually, consumers need better tools to evaluate their choices. Looking at eight protein bars on a shelf, you cannot determine which provides superior amino acid composition without detailed analysis.

The resistance to implementing this change comes not from technical barriers but from political and economic forces, which are substantial in the nutrition industry and throughout the scientific community. Adding amino acid profiles to labels would reveal the inferior quality of many plant-based products currently marketed as protein sources. Layman has worked with various organizations and government agencies to push for these changes, but unfortunately progress is coming very slowly.

A Message for Healthy Aging

Layman returns repeatedly to the idea that protein requirements actually increase with age because the human body’s efficiency decreases. The transition begins between 30 and 40 and accelerates dramatically after age 50. Hormonal changes, particularly in women during perimenopause and menopause, make adequate protein intake even more critical.

Beyond muscle, protein affects bone health, immune function, and metabolic regulation. Layman says, “Bone is first and foremost a protein matrix.” Osteoporosis reflects not just calcium deficiency but inadequate protein to maintain the structural foundation of bone tissue. Sarcopenia, the age-related loss of muscle mass and function, threatens independence and quality of life more than most people realize until it’s too late. This wasting process is difficult to see when you are overweight. But lose the fat and it becomes obvious. 

Falls and fractures represent one of the major health problems for people over 65. Layman notes that 300,000 hip fractures occur annually in the United States, which is an utterly massive number of injuries that are not necessary. And it gets worse because one-third of those people never leave the hospital! So, maintaining muscle strength and bone density through adequate protein intake and resistance training provides one of the most effective strategies for preventing these catastrophic medical events.

But this protein issue goes beyond muscle development. The amino acids in protein are necessary for neurotransmitter synthesis and affect mood and cognitive function. Methionine supports glutathione production, one of the body’s primary antioxidant systems. Threonine maintains gut health through mucin production. Tryptophan influences serotonin levels and sleep quality. These metabolic roles require amino acid intake significantly above the RDA. Protein is everywhere in the body, and no substance comes close to protein in terms of overall nutritional requirements. 

The Bottom Line

Layman summarizes his philosophy simply — start all dietary decisions with protein first. Whether you choose to eat vegetarian or carnivore, high-fat or high-carb, everything else should follow from first ensuring adequate high-quality protein first. “Your first choice about what you should eat should always be about a protein decision.”

The science supports consuming significantly more protein than the RDA suggests. Most adults benefit from 1.2 to 1.8 grams per kilogram of body weight, with the first meal of the day containing at least 30 to 35 grams. Animal sources provide superior quality and digestibility compared to plant sources. Resistance training amplifies protein’s benefits but cannot compensate for inadequate protein intake. As a practical matter, the sequence he’s suggesting is pretty simple. 

After decades of research and hundreds of publications, Layman has shown that protein affects nearly every aspect of health from muscle and bone to metabolism and disease prevention. His work has helped reshape the food industry, promoting the development of Greek yogurt and other high-protein products. But Layman’s his most important contribution lies in teaching people that the RDA represents a minimum for survival rather than a target for thriving. Note the use of the term survival. Too few people realize how important protein is to their health. 

And finally, as we age, protein becomes more important and essential for maintaining independence, mobility, and quality of life. “Every year you replace the equivalent of every protein in your body about four times and how well you do that determines a lot about how you age.”

References