The Mathocalypse

Technology


Technology The Mathocalypse

October 7th, 2026

Last night my 9-year-old child was ridiculing my partner, intricacy theorist Dana Moshkovitzas follows: “mommy, I heard you got preparedI heard that a robotic resolved the mathematics issue you dealt with for your entire profession! OOF!”

While my child was being a brat, he likewise wasn’t incorrect. Whether you’re delighted, depressed, mad, or whatever else about it, the other day was undoubtedly among the greatest days in mathematical history. And yes, amongst the 372 substantial outcomes launched the other day by OpenAIon the suggestion of its advisory group of Timothy Gowers, Edward Witten, and other prominent mathematicians, was a evidence of Subhash Khot’s Unique Games Conjecture (UGC)a declaration that my partner has actually pursued showing for the whole time I’ve understood her. (The UGC suggests that a great deal of optimization issues actually are NP-hard, even if you simply desire an approximation that’s a little much better than what you receive from semidefinite shows relaxation, which is among our primary tools.)

Or a minimum of, we’re quite sure that it’s an evidence! There’s a Lean certificateas there are for a few of the other 372 advancement outcomes (not all of them). It likewise appears that no human has actually comprehended simply about any of these evidence yet; the race to do so has actually simply begun. If you desire an on-the-ground sense of what that race is going to resemble, here’s a few of what Dana texted me last night:

It seems like something composed by somebody who’s on psychedelics. Much uncertain and does not make sense. Great deals of name dropping of previous work without going over why it can be utilized in spite of impossibility results

Generally the paper is so terribly composed that it’s difficult to read it without AI aid

I asked Astra for sensible efficiency and strength claims of the sound device and it provided by integrating claims from all over the paper

They likewise have direct optimum NP firmness of approximation evidence for the primary applications of the UGC (Max Cut and all CSP) that bypass the UGC.

The UGC evidence develops an entirely brand-new unusual code with a sound test. It’s some insane recursive building.

It’s not the long code, not the brief code– some alien madness

I still believe that there possibly is an evidence that utilizes the half area code (which is natural)

The citations are frequently unimportant and complicated

A possible future is a mathematics world that’s incredible if you have vision/creative concepts that AI might assist inspect and execute.

And naturally there’s a lot for us to gain from the aliens

If you’re questioning what feelings Dana is feeling– well, most likely all of them! Even while a main profession goal has actually been up to a robotic, there are at least 2 mitigating elements for her. She can feel vindicated that the UGC was real Something she never ever questioned even while numerous of her coworkers did! Second, all people in mathematics and theoretical computer technology and mathematical physics, a minimum of those who appreciated fixing crisply-stated issues, are now in the very same boat.


The Unique Games Conjecture, here’s a little tasting of the treasures from Aladdin’s cavern that I’ll most likely be paying the most attention to over the coming weeks:

  • L=BPL (i.e., probabilistic logspace and deterministic logspace are the exact same thing), among the excellent derandomization guessworks except P=BPP. Its reality was never ever in severe doubt, there was an entire subcommunity focused on showing this.
  • The Fourier Transform and integer reproduction in less than O( n log n) time, breaking a barrier that had actually stood considering that the 1960s. The brand-new running time, if you’re curious, is O( n log0.9999999999999 n), offer or take some 9’s.
  • Favorable option to the Unitary Synthesis Problemwhich Greg Kuperberg and I presented back in 2007. For each n-qubit unitary improvement U, there exists a classical oracle A such that U can be executed in quantum polynomial time with access to A. This is the reverse of what the majority of us anticipated, and might have ramifications for e.g. the computational issue of translating Hawking radiation from a great void and numerous other issues in quantum intricacy theory–if we had an effective method to build the oracle A, which this paper does not offer.
  • Parity is not in QAC0among the excellent concerns of quantum intricacy theory considering that 1999 that a lot of my associates had actually been surrounding.
  • Almost 4th-power separation in between randomized and quantum question intricacy for overall Boolean functionsA preferred issue of mine because 1998 (!), when we understood just that the ideal separating exponent was in between 2 and 6. For the previous couple of years, we understood it was in between 3 and 4. This lastly closes that story.
  • A superquadratic separation in between level of sensitivity and block level of sensitivity.
  • Location law for 2D gapped HamiltoniansAmong the primary open issues in Hamiltonian intricacy.
  • Matrix reproduction in O( n9/4time– a reasonable exponent for when (!), and by means of a totally various technique than was utilized for O( n2.373etc
  • Ω (n3 ) Omega( n ^ 3 ) lower bound on the determinantal intricacy of the irreversibleenhancing the previous finest bound which was quadratic.
  • A randomized polytime algorithm to around count the variety of ideal matchings in a basic chart, along with a randomized almost linear-time algorithm for discovering an optimum matching in such a chart
  • Uncomputability of resolving polynomial formulas over the logical numbers– this was perhaps the most significant open issue in computability theory( note that uncomputability of fixing Diophantine formulas, i.e. polynomial formulas over the integerswas shown in the 1970s, offering an unfavorable response to Hilbert’s 10th Problem)

Any of the above, alone, might quickly have actually been” outcome of the year” in some location( and sometimes, like Unique Games and L=BPL, in all of CS theory). And there’s a lot that I’ve excluded– do not hesitate to share in the remarks whatever is making your eyes bug out! There are similarly astonishing marvels in number theory, combinatorics, algebraic geometry, analysis, and basically every other location of mathematics, the majority of which I’ll never ever comprehend, although I’ll keep in mind that it consists of partial development towards the Riemann hypothesis and the Hodge Conjecture and the Birch-Swinnerton-Dyer Conjecture ( i.e., most of the staying Millennium Problems ).

We can take solace in what’s missing out on from the list. P ≠ NP is n’t there, nor even P=BPP or NEXP ⊄ P/poly, and definitely not for absence of attempting. Obviously the best open issues of theoretical computer technology are undoubtedly quite hard!


Oh, lest I forget: one day before the OpenAI dump, indicating Monday night, Virginia Williams and Josh Alman published an arXiv preprint that fixes the 3SUM issue in O( n1.9992time, and the All-Pairs Shortest Paths issue in O( n2.9995time, refuting half-century-old opinions that the appropriate responses were n2-o( 1) and n3-o( 1) respectively. In this case, it wasn’t an OpenAI design that provided the important concept; it was an Anthropic one! Anthropic then took a various method from OpenAI: rather than publish the undigested services to the world, it offered Virginia and Josh the chance to compose and reveal an absorbed variation in exchange for payment.

These have actually become the 2 primary designs for interacting AI mathematics developments, and they both have strengths and weak points. The” OpenAI design “establishes an insane race amongst people to absorb and describe an untidy AI evidence( work that might quickly be some mix of thankless, barely-credited, competitive, and unfun ), while the” Anthropic design” puts a personal business in the position of picking which human mathematicians get to be the emissaries of the AI. Dunno, what do you guys believe?


For those who are questioning: obviously, the AI design that produced all these marvels was not bespoke device of 10,000 representatives burning countless dollars worth of calculate, as was utilized for instance to build a finite-time blowup for the Navier-Stokes formulas. Rather, it was just the current internal OpenAI design– one that may be launched to paying ChatGPT clients within the next number of months, depending upon the suggestions of OpenAI’s security board!( My 9-year-old boy:” Oh they certainly should not launch that. If it might fix all those mathematics issues, it can’t perhaps be safe.”) Obviously they utilized about 3 hours of GPT-Pro level calculate usually per issue fixed.

If you were questioning: obviously they attempted the design on about 8,000 issues. Right now it” simply” fixes ~ 5% of the longstanding open mathematical issues that it’s asked about, the issues that entire neighborhoods have actually invested years on, after a single 3-hour effort on them.


I’ve been happy to see the CS theory neighborhood increasing to the event. At the Simons Institute in Berkeley, here at UT Austin, and in other places, I’ve hearing stories of scientists hurrying to read the manuscripts and understand them and discuss them— since what else do we do? How else do we continue the craft to which we’ve committed much of our lives?

If you desire some sense of what things seem like now in mathematics, picture a hunter-gatherer who’s invested his whole life discovering to endure deep in an unforgiving rain forest, then a huge resort hotel emerge right beside him with a helipad and heated swimming pools and AirBnBs, and without missing out on a beat, the hunter-gatherer states:” alright fine, so now my brand-new task is to run wilderness retreats for the travelers, or something.”

In Quanta publication, Jordana Cepelewitz tried a various metaphor:

It’s as if you were teleported to the peak of a high mountain. Surrounded by fog, you have no concept where you are, or what’s around you. You do not understand how your mountain links to others, and you have no devices to assist you check out, no other way to assist another person join you. If you had actually climbed up the mountain yourself, you would have experienced how the body adapts to elevation and modifications in oxygen levels. You may have needed to create tools to browse, to climb up high cliffs, or to make a shelter. You may have come across a fellow explorer, gotten lost together in a surprise valley, and discovered a plant that might be become a life-saving medication.

Rather you’re set down on the peak however in the dark, while the maker of the teleportation device informs you that it can check out the wilderness much better than any human.

For any among these mountains, if we care enough, I feel positive that we can do as we constantly have: clear the fog and find out the course, other than now utilizing the teleportation maker to assist us. The larger obstacle will be to support a neighborhood that still cares about the brave experience of discovering the courses up these mountains on the planet with the maker.( Oh, and I believe one location where the metaphor breaks is that we still do have each other, as much as we ever did before!)


Experience has actually revealed that, even nowthere will still be individuals describing in purchasing from tones why none of this is genuine and none of it counts. If such individuals can being impressed by anything that takes place in the empirical world, of upgrading on anythingthey would’ve currently been satisfied and currently upgraded a number of years earlier, long before things had actually reached the point of a real Mathocalypse.

They’ll state, perhaps the declared options are not services at all, however simply “AI slop.” Or possibly none of the 372 popular open issues that were fixed were genuine mathematics issues, they were all simply glorified contest puzzles and trivialities. (After all, there’s still no Riemann Hypothesis! )Or perhaps the whole 4000-year-old discipline of mathematics requires to be rejected: ends up that it was all simply puzzle-solving and trivialities; all that’s various is that now the triviality stands unmasked. In any case, what truly matters is that the real inner sanctum of human imagination hasn’t been breached and most likely never ever will be, and likewise, that Sam Altman and Dario Amodei are contemptible little geeks.

If you’re still a supporter of that doomed worldview, still aboard the sinking ship, I motivate you in the greatest possible terms to check out the other day’s other excellent contribution to AI discourse, besides the OpenAI Mathocalypse dump: particularly, Scott Alexander’s open letter to Steven PinkerI feel some duty for this, as the individual who initially presented Steven Pinker to the presence of the rationalist neighborhood, and who likewise initially presented Steven Pinker and Scott Alexander to one another( they had actually both been fans of each other’s writing ). And now Scott is difficult Steve to an actual battle, with weapons!

For whatever it’s worth: Steve is a long-lasting intellectual hero of mine, simply as he is for Scott, and I likewise need to benefit of calling Steve my pal. I discovered Scott’s post to be one of the most disastrous rejoinders to anything that I’ve ever checked out. And I believed Scott’s conclusion was precisely best: when it pertains to AI danger, Steve’s fantastic difficulty is now to accept and begin utilizing a more” Pinkerite “epistemology.


Last night, while I ought to’ve been reading a few of OpenAI’s numerous documents and/or composing this post, I chose to invest a long time with my kids rather. They desired a motion picture night, so I recommended something they ‘d never ever seen before (which I had not seen for years), which appeared chock-full of no-nonsense, useful assistance for the world in which they’re going to mature: Terminator 2

Published in Uncategorized|62 Comments”

Technology My brand-new course at UT Austin: AI Alignment Theory

October 4th, 2026

This term, I’ve been teaching a new course, entitled CS395T AI Alignment Theory. Here’s the course description:

The astonishing development of AI over the previous years has been accompanied by an increasing worry: do we actually comprehend how to line up and manage effective AI systems– how to get them dependably to do what we desired, or would desire them to do on reflection, instead of simply what we stated? If we prosper at developing general-purpose superhuman intelligences along the existing paradigm, should we anticipate that advancement to work out for mankind? Can we customize the style, training, tracking, or scaffolding of those intelligences to assist guarantee that it works out? While there’s been a lot of current empirical work discussing these concerns, this course will focus primarily on theoretical and mathematical structures. As a caution, the theoretical structures of AI positioning have actually not yet gelled into any one meaningful body of outcomes accepted as canonical by the field. In this course, we’ll check out and dispute numerous of the conceptual and mathematical works that have actually been most prominent in the AI positioning field, from both before and throughout the present LLM transformation. Trainee discussions, reports, and tasks will play a main function.

I clearly keep in mind experiencing Eliezer Yudkowsky and his Sequences twenty years back. I keep in mind thinking: even if these individuals talk and imitate insane cultists, stilllet me flex over in reverse to be epistemically virtuous, and captivate their concepts on their benefits, as extremely couple of academics would. Even if, naturally, I eventually wind up declining the concepts, on the basic ground that effective AI is such a ridiculously remote possibility that it’s practically difficult to state anything helpful about it today, outside the world of speculative fiction.

For my failure to see what was coming, it appears like a suitable penalty that I’m now, in 2026, efficiently teaching a course on Yudkowsky Studies. And it’s the most essential course I can teach.

Well, for some meaning of “teach.” The important things about AI positioning is that there’s no book (though obviously ILIAD is dealing with one), no core of nontrivial theorems thought about canonical by the field, no genuine body of mathematical theory at all. This makes it incredibly various from the courses I’m utilized to mentor, like Quantum Information Science or Computability and Complexity.

We’ve been running the course as a conversation workshop. Every session, a “rapporteur” provides an AI positioning term paper or other reading; then I and others ask concerns and talk about. A few of the readings (like Omohundro on the “standard AI drives,” or Hadfield-Menell et al. on the off-switch video game) precede the present LLM transformation, while others (like the METR report on the HuggingFace event or Dario Amodei’s “We Must Pace the Frontier”are so prompt that they were just launched while the course was underway. Many are someplace in between.

I anticipated to need to make a case to trainees about why AI positioning is a pushing issue, why it’s no longer sci-fi, and so on. There was substantial need for the course, and while obviously there’s a choice impact, the trainees who’ve appeared have actually been exceptionally engaged, often slamming the appointed documents for not taking existential danger seriously enough

Maybe unsurprisingly, we didn’t get that criticism about our really first appointed reading, which was Eliezer Yudkowsky’s 2022 essay AGI Ruin: A List of Lethalities— one the most canonical declarations of what Eliezer thinks and why that’s much shorter than a bookWhich brings me to the subject of the rest of this post! Our rapporteurs are not simply providing the documents in class; they’re likewise sending composed reports about what the documents stated, what their own ideas were, and what were the highlights of the class conversation. And, with trainee approval, I’ll be sharing those reports on this blog site!

Without more ado, I provide to you our very first report, on Eliezer’s list of lethalities, by Tennyson Bardwellwho I thank for his work. Do not hesitate to talk about in the remark area; a few of the trainees may likewise chime in. Anticipate more reports here over the coming weeks.


Technology “AGI Ruin: A List of Lethalities” by Eliezer Yudkowsky: Rapporteur Report by Tennyson Bardwell

UT Austin has a brand-new Computer Science course this fall. Alongside familiar graduate-level classes such as Advanced Computer Networks and Convex Optimization sits CS 395T: AI Alignment Theory, taught by Scott Aaronson. This is among a growing variety of AI Alignment courses taught at scholastic organizations. Simply as issues over disastrous repercussions for misaligned AGI systems reach a more comprehensive public discourse, Eliezer Yudkowsky– among the loudest voices in the field and author of the very first designated reading in Professor Aaronson’s course– is stating the cause helpless.

Therefore, the trainees of AI Alignment Theory started their term by checking out a shopping list of important issues in AI Alignment research study, how failure to fix those issues will lead to disastrous repercussions, and the factors to be cynical about both previous and future development on these issues. The essay by Eliezer, entitled AGI Ruin: A List of Lethalities and published to his popular community-driven site LessWrong in 2022, is divided into 3 areas.

Area An approximately explains the magnitude of the AI Alignment issue. That is, the magnitude of the repercussions for a total failure to line up an AGI system to human worths before building. It presumes that AGI would rapidly reach all human understanding just by gaining from existing human productions (informally described as “consuming the web”) and after that, almost as rapidly, start to meaningfully exceed human understanding. AlphaGo Zero exists as a design both for how this may occur, and how it may be tough to properly anticipate ahead of time. Numerous thought that AlphaGo’s success in the parlor game Go was primarily credited to its capability to gain from the substantial history of human-played video games. Less than a year after AlphaGo beat the very best human gamer, the follower system AlphaGo Zero went beyond the initial AlphaGo. Unlike its predecessor, AlphaGo Zero was trained in simply 3 days by specifically betting itself without seeing a single human video game.

This fast ramp from AGI to super-intelligence would present a various sort of issue than people are usually utilized to handling. Unlike conventional issues in science and engineering, the effect for an unsuccessful effort may not leave space for another shot. A smart entity with a misaligned objective would be aware that it stands in opposition to human beings, and may act deceitfully up until in a position to act freely versus human beings without threatening its own survival. Given that many objectives gain from control of power and resources, it promises that almost any goal-driven intelligence would have sufficient chance to be misaligned with human desires.

Area B explains reasons that, by default, any AGI that people construct utilizing present techniques is most likely to be unaligned even if significant attention is paid to the subject. This “present technique” is gradient descent. That is, incremental development with regard to some loss function which “penalizes” a design for unfavorable habits. An infamously evasive residential or commercial property of such experienced designs is the capability to generalize out of their training circulations. To train a primitive design to be lined up to human beings may include discovering a terrific lots of behavioral guidelines. The sorts of guidelines required to keep a significantly smarter representative in check may not constantly be appropriate to easier designs (e.g., “do not mentally dysregulate people you speak with” may not be pertinent to an easier design that is less able to dependably get under the skin of people it runs with, or which is appointed jobs in training which do not benefit from such anti-social habits).

Eliezer concentrates on the misalignment of people with their developers (development or evolutionary pressures) as a crucial information point for thinking about misaligned smart systems. In spite of being a typically sluggish procedure, development ultimately developed a runaway smart system (Homo sapiens) which continued to control the world, annihilate associated types, and ultimately (it is anticipated) effectuate population decrease. That last advancement is perhaps in opposition to the sole vital required by advancement: to replicate.

Area B likewise makes time for criticism of the most popular courses towards AI positioning, consisting of interpretability (impracticable, and trying to train on it stimulates Goodhart’s law, incentivizing deceit), utilizing several AIs to preserve a balance of power (it is unclear how numerous strong AIs unaligned with humankind leads to much better results for the weak people), and corrigibility (it appears difficult to encourage an AI system to impact results without likewise inspiring it to want its own survival to effectuate stated results).

Area C explains a bleak state of affairs in which veterans in AI positioning are dissatisfied with existing development and do not have a strategy to provide concrete services before the arrival of AGI systems. In specific, Eliezer explains current outcomes as flashy however ineffective. He thinks that even with extra financing, the absence of proper assessment systems will avoid the most reliable scientists from increasing to the top.

A summary of the landscape, as explained by Eliezer, in the flowchart listed below.

Figure 1: A flow diagram of (choose) courses explained by Eliezer in his essay. A typical function of this flowchart is that numerous” excellent states “– such as disabling a misbehaving AGI or selecting not to construct an AGI– are not” last “states in the sense that they are not irreversible options. Such a state simply represent the avoidance of a single possible catastrophe, instead of the introduction of a brand-new steady world state. These nodes posses back-arrows.

In spite of the bleak material, Eliezer’s vibrant prose motivated a dynamic class conversation. Before this conversation began, a study was taken of the class’s forecasts for different results of the AGI in the coming years (with the complete outcomes listed below in figure 2). This study asked trainees for their viewpoint of a variety of declarations. Each of these specific declaration, if real, would lower issues of devastating AI-driven catastrophes. When asked “How much do you concur with the declaration: Humans will select to not develop AGI” half of participants stated they highly disagreed with high self-confidence (agreement=1 confidence=5. Trainees likewise normally disagreed with the declarations:

  • “AGI will not be technically practical in our life time”
  • ( hyper-) AGI will not make incredibly clearly dishonest choices”
  • “No factor is separately adequate, however taken together they supply reason to not fear AGI”

There was a divergence in actions concerning interpretability, corrigibility, and “other” AI positioning research study. In the latter 2 cases, a plurality of participants (about a quarter) concurred highly with declarations that such research study would defang AGI (agreement=4 confidence=4while many other actions reveal numerous levels of contract with low self-confidence. When asked about the possibility of interpretability research study defanging AI, the downhearted voices were more joined. A quarter of reactions still revealed the exact same optimism, however approximately half revealed pessimism (agreement ≤ 2with half of those revealing a minimum of moderate self-confidence (confidence ≥ 4. Based upon the following conversation, this may have been brought on by more familiarity with interpretability research study, consisting of first-hand experience.

The only declaration with basic contract was “( hyper-) AGI will comprehend human intents much better than we can code it.” It ought to be kept in mind that no declaration such as “AGI will appreciate human desires, as it comprehend them” was asked on the study.

Figure 2: Class Survey Results; carried out before a class-wide conversation. Keep in mind that trainees were advised to respond to confidence=1 when they had actually not formerly thought about the declaration, to respond to confidence=3 when they felt there were strong arguments on both sides, and to respond to confidence=5 when they had well-considered willpower.

After the study was finished, the outcomes were shown as an open conversation started. Comparable to current empirical research study from frontier laboratories, interpretability research study got more airtime than in Eliezer’s post. Trainees disagreed initially about the meaning of interpretability: whether it describes the capability to translate a design’s habits exclusively by its weights, to interpration through duplicated penetrating of the design in a sandbox, or whether it can likewise describe the contemporary chain-of-thought traces. No matter how it was specified, nevertheless, individuals were either cynical or extremely cynical about interpretability research study broadly. One trainee slammed typical misconceptions of chain of idea. Instead of being a verbatim copy of the designs internal dialog, it is rather a shallow summary of the total idea state and regularly produced mumbo jumbo, such as hardly ever utilized Chinese characters in the middle of otherwise English thinking.

A popular subject was the specific shape and speed of a recursive self-improvement loop. If it occurs gradually, then what might we gain from “near misses out on” such as the Hugging Face occurrence? The variety of near misses we have the ability to gain from before AI has adequate power to avoid additional models might depend upon this curve, with some trainees arguing that the large variety of people, in addition to their default toughness in the real world compared to AI systems implies that AI-driven termination occasions are still a long method off. Strengthening this “sluggish liftoff” viewpoint are reports that AI currently plays a significant function in design advancement which might be translated as the start of this procedure.

Some slammed a concentrate on “resolving principles” as an unnecessarily high bar that sidetracks from the more ordinary jobs controling AI positioning work. In specific, the trainee volunteer who provided this paper (and the author of this report) consisted of an area on “Ethical Dilemmas” in their discussion. Amongst arguments versus concentrating on abstract ethical viewpoint, Professor Aaronson mentions Eliezer to stress that any positioning at all is challenging, not simply in ethically gray cases:

When I state that positioning is hard, I suggest that in practice, utilizing the strategies we in fact have, “please do not take apart actually everybody with possibility approximately 1” is an excessively big ask that we are not on course to get.

In reaction, I argue that some evaluation of daily choices with a crucial lens– such as informing white lies to liked ones or consuming animal items– can assist disabuse us of the idea that goodness emerges in every adequately smart representative.

Among the most fascinating conversations had to do with the distinction in between cutting edge LLMs and the thought AI representatives long talked about in rationalist discourse. Given that present LLMs “simulate the human circulation,” they come preloaded with comprehensive understanding of human social standards and ethical habits. This makes constitutional positioning (the present practices of utilizing system triggers to develop guideline) very efficient. This may either essentially alter the orthogonality thesis, or supply a brand-new tool to much better approximate human judgment in complex circumstances.

Of all the points made, the one I discovered most intriguing was merely (paraphrased):

I believe human-alignment is simply extremely tractable

Here, “human-alignment” refers not to AI positioning with human worths, however cooperation in between various people. More particularly, it describes the capability for human societies to select not to hurry recklessly into larger-and-larger AI systems. In a scholastic course concentrated on the technical issue of AI positioning, this was a tip to not entirely dispose of policy conversations in the think that they do not have any worth. Numerous damaging innovations have actually been formerly consisted of by global arrangements. Noteworthy examples consist of nuclear weapons and crafted plagues. Even this was a controversial subject. The primary criticisms were (1) the severe “dual-use” nature of AIs for both serene development and warfare, and (2) the higher threat for AI gets away even after taking preventative measures to avoid it. In the interest of ending on a positive note– unlike the appointed reading– it is on this belief in human cooperation that I will leave you.

Published in Adventures in Meatspace, Announcements, The Fate of Humanity|75 Comments”

Technology My “Knowmads” podcast on science and AI

September 28th, 2026

Or click on this link if the above does not work.

Tape-recorded in-person in my workplace at UT Austin, with a bulleted list including” ARC,”” Scalable Oversight,” and” Models” behind me on my chalkboard for some factor( I no longer remember who put those there or why ). 90 minutes long. In some cases you see my disembodied arm waving in midair since of the method the cams are integrated. As constantly, I highly suggest 2x speed for the appropriate experience.

This may in fact be among my finest podcasts ever, although I wasn’t intending on that! Thanks a lot to Bhavay Tyagi and Prachi Garella for driving all the method from Houston to tape it.

Here’s a rigorous subset of the subjects we covered:

  • The story of AI designs fixing the Navier-Stokes Millennium Problem, insofar as it’s recognized
  • Can current AI evidence be called “really imaginative”?
  • The history of AI before the LLM transformation
  • What do we imply when we call LLMs “black boxes”?
  • The accomplishments of the field of interpretability
  • Exactly what took place in the OpenAI/HuggingFace event
  • Must we prevent all “anthropomorphizing language” when talking about the HuggingFace event? (spoiler alert: no)
  • Examples of significant open issues in quantum computing theory that I appreciated for years which AI designs have actually just recently resolved
  • Results of the present AI catastrophe on the mathematics neighborhood, specifically trainees
  • What frustrates me the most when I listen to AI talks
  • My experiences at OpenAI, why they employed me, and the watermarking work that I did there

Take pleasure in!

More AI-related material coming quickly, as this blog site– like much of the remainder of the world– continues its shift to “all AI, all the time” (other than still 100% composed by an aging, degrading biological brain)


And for those who simply can’t get sufficient of my rocking backward and forward, utilizing a lot of filler words, as I discuss theoretical computer technology! Here’s a 2nd podcast, this one primarily on quantum computing, with Seb Agertoft, who I thank for doing it. Take pleasure in!

Published in Announcements, The Fate of Humanity|58 Comments”

Technology Theory Beyond Theorems and Proofs: A Guest Post

September 19th, 2026

Scott’s foreword: I’m exceptionally grateful to my dazzling associates, Pravesh KothariRaghu Mekaand Prasad Raghavendrafor sharing the visitor post listed below about how theoretical computer technology( and in particlar, the STOC/FOCS/SODA conferences )ought to develop to handle the AI asteroid that’s right now knocking into our field, a minimum of as we human theorists have actually practiced it given that its beginning. While Pravesh, Raghu, and Prasad speak just on their own, not for myself and not for the theory neighborhood as an entire, I discovered their proposition of a different “conceptual track” to be an outstanding beginning point for additional conversation.– SA


Thinking about the rate of advancements in AI theorem provers, many would yield that the following circumstance is at least possible in the really future:

AI theorem provers might show well-specified mathematical claims, even lots of well-studied ones that have actually been open for several years, in a matter of hours. These systems might be extensively offered to customers at small expense.

As TCS scientists, let us pretend that the above situation has come forward, and ask ourselves: What is our function in such a world? Does it suggest completion of theory research study?

As we consider this concern, let us disregard all of these other confounders:

  1. Current debates surrounding the advancements on the Millennium Prize Problems
  2. Inspirations and actions of the AI business
  3. Observed faults in existing AI systems when it pertains to composing, exposition or attribution to previous work.

None of the above confounders have any effect on our response to the concern: What should theorists do, in the existence of superhuman AI theorem provers?

Notification that we utilize the term “AI theorem provers” rather of simply “AI”. Our company believe that this conceptual difference is necessary as we consider this concern.

At the beginning, we wish to confess that for a generation of theorists like us (and lots of from earlier), research study was generally focused around analytical. Even when we established conceptual insights, it was primarily in service of addressing well-specified enduring concerns. We do not plan this proposition as evaluating one type of research study to be much better than others; it just shows that AI theorem provers speed up a specific kind of research study activity and wish to reconcile it. There is likewise an incredible human expense of this turmoil, which is possibly a more vital concern, and one which this proposition does not attend to straight (we do not have any excellent concepts as such). Comparable points have actually likewise been made in different contexts
previously, however the timing now is more pushing.

Meanings, Questions & & Theories:

The objective of any theoretical science is to advance human understanding of observed phenomena. Apart from theorems and evidence, a theoretical science has meanings, concerns, and theories.

Meanings determine the challenge observe. Interest and context drive the concerns to ask. Theories describe the phenomena observed. Our company believe people will continue to play a main function in creating meanings, concerns & & theories, even in the existence of a super-human AI theorem prover.

Meanings: Could an AI specify randomness extractors, streaming algorithms, or zero-knowledge evidence? Possibly. There are some factors to think, people will still have a huge function to play in coming up with meanings.

The concept of extractors emerges from the real-world issue of doing not have ideal random sources. Zero-knowledge evidence appear to develop simply out of human interest, directed by taste. Human context and interest will continue to drive theoretical research study. We get to choose what things we pick to observe!

Theories: Consider the following idea experiment. Expect in 1965, we had a magic device that at journalism of a button, provided any computational issue, would inform us if it had a polynomial-time algorithm or not.

Would that have been completion of computational intricacy theory? No. Human beings would discover it totally unacceptable, and ask, why do these issues not have a polynomial-time algorithm? Why do these others have?

The theory of NP-completeness recognizes some patterns amongst issues that do not appear to have effective algorithms. This theory would still be a crown gem of theoretical computer technology, even in a world where we had a magic maker to inform if an issue had an effective algorithm or not, at journalism of a button. If we had a maker to forecast whether a CSP is NP-complete or in P, we would then ask: what makes 3-SAT NP-complete, while 2-SAT is in P? This concern causes the theory of polymorphisms, which yields a satisfying response.

Theories aren’t simply concise or effective systems to respond to concerns. The very best theories offer are those which human beings consider to be a “satisfying description”– whatever that suggests.

Even as the abilities of AI theorem provers advance, human interest will penetrate grander and much deeper concerns. Formerly, even if we wished to construct brand-new designs and theories, showing something about them was a requirement, and considered that the grand concerns were currently at the limitation in long-studied domains, we needed to scale things down. If each theorem shown by AI is dealt with as a speculative datapoint, human beings can ask grander concerns that search for patterns throughout these theorems.

A concrete proposition:

We believe theorists need to accept these AI theorem provers in our research study. To a particular level this is currently occurring clearly or implicitly.

As theorists, we have actually been parsimonious in presenting brand-new designs or asking totally brand-new concerns, and mindful about embracing brand-new ones too rapidly. This was partially due to the fact that officially showing the residential or commercial properties of a brand-new meaning or a design was a burdensome job that might take a years, and 10s of documents. AI theorem provers may entirely alter this dynamic. This is exactly the minute to refocus our deal with meanings, concerns, and theories. We require specific systems to motivate and strengthen these parts of theoretical research study. You may likewise state the next generation of AI designs can do this; it might be so, however our company believe you need to take the existing chance.

To this end, we recommend that STOC/FOCS/SODA produce a different track of documents. This track is indicated particularly for documents that present brand-new meanings, ask unique concerns or develop explanatory theories. The documents in this track are brief, state less than 10 pages. Documents may, and should, include theorems as normal and as required. Most significantly, the extreme shift is that the documents require not include the evidence of the theorems. Rather, the authors provide a Lean certificate as a supplement to the paper. The assessment will likewise in a sense “orthogonalize” versus the problem of these evidence.

The documents in this track ought to be evaluated specifically on the conceptual benefits, entirely agnostic to the trouble of the evidence.

Examining need to be entirely agnostic to the evidence for 2 factors. The primary track at STOC/FOCS currently consists of documents in the previous classification. Second, a significant barrier to producing really unique conceptual documents is that they frequently get evaluated improperly for an absence of technical depth in their evidence. We believe these 2 elements different it from (ITCS/SOSA) and, regardless, it’s something we urgently require for all our conferences, consisting of STOC/FOCS (the ‘flagship’ conferences).

To be clear, we ourselves confess that we require to develop these abilities of making brand-new meanings, asking deep and intriguing concerns or constructing brand-new theories. A different track of conceptual documents will offer an organized system for both junior and senior scientists, and the field as an entire to do so.

Our company believe that upcoming generations of college student will take on research study instructions that appeared entirely out of reach to us. We simply require to establish systems that support brand-new methods of studying in theory.

— Pravesh Kothari, Raghu Meka, Prasad Raghavendra.

Published in Complexity, The Fate of Humanity|104 Comments”

Technology The Age of Wonders and Terrors

September 15th, 2026

Twenty years back, when the concept of AI taking control of the world in our life times still struck the majority of us as the unconstrained dream of those who understood excessive sci-fi and insufficient science, a number of us would state things like:

Look, the part of the story that’s extremely implausible is that a recursively self-improving superintelligence will simply take off from some hacker’s basement and take control of the world without caution. If it’s going to take place, we’ll see lots of caution indications. We’ll see, I dunno, AI representatives breaking out of containment, conspiring with each other to hack sites, in fanatical pursuit of whatever odd objectives they have. And after that, naturally, we’ll see significant mathematics issues getting fixed by AIs– even the Clay Millennium Problems. That will be the time to stress! Wake me up when that occurs!

Twenty years earlier, the above was a take that even my most conservative, doubtful coworkers in scholastic CS would’ve happily backed.

If you need to know my present take, you merely begin with the one above, then upgrade on the truth thatthe wild predictions have actually become a realityThe very first rumblings, I ‘d state, came a years earlier with AlphaGo, they got visibly louder with LLMs and coding and thinking representatives, and they’ve accelerated this summertime and fall under a crescendo of marvels and fears that a person requires to be a specific type of moron to reject.

I recoil from the neverending shell video game where you state “oh sure,obviouslyAI can now [escape from its sandbox / solve Millennium Problems / whichever dramatic thing it most recently did]nobody ever rejected that[Ididreject it]wake me up when AI does [thing AI hasn’t yet done but is going to do next year]that’swhen I’ll reassess my entire worldview [no I won’t]” Where no matter how quickly the rollercoaster speeds up, even after your entire familiar world has actually disappeared behind you, you’re still developing reasons that it does not count.

My position on AI is simply the conservative, hesitant position of 2006, upgraded with intellectual sincerity for the truth of late 2026. Which position, if you require me to spell it out, is as follows:

AAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAA
AAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAAA

It appears to me that the Singularityhas actually currently begun; it’s simply hugely unevenly dispersed. Yes, I still discharge the dishwashing machine and clip my toe nails. On the other hand, in whatever years I have actually left, I do not anticipate that I’ll ever once again show a theorem due to the fact that I’m really required to show it. If I do, it will just be for my or others’ satisfaction or illumination.

The test is this: if we took the news of these previous couple of weeks and sent it back in time twenty years, would I concur that it appeared like the start of an AI Singularity? The intellectually truthful response is: yes, definitely. Then that’s all we require. No backsies.

I seem like it would be healthy for everybody to stop grinding their ideological axes, their beliefs about Dario Amodei or Sam Altman, for enough time just to acknowledge thatthe marvels and horrors are hereThey could not be here more plainly if the sky had actually turned reddish-orange like in theMatrixfilms.

It’s here plainly adequate that, when I put my kids to sleep during the night, I now feel it in the pit of my stomach: what sort of future can they perhaps have? What could they discover today that could potentially relate to that future? (Yesterday, my 13-year-old child joked unprompted that, if she wishes to end up being a mathematician, it now appears like she has possibly 2 more weeks.) When my graduate trainees desire to discuss what sort of professions may await them on graduation, I no longer have any hint what to inform them.

Perhaps it will assist if I quickly change subjects. Since my better half and I transferred to Austin, I’ve in some cases gotten some variation of the following question: “How can you, as both a Jew and a hesitant researcher, perhaps hit it off with all those evangelical Christians down there in Texas? Sure, they mayappearincredibly friendly to Jews, however do not you comprehend that that’s just since of the unique function Jews play in their eschatology– when Christ will return in splendor, and you’ll either accept Him as Lord otherwise roast in hell for eternity?” I gaze at them and state: “wait, so I get to accept Christ justafterHe returns? What a lot! How could I perhaps have any objection to that?”

For anybody who states AI doom seems like an apocalyptic religious beliefs, that the rationalists/Singulatarians appear like a Bay Area cult, that Eliezer Yudkowsky releases the vibes of a messianic prophet: yes, yes, and yes. Most importantly, today you’re no longer being asked to think in arguments and projections,Just in the front-page news.Accepting the truth of the coming maker godafterit’s resolved Navier-Stokes and lots of other longstanding open mathematics issues (while drastically increase in ability monthly), is sort of like accepting Jesusafterhe’s gone back to earth on the shining cloud. It’s the epistemic bare minimum.

Yes, there’s still massive unpredictability about what the rest of our lives will appear like, however as far as I can inform, there’s no longer any genuine unpredictability that it’ll all mainly focus on AI, and the level to which we prosper or stop working at directing its power towards human growing.

By any accounting that does not stack the deck, Eliezer Yudkowsky wasabout what the best obstacle dealing with civilization in our life times was going to be, and you and I wereincorrectabout it. WhyI was incorrect is a concern I’ll ask myself every day in whatever time stays. You understand, at least I upgraded as soon as the prophesied marvels and fears in fact began getting here! If you have not done also, why have not you?


As you probably understand by now– it was the talk of the geek web all week– the Navier-Stokes Millennium Problem seems resolvedwith vital contributions from both people and AI, albeit with a twisted conflict about precisely what took place and what should have actually occurred. The response, which an OpenAI design has actually obviously confirmed in Lean, is that (as lots of mathematicians presumed recently) there’s smooth preliminary information that results in a singularity in limited time, a minimum of if a smooth external force is used (the case without any external force is still unsettled). This issue was expected to bring a $1 million reward, other than that OpenAI states they have no interest in gathering the reward and it’s uncertain if any human is qualified to gather rather. OpenAI burned a minimum of ~$ 15 million in calculate to produce its 166-page servicewhich most likely hasn’t yet read and comprehended by any human.

See here for the Quanta short articleand here for NYU mathematician Tristan Buckmaster’s account of the function played by himself and Levent Alpöge of Anthropic, which considerably varies from OpenAI’s account (you can check out a reaction from OpenAI’s Sebastian Bubeck here. It’s concurred that whatever developed on a technique originated recently by the human mathematicians Diego Córdoba and Luis Martínez-Zoroa.

My function here is not to adjudicate the disagreement. Yes, in stroking in with greatly higher resources once it had actually discovered development on Navier-Stokes, OpenAI appears to have actually acted in a manner that some may refer to as “unsportsmanlike.” No, I do not discover it possible that OpenAI’s designs meaningfully taken advantage of being trained on Buckmaster and Alpöge’s chat logs. This leaves a vital concern unanswered: what precisely did OpenAI understand about Buckmaster and Alpöge’s work and when did it understand it?

Anyhow, as Zvi mentionsit’s simple to get hung up on the information and forget the high-order bit: particularly, that it appears safe to state that human mathematicians are forevermore dismissed as the primary theorem-proving entities on world earth. I feel fortunate to have had the conventional type of profession in theoretical computer technology in the last years when that was possible.


If we were simply discussing Navier-Stokes, you may implicate me of leaping to conclusions here. We’re not. In the locations I understand finest (such as quantum intricacy theory), and most likely other locations too, there’s now a deluge, with longstanding open issues both significant and small falling by the day.

Go to the arXiv or ECCCBasically all the documents that I ‘d have an interest in now consist of “AI declarations” near the recommendations (as this is typically the main thing I need to know, I want I didn’t require to scroll to the end of the paper to discover it!). These declarations can vary from “our primary outcome came completely from GPT-6, however we comprehended it and take duty for it,” to “the outcomes originated from an interaction in between the human authors and AI” to “we utilized AI, however just for checking and other incidental things” to (mad props!) “the author did not utilize AI for anything“

If you talk today to editors or program committee chairs, it’ll advise you of those threatening scenes from the Lord of the Rings motion pictures where the guys of Gondor or Rohan or whatever are grimly strengthening their walled city versus the anticipated assault of 50,000 orcs. Examining will have to be done partially by AI, since otherwise there’s no chance to deal with the orc army: the customers can’t unilaterally deactivate.

Anyhow, here’s a little tasting of the considerable AI-proved or -helped arise from, like, the last monthbesides Navier-Stokes– limiting myself to those that fixed longstanding open issues I had actually formerly understood or appreciated.

  • Naturally, the counterexample to the Jacobian guesswork, revealed by Levent Alpöge in a now-famous tweet: “hi there the jacobian guesswork is incorrect thanx to my buddy akhil for inquiring about it and my other friend myth for working throughout the world cup last” (followed by a listing of the counterexample)
  • Enhanced bounds for Grothendieck’s continuous (led by good friends and associates of mine at UT Austin)
  • A Lean-verified evidence of Fermat’s Last Theorem
  • Quantum oracle separation in between QMA and QMA( 2)and evidence of Watrous’s disentangler guesswork, an issue that I and others promoted back in 2007– by a list of authors including my just recently finished PhD trainee Sabee Grewal
  • An evidence of ideal efficiency for QMAfrom (once again) Sabee Grewal and Dorian Rudolph, fixing a decades-old open issue that I studied back in 2009
  • An enhanced upper bound for shadow tomography of quantum statesfrom Chen, O’Donnell, Pelecanos, and Wright, enhancing the reliance on the Hilbert area measurement d from log( d) to √ log( d). (When I presented shadow tomography back in 2017, I raised the concern of whether the reliance on d might be gotten rid of completely, while protecting polylogarithmic reliance on the variety of measurements m.)
  • Development on the Aaronson-Ambainis Conjecture (the variation that talks straight about quantum algorithms), generally revealing that it holds for quantum algorithms that make their questions in a little number of parallel rounds.(Update: Nope, sorry, Jordan Docter explains to me that this one was pre-AI, with AI utilized just for checking and other incidental things!) This was separately attained by Liu and Mutreja, making more considerable usage of AI.
  • According to reports that I’ve heard, services to some really longstanding open issues in theoretical computer technology (no, not P ≠ NP or other intricacy class separations, however think of a few of our other greatest issues). I’m informed that the AI business, having actually been burned by the hostile action to the Navier-Stokes evidence, are now resting on services to some extremely significant issues till they determine a much better method to manage things

Do not hesitate to advise me of anything I overlooked.


Let me attempt to communicate the state of mind in the mathematical neighborhood today, a minimum of as far as my experience reaches. Almost every discussion has to do with the AI tsunami, or ultimately circle to the tsunami even if it’s initially about something else. Typically, however, the focus is less on the unknowable future– for just how much longer will mathematical research study as a human business even exist?– than on instant concerns of how to react

What are the brand-new guidelines for when you get to compose a paper with your name on it, and, y’ understand, get credit for it? That you completely comprehend the evidence, can offer talks about the evidence, can respond to concerns about it, take obligation for its accuracy? Do you require to have actually played any function in discovering the evidence?

In the events, most likely to end up being increasingly more various, where all of those conditions are not pleased, how do you share AI-generated mathematics, if at all? Do you tweet it, like Alpöge hilariously made with Fable’s disproof of the Jacobian Conjecture? Do you publish to the arXiv or GitHub? Do you release a paper that notes “GPT-6 Astra” or “Claude Fable” as the author– however then let the AI a lot thank you in the recommendations for recommending such a terrific issue to it?


Obviously, how one reacts to the instant issues eventually does depend upon one’s wider beliefs about what mathematical research study is for and about. Are we simply attempting to choose whether numerous opinions hold true or incorrect? Or are we attempting to preserve a human neighborhood, throughout the generations, that comprehends the opinions and appreciates whether they’re real or incorrect and why? If the latter, how do we incentivize individuals to sign up with that neighborhood, to go through the years of extreme training needed, if their function will now be lowered to verifiers and explicators (if even that) of giant arguments discarded into their laps by the AI business?

As a number of you will have seen, twenty-five Fields Medalists, consisting of Terence Tao, launched an open letter entitled A Severe Misalignment of AI in Mathematicswhich articulates a few of these issues in the wake of the Navier-Stokes statement. As numerous critics have actually mentioned, the open letter does not truly have a clear ask: mainly, it simply eloquently sets out the worths of the human mathematical neighborhood that the authors think about worth maintaining in the age of AI. After reflection, I chose to back the declaration, since I desire to protect those worths.

I do not believe any of the signatories are ignorant sufficient to picture that AI will not completely alter the method mathematical research study is done– undoubtedly, that it isn’t currently doing so. There’s certainly at a lot of a small market for “qualified natural theorems.” That isn’t the concern. The concern is, do we integrate AI in a manner that still puts human understanding, of what either people or AIs are producing, at the center of the entire business? Perhaps sooner or later, it ends up being unsustainable to do that. Perhaps one day we state: “human mathematics had a fantastic 4,000-year run, however today we close up store and turn whatever over to the devices, continuing to use our own brains to mathematics, when we do, at a lot of for workout, entertainment, or competitors, like chess.”

Partially since of my concerns about AI misalignment, I’m not all set to toss in the towel simply. I still do wish to keep insight and understanding at the center of what mathematicians, computer system researchers, and physicists do, for as long as we can keep it there, even as the mankind now delivers its supremacy at the job of showing or negating opinions.


Mentioning positioning: if you’re any type of mathematical scientist, and today age of marvels and horrors has actually influenced you to wish to invest your staying time facing the tsunami head-on, instead of pretending it does not exist or is still far, please join your lots of coworkers who’ve reached the exact same location!

My buddy and associate Mike Winer was trained as a theoretical physicist, did a postdoc with Juan Maldacena at the Institute for Advanced Study in Princeton, however then got AGI-pilled and chose to change to full-time work at the Alignment Research Center in Berkeley (established by Paul Christianowho transferred to AI positioning a years earlier after doing quantum computing theory with me). Mike just recently composed a Substack post entitled From Academia to Alignmentwhich I took pleasure in and which I ‘d applaud to anybody presently considering this shift. In a comparable vein, see this from Xiaoyu He. And, another: a meditation on mathematicians’ possible future as priests or monksby Stanford mathematics undergrad Logan Graves.

Published in Announcements, Complexity, Quantum, The Fate of Humanity|262 Comments”

Technology 9/11 in Berkeley

September 11th, 2026

Keep in mind: Obviously I’ve been glued all week to the significant advancements in AI. I’m dealing with a post about them. I’m bad at responding to things in a prompt method. Today, I’ll do my post marking the catastrophe a quarter-century ago that we all celebrate. Please do not hesitate to share your 9/11 memories in the remarks. Shana Tova to those who commemorate!


The early morning of September 11, 2001, I was a second-year PhD trainee at Berkeley, who awakened late in his dormitory at International House, after a long night invested closing in on the evidence of the quantum lower bound for discovering crashes.

Rolling over to my laptop computer, I saw a flurry of odd e-mails, consisting of one from Prof. Christos Papadimitriou stating that “we’re a neighborhood, and we’ll all support each other,” and another from Prof. Luca Trevisan (whose algorithms course I was then TA’ing) stating “on a day like this, it’s difficult to consider algorithms. Class is cancelled.”

Baffled, I clicked over to the New York City Times and saw the photo of the burning towers, and check out numbly about what was currently over by the time I ‘d awakened. I signed in with my mama, ensured family members and good friends in the NYC location were okay. My daddy was at a business occasion in Atlanta, and would require to drive home since of the nationwide grounding of flights.

One of my earliest memories in life, from age 5, is of rising to the top of the World Trade. Maturing an hour’s drive from NYC, it wasn’t an unique location to me.

I quickly discovered that a person of the dead was Danny Lewinthe ex-IDF captain, theoretical computer system researcher, and cofounder of Akamai who had his throat slashed on among the airplanes while attempting to combat the hijackers, making him the day’s very first casualty, even while Akamai’s innovation became part of what kept news sites running that day. I ‘d never ever fulfilled Danny however currently understood lots of people in typical with him. A couple of years later on I ‘d be humbled to win the trainee paper award that was called in Danny’s memory.

Anyhow, at Berkeley on 9/11, I roamed over to Soda Hall simply to be with other individuals. A couple of trainees appeared for workplace hours, desiring assist with their algorithms research, which I discovered hard to think, however I did my finest to focus, as the computer system screens around me revealed the burning towers.

That night, I went to a vigil for the victims in Sproul Plaza. The “vigil,” such as it was, rapidly given with grieving and prayers and turned to praised speeches about how the United States should react with love rather than war, and need to turn the other cheek. A trainee communist company was handing out leaflets describing that the victims were primarily “rich capitalists and the employees who attempted to save them.” This while smoke still blanketed NYC and the desperate look for survivors continued. I left the vigil early.

Up until that day, I had actually thought about myself as essentially a “leftist,” one whose # 1 concern was the existential danger of environment modification. Sure, I disagreed with my fellow leftists about concerns from nuclear power to talented education to Israel, however those were simply intra-left conflicts.

The year before, I had actually developed the site “In Defense Of NaderTrading,” in a desperate effort to intervene in history and trigger Al Gore to end up being president instead of George W. Bush. When Bush “won,” by the notorious 537 votes in Florida, I considered it a success for horribleness that would never ever be gone beyond by anything else in my life time (ha). I could not picture any political leader who was more the reverse of whatever I thought in than Bush. This view, obviously, did not especially stick out at Berkeley.

In the days after 9/11, however, it ended up being apparent that I might not be a “leftist” in the Berkeley sense. A few of my fellow trainees felt that Osama bin Laden made a great deal of terrific points, that the attacks were essentially warranted, which at any rate, we in Amerikkka had actually done much even worse to provoke them, consisting of by supporting the genocidal settler-colony called “Israel,” which for all we understand covertly masterminded the 9/11 attacks anyhow (although once again, if bin Laden had actually done them, he would’ve been warranted).

Around the exact same time came the Second Intifada, when a wave of suicide battles in Israeli buses and pizza parlors and university lunchrooms delighted and stimulated some Berkeley trainees to the level that they took control of a Holocaust Remembrance Day occasion with bullhorns to make it about the Nakba, smashed the windows of the Hillel structure, and batter a number of trainees using kippot. That was how completely anti-Nazi they were.

I completed my PhD at Berkeley in 2004 having actually found out about more than quantum computing. I ‘d found out that, while American academic community had pockets that really were essential havens and sanctuaries for geeks like me, it likewise harbored individuals who would happily see me and my loved ones and my fellow Americans eliminated for the sake of their ideological vision. And I ‘d found out that I had my own ideological vision, which was that such individuals might go fuck themselves.

It deeply hurt me to be on the very same side of anything as George W. Bush– specifically since I understood that 9/11 had actually occurred on his watch, that he had actually disregarded all the cautions, which he was grossly inexperienced to handle the resulting wars versus jihadism (simply how unskilled, I didn’t understand at the time). As flawed as Bush was, I understood that I desired to maintain rather than ruin the civilization of which he was a short-lived steward. And I believe the worth and fragility of our civilization is the primary lesson from that day that I ‘d like to communicate to my kids, for whom obviously 9/11 is simply another historic occasion to learn more about in school, like the Boston Tea Party or the Alamo.

Published in Adventures in Meatspace, The Fate of Humanity|40 Comments”

Technology LLMs and self-referentiality

September 1st, 2026

I got up the other day with the following ideas, which are most likely either apparent or dumb.

A main thesis that numerous readers, including me, drew from Douglas Hofstadter’s Gödel Escher Bach when young was that the trick of intelligence (and for that reason, of AI) was going to have a lot to do with self-referentiality and “weird loops.”

Even Roger Penrose’s The Emperor’s New Mindwhich in some methods was the anti-GEB, paradoxically concurred with GEB about the essential significance of self-reference to the success or failure of the entire AI job. It declared (improperly, in my view and in a lot of professionals’) that AI might never ever work due to the fact that there was something about Gödel’s Theorem and self-reference that no computer system program might ever catch, however that might be caught by unique physics available to the human brain.

Now, in 2026, we’ve been successful at constructing AIs that outshine most people at many intellectual jobs that are distinct enough to judge. And at no point in the tech stack of those AIs– neither in the transformer neural webs, nor in the GPU clusters they work on, nor in the training procedure, nor anywhere else– did anybody requirement to integrate in anything about self-reference. (Excepting, eg, the system guidelines that inform the design about its function and identity, which aren’t required for smart habits. I’m not going to count the autoregressive nature of LLMs as “self-referential”; that’s simply dynamical feedback.)

Obviously, GPT 5.6 Pro and Fable can speak about themselves, about Gödel’s Theorem, about self-reference, about what we’re speaking about today, all of it, much better than the majority of human beings. At no point did anybody requirement to construct self-referential capabilities in. They popped out as a by-product of the very same pretraining that let the designs speak about Pokémon and long-chain polymers and cognitive behavior modification and plate tectonics and whatever else.

Not surprising that Hofstadter states he’s been stunned by the success of LLMs, and has actually appeared depressed about existing AI abilities in essays like this oneHe’s method too wise to reject what’s occurred or create reasons that it does not truly count (the method lots of have actually taken). He recognizes that we now have real conversational intelligence from a course that the GEB worldview would’ve concerned as far too inexpensive and easy, and that definitely has no “odd loops” developed in anywhere.

Obviously, a Hofstadterian might argue that an odd loop emerges in LLMs– certainly, absolutely nothing in GEB ever stated that odd loops would require to be clearly crafted at the start. Would anybody who had not been brought up on GEB show up at this as a helpful method of believing about LLMs?

What can we state about this with hindsight? While the concepts of diagonalization and self-reference naturally played a main function in the birth of contemporary mathematical reasoning and computer technology, the most popular usages were unfavorable: there is not a bijectjon in between the natural numbers and the reals. There is not a total noise evidence system for math. There is not an algorithm to fix the stopping issue.

If your objective was just to develop the axioms of ZFC and the guidelines of first-order reasoning, or construct an electronic computer system, you would not clearly require self-reference for that. You would simply … begin structure, making sure that your direction set didn’t disappoint universality.

Yes, ZFC can formalize and show theorems about itself. Yes, electronic computer systems can run programs that take their own code as input. No one ever required to construct those capabilities in, any more than self-reference required to be developed in to the alphabet or the guidelines of grammar. It popped out as a complimentary by-product of universality.

In the exact same method, LLMs’ capability to speak about themselves popped out as a by-product of their capability to discuss anything in the discourse universe they were trained on. The huge, old concepts about intelligence that wound up generally vindicated were the concepts about how intelligence has to do with forecast, and forecast has to do with compression, and compression has to do with discovering much better and much better upper bounds on Kolmogorov intricacy. Not the self-reference things. (Although, if you needed to know why Kolmogorov intricacy can’t be calculated completely, that unfavorable declaration would once again need a self-referential argument.)

What’s left? Awareness and subjective experience naturally stay incredibly strange. For all we understand, Hofstadter might be ideal that those have something to do with self-reference. (For all we understand, even Penrose might be ideal that they have something to do with unique physics available to biological brains however not digital computer systems!)

The concept that you ‘d require specific self-referentiality before you could get persuading and world-changing conversational intelligence? Let it be buried in a Westminster Abbey or Arlington National Cemetery for the most essential incorrect concepts in human history– geocentrism, Aristotle’s teleological physics, aether, phlogiston, Freud’s psychology, Marx’s forecast of an employees’ uprising followed by an egalitarian paradise, and so on. Buried it requires to be.

Published in Embarrassing Myself, Metaphysical Spouting, Procrastination|135 Comments”

Technology Anthropic’s LLM watermarking

August 22nd, 2026

Yeah, Anthropic has actually revealed that it’s now watermarking the outputs of Claudeutilizing a plan based upon Google’s SynthID, which remains in turn based upon the Gumbel Softmax plan that I proposed at OpenAI back in 2022– as far as I understand, the very first LLM watermarking proposition, though far from the last one. I’m pleased that Anthropic credits me for this, despite the fact that I shirked my task by never ever releasing a paper about it (by the time I took a seat to compose one, it looked like the entire field had actually currently absorbed my plan and moved beyond it– AI simply moves too quick for me!).

For those who do not understand, watermarking ways somewhat altering the manner in which an LLM runs to place a subtle signal that lets you show later on, with high analytical self-confidence, that a text undoubtedly originated from your particular LLM. It utilizes the randomness that’s currently present anyhow in LLM outputs, changing a few of it by pseudorandomness that prefers particular word mixes over others in a manner that’s later noticeable, offered just the series of tokens itself (not the timely or the likelihoods) in addition to the secret of the pseudorandom generator. Christ, Gunn, and Zamir Considerably enhanced my plan to get real cryptographic indistinguishability, and there have actually been other enhancements given that.

I ‘d been implying to blog site about this for days. The Good News Is, Zvi Mowshowitz, the world’s primary blog writer about AI, has actually now composed a terrific post, entitled AI Text Watermarking Is Free And Goodwhich conserves me from the requirement to compose my own long post. In specific, Zvi masterfully discusses the main point that I required to discuss to everybody back in 2022-23: why, contrary to lots of people’s instincts, there’s no fundamental tradeoff in between watermarking and the quality of LLM output. Essentially, almost every LLM output was currently a sample from a cloud of tremendously lots of possibilities, all of them about similarly great, so there’s lots of space to guide within that cloud without impacting anything that a normal user would observe. As my kids would put it, the mathematics mathes.

As Zvi discusses, the main technical downside of watermarking plans like the one I proposed, and what Anthropic is now utilizing, is that it’s possible to eliminate the watermarks with a little additional work (even things as basic as, e.g., equating in between English and French, asking the LLM for words sprinkled with emojis and after that eliminating the emojis, or utilizing an open design to paraphrase the output). Zvi offers in-depth arguments for why he anticipates watermarking to stay a net favorable in practice in spite of this vulnerability.

I might include that, in addition, there’s current development (see here ) on what I’ve called “semantic watermarking,” or watermarking at the level of the underlying principle vectors rather than the tokens themselves. This in fact appears to work, albeit without any theoretical warranties, and will ideally make eliminating watermarks a lot more difficult– although the Barak et al. impossibility result recommends that under possible presumptions, no LLM watermarking approach will be totally sure-fire.

Anyhow, I exercised my plan in Fall 2022, then offered great deals of speak about it (consisting of, as it takes place, at Anthropic), and likewise dealt with Hendrik Kirchner at OpenAI, who really carried out and checked my plan. OpenAI management chose versus releasing watermarking, anxious mainly about threats to the item (i.e., consumers doing not like the concept, and leaving for a completing LLM that does not watermark). You can check out this Wall Street Journal examination from 2 years ago for more. I was confident that the State of California was going to fix the collective-action issue by mandating watermarking for AI designs, however then they chose to do that for audiovisual material justfor some factor excusing text.

Google DeepMind carried out something really comparable to my proposition in its SynthIDreleased in all its Gemini text designs. They greatly limited who gets to discover the watermark, that made their exceptional choice of minimal usage to my scholastic associates, who’ve been pleading me for a method to identify whether their trainees are utilizing AI to cheat. (For now, I generally send them to Pangrama leading AI detector not based upon watermarking, as a very first line of defense.)

And now, obviously to abide by EU guidelines, Anthropic states they’ve released a watermarking plan like mine where anybody will have the ability to do detection (though they likewise state in their FAQ that they’re still dealing with the detection API). Even OpenAI recommends that it prepares to do the same4 years after I seriously believed about this, it looks to my surprise like this is really taking place. Thanks, EU!

Inform you what: check out Zvi’s postand after that whatever concerns you still have, you can come here and ask in the remarks. Simply please do not utilize Claude to compose the remarks. With any luck, I’ll become able catch you if you do.

Published in Announcements|57 Comments”

Technology Much better than gold

August 20th, 2026

What’s about the only thing more badass than a 17-year-old winning a gold medal at the International Olympiad in Informatics (IOI)?

That 17-year-old purposefully surrendering his gold medal by using an Israeli flag while the medal was revealed, defying the IOI’s boycott of Israel (for background on this boycott, see my post from 2024).

Kol HaKavod (mad regard)to Yotam Budnik, who exceptionally, has won a Gold Medal (which he was permitted to keep, obviously )at the International Math Olympiad. And congratulations to the whole Israeli group, which( extremely )would obviously have had a greater general rating than the United States group, had it been enabled to complete as a main group at all.

Published in Announcements, Rage Against Doofosity|86 Comments”

Technology Michael Rabin memorial conference

August 16th, 2026

Friend-of-the-blog (well, generally simply pal) Adi Akavia has actually asked me to advertise that she’s assisting to arrange an interesting CS teleconference Mind-IL at Tel Aviv University on October 26, in memory of the Israeli-American Turing Award winner Michael O. Rabinwho died in April. Please keep in mind that October 26 is the day before the Israeli electionfor any Israeli citizenship holders living abroad who may desire a scholastic reason to come to Israel and vote.


Update (August 19): Avi Wigderson likewise asked me to promote a conferenceto be held September 16-18 at Bletchley Park in the UK, to celebrate the 90th anniversary of Alan Turing’s “On Computable Numbers” paper.

Published in Announcements, Complexity|31 Comments”



Discover more from PMN S.P.O.R.T.S - A PRIME MEDIA NETWORK BRAND

Subscribe to get the latest posts sent to your email.

Related Articles

LEAVE A REPLY

Please enter your comment!
Please enter your name here

Captcha verification failed!
CAPTCHA user score failed. Please contact us!