
AIが「一生分の仕事」を解いてしまうとき
When AI Solves Your Life’s Work
When AI Solves Your Life’s Work
The AI Daily Brief: Artificial Intelligence News and Analysis
要約
OpenAIが未公開の最先端モデルによる372件の新規数学結果と722本の補足論文を公開し、数学界に大きな波紋が広がっている。リーマン予想やホッジ予想の部分解を含むとされ、数学者の反応は驚嘆から不安、楽観まで様々だ。ホスト(Nathaniel)は、他分野にも同様の変化が及ぶのか、検証が追いつくのかという論点を整理している。冒頭ではOpenAIの売上数字をめぐる混乱やThe Informationの調査、Claude DashboardsとMotionも取り上げた。
- ●OpenAIは内部の最先端モデルで生成した372件の新規結果と722本の補足論文をGitHubで公開した。結果あたりの計算量はChatGPT Pro約3時間分と開示された。
- ●ProofAtlasの未解決問題トップ500のうち90件を完全に解決し、リーマン予想とホッジ予想の部分解も含まれると紹介された。
- ●懸念の中心は検証である。数学者のMaggi氏は、人間が読み・理解できる速度を超えて証明が生産されることの意味を問うた。
- ●行列積の指数が2.25以下になるなど実用に近い成果もある一方、暗号関連の成果が少ない点が指摘され、AI企業が暗号破りの可能性を調べ始めているというAaronson氏の話も紹介された。
- ●Google のPeyman Milanfar氏は、自動化の速度は領域のエントロピーに反比例するとし、数学・コーディングの次に自然科学、金融、医療、法律、文化の順に難しくなると述べた。
- ●ヘッドラインでは、OpenAIの年換算売上が約500億ドルで、報じられた680億ドルは投資家による総額ベースの試算だったとFTが報じ、半導体株が下落した。
章立て
OpenAI売上数字の混乱
FTによると、OpenAIの年換算売上は約500億ドルで、680億ドルという数字は投資家の総額ベース試算だった。NVIDIAやOracleなどの株価が下落し、Anthropic上場への示唆も語られる。
The Informationの購読者調査
AI投資の収益認識、OpenAIの巻き返し、OpenClaw・Claude Code・Codex・Cursorの利用率、オープンウェイトモデルの利用、採用への影響を紹介する。
Claudeの規約変更とDashboards/Motion
Claudeへの過度な虐待的言動を禁じる規約変更と、データを可視化するDashboardsおよびアニメーション生成のMotionを取り上げる。
OpenAIの数学成果の公開
372件の新規結果と722本の補足論文が公開され、数学諮問委員会との協調や推論過程の開示、計算量が説明される。過去のIMO金メダルやNavier-Stokesの流れも整理する。
成果の規模と数学者の反応
トップ500問題のうち90件の解決、リーマン予想などの部分解に対し、驚嘆や「Move 37」との評価が出た。計算コストの急低下も話題になる。
検証の壁と数学という分野の行方
人間の理解が追いつくかという懸念と、学生の動揺や楽観的な見方が紹介される。実用的意義や暗号への影響にも触れる。
次に変わる分野はどこか
経済学などへの波及を示す声と、領域のエントロピーによる難易度の序列を紹介し、今後の注目点を述べて締める。
解説記事
AI業界の最新動向を伝える番組「The AI Daily Brief」の2026年10月9日回は、OpenAIが公開した大量の数学的成果を主題にした。ホストは、これが「分野単位でのAIによる破壊的変化」の最初の実例になり得るのかを問い、数学者たちの反応と他分野への波及の見通しを整理している。
OpenAIの売上数字をめぐる混乱
本編に先立つヘッドラインでは、OpenAIの売上をめぐる報道が取り上げられた。Financial Timesによれば、OpenAIは投資家に対し、9月末時点の年換算売上が約500億ドルだったと伝えた。先月広く報じられた680億ドルは、OpenAI自身ではなく一投資家の試算だったという。番組によると、Anthropicが第三者経由の販売分も含む総売上で数字を示すのに対し、OpenAIは分配後の純売上を使う。その投資家は両社を同じ基準で比べようとしたとされる。この報道を受けて半導体株は下落し、NVIDIAは3%、Oracleは5.5%の下げだった。ホストは、Anthropicが上場して監査済み財務が出れば、非標準的な会計のため投資家向けの数字より低くなるだろうと指摘した。
調査が示す採用状況と新機能
The Informationの購読者調査では、AI投資の回収について35%が「支出の数倍のリターン」と回答した。サービス別では、夏の調査でClaudeに抜かれていたOpenAIが再びトップに戻ったという。エージェント型コーディングツールではClaude Codeが45%、Codexが33%、Cursorが23%だった。オープンウェイトモデルは72%が何らかの形で利用する一方、中国製の利用は21%にとどまる。採用への影響は小さく、半数近くが「変化なし」と答えた。ホストは回答者が先進的な層に偏る点を断っている。
Anthropicは利用規約を改め、持続的で不必要に虐待的な言動を禁じた。極端な場合に限る方針だとされる。BoxのAaron Levy氏は、人間の無礼な対話が将来のモデルの学習データになるのを避ける意味で妥当だという見方を示した。機能面では、CRMや会計ソフトのデータからダッシュボードを生成するClaude Dashboardsと、データを短いアニメーションにするMotionが発表された。アニメーションは動画ではなくコードで書かれるため、文言や数値を簡単に修正できる。
OpenAIの数学成果とその規模
本題は、OpenAIがGitHubで公開した372件の新規結果と722本の補足論文である。未公開の内部モデルが生成したもので、先月のNavier-Stokes解法と同じモデルの可能性が高いと番組は述べる。新設のMathematics Advisory Committeeとの調整のもと、10問分の推論過程が公開され、1結果あたりの計算量はChatGPT Pro約3時間分と開示された。
ProofAtlasが作成した未解決問題トップ500のうち90件が完全に解かれ、リーマン予想とホッジ予想の部分的な解も含まれるという。Rutgers大のAlex Kontorovich氏は、人間が行ったなら即座にフィールズ賞だという趣旨の投稿をした。Anthropic研究者は「数学史上最も重要な瞬間」と呼び、元OpenAI研究者は数学における自分の「Move 37」の瞬間だと表現した。Navier-Stokesでは約1万エージェントで88時間かかったとされ、効率化の速さにも注目が集まった。
検証できるのか、数学者はどうなるのか
最大の論点は、人間が内容を理解し検証できるかである。Francesco Maggi氏は、数学は正しい命題の蓄積ではなく、理解され、結びつけられ、再利用されて成り立つと述べ、生成速度が吸収速度を大きく超えたときの意味を問うた。物理学者のSteve Su氏は、機械が発明した概念の網を人間が把握できなくなり、数学や物理が解釈可能性を失う可能性に言及した。
数学を学ぶ学生の動揺も紹介された。ある大学院生は、取り組みたかった課題の大半が無力化されたと語った。一方、元ワシントン大教授のShriram Kannan氏は、数学者は新しい世界にいずれ適応して喜ぶはずで、これは「自分は数学の世界的な担い手だ」というアイデンティティへの一時的な衝撃にすぎないと述べた。
なお、多くの結果は純粋理論的で実用性は限定的だが、行列積の指数が2.25以下に改善された成果は、AI推論の基盤となる計算に関わる点で注目された。また、暗号分野の成果が目立って少ないという指摘があり、Scott Aaronson氏はAI企業が暗号プロトコルを破れるか内々に調べ始めていると書いた。
次に変わるのはどの分野か
Wharton校のEthan Mollick氏は、経済社会学の分野でほぼ自律的な研究が一流誌水準に達しつつあると述べた。ただし数学は検証可能で強化学習に適した問題だとの指摘もある。Peyman Milanfar氏は、自動化の速度は領域のエントロピーに反比例するとし、数学とコーディングの次に自然科学、金融、医療、法律、文化・芸術・経営の順に難しくなると整理した。ホストは、どの分野にどう波及するのか、あるいは人間や制度の慣性が歯止めになるのかを今後の注目点とした。
まとめ
今回の話は、成果の真偽の確認と、数学という営みの意味づけの両面で課題が残ることを示している。各結果の正否は、先月のNavier-Stokesを含め、まだ数学者による検証の途上にある。日本の研究者や技術者にとっても、AIが出した成果をどう検証し、人間の理解や文化にどう結びつけるかは共通の課題になり得る。売上報道に表れたように、AI業界の数字や期待は市場を大きく揺らす。主張と検証済みの事実を分けて読む姿勢が、これまで以上に求められる。
文字起こし(英語・自動生成)
So far, the idea that AI would upend and undermine entire industries hasn't really come to fruition. Certainly, we're seeing how people do pretty much everything change. Coding is perhaps the most changed, and yet demand for coders seems to be going up. With a new set of mathematics results from OpenAI, however, some are asking, is this the first field-level AI disruption actually happening in practice? The AI Daily Brief is a daily podcast and video about the most important news and discussions in AI. All right, friends, quick announcements before we dive in. First of all, thank you to today's sponsors, KPMG, Harbor, Robots and Pencils, and Blitzy. To get an ad-free version of the show, go to patreon.com slash AI Daily Brief, or you can subscribe on Apple Podcasts. And if you want to learn more about sponsoring the show, send us a note at sponsors at aidailybrief.ai. Uh-oh. Financial figures for the AI industry are being called into question as OpenAI's revenue numbers are revealed to be significantly lower than previously reported.
The Financial Times reports that OpenAI has told investors that they hit roughly $50 billion in annualized revenue at the end of September. That's a fairly big gap from the $68 billion that was widely reported last month. So is this just misreporting or something else going on? Well, sources said that the $68 billion whisper number came not from OpenAI itself, but from an OpenAI investor. Apparently what was going on is that that investor was trying to get an apples-to-apples comparison with Anthropic. Anthropic numbers are always quoted as gross revenue before revenue sharing. In other words, when Anthropic tokens are sold through a third party like Amazon or Microsoft, Anthropic's including that in the total revenue, even though by the terms of their deal, some big chunk of that revenue is going straight to the pockets of those third parties. The logic, I imagine, is to try to provide an overall number of the total expressed demand in the form of anthropic tokens sold. And that is a useful number to know. However, now that these companies are going public, there's not really a convention that's likely to fly in public markets.
Meanwhile, OpenAI has, for their part, always used net revenue after those revenue splits. To make matters worse, the FT noted that the OpenAI investor calculated gross revenue based on a short time frame. It's unclear whether they extrapolated annualized revenue from a month, a week, or even a day. The news was not all bad here. OpenAI also told investors that they achieved 77% run rate growth in the third quarter and 107% growth in their enterprise business, but even that incredible growth was completely overshadowed by the $20 billion gap in revenue. In many ways, Wall Street reacted as if OpenAI had missed on revenue during an earnings report. Semiconductor stocks were down significantly after the FT published their report, with NVIDIA down 3%, Oracle down 5.5%, and the long tail of NeoCloud stocks being even larger drops. And yet, even if we recognize that OpenAI is not yet a public company, had not shared the misreported number, and did not in fact miss projections on an earnings report, the reaction still brings up some big questions around the Anthropic IPO. Once Anthropic's audited financials
are available, they are presumably going to be much lower than the numbers Anthropic has been sharing with private market investors by the nature of using non-standard accounting. Ultimately, the narrative battle is growing more intense as Anthropic heads towards IPO. Ahmed is investing wrote, I don't know all the facts about this OpenAI revenue story, but what's becoming apparent once again is that this entire semi-trade is levered against OpenAI and Anthropic and its fragile AF. Any narrative on slowing growth and the street loses their mind. Next up, some interesting results from the information's latest subscriber survey. In terms of audience, the information has a very tech-forward tech industry type of reader base. They found that two-thirds of their subscribers work at an organization that develops AI applications. Perceived return on investment seems to be rising rapidly, with 35% saying their organization is returning multiples on their AI spend. Another 17% said the returns are positive but less than hoped. 30% said it was too early to tell, and 8% said returns were neutral, leaving just 8% who said AI spend had been a net negative. The split between AI services is shaken up again, with OpenAI back on top.
During the last survey conducted during the summer, Claude had overtaken ChatGPT and Google Gemini was also gaining. In this edition, usage of Claude and Gemini both fell, as the information's readers signed back up with OpenAI. These three vendors are still in the clear lead, with between 65% and 78% of readers using their services. The next grouping includes Grok, Microsoft Copilot, and Perplexity, each between 25% and 29%. The survey didn't include a specific section on recently released personal agent use, but the information did find that OpenClaw use is still at 12%. Fascinatingly, this is the high watermark for OpenClaw among subscribers of the information so far this year, which means despite that hype train having died down quite a long time ago, the product itself continues to resonate. Among other agentic coding tools, 45% use ClaudeCode, 33% use Codex, and cursor usage has doubled in recent months to reach 23%. The information noted huge changes from a year ago. This time last year, 59% of readers said that they had never used an agent. The use of open-weight models is also climbing,
with 72% of survey respondents now using open-weight models with some regularity. 20% said that open-weight models now drive the majority of their organization's workload, 10% said usage was about even with closed models, and 42% said open-weight models are deployed to a minority of workloads. Interestingly, although 46% of respondents said that they were using open-weight models for their work, only 21% said that they were using Chinese open-weight models, suggesting that most organizations, even these very forward startup-type organizations, are still a little hesitant to rely on foreign models. Finally, the information found little evidence that AI adoption has shifted hiring patterns. Only 22% said that they were hiring less, 10% said that they were hiring more, and 21% said that this was too early to tell. This left almost half saying that AI hasn't altered their hiring practices at all. Of course, this is all just a small snapshot of a highly enfranchised set of AI users, but still provide some really interesting patterns nonetheless. One update that is getting a lot of chatter on social media In a change to their terms of service Anthropic will now ban users for being too mean to Claude According to the new policy users are prohibited from quote sustained and needless abusive or cruel behavior
Anthropic said that the policy is only meant to apply in extreme cases and should not interfere with, quote, common versions of user frustration and pushback. Thank goodness, because the number of times I find myself asking Claude or GPT what the heck is wrong with it in much more colorful language would most certainly get me banned. Now, the discussion from there mostly went into, AI model welfare and AI model consciousness, which is way beyond the scope of the headlines. But for a takeaway from the AI realist perspective, Box's Aaron Levy writes, This sounds weird, but it is actually probably a good policy. Even if you don't believe AI is conscious, I don't, it stands to reason that you don't want future models trained on endless content of humans being rude to models. The models only understand the data they've been trained on, or what they run into in their interactions. So if you want safe and aligned models, we probably want AI to have lots of good interactions in their training data. For those just here for the features, a more interesting update for Claude was the introduction of Claude dashboards and Claude Motion. Dashboards allow users to connect Claude to a dataset such as a CRM or accounting software and generate a custom dashboard for visualization of the data.
Users can also query the data in natural language to gather further insights. The dashboards are also configured to automatically update as new data comes in. Now, this kind of functionality has been technically possible for a while, but Anthropic is packaging it as a simplified feature accessible through simple prompts. Motion, meanwhile, allows users to transform the data into short animations suitable for presentations. The animations are written in code rather than video frames, so you can easily modify words and numbers without needing to start over with each edit. This seems to be the practical use case for the flashy animations highlighted in the release of Opus 5.5. Along with OpenAI's introduction of Intelligent UI earlier this week, we're seeing a massive expansion in how AI can be used to communicate and interact with data. The folks in Anthropic, meanwhile, suggest pushing this feature as hard as you can, with Robert Bayh posting, Try giving it bigger briefs than you think it can handle. It will seriously impress you. And indeed, early users are certainly impressed by what they see. Drew Fallon of Iris Finance made an animated product demo, commenting, A year ago, this would cost tens of thousands of dollars. Insane.
If you haven't yet decided what you were going to spend your hacking time on this weekend, Claude Motion seems like a pretty good place to play. For now, though, that is going to do it for today's headlines. Next up, the main episode. A new study from KPMG in the University of Texas at Austin found that when people work with AI, similar skills don't guarantee similar outcomes. Researchers studied more than 500 early career professionals and found that the best performers consistently amplified the value of AI by guiding, evaluating, and refining its outputs. These top performers, called AI amplifiers, weren't defined by what they knew alone, but by how they worked with AI. Learn more about what separates AI amplifiers from everyone else at kpmg.com slash us slash AI amplifiers. Every episode, I talk about the competition between OpenAI, Anthropic, SpaceX AI, Google, and Meta. And if you've been listening for a while, you might have a favorite. Maybe you think OpenAI and Anthropic can stay ahead,
or perhaps Meta's open source strategy can win out. Whatever your view, every AI lab creates a different investment opportunity. Harbor Capital Advisors AI Lab Ecosystem ETF Suite lets you invest in the ecosystem behind the AI Lab you believe in. Search Harbor AI Lab Ecosystem ETFs wherever you invest or follow at Harbor Capital on X to learn more. Visit harborcapital.com for a prospectus containing investment objectives, risks, fees, expenses, and other important information. Read and consider it carefully before investing. Risks include principal loss and artificial intelligence-related risks. Harbor ETFs are distributed by Foresight Fund Services, LLC. Harbor is not affiliated with AI Daily Brief, and the funds are not affiliated with, sponsored by or endorsed by any AI lab. This is a paid advertisement and not personalized investment advice. Investing involves risk, including possible loss of principle. The best teams don't have a single star carrying everyone else. They know their own strengths and each other's weaknesses and play to both. That's the team Robots and Pencils has built on purpose. Nobody there is grinding through busy work to pad a headcount number. People come for the hard problems and they stay because everyone around them is leveling up at the same time. In a market full of companies
that are just trying to hire fast, that's worth a look. Check out robotsandpencils.com slash careers. Blitzy's understanding of massive code bases unlocks autonomous security fixes, modernization, and new features. So what happens when there's no legacy code at all? Greenfield is supposed to be the easy part. Clean slate, no technical debt. But even Greenfield moves at human speed one sprint at a time. Blitzy changes the unit of work from the developer to the project, autonomously planning, building, testing, and validating entire applications from scratch. Hundreds of thousands of lines of production-ready code. One Blitzy customer stood up a brand new application, 534,000 lines of code, compressing a 65-week roadmap into two weeks. Another shipped an entire application with no front-end engineer. Legacy or Greenfield, the answer is the same. Software at the speed of compute. Build what's next at Blitzy.com. That's B-L-I-T-Z-Y dot com. Welcome back to the AI Daily Brief. Earlier this week, OpenAI released a tome of pioneering mathematical proofs that completely
upended the academic field. The results were published on the GitHub repo and included 372 novel results and 722 supporting papers. OpenAI said that the results were generated by an internal frontier model that hasn't been released to the public, likely the same model that produced the Navier-Stokes solution last month, which kicked off a massive round of discussion. This time around, OpenAI coordinated the release with their newly appointed Mathematics Advisory Committee. Some of their suggestions included releasing reasoning traces alongside the proofs, which OpenAI has done for a sample of 10 problems. The committee also suggested it would be useful to know the resources committed to each problem. In that regard, OpenAI disclosed that the average results used to compute equivalent to roughly three hours of ChatGPT Pro thinking. But the real story here is how this release changes the field of mathematics. in one fell swoop OpenAI has solved a significant chunk of the most difficult outstanding problems in the field we did have a few progressive warning shots over the past year last summer models from Google and OpenAI achieved gold level performance in the International Math Olympiad the top contest for high school mathematicians Then earlier this year models from OpenAI and Anthropic disproved several Erdos conjectures
which are famous in long-standing problems in geometry. Then last month, OpenAI released a proof for the Navier-Stokes equation, an almost 200-year-old problem in fluid dynamics. Navier-Stokes is one of seven Millennium Prize problems which have a million-dollar prize attached. Only one has been solved by a human, and nearly making significant progress on one is enough to guarantee a Fields Medal, often compared to a Nobel Prize. If you are currently in your head hearing Robin Williams and Stellan Skarsgård argue about a Fields Medal, in Goodwill Hunting, you are not alone. In any case, the Navier Stokes problem was one of the most significant scientific results produced by an AI model. And yet, it was just a tiny precursor for what OpenAI released this week. The best way to understand the scope of OpenAI's proofs is to refer to a list of the top 500 open problems in mathematics produced by ProofAtlas. This is an AI-generated list, but it's a decent representation of the most well-known problems in the field. OpenAI fully solved 90 of the top 500 problems. There's also dozens of partial proofs that would qualify as significant contributions to the field.
Among them were partial solutions to two additional Millennium Prize problems, the Riemann Hypothesis and the Hodge Conjecture. For the purposes of this episode today, I'm not going to get deep into what these problems actually are, as they all involve postgraduate-level theoretical mathematics that is basically impenetrable to a layman. Just to make that point, the Riemann hypothesis is arguably the most important problem in number theory. It states that the Riemann zeta function, which concerns the distribution of prime numbers, has zeros only at even integers and complex numbers with a real component of one half. Instead, it's easier to gauge how big a moment this is by observing how professional mathematicians reacted. Alex Konturovich, a Rutgers professor, wrote, Quasi-Riemann hypothesis? Are you kidding me? If a human did this, it would be an instant fields medal. No questions asked. Indeed, most of the initial takes came from mathematicians that have been relatively positive about AI's contribution to the field. Darren Litt from the University of Toronto framed this moment as being about mathematicians shifting from working on problems for decades to verifying proofs produced by AI. He commented,
Fun! Looks like mathematicians have a lot of exciting work to do. If I understand correctly, one of the results is a very special case of a conjecture of mine. which is also a consequence of stronger work in progress by a student of mine. Mathematician Micah Warren wrote, Historically speaking, in baseball terms, today would be the day that Barry Bonds blasted his 756th home run. If you would have told Bonds this in 1987, it would have been unbelievable. But in 2007? Still, there was a sense of awe from those who had worked at the intersection of AI and mathematics over recent years. Anthropic researcher Levin Dalpoji, who had been working on Navier-Stokes before OpenAI released their result, called this Obviously the most significant moment in mathematical history. Former OpenAI researcher Acer said the quasi-Riemann result was his personal Move 37 moment for mathematics. Move 37 refers to AlphaGo's unintuitive move that helped it defeat Go world champion Lee Seidel in 2016.
Even for those without mathematical training, the results were fairly mind-blowing. Responding to his first read-through the results, Chubby on X wrote, The list is absurd. A zero-free half-plane for the zeta function, which is the first result of its kind in over a century. Hilbert's tenth problem over the rationals, the Hodge conjecture for CM Abilene varieties, irrationality at Catalan's consent, and dozens more. Any one of these would normally be a career. But the number that many aren't seeing is the following. It's three. That's the average hours of ChatGPT Pro Compute per result. A month ago, Navier-Stokes took them around 10,000 agents in 88 hours. That efficiency gain is within just a few weeks. Math Twitter obviously is shocked. Again, this is literally the intelligence explosion happening right now. 2027 will be the year of superintelligence. I'm now convinced of that. Still, one big wrinkle in the release is the difficulty in understanding exactly what's been published. The Navier-Stokes result published last month is still in the process of being verified by mathematicians, and this is hundreds of new results to pour over.
Mathematician Francesco Maggi pointed out the obvious issue with high-volume AI mathematics, writing, Mathematics does not simply grow by accumulating correct statements. Results have to be understood, connected, explained, challenged, reused. Until that happens, they risk remaining lettera morta. Now imagine adding 400 more AI-generated proofs. For each of them, until some human mathematician picks it up, studies it, and connects it to the existing mathematical culture, isn't it in essentially the same position? So my question is genuinely, what is the intended mathematical value of releasing hundreds of proofs at once? What happens if there simply aren't enough people willing or able to read them? To be clear, I think mathematicians should discover mathematics by any means necessary, including scavenging through artificial alien-looking mathematical material, since there may be extraordinary things to find there. But humans are not machines, as attention, understanding, taste, and mathematical culture are scarce resources. What happens if the rate at which mathematics is generated becomes much greater than the rate at which mathematicians can absorb it? Physicist Steve Su took it a step further, asking, What happens when a single day of AI research produces more mathematics than humanity can
absorb in a century? Imagine millions of superhuman research agents at work. Formal systems may verify the proofs, but humans won't have enough context to understand the underlying web of machine-invented concepts. In his view, we could quickly reach a period where fields like mathematics and physics lose interpretability. At that point, he continued, AI is no longer a scientific tool. humans are receiving selected explanations from a much larger intellectual civilization they cannot independently grasp. Another big strand of the conversation was speculation on what this could mean for mathematics as a field of academic study. Dan McCadier wrote, OpenAI is downplaying this for PR reasons. If you're a mathematician, you must feel like a nuclear bomb hit and you're at ground zero. Reasoning models are two years old. In that time, they went from incapable of basic arithmetic to solving problems humans couldn't solve for decades. Math is only the beginning. AI will revolutionize the entirety of the human scientific endeavor. Most of us don't appreciate what that means. And you can feel a bit of that shell shock that Dan is talking about, especially among mathematic students still trying to find their place in the world Mostly harmless grad student asked What can I pivot to now that AI has taken away math They later explained that this release quote
kind of nukes basically all the problems and projects I'd want to contribute to or work on in my program. And yet, if there were many lamenting over the apparent end of a field, some mathematicians were incredibly optimistic. Shriram Kanaan, a former math professor at the University of Washington, wrote, It's a great day for humanity and math. If you asked a mathematician anywhere from 5000 BC to January 2026 whether they would like a math genie to exist, 100% of mathematicians would have said yes. Mathematicians give up significant opportunity costs to help advance human knowledge. It's a no-brainer that they will eventually adjust to this new world and be glad this happened. It's just an extreme but temporary shock to I am world class at math as an identity. Another very different topic of conversation was how results in any of these theoretical mathematics problems apply to the real world. The Erdos problems are mostly just intellectual exercises with no clear application. Even the Navier-Stokes proof, while very impressive, doesn't really have an obvious impact in the real world. Engineers working on fluid dynamics problems like aircraft design get by just fine using a simplified version of Navier-Stokes.
Most of the problems that OpenAI solved in this batch are similar. Very impressive, long-standing, but ultimately not all that practical. Still, there were a few examples that could make an impact in the real world. One was a dramatic improvement to the theoretical efficiency of matrix multiplication, the math that underpins AI inference. Cornell math professor Stephen Strogatz commented, Many staggering results here, but this is a particularly amazing one. The exponent for matrix multiplication is no more than 2.25. The previous world record had been something like 2.37. This leap in progress is like Bob Beeman's long jump. For some, what was most interesting was not what was there, but what was missing. Bitcoin security researcher Justin Drake noted a, quote, striking underrepresentation of cryptographic breakthroughs among the 722 mathematical results OpenAI published. I've witnessed firsthand the U.S. government censoring academic quantum cryptanalysis results. Backroom interventionism is my base case. University of Texas computer science professor Scott Aronson added on his blog, My sources tell me that AI companies have now started, gingerly and discreetly,
investigating whether their latest internal models can break important cryptographic protocols and primitives. If they can, then it would certainly be nice to get ahead of things before the rest of the world figures out the same. In other words, if OpenAI's internal model can break major cryptographic schemes, that's not just a problem for Bitcoin. It's a problem for the cryptography that underpins pretty much all of the digital world. Still, with all of this, maybe the most resonant conversation was whether we can expect the same thing to happen to other fields. Arpit Gupta, an associate professor of finance at NYU Stern, wrote, You're deluding yourself if you think AI advances like in math aren't coming to economics and other social sciences. Warden economics professor Ethan Malik is already starting to see it, commenting, After a brief but institution-eroding slop science era, it increasingly looks like we are going to have two revolutions from AI. One, everything ever published will be re-read and re-judged in ways that human scientists never anticipated. it. Two, novel discoveries will start to come fast. I am seeing rapid increases in the ability
of AI to do novel work in my field of economic sociology, with nearly autonomous research getting to top journal level. Still, a big question is whether all other fields are able to be approached in the same way as mathematics. By and large, these math proofs were not generated by reasoning into novel ideas. They were largely within the domain of problems that can be solved with big data approaches coupled to pattern recognition. It would be reductive to call them brute force, but these are all verifiable problems that fit neatly into the reinforcement learning paradigm. Pedro Domingos, a computer science professor at the University of Washington, summed it up by commenting, mathematics is a classic example of more of X paradox, hard for humans, easy for machines. Unpacking his roadmap for where this is all going, Google distinguished scientist Payman Milenfarr wrote, speed of automation is in inverse proportion to the entropy of the subject domain. Math and coding are structured and orderly, which is why AI is able to do them so well. In order of difficulty, the hard sciences are next.
They may be governed by strict physical laws, but are complicated by dependence on noisy measurements and laboratory experimentation. Financial markets are more difficult still. They are complex, adaptive systems subject to human psychology. While past data is abundant, the underlying statistics are non-stationary and highly variable. Clinical diagnostics and decision-making in medicine are far more difficult than most people realize. It's very high stakes, it requires synthesizing very fragmented signals, and dealing with idiosyncratic human behavior in biology. It's a nightmare of incomplete information. The law, jurisprudence, and legal interpretation operate on structured text and precedent most of the time, but rely heavily on ambiguous language, evolving societal norms, persuasive rhetoric, and human judgment. They won't be automatic anytime soon. Culture, art, and strategic leadership are an odd mix, but I think all of them are quite safe from automation. They are the pinnacle of disorder and human creativity. They rely on tacit knowledge, intuitive perceptions of the world, emotional intelligence, understanding of cultural consensus, and navigating complex moral trade-offs.
So rejoice! We'll have artists, lawyers, doctors, and CEOs for a long time yet. In fact, I think from here, there are two incredibly fascinating paths to watch. One is what Payman was speculating on, which fields this sort of disruption comes to in what ways, or alternatively shows us that there are human or systemic inertia road bumps that make this disruption not as inevitable in other areas. But the second thing that I think will be interesting to watch is what mathematicians do with all of this. To use the words of the previous commenter, what happens when every mathematician has a mathematical genie? Does it negate all of their hard work and learning? Or does it allow them to do things that are undreamed of as of yet? Over the coming months, that is what we will begin to see. For now, that is going to do it for today's AI Daily Brief. I appreciate you listening or watching as always. And until next time, peace.
番組の概要欄(原文)
<p>OpenAI’s latest mathematical results are forcing researchers to confront what happens when AI can tackle problems that once defined entire careers. With hundreds of proposed proofs awaiting scrutiny, the debate spans scientific discovery, professional identity, and which fields could face similar disruption next. In the headlines: confusion over OpenAI’s revenue numbers rattles markets, a new survey tracks shifting AI adoption, and Anthropic introduces Claude Dashboards and Motion.</p><p><strong>Brought to you by:</strong></p><p><strong>KPMG</strong> – Research from KPMG and the University of Texas at Austin shows the highest-impact AI users treat AI like a reasoning partner — and those skills can be taught at scale. Learn more at <a href="https://kpmg.com/us/Sophisticated" rel="ugc noopener noreferrer" target="_blank">https://kpmg.com/us/Sophisticated</a></p><p><strong>Granola</strong> - The AI notepad for people in back-to-back meetings. Try it free <a href="http://granola.ai/brief" rel="ugc noopener noreferrer" target="_blank">granola.ai/brief</a> </p><p><strong>Harbor - </strong>Invest in the AI ecosystem. <a href="https://www.harborcapital.com/aidaily" rel="ugc noopener noreferrer" target="_blank">https://www.harborcapital.com/aidaily</a></p><p><strong>Section</strong> - Section turns AI investment into workforce transformation and ROI - <a href="https://www.sectionai.com/" rel="ugc noopener noreferrer" target="_blank">https://www.sectionai.com/</a></p><p><strong>Blitzy - </strong>Want to accelerate enterprise software development velocity by 5x? <a href="https://blitzy.com/" rel="ugc noopener noreferrer" target="_blank">https://blitzy.com/</a></p><p><strong>Robots & Pencils</strong> - Cloud-native AI solutions that power results <a href="https://robotsandpencils.com/" rel="ugc noopener noreferrer" target="_blank">https://robotsandpencils.com/</a></p><p>The AI Daily Brief helps you understand the most important news and discussions in AI. </p><p><strong>Newsletter: </strong><a href="https://aidailybrief.beehiiv.com/" rel="ugc noopener noreferrer" target="_blank">https://aidailybrief.beehiiv.com/</a></p><p><strong>Interested in sponsoring the show? </strong>sponsors@aidailybrief.ai</p><p><br /></p>
関連エピソード

AIを形づくる5つの論争:収益とインフラ、利用層、主権、規制、データセンター
The AI Daily Brief: Artificial Intelligence News and Analysis

エージェントはどう決断するのか:Goodfireのエリック・ビゲロウが語るクリティカルトークン、相転移、文脈内学習
"The Cognitive Revolution"

進化は「ランダム探索」ではない:Akarsh Kumar が語る人工生命・ASAL・Core War
Machine Learning Street Talk (MLST)

Beam:米国発の大規模オープンモデル ― ReflectionAI共同創業者兼CEO Misha Laskinが語る
No Priors: Artificial Intelligence | Technology | Startups

OpenAIのセキュリティ担当「モデルの制御は今や地獄」――AI Explained クロスポスト
80,000 Hours Podcast

ChatGPTの生みの親が語る、AIは世界をどう理解するのか|Ilya Sutskever
Eye On A.I.