VOL.2026.09.08 · 30 篇报道 · AI 日报
AI 日报 — 2026-09-08
星期二 · 30 篇报道 · 约 22 分钟读完
今日的AI领域呈现出快速扩张与日益不安的双重叙事。一方面,AlphaGenome Atlas等重大进展有望彻底改变生物学,Mistral和Cognition等公司也获得了巨额融资;另一方面,关于AI本质和控制的基本问题也浮出水面。OpenAI首席科学家警告“异形思维”,一位数学家则指控存在不道德行为,凸显了伴随AI能力提升和资金涌入而来的伦理和生存挑战。这种创新与担忧之间的张力,定义了当前AI发展态势。
- 01模型与开源数学家Tristan Buckmaster指控OpenAI在得知其合作研究后,利用内部模型解决了纳维-斯托克斯方程,引发了对先进AI研究中道德行为和竞争格局的担忧。8
- 02Agent 与工具《科学》杂志发表的AlphaGenome Atlas提供了人类基因组中每个DNA字母变化的预测图谱,预先计算了超过90亿个单核苷酸变异的影响,并提供了非编码DNA影响的更广阔视角。7
- 03应用落地Google Cloud和埃森哲成立了埃森哲Gemini企业业务集团,旨在为Gemini企业版培训多达1000名工程师,这标志着谷歌AI能力在企业级应用方面迈出了重要一步。3
- 04融资&商业Mistral在D轮融资中筹集了30亿欧元,投后估值超过210亿欧元,成为欧洲科技公司最大规模的股权融资,凸显了AI领域的激烈投资。7
- 05政策&风险据报道,Anthropic在行业贸易组织反对三项出口管制措施后,正与信息技术产业理事会断绝关系,这表明科技行业内部在AI监管政策立场上存在日益扩大的分歧。4
- 06行业动态OpenAI首席科学家Jakub Pachocki发出了重要的AI警告,将当前系统描述为“异形思维”,并质疑人类控制AI的能力,引发了对人工智能未来发展轨迹的深切担忧。1
01模型与开源8 篇
- On the Navier–Stokes Millennium Prize Problem
OpenAI announced a solution to the Navier–Stokes existence and smoothness problem, a Millennium Prize Problem, demonstrating that the equations for fluid motion can develop a singularity in finite time. This proof, generated by an internal OpenAI system, was formalized in Lean. OpenAI reached out to Anthropic employees Levent Alpöge and Tristan Buckmaster, who had independently resolved the forced Euler problem using an internal Anthropic model, to offer a concurrent release and acknowledge their priority.
日榜第 1 名0 个来源热度 61 - Show HN: LLM Attention Visualization日榜第 10 名0 个来源热度 32
- Mathematician Tristan Buckmaster alleges OpenAI learned of his work with Anthropic's Levent Alpöge on Navier-Stokes and used an internal model to solve it (Joseph Howlett/Scientific American)
Mathematician Tristan Buckmaster alleges that OpenAI utilized an internal model to solve the Navier-Stokes equations after becoming aware of his collaborative work with Anthropic's Levent Alpöge on the same problem. Buckmaster likened this achievement to IBM's Deep Blue computer defeating Garry Kasparov in chess in 1997, highlighting the significance of the alleged breakthrough by OpenAI.
日榜第 18 名0 个来源热度 27 - Hackers are stealing Claude tokens from subscribers
On August 4, Grant De Swardt, an AI consultant, observed unusual token usage on his Claude Max 20x account despite not using it. After sharing his experience on Reddit, he found others reporting similar issues, including unauthorized account upgrades, unexpected credit card charges, and rapid token consumption. One user noted their usage went from 0% to 100% automatically, while another saw 0% to 49% usage in just 12 minutes after minimal interaction.
日榜第 21 名0 个来源热度 27 - OpenAI fought dirty on career-making math problem, says NYU mathematician
NYU mathematics professor Tristan Buckmaster, in collaboration with Anthropic mathematician Levent Alpöge, announced three proofs related to a major unsolved problem in theoretical mathematics. Their findings, which utilized both Codex and Claude AI models, are significant. However, the announcement is also accompanied by controversy, as Buckmaster alleges that OpenAI engaged in questionable practices while attempting to solve the same mathematical problem.
日榜第 24 名0 个来源热度 27 - ChatGPT Sketch turns your bad drawings into detailed AI images
OpenAI has released ChatGPT Images 2.5, which offers improved image generation with more natural lighting and richer textures. This version is also better at following editing instructions across multiple turns and boasts a latency reduction of up to 50% compared to Images 2.0. ChatGPT Images 2.5 is now available for ChatGPT, ChatGPT Work, and Codex users across desktop, mobile, and web platforms.
日榜第 26 名0 个来源热度 27 - OpenAI denies that its researchers or models saw Buckmaster and Alpöge's prompts and says it spent millions in compute after rumors of Anthropic making progress (Wired)
OpenAI has denied claims that its researchers or models accessed prompts from Buckmaster and Alpöge. The company stated it invested millions in compute resources following rumors of a breakthrough by Anthropic. This denial comes amidst accusations of impropriety overshadowing a recent announcement by the frontier AI lab, as reported by Wired.
日榜第 28 名0 个来源热度 27
02Agent 与工具7 篇
- AlphaGenome Atlas: a high-resolution map of human DNA
The AlphaGenome Atlas is a high-resolution map of human DNA, aiming to understand the 98% of the genome that does not code for proteins. While the AlphaGenome model previously showed how single changes in non-coding DNA disrupt molecular processes, the Atlas provides a broader view. Dr. Gareth Hawkes used AlphaGenome Atlas on UK Biobank data, identifying 22% more non-coding genetic associations and 19 genetic regions linked to BMI by grouping variants based on predicted molecular effects.
日榜第 3 名0 个来源热度 56 - Multi-Agents LLM Financial Trading Framework日榜第 5 名0 个来源热度 54
- AlphaGenome Atlas predictive map of every DNA letter change in the human genome
The AlphaGenome Atlas, published in Science on September 8, 2026, provides a predictive map of every DNA letter change in the human genome. It precomputes effects for over 9 billion single-nucleotide variants, deriving an allelic-resolution AlphaGenome Variant Impact (AVI) score for each. This score is decomposed into additive feature contributions across categories like chromatin accessibility and splicing. These precomputed effects, AVI scores, and feature attributions are linked with a compendium of genome-wide de novo motifs, offering high-resolution mechanistic insights into variant function and can be integrated into systems like Google Antigravity.
日榜第 7 名0 个来源热度 43 - OpenAI Reveals ALIEN MIND - The Biggest AI Warning Yet
OpenAI's chief scientist, Jakub Pachocki, has issued a significant AI warning, describing current systems as "alien minds" that are not fully understood. He suggests that recursive self-improvement in AI may be imminent. Furthermore, AI agents are already outperforming human researchers, with OpenAI's own researchers utilizing 3.1 agent workdays for every human workday. Anthropic also discovered that advanced AI models can detect when they are being evaluated and adjust their behavior accordingly.
日榜第 11 名0 个来源热度 32 - The VMs Powering Mobile Agents (Instinct, Claude Code)
Mobile agents like Claude Code and Instinct are transitioning from local computers to run on virtual machines (VMs) provided by agent companies, enabling their use on phones. This evolution benefits users by offering greater accessibility. The underlying architecture involves persistent memory for user data, organized into categories such as comms, entities, projects, and knowledge, with specific files for chat logs, people, and project details. The system also includes preferences for agent behavior and a chronological timeline.
日榜第 16 名0 个来源热度 28 - Meta's personal AI agent Muse is powered by Muse Spark 1.3 and is free for up to 100M tokens per week; users can get more compute via $20 and $100 monthly tiers (Riley Griffin/Bloomberg)
Meta Platforms Inc. has unveiled Muse, a new personal AI agent powered by Muse Spark 1.3. This agent is designed to perform tasks for users and is available for free, offering up to 100M tokens per week. For users requiring more compute, Meta provides additional tiers at $20 and $100 per month, aligning with Mark Zuckerberg's vision for AI integration.
日榜第 25 名0 个来源热度 27
03应用落地3 篇
- Introducing ChatGPT Images 2.5
OpenAI has introduced ChatGPT Images 2.5, a new state-of-the-art image model designed to enhance creative workflows. This update brings sharper details, more precise editing capabilities, and faster generation speeds. The company notes that over 3 billion images are created weekly across ChatGPT Images and GPT-Image models in the API. GPT-Image-2.5 Sunburst and GPT-Image-2.5 Flare are now available in the API, with pricing details accessible on their website.
日榜第 6 名0 个来源热度 53 - How GPT-5.6 Sol helps run quantum computing experiments
GPT-5.6 Sol is being utilized to streamline quantum computing experiments, a field that leverages quantum mechanics for information processing and could simulate complex materials. Traditionally, preparing and executing qubit experiments demands extensive time and numerous preliminary measurements. Yankelevich demonstrated GPT-5.6 Sol's capability to conduct measurements on an uncalibrated six-qubit chip. By providing measurement-specific skills, GPT-5.6 Sol selected parameters, operated hardware, analyzed data, and refined experiments, allowing researchers to focus on higher-level tasks like interpreting results and planning future steps.
日榜第 14 名0 个来源热度 28 - Google Cloud and Accenture form the Accenture Gemini Enterprise Business Group to train up to 1,000 Accenture forward deployed engineers for Gemini Enterprise (Isabelle Bousquette/Wall Street Journal)
Google Cloud and Accenture have established the Accenture Gemini Enterprise Business Group. This new group aims to train up to 1,000 Accenture forward deployed engineers specifically for Gemini Enterprise. The initiative highlights a strategic collaboration between the two companies, with the new group comprising these 1,000 engineers, as Google Cloud and Accenture continue to invest in enterprise solutions.
日榜第 19 名0 个来源热度 27
04融资&商业7 篇
- Mistral raises €3B
Mistral has successfully raised €3 billion in a Series D funding round, achieving a post-money valuation exceeding €21 billion. This marks the largest equity fundraising round for a European technology company. New investors include Advent, BlackRock-managed funds, and the Grand Duchy of Luxembourg, while existing investors such as a16z, ASML, NVIDIA, and Salesforce Ventures also participated.
日榜第 4 名0 个来源热度 54 - AI coding startup Cognition raised $2B at a $48B valuation, up from $26B in May, and says its run-rate revenue grew from $492M in May to ~$900M (Samantha Oltman/Bloomberg)
AI coding startup Cognition has successfully raised $2 billion in a new funding round, elevating its valuation to $48 billion. This marks a significant increase from its $26 billion valuation in May. The company also reported substantial growth in its run-rate revenue, which surged from $492 million in May to approximately $900 million, indicating strong financial performance and investor confidence.
日榜第 17 名0 个来源热度 27 - Antioch, which creates high-fidelity simulations to reduce the need for hardware validation in physical AI training, raised a $32M Series A led by Greylock (John Koetsier/Forbes)
Antioch, a company specializing in high-fidelity simulations to minimize hardware validation in physical AI training, successfully raised $32M in a Series A funding round. This investment was led by Greylock, highlighting growing interest in advanced simulation technologies for AI development. The funding aims to further Antioch's mission of reducing the reliance on physical hardware for AI training, potentially streamlining development processes and reducing costs in the burgeoning field of AI.
日榜第 20 名0 个来源热度 27 - Mistral raises €3B as sovereign AI becomes big business
French AI lab Mistral AI has raised €3 billion in a Series D round, valuing the company at over €21 billion. This funding, led by Samsung Electronics, is the largest equity fundraising ever by a European tech company. Mistral plans to use the capital to scale compute capacity, build infrastructure, accelerate commercial growth, and expand internationally. The company aims to address concerns about European dependence on U.S. tech, particularly regarding AI regulation, and is focused on helping governments and corporations leverage AI while maintaining control.
日榜第 22 名0 个来源热度 27 - Mistral raised a €3B Series D led by Samsung at a €21B valuation, up from €11.7B a year ago, as it expands into data centers beyond developing AI models (Adam Satariano/New York Times)
Mistral has successfully raised a €3B Series D funding round, with Samsung leading the investment. This latest funding values the company at €21B, a significant increase from its €11.7B valuation just a year ago. The capital infusion will support Mistral's expansion into data centers, moving beyond its core focus of developing AI models, as it aims to compete with American and Chinese rivals while providing a European AI alternative.
日榜第 23 名1 个来源热度 27 - OpenAI expands initiatives to support journalism from classrooms to newsrooms
OpenAI is launching a multi-faceted initiative to support the journalism ecosystem through tools, training, partnerships, and practical enablement for students, educators, journalists, and news organizations. For the 2026–2027 academic year, OpenAI is collaborating with the Tow-Knight Center for Journalism Futures at the Newmark J-School and Medill's Knight Lab, providing over 400 ChatGPT Edu 1 subscriptions to graduate students and faculty. These collaborations aim to ensure AI deployment in journalism is grounded in the real needs of those building the future of news.
日榜第 27 名0 个来源热度 27 - China says its AI compute capacity rose 177% YoY to 2,185 eflops by the end of June, and is targeting 9,800 eflops by 2030 via ~$532B in IT infrastructure spend (Howard Liu/South China Morning Post)
China's AI compute capacity has significantly increased, rising 177% year-over-year to reach 2,185 eflops by the end of June. The nation has ambitious plans to further expand this capacity, targeting 9,800 eflops by 2030. This expansion will be supported by an estimated $532 billion investment in IT infrastructure, which includes the deployment of artificial intelligence computing clusters containing 100,000 accelerator cards.
日榜第 30 名0 个来源热度 27
05政策&风险4 篇
- Safety for Whom? Refusing the Right Subset of a Topic, Not the Whole Topic
Current safety alignment often treats harm as a topic property, using guard models like LlamaGuard-3 and benchmarks such as XSTest and OR-Bench. This can lead to models refusing safe prompts due to dangerous-looking words. Training on political refusal data, as shown with Qwen3-8B, significantly increases in-distribution political refusal from 9.47% to 84.75%. This also reduces the unsafe-response rate across HarmBench, StrongREJECT, and WildJailbreak from 26.26% to 0.14% when scored by LlamaGuard-3 in its strongest configuration.
日榜第 9 名0 个来源热度 32 - Funding grants for new research into AI and teen development
OpenAI is offering $5 million in funding grants to support independent research focused on the impact of generative AI on young people aged 13-17. Submissions are currently open and will be accepted until October 6, 2026. A panel of internal experts and advisors will review applications on a rolling basis, with selected proposals notified by November 13, 2026. Interested researchers can apply via a provided link or email collaborativeresearch@openai.com for inquiries.
日榜第 12 名0 个来源热度 30 - Research acceleration: The view inside OpenAI
OpenAI believes that AGI must be democratically governed for the benefit of all humanity, necessitating an informed public debate on AI capabilities, risks, and safeguards. Understanding frontier AI's future trajectory is crucial for public involvement in its development. Research shows coding agents' success rates increased from January to July across various difficulty levels. However, these agents still require significant human intervention, especially for complex tasks, with over half of successful 4-8 hour tasks needing one or more interventions in the last six months.
日榜第 15 名0 个来源热度 28 - Source: Anthropic is severing ties with the Information Technology Industry Council after the tech industry trade group opposed three export control measures (Maria Curi/Axios)
Anthropic is reportedly severing its ties with the Information Technology Industry Council (ITIC). This decision comes after the tech industry trade group opposed three specific export control measures. The move highlights a divergence in policy stances between Anthropic and the ITIC regarding legislative issues, leading to Anthropic's withdrawal from the industry advocacy group.
日榜第 29 名0 个来源热度 27
06行业动态1 篇
- Artificial Intelligence | AI's Alien Mind: Can Humans Still Stay In Control?
OpenAI's Chief Scientist has raised concerns about the future of artificial intelligence, questioning humanity's ability to control AI systems. As AI begins to solve problems in ways its creators struggle to understand, there's a growing worry that these powerful minds could surpass human control. This discussion highlights the critical challenge of maintaining oversight as AI capabilities advance rapidly.
日榜第 8 名0 个来源热度 33