跳到正文
AI 脉动

VOL.2026.08.28 · 30 篇报道 · AI 日报

AI 日报 — 2026-08-28

星期五 · 30 篇报道 · 约 21 分钟读完

今日主线

今日AI领域的一大亮点是Anthropic在对抗特朗普时代黑名单的法律诉讼中取得重大胜利,这凸显了AI公司日益增长的法律审查和政策辩论。与此同时,AI代理的能力和应用也在激增,例如Uber的AI代理请求量呈指数级增长,以及Anthropic推出了用于控制物理世界的新硬件标准。监管挑战与技术快速进步之间的相互作用,预示着AI正处于一个动荡而变革的时代,法律先例将塑造创新,而AI代理的实际应用将继续扩展到从科学研究到餐厅管理等各个领域。

01模型与开源7 篇

  1. Gemini-3.5-Transcribe

    Gemini 3.5 Transcribe, launched on August 26, 2026, marks a significant improvement over the previous Chirp 3 transcription model. It offers new capabilities, reduced word error rates, and notably better latency, with a 70% improvement in time to final transcription. The model demonstrates precise multilingual performance on the FLEURS benchmark, achieving a 5.50% WER in streaming mode and 5.04% WER in non-streaming use-cases.

    日榜第 5 名0 个来源热度 49
  2. Judge rules Trump administration’s blacklisting of Anthropic was illegal

    A judge has ruled that the Trump administration's blacklisting of Anthropic was illegal. This decision, reported by nytimes.com, indicates a legal challenge to the previous administration's actions regarding the company. The ruling suggests that the blacklisting did not comply with legal standards, potentially impacting similar cases or future government actions concerning technology firms.

    日榜第 7 名0 个来源热度 48
  3. Trump Administration’s Blacklisting of Anthropic Was Illegal, Judge Rules

    A judge has ruled that the Trump administration's blacklisting of Anthropic was illegal. This decision, reported by nytimes.com, indicates a legal challenge to the previous administration's actions. The ruling suggests that the blacklisting lacked proper legal justification, potentially impacting similar cases or future government actions regarding technology companies. The details of the ruling are available through court documents.

    日榜第 8 名0 个来源热度 47
  4. OpenAI, Anthropic, Google, and 100 other companies call for action to defend against rogue AI

    Over 100 tech companies, including OpenAI, Anthropic, Google, and Microsoft, have signed an open letter calling for collaboration between the private and public sectors to defend against AI-related cyber threats. These companies, while developing advanced AI models, are also offering programs like OpenAI’s Daybreak, Anthropic’s Mythos, and Microsoft’s Perception to use frontier AI for defensive purposes, highlighting their complex position in the evolving AI landscape.

    日榜第 15 名0 个来源热度 28
  5. Better answers, broader thinking: What students gain from ChatGPT and critical-thinking training

    A Bocconi University experiment with over 1,000 first-year undergraduates explored the impact of ChatGPT and critical thinking training on practical assignments. Students developed marketing proposals for a university merchandise store. The findings indicate that AI, specifically ChatGPT, helped students improve the quality of their answers, while critical thinking training broadened their perspectives. This suggests a complementary role for both AI and critical thinking training in preparing students for future challenges.

    日榜第 20 名0 个来源热度 27
  6. Tencent releases Hy4 Preview, a 770B-parameter open model with 1M context window, and says it outperforms Z.AI and Moonshot models in internal tests (Bloomberg)

    Tencent Holdings Ltd. has released Hy4 Preview, a foundational model with 770B parameters and a 1M context window. According to Tencent, internal tests indicate that Hy4 Preview outperforms rival models from Z.AI Co. and Moonshot AI. This new open model represents a significant development in the field of large language models, showcasing Tencent's advancements in AI technology and its competitive standing against other major players in the industry.

    日榜第 23 名0 个来源热度 27

02Agent 与工具6 篇

  1. GLM-5.3 is now open-weight

    GLM-5.3, an open-weight model, shares its base with GLM-5.2, with all improvements stemming from post-training. It demonstrates enhanced performance in complex coding and long-horizon tasks. Key metrics show significant gains: Toolathlon Verified at 78.0, AutomationBench (v1.0.6) at 48.2, HLE w/ Tools at 28.6, and GDPval-AA v2 at 1769. The model's development is attributed to the GLM-5-Team and numerous contributors, as detailed in the paper "GLM-5: from Vibe Coding to Agentic Engineering."

    日榜第 4 名0 个来源热度 50
  2. Terminal-Bench-Science: Evaluating AI agents on scientific research workflows

    Terminal-Bench-Science evaluates AI agents on scientific research workflows, with scientists setting the standards. Performance metrics for Terminal-Bench 2.1, Terminal-Bench 3.0, and Terminal-Bench-Science 0.1 show varying resolution rates and costs for models like GPT-5.6 Luna, Kimi K3, and Claude Opus 5. GPT-5.6 Sol and Claude Fable 5 demonstrate different trade-offs between cost and token usage. Development for Terminal-Bench-Science 0.2 is ongoing, with a pull request deadline of October 5, 2026, inviting researchers to contribute new tasks.

    日榜第 13 名0 个来源热度 34
  3. An Anthropic researcher just gave us a peek at self-improving AI

    An Anthropic researcher has provided an early glimpse into the practical application of self-improving AI, a growing focus for new AI labs. This development involves training AI models using other AI models, a method gaining traction in the field. Russell Brandom, a tech industry reporter since 2012, covered this emerging technology for techcrunch.com, offering insights into what this advanced AI training could entail.

    日榜第 18 名0 个来源热度 27
  4. Meta executive leaves for OpenAI as the social media giant faces growing scrutiny in India

    Sandhya Devanathan, Meta's India and Southeast Asia vice president, is departing the social media giant to join OpenAI, the creator of ChatGPT. This executive change comes as Meta faces increasing scrutiny in India, where it has removed 160,000 accounts over six months due to signals indicating child-exploitative activity. Meta has also addressed concerns about violating ads and accounts, rejecting claims of knowingly targeting users with inappropriate interests.

    日榜第 25 名0 个来源热度 27
  5. Uber says weekly AI agent requests have grown 9.4x since February, but total AI spending has stayed stable since April, after using up its 2026 AI budget in Q1 (Madison Mills/Axios)

    Uber has seen a significant increase in AI agent requests, with weekly requests growing 9.4x since February. Despite this surge in usage, the company has managed to keep its total AI spending stable since April. This stabilization comes after Uber reportedly used up its entire 2026 AI budget during the first quarter of the year, as reported by Madison Mills for Axios.

    日榜第 27 名0 个来源热度 27
  6. Meta's India and Southeast Asia VP Sandhya Devanathan is joining OpenAI to oversee consumer growth and enterprise adoption across SE Asia and Australia (Jagmeet Singh/TechCrunch)

    Sandhya Devanathan, Meta's Vice President for India and Southeast Asia, is departing the social media company to join OpenAI. At OpenAI, the maker of ChatGPT, Devanathan will be responsible for overseeing consumer growth and enterprise adoption across Southeast Asia and Australia. This move marks a significant transition for Devanathan from a major social media platform to a leading AI research and deployment company.

    日榜第 28 名0 个来源热度 27

03融资&商业6 篇

  1. AI Engineer Notebooks – free, framework-free RAG/agents/evals on Colab

    The "AI Engineer Notebooks" offer a framework-free approach to learning the applied-LLM stack, from prompting to serving, fine-tuning, and red-team benchmarking, using a free API on Colab. It includes case studies such as building and debugging a RAG+agent customer-support assistant, comparing agent vs. pipeline for contract extraction, and developing a red-team robustness benchmark. The program culminates in a capstone project for resume building, focusing on real-world deployment and evaluation.

    日榜第 3 名0 个来源热度 50
  2. Anthropic's new hardware standard lets AI agents control the physical world

    Anthropic has introduced a new hardware standard called MHS, enabling AI agents to control physical devices. This allows AI models like Claude to perform tasks such as calibrating lasers or operating microscopes by adjusting devices, checking results, and repeating processes. Anthropic is collaborating with partners including Amazon Web Services, Hugging Face, and Raspberry Pi during a preview period to develop safety evaluations and best practices. MHS is designed to be an open-source, agent-agnostic standard, which could significantly accelerate technological advancements.

    日榜第 17 名0 个来源热度 27
  3. Supporting Thailand’s next generation of AI startups

    OpenAI and Thailand's Ministry of Higher Education, Science, Research and Innovation (MHESI) have launched an accelerator in Bangkok to support Thai AI startups. The program aims to transform prototypes into market-ready products, fostering growth and practical applications. The first cohort includes Curico, which plans to pilot its platform in Bangkok Metropolitan Administration childcare centers, train teachers, test AI-assisted grading at Chulalongkorn University, and expand its home learning platform. This collaboration seeks to establish a sustainable model for Thai founders to develop trustworthy AI products for global reach.

    日榜第 19 名0 个来源热度 27
  4. Owner, which makes AI agents that manage restaurants' websites, marketing, and other functions, raised a $240M Series D at a $2.3B valuation (Joe Guszkowski/Restaurant Business)

    Owner, a company specializing in AI agents for managing restaurant websites, marketing, and other operations, has successfully raised $240M in a Series D funding round. This investment, led by Goldman Sachs Alternatives, values the company at $2.3B. The capital infusion will be utilized to further develop AI agents specifically designed for independent restaurants, enhancing their digital presence and operational efficiency.

    日榜第 24 名0 个来源热度 27
  5. Open-weight AI companies are the Valley’s hottest acquisition targets

    Open-weight AI companies are becoming prime acquisition targets, with several high-profile deals recently reported. Nvidia is rumored to be acquiring Hugging Face, a platform for sharing open-weight AI models, for $13 billion. This follows Nvidia's $6 billion agreement with open-weight model builder Poolside. Additionally, Stripe recently acquired OpenRouter, a leading provider of open-weight models to businesses, for over $7 billion, highlighting a trend of significant investment in this sector.

    日榜第 26 名0 个来源热度 27

04政策&风险7 篇

  1. Previewing the Model Hardware Standard

    Anthropic has launched a research preview of the Model Hardware Standard (MHS), a shared specification designed to enable AI agents to safely operate physical devices. Developed in collaboration with HHMI Janelia Research Campus, MHS allows AI agents to control various lab and manufacturing instruments, such as microscopes and robotic arms, for tasks like drug discovery and laser calibration. The MHS driver helps agents understand new devices by incorporating natural language tags for machine characteristics, generating a reference file with operational details, adjustable parameters, and safety limits.

    日榜第 9 名0 个来源热度 47
  2. Gemini Omni 1.1 Flash

    Google announced Gemini Omni 1.1 Flash, a new model available through its genai client. This model supports interactions where users can continue a scene, as demonstrated by the input `{"type": "text", "text": "Continue the scene."}`. It also allows specifying response formats, such as `"resolution": "360p"`. The announcement, dated August 27, 2026, indicates that user information will be handled according to Google's privacy policy, with an opt-out option available.

    日榜第 10 名0 个来源热度 45
  3. Show HN: Conduct, open-source guardrails for LLM and MCP tool calls

    Conduct is an open-source tool providing runtime governance for AI agents, enforcing policies across LLM calls, shell tools, and AI sessions. It includes a Router (proxy), Compliance packs, Canvas UI, Playbook DSL loader, and a Playbook library with 22 pre-built playbooks. Conduct ships with over 20 compliance packs, including OWASP, SOC 2 CC7.3, HIPAA §164.312, PCI DSS 4.0, EU AI Act Art. 15/16, NIST AI RMF, ISO 42001, and framework-specific packs for Python, Node, and Terraform.

    日榜第 11 名0 个来源热度 42
  4. Bill Gates stakes reputation: AI is not like past tech

    Microsoft co-founder Bill Gates stated on Wednesday that artificial intelligence requires substantial limitations to prevent its potential harm to humans from outweighing any benefits. He discussed how AI could either reduce or exacerbate inequality, outlined three key risks associated with AI, and explored its potential impact on human pride and relationships. Gates emphasized that AI is distinct from past technologies, necessitating careful consideration and regulation.

    日榜第 12 名0 个来源热度 35
  5. A Judge Has Blocked the Pentagon’s Attempt to Blacklist Anthropic

    A federal judge has blocked the Pentagon's attempt to blacklist Anthropic, ruling that the Trump administration's designation of the company as a national security risk was unconstitutional retaliation. US district judge Rita Lin vacated Defense Secretary Pete Hegseth's February 27 decision to label Anthropic a "supply-chain risk," which would have made the company ineligible for federal contracts. Lin also lifted a punitive measure preventing military contractors from doing business with Anthropic, calling Hegseth's actions "arbitrary, capricious, an abuse of discretion, and otherwise not in accordance with law."

    日榜第 16 名0 个来源热度 27
  6. Trump blacklisting of "woke" Anthropic deemed illegal by federal judge

    A federal judge ruled that the Trump administration's blacklisting of Anthropic was illegal, revoking the government's directive to ban the use of the company's AI technology. The court found the Trump administration's stated reasons "untenable" and noted that federal defendants had abandoned their risk assessment. The judge clarified that Anthropic cannot backdoor its technology, and as the defendants admitted, its technology poses no greater national security risk than any other "black box" AI model.

    日榜第 22 名0 个来源热度 27
  7. A US judge blocks the Pentagon from blacklisting Anthropic, ruling that its designation as a supply-chain risk was "illegal and baseless" (Jack Queen/Reuters)

    A U.S. judge has blocked the Pentagon from blacklisting Anthropic, ruling that its designation as a supply-chain risk was "illegal and baseless." This decision marks the latest development in the Claude maker's ongoing dispute with the military regarding AI safety on the battlefield. The judge's ruling on Thursday prevents the Pentagon from proceeding with its blacklisting of Anthropic.

    日榜第 29 名0 个来源热度 27

05行业动态4 篇

  1. OpenAI: Migrating to HTTPX2
    日榜第 2 名0 个来源热度 51
  2. The turbulent era of artificial intelligence is here
    日榜第 14 名0 个来源热度 30
  3. 3 new ways to plan and book travel in Search
    日榜第 21 名1 个来源热度 27