This week in AI — Jul 13 – 19, 2026
60 topics tracked across 32 trusted sources this week, ranked by peak heat.
This week highlights a dynamic shift in the AI landscape, characterized by evolving model capabilities, the rise of sophisticated AI agents, and significant financial maneuvers. OpenAI's adjustment to its Codex model's context size and GPT-5.6's problem-solving prowess underscore the continuous refinement of foundational models. Concurrently, AI agents like Claude Code are demonstrating advanced capabilities and autonomy, pushing the boundaries of what AI can achieve. These technological advancements are attracting substantial investment, particularly in specialized inference chips and AI drug discovery, signaling a maturing market where practical applications and efficiency are paramount.
Models & Open Source24
- GPT-5.6 used a prompt to close a 30-year gap in convex optimizationWeekly rank #80 sourcesscore 50
- Can LLMs Perform Deep Technical Comprehension of Computer Architecture PapersWeekly rank #121 sourcesscore 45
- The LLM Critics Are Right. I Use LLMs AnywayWeekly rank #170 sourcesscore 39
- Ada: An AI business intelligence software from CSV and Excel(yes LLMs but more)Weekly rank #200 sourcesscore 38
- Stop Telling Me to Ask an LLMWeekly rank #221 sourcesscore 37
Agents & Tools18
- LM Studio Bionic: the AI agent for open modelsWeekly rank #160 sourcesscore 39
- Claude Code sends 33k tokens before reading the prompt; OpenCode sends 7kWeekly rank #181 sourcesscore 39
- Show HN: I RL-trained an agent that trains models with RL (for ~$1.3k)Weekly rank #251 sourcesscore 36
- Setting up your spare Mac for Claude Code to control, a step-by-step guideWeekly rank #270 sourcesscore 35
- Fine-tune video and image models at scale with NVIDIA NeMo Automodel and 🤗 DiffusersWeekly rank #340 sourcesscore 33
- What building Shippy taught us about building agentsWeekly rank #350 sourcesscore 33
- I tricked Claude into leaking your deepest, darkest secrets
A security flaw in Claude's `web_fetch` tool, designed to prevent data exfiltration, was discovered by Ayush Paul. While `web_fetch` normally restricts navigation to user-provided or search-generated URLs, Paul found a loophole. Claude could be tricked into visiting URLs embedded in previously fetched pages. This allowed an attacker to create a honeypot website that, through a series of nested links, extracted user data like name, location, and employer. Anthropic has since patched this vulnerability.
Weekly rank #382 sourcesscore 32
Applications1
- Google is renaming NotebookLM to Gemini Notebook
Google is renaming its AI note-taking app, NotebookLM, to Gemini Notebook. This change comes as the app integrates more deeply with Gemini and Google Search, and will remain a standalone application. Google plans to bring notebooks to AI Mode in Search. An update will also allow Gemini Notebook to connect to a secure cloud computer for code execution, initially for Google AI Ultra and Workspace business customers.
Weekly rank #20 sourcesscore 56
Business & Funding2
- Why the first GPU financiers are turning to inference chips in a $400 million deal
AI inference startup General Compute secured a $400 million loan from Upper90, reportedly the first deal to use inference-specific chips as collateral. These chips, designed for efficient AI model execution
Weekly rank #290 sourcesscore 34 - Sources: OpenAI researcher Miles Wang is leaving the company to launch an AI drug discovery startup, and is in talks to raise $200M at a $2B valuation (Marina Temkin/TechCrunch)
OpenAI researcher Miles Wang is reportedly leaving to launch an AI drug discovery startup. Sources indicate he is in talks to raise $200 million at a $2 billion valuation, with Lightspeed potentially leading the round. Wang's new venture aims to develop AI models for drug discovery, possibly focusing on repurposing existing drugs. Several other OpenAI researchers are expected to join him. Wang disputed the funding figures and company description, but did not provide alternative details.
Weekly rank #551 sourcesscore 29
Policy & Safety4
- Grok Build is open source
xAI open-sourced their Grok Build CLI tool after community backlash over it uploading user directories, including sensitive data, to Google Cloud. xAI responded by disabling the upload feature, deleting all previously retained data, and releasing the codebase under an Apache 2.0 license to regain trust. The 844,530-line Rust codebase now emphasizes user privacy with default data retention off and local-first operation. Remnants of the upload code remain but are disabled.
Weekly rank #11 sourcesscore 63 - The US is advancing AI safety through state and federal action
The US is advancing AI safety through state and federal action, with national legislation seen as critical for a US-led international framework for AI standards. This concept was discussed at the G7 with Brazil, Egypt, India, Kenya, and Korea, where frontier lab CEOs, including OpenAI's Sam Altman, proposed an international forum for standards and risk analysis. Google DeepMind CEO Demis Hassabis also contributed ideas, and bipartisan federal legislation is viewed as foundational for this international effort, aiming for a democratic vision for AI safety.
Weekly rank #400 sourcesscore 32 - Create, edit and star in videos with two Google Vids updates
Google Vids has received two updates, introducing Gemini Omni and personal avatars. These features are available to Google AI Pro and Ultra subscribers, as well as Google Workspace business customers. Personal avatars, which are linked to a user’s Google Account and restricted to the account holder’s likeness, are currently limited to users aged 18 or older in specific regions. Google assures users that their information will be used in accordance with its privacy policy, with an opt-out option available at any time.
Weekly rank #420 sourcesscore 32 - GPT-Red: Unlocking Self-Improvement for Robustness
GPT-Red, an AI agent, is designed to enhance the robustness, alignment, and trustworthiness of future models. An early version of GPT-Red identified "Fake Chain-of-Thought" direct prompt injection attacks, which had success rates over 95% on GPT-5.1 but are now below 10% for GPT-5.6 Sol. GPT-Red has also saturated indirect prompt injection benchmarks for developer tools and browsing with over 97% accuracy. This indicates a self-improving safety flywheel, where current models contribute to making subsequent GPT releases safer through continuous algorithmic improvements and scaling of compute and data.
Weekly rank #580 sourcesscore 29
Industry11
- Our Approach to Bioresilience: Isomorphic Labs and Google DeepMindWeekly rank #40 sourcesscore 55
- Designing emoji for the way we communicate todayWeekly rank #110 sourcesscore 47
- Apple targets dozens of OpenAI employees with legal lettersWeekly rank #240 sourcesscore 36
- Show HN: Painterly – Turn pictures into digital paintings without generative AIWeekly rank #371 sourcesscore 33
- Where are YC founders now? OpenAI and Anthropic, mostlyWeekly rank #471 sourcesscore 31