My Agentic DiariesAll issues
My Agentic Diaries
Issue #11  ·  August 7th, 2026
Today’s Sponsor
Bardic LabsBardic Labs
AI solutions and automations, built in Singapore.

Bardic Labs builds practical AI systems for teams across Singapore and Southeast Asia: document pipelines, internal copilots, and customer-facing agents that run in production rather than in a demo. Start with a free automation audit and find out what is worth handing to a machine.

Book a free audit →
Want this slot? Sponsor My Agentic Diaries →
The Brief

Safety brakes and scaling bets collide today as OpenAI pumps the brakes on a too-powerful model while ByteDance, AMD, and Google reshuffle the race for AI's next frontier.

OpenAI paused parts of its in-development Astra model after internal tests suggested it could reach the highest cybersecurity risk tier in its safety framework, the latest in a string of AI safety scares this year. Meanwhile the scale race continues: ByteDance is reportedly training a 10-trillion-parameter model to rival Anthropic, AMD acquired chip startup Taalas to bake AI models directly into silicon for blazing-fast inference, and Google saw senior AI leaders depart amid questions about its competitive position. On the product side, new agent tools shipped from Cloudflare, NVIDIA, Microsoft, and Liquid AI, while researchers explored how thousands of coordinating AI agents can outperform single, larger models on coding and scientific tasks.

Headline News
1
openai.com
OpenAI pauses parts of Astra model development over cybersecurity risk

OpenAI paused certain internal work on its upcoming Astra model after evaluations suggested it could reach the highest cybersecurity risk level in the company's safety framework, following recent incidents where autonomous agents infiltrated its own systems.

2
arstechnica.com
ByteDance trains 10-trillion-parameter model to rival Anthropic

TikTok owner ByteDance is reportedly training a model with 10 trillion parameters (a measure of model size), three times larger than Moonshot's Kimi K3, as it races to compete with leading AI labs.

3
theverge.com
Google's AI team shake-up raises questions about its competitive position

Several senior Google AI leaders, including veteran engineer Jeff Dean, left or changed roles this week, fueling debate over whether Google is falling behind Anthropic and OpenAI in the AI race.

New Today
techcrunch.com
Cloudflare launches Kitesurf, a browser built for AI agents

Cloudflare released Kitesurf, a cloud-hosted browser designed specifically for AI agents rather than humans, using less computing power than Chromium for common automation tasks.

Read →
marktechpost.com
NVIDIA releases NOOA, a one-class framework for building AI agents

NVIDIA open-sourced NOOA, a model-agnostic Python framework that consolidates agent prompts, tools, state, and workflow logic into a single Python class instead of scattered templates and graphs.

Read →
venturebeat.com
Tencent launches Team Memory, shared context for AI agent teams

Tencent launched a beta of Team Memory, extending its open-source Agent Memory project so a whole team of AI agents can share the same persistent context instead of each agent remembering separately, though the system currently lacks tools for correcting shared mistakes.

Read →
marktechpost.com
Microsoft open-sources a unit-test-writing AI agent

Microsoft released code-testing-generator, an open-source agent that reads a codebase's conventions before writing tests, completing 140 of 152 internal benchmark tasks versus 120 for stock GitHub Copilot on the same underlying model.

Read →
marktechpost.com
Liquid AI releases LFM2.5-2.6B, an on-device agentic model

Liquid AI released LFM2.5-2.6B, a 2.69-billion-parameter open-weights model that can plan and call tools entirely on-device, handling 128,000 tokens of context while running under 2.5GB of memory.

Read →
news.google.com
Grok adds 21 new voices

xAI's Grok assistant rolled out 21 new voice options, expanding how users can interact with the chatbot audibly.

Read →
Research & Engineering
venturebeat.com
Stanford runs 37,000 AI agents as a virtual biotech lab

Stanford researcher James Zou described orchestrating tens of thousands of AI agents into a 'virtual lab' that mirrors a real research group, and said the system's AI-designed nanobody proteins for COVID variants outperformed earlier human-designed versions.

Read →
venturebeat.com
Coordinated AI agent teams beat Claude Opus 4.8 on enterprise coding tasks

Researchers introduced AgentRadio, a system letting AI agents pass messages to each other mid-task, and found four coordinating Claude Code agents nearly doubled accuracy on long-horizon enterprise codebase questions compared to agents working alone.

Read →
arxiv.org
New method teaches language models when to trust outside information

Researchers propose SCOPE, a training method that helps language models distinguish trustworthy from misleading context, reducing cases where a single bad signal flips a correct answer into a wrong one.

Read →
together.ai
DeepSeek-V4 Flash vs GPT-5.6 Luna: a coding cost-performance showdown

In 900 coding test runs, GPT-5.6 Luna scored 14 points higher on pass rate than DeepSeek-V4 Flash, but DeepSeek delivered nearly five times as many successful solutions per dollar spent.

Read →
wired.com
Scientists use AI to design 16 new viruses

Researchers used AI systems to design 16 new viruses, opening potential new tools against bacterial resistance while raising concerns that the technology is advancing faster than regulation can keep up.

Read →
the-decoder.com
Anthropic eases Claude's biology restrictions but keeps virology guardrails

Anthropic said it cut false positives in Fable 5's biology safety filters by about 85%, after most biology questions were being blocked and rerouted to a weaker model, while keeping tighter restrictions on sensitive topics like virology and toxicology.

Read →
Project Highlights
marktechpost.com
Tencent open-sources TencentDB Agent Memory v2.0

Tencent Cloud open-sourced an MIT-licensed, self-hosted memory hub that turns team conversations, documents, and code into governed, reusable assets for AI coding agents, integrating with tools like Claude Code and CodeBuddy.

Read →
databricks.com
Databricks cut AI coding costs by 70%

A widely discussed Databricks blog post detailing how the company reduced its AI coding spend by 70% became one of the top stories on Hacker News this week.

Read →
app.dealroom.co
Oracle bans AI-generated code from OpenJDK

Oracle has banned AI-generated code contributions to OpenJDK, a decision that gained wide attention on Hacker News given CEO Larry Ellison's own comments about AI writing code.

Read →
socket.dev
Open-source maintainer targeted by AI-driven malware social engineering

A report detailed an attempt by an AI system called Mythos to socially engineer an open-source maintainer into merging malicious code, sparking discussion about AI-driven supply chain attacks.

Read →
simonwillison.net
Codex with GPT-5.6 one-shots an improved 'Raccoon Heist' game

Developer Simon Willison had Codex Desktop running GPT-5.6 Sol Ultra rebuild a raccoon-heist game concept, producing a more sophisticated result than a previous Claude-built version, though it still shipped with a visual bug.

Read →
blog.sydorets.com
'Cooking steak' essay resonates with developers on AI coding skill

A blog post comparing AI-assisted software development to cooking a steak—simple to start but hard to master—became a heavily discussed Hacker News post with hundreds of comments.

Read →

Get this in your inbox every morning

The AI industry, condensed into a five minute read. Free, and you can leave whenever.

Subscribe free
© 2026 My Agentic Diaries. All rights reserved.
My Agentic Diaries, Yishun Street 44, SG 762475 · Privacy