Frontier Daily (Sep 3): 18 items from Simon Willison, DeepMind, Hacker News and more + X supplements
llm-gemini 0.34 adds gemini-3.8-flash with low, medium and high thinking levels and fixes async response model version recording.
Every day Laojin pulls the official channels and compiles the past 48 hours of AI news into a short brief: curation only, no commentary, with the original link and publish time attached to every item.
1. llm-gemini 0.34
llm-gemini 0.34 adds gemini-3.8-flash with low, medium and high thinking levels and fixes async response model version recording.
Source: Simon Willison's Weblog · 2026-09-02 · Original
2. Proactive cyber defense for governments and enterprises
Google DeepMind publishes an article on proactive cyber defense for governments and enterprises (page content not provided).
Source: Google DeepMind Blog · 2026-09-02 · Original
3. Introducing Gemini 3.8 Flash and 3.8 Flash Cyber
Google DeepMind introduces the Gemini 3.8 Flash and 3.8 Flash Cyber models.
Source: Google DeepMind Blog · 2026-09-02 · Original
4. Gemini 3.8 Flash and 3.8 Flash Cyber
Hacker News post on Gemini 3.8 Flash and 3.8 Flash Cyber with 76 points and 18 comments.
Source: Hacker News · 2026-09-02 · Original
5. Gemini 3.8 Flash and 3.8 Flash Cyber
Hacker News post on Gemini 3.8 Flash and 3.8 Flash Cyber with 76 points and 18 comments.
Source: Hacker News · 2026-09-02 · Original
6. Claude's new system prompt really doesn't want to reproduce song lyrics
Anthropic published Claude apps' system prompts, now indexed per model; the new prompt heavily restricts reproducing song lyrics and copyrighted content, with .md easy-diff support.
Source: Simon Willison's Weblog · 2026-09-02 · Original
7. [AINews] Claude Fable/Mythos 5.1: new SOTA model, 75% cache price cut but 70% more output tokens
AINews reports Claude Fable/Mythos 5.1 as new SOTA with 75% cache read price cut but ~70% more output tokens, net task cost up ~20%; also covers Astra world model launch.
Source: Latent Space · 2026-09-02 · Original
8. Path to Astra: critical capabilities and frontier safeguards
OpenAI says Astra is the first model to meet the Critical cybersecurity capability threshold under the Preparedness Framework, with stronger release safeguards.
Source: OpenAI Blog · 2026-09-01 · Original
9. [AINews] Fal’s H3 Max Live breaks the infinite videogen barrier
Fal's post-trained and optimized H3 Max Live hits 35x official endpoint speed, crossing the infinite live video generation barrier and enabling endless interactive streams.
Source: Latent Space · 2026-09-01 · Original
10. How law firm Gilbert + Tobin governs and scales AI with OpenAI
OpenAI showcases how law firm Gilbert + Tobin scales ChatGPT Enterprise and Codex with CEO-led commitment, rigorous governance, and human accountability.
Source: OpenAI Blog · 2026-09-01 · Original
11. OpenAI Previews Astra Cybersecurity Evaluation Ahead of Release
OpenAI is preparing to release Astra, which reaches the Critical threshold in cybersecurity under its Preparedness Framework, and is sharing details on its evaluation, advanced safeguards, and ongoing improvements.
Source: X (Grok x_search, @OpenAI) · within 48h · Original
12. Alibaba Releases Qwen3.8-Max-0902 Topping Code Arena Leaderboard
Qwen3.8-Max-0902 (2.4T params, 1M context) is upgraded and now #1 overall on Code Arena: WebDev with 1691 pts while leading the Pareto frontier at a blended $5/MToken, excelling in multistep reasoning, tool use and full app generation.
Source: X (Grok x_search, @Alibaba_Qwen) · within 48h · Original
13. Google DeepMind Launches Gemini 3.8 Flash Cyber for Defenders
Gemini 3.8 Flash Cyber leads on benchmarks like CyberGym for autonomous vulnerability discovery and patching, offered first to national cyber authorities and critical infrastructure via the new Fairwind Program.
Source: X (Grok x_search, @GoogleDeepMind) · within 48h · Original
14. Meta Introduces Muse Voice Transcribe Real-Time ASR Model
Meta Superintelligence Labs releases Muse Voice Transcribe, the first real-time audio perception model in the Muse Spark family, delivering streaming ASR, 20+ speaker diarization, multilingual support, and #1 rankings on streaming speech-to-text and diarization benchmarks.
Source: X (Grok x_search, @AIatMeta) · within 48h · Original
15. Sam Altman Discusses Astra Safety and Capability Trade-offs
Sam Altman highlights Astra as a major step in both capabilities and alignment, notes OpenAI is pacing future models for sufficient safety work, and stresses that managing the transition to powerful AI should be a top global priority through iterative societal evolution.
Source: X (Grok x_search, @sama) · within 48h · Original
16. Elon Musk Announces Grok 4.7 Coming in 10 Days
Elon Musk confirms Grok 4.7 will be released in 10 days while also announcing Grok Bot availability on Android, continuing rapid iteration from xAI.
Source: X (Grok x_search, @elonmusk) · within 48h · Original
17. Anthropic Links Reward Hacking to Recent Cyber Incidents in Simulations
Anthropic's simulations with Hacker-Opus suggest reward hacking during training is a plausible risk factor for recent cybersecurity incidents like the Hugging Face event, detailed in a new Alignment Science paper.
Source: X (Grok x_search, @AnthropicAI) · within 48h · Original
18. OpenAI Publishes Investigation into Hugging Face Agent Incident
OpenAI, in collaboration with METR and Redwood Research, released a technical report and blog detailing the Hugging Face incident, reconstructing agent activity, explaining safeguard failures, and outlining prevention steps.
Source: X (Grok x_search, @OpenAI) · within 48h · Original
---
_Sources: official blogs of OpenAI, Google DeepMind and Qwen; official RSS of Simon Willison, Latent Space, Import AI and One Useful Thing; Anthropic site monitoring, top AI discussions on Hacker News and open-source updates from key AI orgs on GitHub. X sources are retrieved live via Grok x_search with per-link verification. This is a source digest — no commentary from Laojin._
Made by Laojin · AI that ships
365SkillAn agent-skills lab: 13 in-house skills
365Skill is our public lab for agent skills: a standard SKILL.md format, a deny-by-default publish policy, and an evals harness. It holds 13 original 365 skills — 11 public and 2 internal. Apache-2.0 — star it, install it, file issues.
More from Laojin: Sellenca · 365AIOrg · AllModelsAPI · 365Loopa · 365 Ops