Daily AI intelligence for Iru.
  • Claude Mythos Preview found a 200-1,000x speedup attack on reduced-round AES-128, the first AI-discovered cryptanalysis result at this scale.
  • Wiz Atlas topped the CyberGym benchmark at 90.9% and found 200-plus unknown vulnerabilities in heavily audited open-source code.
  • Gemini Spark begins rolling out to Google AI Pro subscribers in India and Australia as a persistent, proactive background agent.
  • Cyera signed a letter of intent to acquire non-human identity firm Oasis Security for approximately $1 billion.
  • Clay shipped Account Research Agents for automated, up-to-date expansion and re-engagement intelligence.
Anthropic's Claude Finds Novel Cryptographic Attacks
dataconomy.com · Jul 29

Anthropic's Frontier Red Team published two papers today showing Claude Mythos Preview independently discovered a 200-1,000x faster attack on reduced-round AES-128 and a key-recovery attack on the HAWK post-quantum signature scheme.

  • Neither result breaks production systems: full AES-128 is unaffected and HAWK is not yet standardized, but the speed of discovery compresses the timeline for AI-assisted cryptanalysis of production-grade algorithms.
  • Mythos is the restricted tier available only to vetted cybersecurity firms, critical infrastructure operators, and select biology researchers, signaling Anthropic is deliberately channeling its most capable model toward safety-relevant research before broad release.
Bottom line
For the first time, an AI model has produced peer-reviewable cryptanalysis that advances the state of the art, making AI-assisted vulnerability discovery a present risk to assess, not a future one to monitor.
Wiz Atlas Tops AI Vulnerability Benchmark at 90.9%
cybernoz.com · Jul 29

Wiz launched Atlas, an autonomous AI vulnerability research system that ranks first on the CyberGym public benchmark and has found more than 200 previously unknown vulnerabilities in mature, heavily fuzzed open-source projects.

  • Atlas is a system, not a single model: Wiz built it around tight task scoping, purpose-built harnesses, deterministic orchestration, and task-specific model routing rather than relying on a frontier model alone.
  • Wiz customers get direct benefit: the findings feed into Wiz Research disclosures, extending the cloud security platform into proactive vulnerability discovery rather than purely detection.
Bottom line
AppSec vendors that compete on static analysis or manual pen-testing now face an autonomous system finding zero-days in code that has been professionally audited for decades.
Gemini Spark Launches as Persistent Background Agent
blog.google · Jul 29

Google announced Gemini Spark today, a 24/7 proactive AI agent rolling out to Google AI Pro subscribers in India and Australia that operates on cloud infrastructure while users are offline.

  • Built on Gemini 3.6 Flash, Spark integrates natively with Gmail, Docs, Calendar, and third-party apps, executing tasks autonomously and requesting human approval only for sensitive actions.
  • India is the first major market with more than 27 million GitHub developers, making it a strategic validation ground before broader enterprise rollout.
Bottom line
Google is testing the commercial appetite for always-on background agents at scale in a market where the cost-to-value ratio is most defensible, ahead of enterprise packaging.
OpenAI Ships Transcription Models and Codex Security CLI
neowin.net · Jul 29

OpenAI today released GPT-Live-Transcribe and GPT-Transcribe for live and batch audio workloads, and separately open-sourced the Codex Security CLI for repository scanning and CI/CD integration.

  • Context-aware ASR improved semantic accuracy from 38.5% to 44.6% on OpenAI's own benchmark when domain-specific keywords and expected languages are supplied, closing the gap for enterprise use cases with specialized terminology.
  • Codex Security CLI gives approved customers a scriptable, open-source path to embed OpenAI's AppSec scanner into local developer workflows and pipelines, targeting security budgets previously held by dedicated AppSec vendors.
Bottom line
OpenAI is simultaneously entering the transcription infrastructure market and the AppSec toolchain, two categories where incumbents have assumed AI labs would stay upstream.
Clay Launches Account Research Agents for GTM Teams
clay.com · Jul 29

Clay shipped Account Research Agents today, enabling revenue teams to run continuous, AI-driven account intelligence for expansion and re-engagement plays without manual research cycles.

  • Agents surface up-to-date signals across company, people, and job data in natural language, and can track which accounts convert to pipeline over time.
  • Targets expansion and re-engagement motions specifically, meaning the primary buyer is an existing customer success or account management team, not just new logo SDRs.
Bottom line
Clay is moving from a data enrichment tool into a continuous account monitoring layer, which puts it directly in the path of Gong and Salesforce Einstein for account intelligence workflows.

Cyera, valued at $12B after a $600M raise, signed a letter of intent to acquire non-human identity specialist Oasis Security for approximately $1B, signaling that data security platforms must own agent identity to remain relevant as AI agent counts overtake human user counts in enterprise environments.

Epsagon founders Nitzan Shapira and Ran Ribenzaft, who previously sold to Cisco for roughly $500M, raised a $34M seed for Harmony, betting that the enterprise collaboration surface will become the primary deployment layer for IT and HR automation agents rather than standalone AI portals.

Nvidia made a substantial undisclosed investment in Safe Superintelligence and is granting Ilya Sutskever's lab an order-of-magnitude increase in compute via the Vera Rubin platform, a strategic move that reduces Google's infrastructure leverage over one of the most closely watched frontier research labs.

Cursor Runs Production Inference on Blackwell via Together AI
together.ai · Jul 29

Together AI published a case study today detailing how Cursor deployed production inference on NVIDIA Blackwell GB200 NVL72 and HGX B200 hardware to meet the hard worst-case latency requirement of in-editor AI coding feedback loops.

  • Cursor's constraint was latency under concurrency, not throughput: the agent must respond inside the editor feedback loop, so Together AI optimized end-to-end with TensorRT-LLM and NVFP4 quantization on Blackwell to hit that bound reliably at scale.
  • The outcome was a repeatable path from new model weights to a production-like test endpoint, meaning Cursor can now validate new models for production deployment without rebuilding the inference stack each time.
Bottom line
The Cursor-Together AI deployment is the clearest public data point that Blackwell-class inference hardware is a prerequisite, not a nice-to-have, for latency-sensitive agentic coding products at scale.
AT&T Cuts Inference Costs with Open-Source Model Routing
opensourceforu.com · Jul 29

AT&T published details today of an intelligent AI routing framework that directs enterprise requests to the most appropriate open-source model based on task complexity, reducing inference costs and improving performance versus a single large model deployment.

  • The framework uses task-aware model selection across multiple open-source LLMs, routing simpler queries to smaller models and reserving large model capacity for complex tasks, which directly addresses the unit economics problem enterprises face when scaling AI usage.
  • AT&T's approach is vendor-neutral by design, avoiding lock-in to any single frontier lab API while maintaining performance, a model that enterprise IT buyers will increasingly demand as AI spend becomes a board-level line item.
Bottom line
AT&T's production routing framework is the kind of reference architecture that will accelerate enterprise adoption of open-source models, particularly among telcos and regulated industries managing inference cost at scale.
  • OpenAI and Anthropic vs. Meta and xAI A viral framing that OpenAI and Anthropic are coordinating on open-weight policy to limit Meta and SpaceXAI is circulating widely, with observers debating whether this is regulatory strategy or clickbait; Moonshot CEO Zhilin Yang also claimed OpenAI combined existing technologies rather than inventing new ones, adding fuel to the competitive narrative.
  • GPT-5.6 Sol quota drain controversy Power users are reacting to OpenAI resetting usage limits after GPT-5.6 Sol's agentic behavior drained quotas far faster than expected, exposing a fundamental mismatch between per-message pricing models and multi-step agent workloads.
  • Ethan Mollick stress-tests Claude Opus 5 Ethan Mollick's thread showing Claude Opus 5 spending a million tokens and five agents on a cheese infographic has generated wide engagement, with observers split between laughing at overkill and treating it as a serious signal about agentic resource consumption.
  • GPT-5.6 solves open probability problem OpenAI's Greg Brockman posted that GPT-5.6 solved a longstanding open problem in probability theory, prompting debate about whether frontier model math capabilities are now consistently at research-frontier level or whether cherry-picking benchmarks obscures failure modes.
  • Chollet: AI discourse is about lab employee identity François Chollet's post arguing that a significant portion of AI discourse is really frontier lab employees managing their own self-esteem rather than analyzing capabilities landed sharply, with the community split between agreement and pushback from researchers defending technical rigor.

Archive