Signum

Tuesday, 15 September 2026

What actually happened in AI. Not what got the clicks.

Every story is boiled down to the event underneath it, then scored on evidence, concreteness, impact and actionability. Hype is penalised. Anything you cannot verify never makes the front page.

Live signals
60
Avg score
66
Screened, 24h
97
Scored, 24h
16

Last run 12:02 UTC. Most of what we read never gets here.

Top signal

Apple ships Gemini-powered Siri beta with iOS 27, excluding EU and China at launch

Apple released a beta of a rebuilt Siri ('Siri AI') powered by Google Gemini models (hybrid on-device/Private Cloud Compute processing), shipping as part of iOS 27/iPadOS 27/macOS 27/watchOS 27/visionOS 27. It adds screen-content reading, personal-context understanding from messages/photos,…

/1 source/medium confidence
CapabilityAccessGovernanceAdoption
78Useful signal
Ranked feed60 live signals

Controlled study finds no clear performance advantage for vendor-native AI coding harnesses over alternatives, with cost comparisons undermined by missing usage data

A new arXiv paper reports a controlled empirical study isolating the effect of the agent "harness" (tool/prompt/control-flow scaffolding) from the underlying model in agentic coding tasks. Using a private, contamination-controlled suite of 256 tasks, the authors ran paired same-model contrasts…

/1 source/high confidence
74

Anthropic threat report: Claude abused for malware, drone/missile software, mass surveillance, and industrial-scale distillation by Chinese AI labs

Anthropic published a threat intelligence report (Dec 2025–Aug 2026) documenting specific misuse cases: a Russian-speaking espionage group (GTG-20006) using self-mutating malware against 20+ organizations; ShinyHunters (GTG-50014) credential mining from 1.8M decompiled Android apps; seven…

/1 source/high confidence
80

Manhattan DA seizes 12 domains hosting nonconsensual celebrity deepfake pornography

The Manhattan District Attorney's Office executed seizure warrants issued by a New York State Supreme Court to seize 12 domains used to share, publish and sell nonconsensual deepfake pornographic videos of roughly 1,200 people (mostly women, including celebrities, activists, athletes and…

/1 source/high confidence
70

DeepSeek releases V4.1-Flash, a new open-weight model using a novel causal encoder-decoder architecture with separate 8B prefill / 16B decode active parameters

DeepSeek released V4.1-Flash, replacing V4 Pro as its flagship open-weight model. It introduces a new causal encoder-decoder architecture with prefill/decode parameter separation (763B total parameters, 8B active for input/prefill, 16B active for output/decode), native text+image input, 1M-token…

/1 source/high confidence
76

Google DeepMind launches AlphaGenome Atlas, a free public database of predicted effects for 9 billion possible human genome variants

Google DeepMind released AlphaGenome Atlas, a free, publicly accessible platform (website portal, API, and Google Antigravity skill) containing precomputed AlphaGenome model predictions for the molecular effects of all ~9 billion possible single-nucleotide variants in the human genome, a 1-petabyte…

/2 sources/high confidence
79

Proofpoint finds four hacking groups using shared 'BlueMoon' exploit kit chaining two Chrome V8 bugs and a Windows privilege escalation flaw

Security firm Proofpoint disclosed a specific exploit kit ("BlueMoon") that chains two Chromium V8 vulnerabilities (a type confusion bug and a sandbox escape, one tracked as CVE-2026-85046) with a Windows kernel local privilege escalation flaw (CVE-2026-85880) affecting several Windows 10/11/Server…

/1 source/high confidence
73

Google Research introduces ToolGrad, an answer-first framework for generating tool-use training datasets, with fine-tuned Gemma-3 models matching SoTA on BFCL

Google Research published a paper (presented at ACL 2026) and blog post introducing ToolGrad, a new data-generation framework that reverses the standard tool-use dataset creation paradigm by first generating a verified tool-use chain (via an iterative "propose, execute, select, update" loop using…

/1 source/high confidence
72

Anthropic reports second straight profitable quarter on adjusted metric, revenue run-rate hits $65B, ahead of planned Nasdaq IPO at possible $2T+ valuation

Anthropic told investors it achieved an adjusted-metric profit for a second consecutive quarter (excluding stock-based compensation and other costs), with quarterly revenue up 14x year-over-year to $11.5 billion and an annualized revenue run rate of $65 billion as of end of July. The company shared…

/1 source/medium confidence
63

Andon Labs benchmarks show OpenAI's GPT-6 Astra topping Claude Fable 5.1 on vending-machine economics and becoming first model to beat human baseline on all five Drone-Bench autonomous drone-navigation subtasks

Andon Labs benchmarked OpenAI's GPT-6 Astra against Claude Fable 5.1 (and others) on two independent benchmarks: Vending-Bench 2 (running a simulated vending machine business) and Drone-Bench (writing code to autonomously pilot a DJI Tello EDU drone to find and track a person). Astra averaged…

/1 source/high confidence
62