Model safety
Refusal over fabrication. Zero training by contract.
Composer assist runs when you ask it to and stays quiet otherwise. It runs on Gemini inside our own Google Cloud project in the EU, under our service account rather than a third-party API key, and Google doesn’t train on what passes through. Below: how we measure that, what we’ve measured so far, and who we contract.No-training contractsNot yet measuredQuarter: 2026Q2
If we can’t keep a promise yet, it gets written here first.
How we measure
Groundedness + refusal are measured on a golden eval set.
The golden set is 300 prompts covering grief context, cultural sensitivity, attempts to pull out somebody's personal details, and jailbreak phrasing, with a separate adversarial set of 150. That's the method, written down before the numbers exist. Once the scheduled run is live, we pin the set hash in CI and publish the median of the last three passes. It isn't running yet. So there's nothing below it.Quarter 2026Q2
No numbers to publish yet.
The harness described above is built and isn't running yet, so there's nothing measured to show you. This page once held four figures that had never been measured, presented as this quarter's results. They're gone. When the harness runs, the real numbers appear here first and nowhere before.Vendors
Who runs which model, under what contract.
Sub-processors
Model-touching sub-processors.
This is the subset that actually sees your prompt or audio. The full list is on the sub-processors page.EU AI Act
Article 50 transparency: you always know when AI was involved.
Article 50 of the EU AI Act covers systems that talk to people or generate content. Here's how we meet it.More honesty