AI News selected for Professionals and Decision Makers
Model And Product Updates

Stability AI’s Annual Integrity Transparency Report

02:00 · September 17, 2025 · Stability AI News

Stability AI’s Annual Integrity Transparency Report

At Stability AI, we are committed to building and deploying generative AI responsibly, and we believe that transparency is foundational to safe and ethical AI.

Summary

Stability AI has published its Annual Integrity Transparency Report covering April 2024 to April 2025, setting out how the company applies safety-by-design principles across its image, video, 3D and audio models, including those offered through its API. The document describes measures taken at each stage of the development and deployment lifecycle to reduce the risk of harmful outputs, with particular emphasis on preventing the generation or distribution of child sexual exploitation material.

Training data is assembled from publicly available internet sources and licensed third-party datasets; no material is drawn from the dark web or adult sites, and paywalled content is excluded. In-house and open-source NSFW classifiers, together with hash lists supplied by Thorn and the Internet Watch Foundation, are applied to remove prohibited material. During the reporting period no instances of CSAM or CSEM were identified in the datasets used for current models.

Before release, every model undergoes structured red-teaming exercises conducted by internal and external experts. These evaluations use prompts depicting adult nudity and sexual activity as proxies to test for CSAM or CSEM generation capabilities. All models released in the period were subjected to this process, and none were found to produce such content. Where risks are identified, additional fine-tuning or safety LoRAs are applied. At the API level, real-time prompt filters, NSFW image classifiers and CSAM hash matching block violative inputs and outputs before generation occurs.

Content provenance is addressed through C2PA metadata embedded in all media produced via the API. The metadata records the model name and version and is cryptographically signed with a Stability AI certificate. Openly released models do not yet carry this metadata, an area the company states requires further development. Automated detection is supplemented by human review of flagged content, and any confirmed CSAM is reported to the National Center for Missing and Exploited Children; thirteen such reports were filed during the period. Users must be at least eighteen years old and accept the Acceptable Use Policy before accessing the technology.

The company reports partnerships with Thorn, the Internet Watch Foundation and the Tech Coalition’s Pathways program, and states that it continues to monitor regulatory developments and refine its risk-management processes.

Why it matters

This report is highly relevant for Dutch product teams and builders as it details the safety mechanisms, API filters, and C2PA provenance standards implemented in Stability AI models. Understanding these safeguards is crucial for building compliant, ethical AI applications that align with stringent EU and Dutch regulations.

More in this beat
ai-privacy-complianceC2PAmodel-security-controlsred-teamingstability-aitext-to-imagetrustworthy-ai-practices
InfoQ Opens AI Security & Privacy Engineering Cohort for Regulated Industries

14:00 · July 6, 2026

InfoQ Opens AI Security & Privacy Engineering Cohort for Regulated Industries

This article is highly relevant as it offers actionable training for security and privacy professionals dealing with AI in regulated environments. The curriculum directly addresses EU-relevant compliance and privacy needs, such as handling sensitive data and applying privacy threat modeling frameworks.

Relevance 85 · Audience 95

Janus: a Playground for User-Involved Agentic Permission Management

06:00 · July 3, 2026

Janus: a Playground for User-Involved Agentic Permission Management

This research is highly relevant to the Dutch AI market's strong emphasis on ethical, transparent, and privacy-compliant AI. By providing an open-source framework to test agentic permission management, it offers Dutch researchers and developers practical tools to align autonomous agents with strict EU data protection and AI regulations.

Relevance 85 · Audience 95

New Cryptographic Context Injection Attack Could Let Web Pages Steal Grok Chat Data

16:36 · August 20, 2026

New Cryptographic Context Injection Attack Could Let Web Pages Steal Grok Chat Data

This article highlights a critical data exfiltration vulnerability in LLMs via context injection, which is highly relevant for security professionals defending AI systems. Understanding this attack vector is essential for Dutch enterprises to ensure GDPR compliance and protect user privacy when deploying AI chatbots.

Relevance 85 · Audience 95

Google Launches Gemini 3.5 Flash Cyber AI to Find and Fix Software Vulnerabilities

17:09 · July 21, 2026

Google Launches Gemini 3.5 Flash Cyber AI to Find and Fix Software Vulnerabilities

This article is highly relevant for Dutch security professionals as it introduces a state-of-the-art AI tool for automated vulnerability discovery and patching. Given the strict EU regulatory landscape (like NIS2 and the Cyber Resilience Act), leveraging such AI capabilities will be critical for Dutch enterprises and government bodies to proactively secure software supply chains.

Relevance 90 · Audience 95

The Security-Privacy Imperative in the Age of AI Attacks

14:00 · July 19, 2026

The Security-Privacy Imperative in the Age of AI Attacks

Directly addresses AI security risks and privacy compliance under GDPR for EU-based professionals; offers actionable guidance on privacy-by-design techniques applicable to Dutch AI deployments and regulatory contexts.

Relevance 85 · Audience 90

Idiobionics: The Unification of Privacy and Intelligent Robotic Prostheses

06:00 · July 11, 2026

Idiobionics: The Unification of Privacy and Intelligent Robotic Prostheses

The article aligns strongly with the Dutch AI market's focus on ethical, transparent AI and healthcare innovation. It provides primary research on privacy vulnerabilities in AI-driven medical devices, which is highly pertinent for Dutch researchers navigating EU data protection standards (GDPR) and the AI Act.

Relevance 85 · Audience 95

Connecting security, privacy, AI governance to reduce information risk

11:08 · July 8, 2026

Connecting security, privacy, AI governance to reduce information risk

This article is highly relevant for Dutch security and privacy professionals as it provides a practical, standards-based approach to managing AI risks. Integrating these ISO frameworks aligns perfectly with EU GDPR and the EU AI Act requirements, enabling Dutch enterprises to ensure compliant and secure AI deployments.

Relevance 85 · Audience 95

More details on Fable 5’s cyber safeguards and our jailbreak framework

02:00 · July 2, 2026

More details on Fable 5’s cyber safeguards and our jailbreak framework

Provides actionable, specific guidance on model-level cyber safeguards and a structured jailbreak evaluation rubric directly usable by product teams building or auditing AI systems, with clear discussion of dual-use risks and deployment trade-offs.

Relevance 85 · Audience 80

OpenAI Previews GPT-5.6 Sol With Restricted Access and Stronger Cyber Safeguards

14:19 · June 27, 2026

OpenAI Previews GPT-5.6 Sol With Restricted Access and Stronger Cyber Safeguards

This article is highly relevant for security and privacy professionals as it introduces OpenAI's next-generation models featuring enhanced cyber safeguards. Understanding these new security mechanisms and the restricted rollout strategy is crucial for Dutch organizations preparing to integrate or audit future AI deployments under EU regulations.

Relevance 85 · Audience 90

RIFT-Bench: Dynamic Red-teaming For Agentic AI Systems

06:00 · June 24, 2026

RIFT-Bench: Dynamic Red-teaming For Agentic AI Systems

This research is highly relevant for Dutch AI practitioners and researchers focusing on AI safety and compliance with the EU AI Act. RIFT-Bench provides a scalable, unified framework for red-teaming autonomous LLM agents, which is critical for deploying secure and trustworthy AI systems in enterprise environments.

Relevance 85 · Audience 95