Currently free during beta - premium features coming soon. Subscribe now to lock in early access.
AI_SAFETY

EU Regulatory Changes

1550 changes tracked across 24 compliance frameworks including DORA, NIS2, GDPR, EU AI Act, Cyber Resilience Act, and more.

All DORA NIS2 GDPR CSRD MaRisk ISO27001 EU_AI_ACT CRA DSA DMA eIDAS2 SOC2 PCI_DSS HIPAA ISO42001 AMLD6 PSD3 DATA_ACT GPSR CER EUDR CVE BREACH AI_SAFETY
arXiv: InkShield: Writing Style Protection Against Unauthorized Handwriting Mimicry
This publication introduces InkShield, a novel technical framework designed to protect individuals from unauthorized handwriting mimicry by AI systems. The paper details a method that subtly alters...
Read analysis →
arXiv: What Does It Take to Detect an AI Agent? Minimal Feature Sets for Behavioral Detection under Browser Automation
A new preprint from arXiv, titled "What Does It Take to Detect an AI Agent? Minimal Feature Sets for Behavioral Detection under Browser Automation," published on July 29, 2026, presents a research ...
Read analysis →
arXiv: Defending Against Backdoor Attacks via Alignment Checking in Model-Contrastive Federated Learning
This paper, published on arXiv, proposes a new defensive technique called alignment checking for detecting backdoor attacks in federated learning systems. Backdoor attacks occur when malicious part...
Read analysis →
arXiv: ToxScreen: Detecting Whether an LLM Has Been Poisoned
A new preprint titled ToxScreen: Detecting Whether an LLM Has Been Poisoned has been published on arXiv, proposing a method to identify whether a large language model has been deliberately compromi...
Read analysis →
arXiv: Before Agents Speak: Pre-hoc Failure Risk Inference in Multi-Agent Systems
This paper, published on arXiv, introduces a novel framework called "Pre-hoc Failure Risk Inference" for multi-agent AI systems. Rather than detecting failures after they occur, the framework aims ...
Read analysis →
arXiv: Forecasting Trajectory-Level Safety Risks in Black-Box Multi-Turn Interactions
arXiv: SecRespond: Benchmarking AI Agents for Real-World Post-Compromise Incident Response
arXiv: Verifiable Random Sampling
arXiv: FARI: Robust One-Step Inversion for Watermarking in Diffusion Models
arXiv: Not In My Git Yard: Catching Backdoors at Commit and Release Time
arXiv: Graph Is the Verifier: Agentic Reinforcement Learning for Interprocedural Vulnerability Detection
arXiv: Fingerprint-Driven Automation: Coupling Reconnaissance with POC Verification
arXiv: Borrowed Strength: Best-of-N Search over a Code EncodingBreaks Self-Check Jailbreak Defenses
arXiv: Guarding Organizations Against Malware Risk: A Novel Graph-Based Malware Detection Method
arXiv: CDN Tsunami: Exploiting HTTP/3-HTTP/1.1 Conversion for DoS Attacks
arXiv: Recover, Decode, Reguard: Guard-Agnostic Defense Amplification againstEncoded VLM Jailbreaks
arXiv: Explicit Separations for One-Query Unitary Synthesis
arXiv: Collusion with Competitive Marginals: Price-Level Audits Are Blind by Construction
arXiv: QUIC-TRIP: A Triple-Redundant Journey Toward Secure Substation Communications
arXiv: A Controlled Candidate-Set Benchmark for Offline Satellite-Security Plan Decomposition