Currently free during beta - premium features coming soon. Subscribe now to lock in early access.
AI_SAFETY

EU Regulatory Changes

1476 changes tracked across 24 compliance frameworks including DORA, NIS2, GDPR, EU AI Act, Cyber Resilience Act, and more.

All DORA NIS2 GDPR CSRD MaRisk ISO27001 EU_AI_ACT CRA DSA DMA eIDAS2 SOC2 PCI_DSS HIPAA ISO42001 AMLD6 PSD3 DATA_ACT GPSR CER EUDR CVE BREACH AI_SAFETY
arXiv: On Understanding, Identifying, and Mitigating Vulnerabilities in Agentic Large Language Models
This publication, dated August 11, 2026, is a technical research paper from arXiv that analyzes security vulnerabilities specific to agentic large language models (LLMs)—systems that can autonomous...
Read analysis →
arXiv: Synthesizing Probabilistic Saturating Counters with Differentially Private Formal Guarantees
This publication introduces a new method for designing probabilistic saturating counters, which are hardware components used to track event frequencies, with built-in differential privacy guarantee...
Read analysis →
arXiv: Beyond Detection Accuracy: Measuring Explanation Cost, Stability, and Utility for Resource-Aware IoT Intrusion...
arXiv: Logit-Boundary Geometric Belief Interfaces and Sparse Sheaf-Enclave Protocols: A Self-Contained Substrate for ...
arXiv: From Prompt Injection to Web Exploitation: Revisiting Classic Vulnerabilities in LLM-Integrated Applications
arXiv: Withholding the Completing Chunk: Deterministic Pair-Completion Guardrails for Streaming LLM Output
arXiv: Generating Attacks for LLMs with GFlowNets
arXiv: MarkNull: Model-Agnostic Watermark Removal in AI-Generated Images via On-Manifold Latent Manipulation
arXiv: Generative AI for Encrypted Traffic Analysis: Synthetic Dataset Generation and Classifier Evaluation
A new academic paper, published on arXiv on August 10, 2026, proposes using generative AI to create synthetic datasets for training and evaluating machine learning models that analyze encrypted net...
Read analysis →
arXiv: ColluSkill: Adversarial Cross-Skill Composition for Evading Agent Skill Scanners
The publication introduces ColluSkill, a novel adversarial technique that demonstrates how malicious actors can evade AI agent safety scanners by composing multiple benign skills in sequence to exe...
Read analysis →
arXiv: Full-Key Recovery and Forgery from One MQOM v2.1 Signature
A new academic paper, titled "Full-Key Recovery and Forgery from One MQOM v2.1 Signature," has been published on arXiv. The paper demonstrates a practical cryptographic attack against the MQOM v2.1...
Read analysis →
arXiv: Activation Probes Surface Code-Security Signals that the Model's Output Misses
A new research paper, published on arXiv, demonstrates that analyzing a large language model's internal activations can reveal whether it is generating insecure code, even when the model's final ou...
Read analysis →
arXiv: Measuring the Wrong Thing: Internal Harmfulness Scores Anti-Rank Successful Jailbreaks
A new research paper, published on arXiv, challenges the reliability of internal harmfulness scores used to evaluate AI safety. The study demonstrates that these scores, which are often used to ran...
Read analysis →
arXiv: From Runnable to Verifiable: An Independent Reproducibility Study of LLM/Agent-Driven Vulnerability Validation...
This paper, published on arXiv, presents an independent reproducibility study of large language model (LLM) and agent-driven tools designed to automatically validate software vulnerabilities. The s...
Read analysis →
arXiv: Dual-Adversarial Safety Alignment: Cultivating Intrinsic Threat Comprehension in LRMs
A new academic paper, published on arXiv, proposes a novel technical method called Dual-Adversarial Safety Alignment to improve the safety of large reasoning models (LRMs). The paper argues that cu...
Read analysis →
arXiv: RangeFactory: Scalable Construction of Multi-Hop Cyber Ranges
The publication introduces RangeFactory, a technical framework for automatically generating multi-hop cyber range scenarios, which are realistic training environments for simulating complex cyberat...
Read analysis →
arXiv: STAIR: Effective Incident Response Using an End-to-End Agentic Planning Framework
A new research paper, titled STAIR, proposes an end-to-end agentic planning framework designed to improve incident response for AI systems. Published on arXiv, this is not a regulatory mandate but ...
Read analysis →
arXiv: Sound Enforcement of Dynamic Release Information Flow Policy-Full Version
A new academic paper, titled Sound Enforcement of Dynamic Release Information Flow Policy, has been published on arXiv. It proposes a technical framework for enforcing dynamic information flow cont...
Read analysis →
arXiv: ANTMAN: An Efficient and Interpretable RTL-Level Run-Time Detection Framework for Stealthy Branch Predictor At...
arXiv: ActBench: Self-Evolving Benchmark of Behavioral Safety in Cowork Agents