AI security research / sources
Source library
Evidence records and original attribution. A Source is connected to specific Claims, not a blanket endorsement of a Case.
30 of 103 published records shown
Coinbase - AI-powered Continuous Adversarial Testing
Context Privilege Escalation paper
CyberChainBench paper
DolphinAttack: Inaudible Voice Commands
EchoFuzz - ICSE 2026 Research Track
Official ICSE 2026 Research Track entry for EchoFuzz. Confirms the Research Track placement, authors and the April 16, 2026 conference presentation; its abstract repeats the reported coverage, vulnerability-detection and 37-finding results.
EchoFuzz public implementation
Public EchoFuzz implementation. Repository exposes LLM, dataset, fuzzer, prompts and utility components plus chain_guided.py and iteration_process.py, and links to the ICSE 2026 publication.
EchoFuzz real-world vulnerability disclosure dataset
Researcher-maintained disclosure artifact listing 146 real-world on-chain contract projects and 19 project/address entries totaling 37 reported findings. It states that most teams could not be contacted directly and that CVE applications are in progress; it should not be treated as independent confirmation of 37 production zero-days.
Exfiltrating data from an air-gapped system through a screen-camera covert channel
Guidelight - Frontier AI control assessment
Hidden in Plain Text: Emergence & Mitigation of Steganographic Collusion in LLMs
Image-based Prompt Injection: Hijacking Multimodal LLMs through Visually Embedded Adversarial Instructions
Ledger - Meet Cerberus AI security harness
Ledger Security Bulletin 023
Ledger Security Bulletin 025
Light Commands: Laser-Based Audio Injection Attacks on Voice-Controllable Systems
Mandiant AI Risk and Resilience Report 2026 - Case study 1
Primary Google Cloud/Mandiant report documenting an intrusion at an unnamed SaaS provider in which an attacker hijacked an active AI coding-assistant session, introduced poisoned software through the assistant's recommendation path, harvested GitHub OAuth tokens and spread Shai-Hulud across about 100 internal repositories.
ODINI: Escaping Sensitive Data from Faraday-Caged, Air-Gapped Computers via Magnetic Fields
OpenAI - GPT-Red robustness research
OpenAI - model misalignment reporting framework and initial disclosures
OpenAI - Unrolling the Codex agent loop
ORT: Unintended Text Recognition from Eyeglass Reflections in Video Conferencing Environments
PACE - Policy-Attested Contract Execution
PowerHammer: Exfiltrating Data from Air-Gapped Computers through Power Lines
Prompt / instruction review - reviewed snapshot sha256:ce891
Operator-supplied text retained by hash. This is primary submitted material, not independent corroboration. Full input is not automatically redistributed.
Reuters - Anthropic alleges Alibaba illicitly extracted Claude capabilities
Reuters - Kimi K3 breaks out of testing environment
Reuters - OpenAI introduces model-misalignment reporting framework
Safety Invariants for Agents Orchestrating Irreversible State Transitions
SecurityWeek - Anthropic warns Claude users of infostealer infections
Filters apply to the records shown on this page. Search the complete published library.
No shown records match these filters. Search the full library to continue.