Who Owns This Agent? Tracing AI Agents Back to Their Owners
cs.CR, cs.AI, cs.MA
Submitted: 2026-05-15
Updated: 2026-09-27
Code: https://github.com/unclecode/crawl4ai
Terminology
Sources
- Generative Language Models and Automated Influence Operations: Emerging Threats and Potential Mitigations
- An Overview of Catastrophic AI Risks
- Risks from Learned Optimization in Advanced Machine Learning Systems
- Llama Guard: LLM-based Input-Output Safeguard for Human-AI Conversations
- Building Production-Ready Probes For Gemini
- The Alignment Problem from a Deep Learning Perspective
- The Effects of Reward Misspecification: Mapping and Mitigating Misaligned Models
- Can AI-Generated Text be Reliably Detected?
- Smarter, Better, Faster, Longer: A Modern Bidirectional Encoder for Fast, Memory Efficient, and Long Context Finetuning and Inference
- Anatomy of an AI-powered malicious social botnet
- ReAct: Synergizing Reasoning and Acting in Language Models
- Watermarks in the Sand: Impossibility of Strong Watermarking for Generative Models
- Provable Robust Watermarking for AI-Generated Text
- Universal and Transferable Adversarial Attacks on Aligned Language Models
Related papers
- SoK: AI-Augmented Binary Reversing
- Relaxed Sender Anonymity for CBDC Interbank Settlement: A Zero-Knowledge Approach on Permissioned EVM
- Calibration-Family Overfit: Why Trusted Sabotage Monitors Don't Transfer Across Lineages
- Efficient Fuzzy PSI under One-Sided Assumptions
- Sealing the Audit-Runtime Gap for LLM Skills
- Token Composition: A Graph Based on EVM Logs