arXiv preprint · 2026
Christoph R Landolt, Tobias Lorenz, Marta Kwiatkowska, Mario Fritz
Models responsible AI release as a bilevel Stackelberg game, showing that defender welfare depends on the capability gap rather than the shared capability level — so for dual-use models the decisive lever is the sequencing of access, not the deploy-or-withhold threshold.
CyCon 2026 — 18th International Conference on Cyber Conflict, NATO CCDCOE · 2026
Christoph R Landolt, Julian Jang-Jaccard, Valentin Mulder, Roland Meier, Christoph Würsch, Mario Fritz
Implements autonomous offensive red-team agents with deep multi-agent RL in the NASim environment, showing that MARL agents learn coordinated attack patterns such as lateral movement and capture the flag, while identifying training stability at scale as the primary bottleneck.