papersSEP 10 04:00 UTC
Paper studies measurable independence when AI models build, defend and test software
A new arXiv paper examines the situation where a single family of generative AI systems takes on three roles in software work: authoring application code, protecting and monitoring it, and hunting it for exploitable weaknesses. The authors introduce the notions of measurable independence and bounded autonomy to evaluate how separate these roles really are and how much freedom such systems should be granted. The work pushes back against the assumption that complete autonomy for these multi-role models is the right default.