AI Safety: Not Optional, Not Later

📅 2026-09-09
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
论文提出一种结合模型级监督和系统级控制的安全设计架构,以解决多层AI安全故障问题。
📝 Abstract
Incidents show that AI safety failures often arise across multiple layers. We present a safety-by-design assurance architecture combining model-level supervision, such as Scientist AI, with system-level controls over scaffolds and harnesses, independent verification, monitoring, and evidence infrastructure, supported by governance for accountability and evidence interoperability.
Problem

Research questions and friction points this paper is trying to address.

AI Safety
Safety Failures
System-level Controls
Innovation

Methods, ideas, or system contributions that make the work stand out.

safety-by-design
model-level supervision
system-level controls
governance for accountability