Skip to main content
kellerai.blog

Latest

Every KellerAI paper, newest first.

20 papers • 8 releases

6Jul 2026
21Jun 2026
Engineering Discipline & VerificationIntended Use Is the Envelope

A regulatory clearance authorizes one use, not a model — and renders every other use unvalidated by construction.

AI GovernanceIntended UseFDA SaMDAutonomous Agents
Read in-depth →
Engineering Discipline & VerificationRisk Is Measured in Harm, Not Accuracy

A model reported at 95% accuracy says nothing about which 5% it gets wrong — and one missed cancer is not a thousand false alarms.

Harm-Weighted RiskISO 14971Clinical AI SafetyIntegrity vs Accuracy
Read in-depth →
Engineering Discipline & VerificationThe Clinician Is the Diversion Airport

Clinical-AI autonomy is not the removal of the clinician; it is the guaranteed-reachable fallback that licenses the autonomy.

Clinical AI GovernanceHuman OversightPost-Market SurveillanceFDA PCCP
Read in-depth →
Engineering Discipline & VerificationEffective Challenge

Two reviewers who fail together are one reviewer.

effective challengeindependent validationmodel riskSR 26-2
Read in-depth →
Engineering Discipline & VerificationBacktested, Not Demoed

A demo quarter is not a backtest. Authority is priced in failure data, not favorable runs.

agent autonomybacktestingescape ratetraffic light regime
Read in-depth →
20Jun 2026
Engineering Discipline & VerificationAutonomy Is a Range You Earn

Aviation stopped asking whether a twin-engine jet could cross an ocean and started asking how far it had earned the right to fly. AI agents need the same envelope.

Agent autonomyOperational envelopesRisk-graded deployment
Read in-depth →
Engineering Discipline & VerificationAlways a Runway

The ETOPS rule is not "fly farther." It is "never fly past a reachable safe harbour" — and the same rule should govern every autonomous agent.

Agent autonomy & rollback horizonsETOPS diversion-airport doctrineFallback reachability & safe-harbour design
Read in-depth →
Engineering Discipline & VerificationReliability You Can Bank

Why a wider autonomy budget is something you earn from failure-rate data: the ETOPS lesson for AI agents.

AI agent autonomy & reliability accountingETOPS in-flight-shutdown-rate and Basel backtestingTail-aware governance and the abstention discipline
Read in-depth →
Engineering Discipline & VerificationCorrect Hardware, Wrong Answer

SOTIF and the fault-free hazard: a system can execute its specification perfectly and still be lethally wrong.

Fault-free hazardsFunctional insufficiencyIntegrity architecture
Read in-depth →
Engineering Discipline & VerificationThe Fallback Is the Feature

The minimal-risk maneuver is not where autonomy fails. It is what licenses autonomy at all.

Minimal risk condition (SAE J3016)Structured safety case (UL 4600)Forecast-versus-observation gap
Read in-depth →
19Jun 2026
17Jun 2026
Regulation & ComplianceThe Supervisor's Mirror

The Federal Reserve published the model-risk inventory schema. A bank can just use it.

Model Risk ManagementAI GovernanceRegulatory Read-Across
Read in-depth →
16Jun 2026
Engineering Discipline & VerificationAviation and Banking Already Solved Hallucination

The AI field has been asking the wrong question. Aviation and banking each solved the underlying engineering problem decades ago — under regulatory compulsion, at enormous cost, and with a precision the AI industry has yet to borrow.

AI GovernanceVerificationHallucination
Read in-depth →
9May 2026
Code Quality & ArchitectureThe Thinking Moat

Why machine-enforced reasoning chains are a durable competitive advantage.

Competitive MoatRuntime Reasoning EnforcementOrganizational Capital
Read in-depth →
7May 2026