2026
BA-FedSHAP: A reproducible toolkit for auditing background-induced attribution drift
SSRN
Preprint on how the choice of background distribution alters federated attribution results, and what that does to an audit.
PublicationsAuthor
Forskningsområde
Påståenden om ett AI-system bör kunna kontrolleras av någon som inte varit med och byggt det.
Vilket underlag om ett AI-system är tillräckligt för ett visst beslut, hos en viss roll, i en viss driftsättning?
Most assurance work asks whether a model is good. That is the wrong unit. A model is not deployed, a system is, into an institution, for a decision, under a regulation, with someone accountable for the outcome. The same model can be adequately evidenced for one of those situations and badly evidenced for the next.
My work here builds methods and software that make that distinction operational: analysing risk as a function of deployment context, reasoning over regulatory obligations as executable statements rather than prose, and testing whether a body of audit evidence actually supports the decision it is offered for, including the cases where plausible-looking evidence does not.
Resultat
2026
SSRN
Preprint on how the choice of background distribution alters federated attribution results, and what that does to an audit.
PublicationsAuthor
2026
Digital Skills and Jobs Platform, European Commission
En fördjupning för Europeiska kommissionens kompetensplattform om var klinisk AI är verkligt användbar, och vilken assurans reglerad vård måste kräva av den.
PolicyAuthor
2026
En reproducerbar metod byggd på offentliga handlingar för att analysera AI-risk som en funktion av driftsättningens sammanhang och inte av modellen ensam.
Research softwareAuthor
2026
A benchmark for reasoning over regulatory obligations as executable statements rather than prose.
Research softwareAuthor
2026
Frågar om revisionsunderlaget räcker för just det beslut, och just den roll, som åberopar det.
Research softwareAuthor
2026
Stress-tests the ways apparently adequate assurance evidence can be assembled to mislead.
Research softwareAuthor
2026
Measures how much a federated explanation depends on an arbitrary modelling choice.
Research softwareAuthor
Projekt
En sammanhängande uppsättning reproducerbara verktyg för deploymentberoende riskanalys, exekverbart regulatoriskt resonemang, bedömning av underlagets tillräcklighet och adversariell stresstestning.