2026
BA-FedSHAP: A reproducible toolkit for auditing background-induced attribution drift
SSRN
Preprint on how the choice of background distribution alters federated attribution results, and what that does to an audit.
PublicationsAuthor
Domaine de recherche
Les affirmations portant sur un système d'IA devraient pouvoir être vérifiées par quelqu'un qui n'a pas participé à sa construction.
Quelles preuves relatives à un système d'IA suffisent à une décision donnée, prise par une fonction donnée, dans un déploiement donné ?
Most assurance work asks whether a model is good. That is the wrong unit. A model is not deployed, a system is, into an institution, for a decision, under a regulation, with someone accountable for the outcome. The same model can be adequately evidenced for one of those situations and badly evidenced for the next.
My work here builds methods and software that make that distinction operational: analysing risk as a function of deployment context, reasoning over regulatory obligations as executable statements rather than prose, and testing whether a body of audit evidence actually supports the decision it is offered for, including the cases where plausible-looking evidence does not.
Productions
2026
SSRN
Preprint on how the choice of background distribution alters federated attribution results, and what that does to an audit.
PublicationsAuthor
2026
Digital Skills and Jobs Platform, European Commission
Une analyse approfondie pour la plateforme de compétences de la Commission européenne : où l'IA clinique est réellement utile, et quelle assurance les soins réglementés doivent en exiger.
PolicyAuthor
2026
Une méthode reproductible, fondée sur des documents publics, pour analyser le risque de l'IA en fonction du contexte de déploiement et non du seul modèle.
Research softwareAuthor
2026
A benchmark for reasoning over regulatory obligations as executable statements rather than prose.
Research softwareAuthor
2026
Demande si les preuves d'audit suffisent à la décision précise, et à la fonction précise, qui s'y appuie.
Research softwareAuthor
2026
Stress-tests the ways apparently adequate assurance evidence can be assembled to mislead.
Research softwareAuthor
2026
Measures how much a federated explanation depends on an arbitrary modelling choice.
Research softwareAuthor
Projets
Un ensemble cohérent d'outils reproductibles pour l'analyse du risque conditionnée au déploiement, le raisonnement réglementaire exécutable, l'évaluation de la suffisance des preuves et les tests de robustesse adverses.