Ricerca
Metodi aperti, prove verificabili
Progetti di ricerca finanziati, pubblicazioni sottoposte a revisione paritaria e software open source — tutto ciò che affermiamo può essere riprodotto.
Anonimizzazione e pseudonimizzazione basate sul rischio per Nestlé
Sfruttare i dati di eventi e i dati longitudinali nell'industria e nel settore sanitario attraverso tecnologie che preservano la privacy
Condivisione dei dati che preserva la privacy per la popolazione sintetica delle FFS
Questa è una selezione. Elenco completo delle pubblicazioni: Google Scholar · ORCID.
Pacchetti R open source su CRAN — stima del rischio, anonimizzazione e metodi di sintesi delle popolazioni.
Visualizza le statistiche di download CRAN · dal2010
RecordLinkage
R package for probabilistic and deterministic record linkage — deduplication and entity resolution across data sources, with comparison, classification, and evaluation tools.
CRAN · dal—
robSynth
Robust synthetic microdata generation when the training data are contaminated — a drop-in alternative to synthpop using MM-estimators, weighted logistic regression, and tree-based robust methods so outliers and coding errors don't propagate into the synthetic output.
CRAN · dal2007
sdcMicro
R package for statistical disclosure control of microdata — risk estimation, anonymization methods, and utility measurement. Peer-reviewed in the Journal of Statistical Software.
CRAN · dal2014
simPop
Simulation of synthetic populations from survey and census data.
CRAN · dal—
synvey
Design-aware, robust synthetic data generation — the successor to robSynth. Replaces the conditional models in sequential synthesis with MM-estimators and Huber-weighted logistic regression, keeping the synthetic data close to the clean data-generating process even when the source is contaminated.