18–21 May 2026
Europe/Warsaw timezone

Contribution List

245 out of 245 displayed
  1. Tim Morris (Novartis Pharmaceuticals UK Ltd)
    19/05/2026, 09:00

    Simulation studies involve drawing random numbers to understand the properties and behaviour of statistical methods. Statisticians have been using simulation studies since before computers existed (e.g. ‘Student’ in 1908). However, when it comes to simulation studies, we are largely self-taught. It is often hard understand a simulation study, or even its objective. Indeed, the rationale for...

    Go to contribution page
  2. David Solano (Research Group on Statistics, Econometrics and Health (GRECS). University of Girona)
    19/05/2026, 10:00

    Outlier detection in functional time series is challenging due to temporal dependence and the coexistence of magnitude, shape, and partially contaminated anomalies. Existing methods often assume independence or rely on model-based approaches, such as the Standard Smoothed Bootstrap on Residuals (SmBoR), which may perform poorly under model misspecification. Model-free alternatives, such as the...

    Go to contribution page
  3. Manuel Moreno (universitat de girona)
    19/05/2026, 10:00

    Spatial analyses in epidemiology often rely on accurate geolocation of individuals to estimate spatially structured health outcomes. However, routinely collected surveillance data frequently lack precise residential coordinates, introducing positional uncertainty that can bias spatial inference. This study examines the impact of uncertainty in patient location on the estimated spatial...

    Go to contribution page
  4. Agnieszka Kubik-Komar (University of Life Sciences in Lublin)
    19/05/2026, 10:00

    In automatic object detection, reliably counting objects remains challenging, particularly in scenarios with densely packed objects, overlapping instances, large scene variability, or multi-class cases.
    Common evaluation metrics for object detection are based on Intersection over Union (IoU) and do not directly measure the correctness of the number of detected objects. Consequently, a model...

    Go to contribution page
  5. Urszula Bronowicka-Mielniczuk (University of Life Sciences in Lublin/Department of Applied Mathematics and Computer Science)
    19/05/2026, 10:00

    Missing data is one of the most persistent challenges in environmental monitoring, undermining the reliability of analyses and limiting effective resource management. This issue is particularly critical under European regulations such as the Nitrates Directive (91/676/EEC), which requires accurate monitoring of nitrate concentrations in groundwater to protect ecosystems and public health. Yet,...

    Go to contribution page
  6. Moritz Pamminger (Medical University of Vienna, Center for Medical Data Science, Institute of Clinical Biometrics)
    19/05/2026, 10:00

    When addressing a particular research question using observational data, many decisions must be made during the conceptualization of the statistical analysis plan. This multiplicity of analysis strategies is a well-known problem that leads to high variation in research findings and associated low replicability, since each decision can lead to different results, even if each decision on its own...

    Go to contribution page
  7. Agnieszka Król (AstraZeneca)
    19/05/2026, 10:00

    Patient reported outcomes (PROs) are routinely used in randomized clinical trials (RCTs) to capture patients’ health status. Symptom-related PROs represent patients’ subjective perception of their health and are often collected multiple times during a clinical trial. For instance, in COPD, breathlessness or cough scores are captured using a small-range ordinal scale (0-4), representing...

    Go to contribution page
  8. Nasrin Salimian (Pardis Specialized Wellness Institute)
    19/05/2026, 10:00

    Background. Managing patients with multiple chronic conditions is a major challenge in modern health systems, particularly when exercise and lifestyle interventions are delivered in real-world settings. Robust statistical and machine-learning models require carefully designed data structures that capture the complexity of patients’ trajectories, comorbidities, and treatment exposures. In this...

    Go to contribution page
  9. Miglė Gervytė (Institute of Data Science and Digital Technologies, Faculty of Mathematics and Informatics, Vilnius University)
    19/05/2026, 10:00

    Spatial transcriptomics (ST) is a methodological suite that facilitates the in situ, high-resolution measurement of the transcriptome across a designated tissue section. By integrating transcriptional data with spatial coordinates, ST techniques enable the elucidation of key biological phenomena, including cell-type-specific gene regulatory networks, the spatial patterning of cellular...

    Go to contribution page
  10. Zita Zarándy (Center for Molecular Fingerprinting, Semmelweis University)
    19/05/2026, 10:00

    Non-communicable diseases (NCDs) impose the largest global burden of morbidity, premature mortality, and healthcare expenditure. To shift from reactive to preventive care, early detection of pre-symptomatic molecular changes is essential. We propose a statistical framework for identifying the most sensitive and robust early molecular predictors of prevalent NCDs — including cardiovascular...

    Go to contribution page
  11. Tina Lang (Bayer AG)
    19/05/2026, 10:00

    Preclinical experiments form the empirical foundation of translational medicine by assessing the feasibility, safety, and efficacy of new therapeutic approaches. Yet, unlike the highly regulated standards of clinical trials, preclinical research often exhibits substantial methodological heterogeneity, leading to concerns about reproducibility, bias, and the robustness of conclusions. These...

    Go to contribution page
  12. Hisashi Noma (The Institute of Statistical Mathematics)
    19/05/2026, 10:00

    Conventional two-stage procedures for binary-outcome meta-analysis use fixed plug-in estimates of within-study variances and depend heavily on large-sample normal approximations. These assumptions are often untenable and can lead to inaccurate inference, especially in sparse settings. Likelihood-based random-effects models, including the binomial–normal and the hypergeometric–normal (HGN)...

    Go to contribution page
  13. Stanislaw Leskow (Warsaw School of Economics (SGH))
    19/05/2026, 10:00

    Physiological monitoring often generates data characterized by strong cyclostationarity (circadian rhythm) and sensor artifacts – irregular noise. Conventional models (e.g. ARIMA) often fail to capture the time-varying dependencies or conflate behavioral rhythms with noise. We propose a signal processing framework adapted for digital health data: the Fraction-of-Time (FOT) probability...

    Go to contribution page
  14. Laura Kohlhas (Cogitars)
    19/05/2026, 10:00

    Meaningful prediction of when the target number of events will be reached is essential for both sample-size determination and operational planning in event-driven clinical trials. In oncology studies, progression-free survival (PFS) based on RECIST assessments is one of the most commonly used endpoints. Tumor evaluations for determining progression are typically performed at pre-scheduled...

    Go to contribution page
  15. Tomke Eiben (University of Applied Sciences and Arts, Hannover, Germany)
    19/05/2026, 10:00

    The introduction of standardized reporting guidelines has long been a response to inadequate study descriptions, starting with the CONSORT statement for clinical trials in the 1990s (Begg et al., 1996). One major approach to improve transparency and methodological rigor has been the introduction of standardized reporting guidelines such as the ARRIVE guideline (Percie du Sert et al., 2020)....

    Go to contribution page
  16. Katarzyna Jagoda (University of Plymouth)
    19/05/2026, 10:00

    Title: Joint modelling of general and mental health using copula models: a simulation-based evaluation for COVID-19 health research.

    COVID-19, the disease caused by the SARS-CoV-2 coronavirus, led to a global pandemic that began in December 2019. In the UK, government-mandated lockdowns were imposed to reduce the spread of the disease and understanding the impact of these actions on the...

    Go to contribution page
  17. Dorota Domagała (University of Life Sciences in Lublin)
    19/05/2026, 10:00

    One of the most serious effects of globalisation and human activity on the environment is air pollution. Nitrogen dioxide is particularly harmful to human health. Monitoring its content in the air over a long period of time allows trends to be assessed, and appropriate measures to be taken to improve air quality. In Poland, the Chief Inspectorate for Environmental Protection and its regional...

    Go to contribution page
  18. Dorota Domagała (University of Life Science)
    19/05/2026, 10:00

    This study evaluated the performance of the XGBoost method for imputing missing values in air quality data. The analysis used complete measurements of PM2.5, PM10, SO₂, NO, NO₂, and C6H6 recorded in Lublin in January 2020. To simulate missing data, 15%, 20%, and 25% of observations were randomly removed from each variable and imputed using XGBoost trained on the remaining data. Additionally,...

    Go to contribution page
  19. Małgorzata Graczyk (Poznań University of Life Sciences)
    19/05/2026, 10:00

    The present study investigates the spatial variability of alder (Alnus) pollen concentrations across different regions of Poland during the period 2001–2020. The primary objective was to identify and classify areas within the country that exhibit similar levels of alder pollen occurrence. The analytical results enabled the delineation of zones characterized by comparable concentrations of the...

    Go to contribution page
  20. Paweł Kurasiński (University of Life Sciences in Lublin City: Lublin)
    19/05/2026, 10:00

    This paper presents an analysis of the impact of data clustering on the accuracy of Weibull distribution parameter estimation in strength tests of mineral fertilizer granules. Two approaches are compared: traditional clustering into fixed-width intervals and optimal clustering, derived from a correctly constructed Fisher information matrix for clustered data. Maximum likelihood estimators for...

    Go to contribution page
  21. Sook-young Woo (Samsung Medical Center)
    19/05/2026, 10:00

    Selecting clinically meaningful cutoff values for continuous prognostic variables is challenging when the association with risk is U-shaped and competing risks are present. We propose a C index-based method to estimate an optimal pair of cutoff values (c_1,c_2) by directly targeting discriminative accuracy. Our approach first fits a smoothing spline to the log relative hazard from the Fine and...

    Go to contribution page
  22. Gesa Richter (Department of Periodontology, Oral Medicine and Oral Surgery, Charité – Universitätsmedizin Berlin, Berlin, Germany)
    19/05/2026, 10:00

    Genetic susceptibility plays a particularly important role in early-onset (EO) and severe periodontitis (PD). The genetic risk remains largely unexplained, due to limited sample sizes and heterogeneous phenotypes in genome-wide association studies (GWAS). This study investigates whether current GWAS data can be used to construct a polygenic score (PGS) capturing genetic susceptibility to...

    Go to contribution page
  23. Monika Różańska-Boczula (University of Life Sciences in Lublin)
    19/05/2026, 10:00

    The riverine ecosystems of the Biała and Czarna Lada River valleys are undergoing progressive degradation, accompanied by the spread of invasive plant species. To identify the factors driving these processes, we combined predictive modelling with multivariate analyses. To estimate the odds of habitat invasion and degradation, we used logistic regression and classification trees, which revealed...

    Go to contribution page
  24. Dominic Enders (Institute of Biostatistics and Clinical Research, University of Münster)
    19/05/2026, 10:00

    Introduction
    Rapid results from antimicrobial susceptibility testing (AST) are essential to guide the antimicrobial therapy of critically ill patients. Recent developments have revealed that readily available matrix-assisted laser desorption-ionization-time of flight (MALDI-TOF) mass spectrometry data, which is routinely used for bacterial species identification, can also be used to predict...

    Go to contribution page
  25. Christoph Gerlinger (Bayer AG, Clinical Statistics and Analytics, Berlin and Department of Gynecology, Obstetrics and Reproductive Medicine, University Medical School of Saarland)
    19/05/2026, 10:00

    Realise D is a public-private partnership of almost 40 partners from academia, regulatory bodies, clinical research institutes and hospitals, patient organizations, pharmaceutical companies, methodologists, and European Research Infrastructures. Realise D is part of the European Union’s Innovative Health Initiative and funded jointly by the EU and industry. The project started officially in...

    Go to contribution page
  26. Giulia Varvarà (University of Rennes)
    19/05/2026, 10:00

    In oncology trials, tumour-based endpoints, like Progression-Free Survival (PFS), Disease-Free Survival (DFS) or Relapse-Free Survival (RFS), are widely accepted. However, their use is more controversial compared with Overall Survival (OS), due to the subjectivity of tumour assessment, and their sensitivity to censoring rules, as they are more prone to obtaining different results depending on...

    Go to contribution page
  27. Karolina Majdak (Jagiellonian University Medical College, Chair of Epidemiology and Preventive Medicine, Department of Medical Sociology)
    19/05/2026, 10:00

    The aim of the study was to verify changes in the association between place of residence and depression prevalence before and after the onset of COVID-19 pandemic. Second objective of the study was to identify indicators of social capital as determinants of the prevalence of depression among people aged 50 years or older living in rural and urban areas.

    The study included data from two...

    Go to contribution page
  28. Fatuma Hassen (Department of Medical Laboratory Sciences, College of Health Sciences, Addis Ababa University, Addis Ababa, Ethiopia.)
    19/05/2026, 10:00

    Purpose: Globally, breast cancer is the most commonly diagnosed cancer and the leading cause of cancer-related deaths in women. The purpose of this study was to determine the survival of breast cancer patients and associated factors.
    Methods: This study was done among breast cancer patients treated at the Oncology Center of Tikur Anbessa Hospital, Addis Ababa, Ethiopia. Clinical data were...

    Go to contribution page
  29. Michalina Gajdzica (Jagiellonian University - Medical College, Chair of Epidemiology and Preventive Medicine)
    19/05/2026, 10:00

    The growing role of social network sites in building and maintaining social relationships, generating social support, and providing information implies the need to develop a tool for measuring online social networks, intended for use in population studies related to health and quality of life. Models based on Item Response Theory are widely used in test development and has proven advantages...

    Go to contribution page
  30. Syntia Souza (UFRPE)
    19/05/2026, 10:00

    Autism Spectrum Disorder (ASD) is a neurodevelopmental condition characterized by difficulties in social interaction, communication, and restricted and repetitive behavior patterns. In Brazil, according to the Instituto Brasileiro de Geografia e Estatística (IBGE, 2022), approximately 2.4 million individuals have been diagnosed, and this number may reach six million when unidentified cases are...

    Go to contribution page
  31. Jovin Tibenderana (St Francis University College of Health And Allied Sciences)
    19/05/2026, 10:00

    Background: Globally about 9 million neonates are diagnosed with birth asphyxia yearly. In Tanzania 40.6% of all neonatal deaths are attributed to birth asphyxia. There is scarcity of evidence on predictors of in-hospital survival among asphyxiated neonates in Tanzania, therefore study aimed to determine trends and predictors of survival among neonates who sustained birth asphyxia at...

    Go to contribution page
  32. Beth McDougall (Department of Statistics, Phastar)
    19/05/2026, 10:00

    Background:
    Oncology trials remain the largest sector of global drug development. However, their complexity, resource demands, and modest success rates underscore the need for more efficient, patient-centred, and methodologically innovative designs.
    Objective:
    To provide a longitudinal assessment of interventional oncology trial characteristics and design trends from 2000 through 2025...

    Go to contribution page
  33. Johanna Ledoux (University of Zurich, Department of Biostatistics)
    19/05/2026, 10:00

    The win ratio statistic has gained prominence as an interpretable method for analyzing composite endpoints in clinical trials, typically with a superiority objective. The use of the win ratio requires simulation to estimate the necessary sample size (1). Adapting win statistics to non-inferiority trials and incorporating covariate adjustment remain unresolved methodological challenges...

    Go to contribution page
  34. Simon Wandel (Novartis Pharma AG)
    19/05/2026, 10:45

    The pharmaceutical industry has a long history with Bayesian statistics. Already in 1986, Racine, Grieve, Flühler and Smith wrote an article entitled Bayesian Methods in Practice: Experiences in the Pharmaceutical Industry[1], highlighting four typical applications they encountered at that time. Since then, Bayesian methods have been applied to many more problems in the pharmaceutical...

    Go to contribution page
  35. Max Menssen (University Medical Center Göttingen, Department of Medical Statistics, Göttingen, Germany)
    19/05/2026, 10:45

    Bootstrap calibration grounds on a simple idea: Based on a bootstrap sample, one can compute the bootstrap coverage probability of the desired interval. Then, one can alternate the intervals limits until the bootstrap coverage probability approaches the nominal level, e.g. by alternating the α-level used for interval calculation. Finally, the desired interval is calculated replacing the...

    Go to contribution page
  36. Konstantin Emil Thiel (Paracelsus Medical University City: Salzburg)
    19/05/2026, 10:45
    oral presentation

    Analysis of covariance (ANCOVA) assesses the effect of a group factor on a response while accounting for covariate information. We propose a nonparametric ANCOVA based on Mann-Whitney effects, specifically designed for randomized trials. Unlike classical ANCOVA, our approach does not rely on distributional assumptions or metric-scale data; Ordinal measurements (such as Likert-scale items) are...

    Go to contribution page
  37. Ann-kathrin Ozga (Institute for Medical Biometry and Epidemiology, University Medical Center Hamburg-Eppendorf)
    19/05/2026, 10:45
    oral presentation

    Introduction:
    In clinical trials time to event endpoints like time to death, time to hospitalization or time to myocardial infarction are often or primary interest. Although multiple events might be observed per individual, only the time to the first occurring event is considered in primary analysis. One reason for this could be that guidelines recommend analyzing the data using the same...

    Go to contribution page
  38. Markus Neuhäuser (Koblenz University of Applied Sciences)
    19/05/2026, 10:45

    This talk will present the closed testing procedure as it was first introduced by Marcus et al. (1976). We discuss further developments and early contributions published in the following years. A special focus is given to conferences as the ones in Oberwolfach, Bad Ischl and especially Gerolstein. We also report from the first International Conferences on Multiple Comparison Procedures (MCP)...

    Go to contribution page
  39. Sathish Ravindranth (TU Dortmund University)
    19/05/2026, 10:45
    oral presentation

    Aging is the dominant risk factor for neurodegenerative and systemic diseases, yet its molecular signatures remain obscured within high-dimensional, noisy, and strongly correlated proteomes. To address this challenge, we introduce the Protein Risk Score (ProtRS) framework—a systematic evaluation framework for ProtRS modeling that assesses how different multivariate approaches extract...

    Go to contribution page
  40. Franz König (Medical University of Vienna)
    19/05/2026, 11:03

    The strict control of the studywise Type I error rate has long been a cornerstone of confirmatory clinical trials. Closed testing and adaptive designs are two influential ideas in modern trial methodology, yet they emerged from different motivations: one from the need to rigorously control multiplicity when testing multiple hypotheses, the other from the desire to build flexibility into study...

    Go to contribution page
  41. Qiong Wu (University of Marburg)
    19/05/2026, 11:03
    oral presentation

    Polygenic risk scores can be used to model the individual genetic liability for human traits. Current methods primarily focus on modeling the mean of a phenotype neglecting the variance. However, genetic variants associated with phenotypic variance can provide important insights to gene-environment interaction studies. To overcome this, we propose snpboostlss, a cyclical gradient boosting...

    Go to contribution page
  42. Levin Wiebelt (Charité - Universitätsmedizin Berlin)
    19/05/2026, 11:03
    oral presentation

    A common goal in medical research is to estimate a difference between treatment groups and quantify its uncertainty, or to infer a population-level difference. The most commonly used nonparametric group difference measure is the Mann-Whitney (MW) effect. It applies to a broad range of outcomes, including skewed, heteroskedastic and ordinal distributions, since it does not assume a parametric...

    Go to contribution page
  43. Yang Han (University of Manchester)
    19/05/2026, 11:03

    Multiple-use prediction and calibration for all future values play a valuable role in many areas including health and medical research. Simultaneous tolerance bands (STBs) can be used for these purposes. Motivated by real-world problems in health research, this study focuses on the construction of exact STBs for multiple regression over any given rectangular covariate region and for polynomial...

    Go to contribution page
  44. Daniele Giardiello (Bicocca Bioinformatics Biostatistics and Bioimaging B4 Center, School of Medicine and Surgery, University of Milano-Bicocca,)
    19/05/2026, 11:03
    oral presentation

    Researchers in biomedical research often analyse data that are subject to clustering. Independence among observations are generally assumed to develop and validate risk prediction models. For survival outcomes, the Cox proportional hazards regression model is commonly used to estimate an individual’s risk at fixed time horizons. The stratified Cox proportional hazards and the shared gamma...

    Go to contribution page
  45. Katja Ickstadt (Department of Statistics, TU Dortmund University)
    19/05/2026, 11:15

    The Bayesian approach in general has a lot to offer in times of Machine Learning (ML) and Artificial Intelligence (AI). The Bayesian framework itself offers a learning environment, where the prior, and, subsequently the posterior distributions can be updated sequentially, and where human expertise can be incorporated. The approach allows for uncertainty quantification of all quantities of...

    Go to contribution page
  46. Stephen Schüürhuis (Institute of Biometry and Clinical Epidemiology, Charité - Universitätsmedizin Berlin)
    19/05/2026, 11:21
    oral presentation

    In the context of a two-group comparison, when the assumption of equal variances between groups is doubtful or the data may be skewed or ordinal, the classical t-test and an effect measure parameterized in terms of means may no longer be suitable. In such cases, it appears more appropriate to formulate the problem as the nonparametric Behrens-Fisher problem of testing H0: θ = 1/2, where θ =...

    Go to contribution page
  47. Duoerkongjiang Alidan (Institute of Medical Biometry and Epidemiology; University Medical Center Hamburg-Eppendorf (UKE))
    19/05/2026, 11:21
    oral presentation

    Accurate analysis of multiple time-to-event endpoints is a persistent challenge in clinical research, where patients may experience several recurrent non-fatal events alongside a competing fatal event. Conventional survival analysis approaches, such as time-to-first-event analyses or the Cox proportional hazards model, often neglect recurrent events or assume independence between event types,...

    Go to contribution page
  48. Lingjiao Wang (Warwick Medical School, University of Warwick)
    19/05/2026, 11:21

    Background: The stability of a drug product over time is a critical property in pharmaceutical development. A key objective in drug stability studies is to estimate the shelf-life of a drug, involving a suitable definition of the true shelf-life and the construction of an appropriate estimate of the true shelf-life. Simultaneous confidence bands (SCBs) for percentiles in linear regression are...

    Go to contribution page
  49. Werner Brannath (University of Bremen)
    19/05/2026, 11:21

    In this talk, we will explore the relationship between the closed testing principle for multiple tests with family-wise error rate (FWER) control and the partitioning plus projection principle for constructing simultaneous confidence intervals. Starting with the simple observation that a multiple test with FWER control is formally equivalent to a one-sided simultaneous confidence interval for...

    Go to contribution page
  50. Ngoune Darwin (TU Dortmund University)
    19/05/2026, 11:21
    oral presentation

    Regularized Multi-Omics Regression Modeling for Transcriptomic–Proteomic Integration in Mice with induced liver Damage.

    Toxicological compounds exert complex effects on tissues and organisms, which can be investigated using genomic, transcriptomic, and proteomic data. A central challenge lies in understanding the relationship between RNA and protein levels. While these are expected to be...

    Go to contribution page
  51. Gerhard Nehmiz (Consultant to Boehringer Ingelheim Pharma GmbH&Co. KG)
    19/05/2026, 11:35

    Lessons learned in the last 25 years

    Gerhard Nehmiz, consultant for Boehringer Ingelheim Pharma GmbH & Co. KG, Biberach, Germany

    gerhard.nehmiz.ext@boehringer-ingelheim.com

    The Working Group "Bayes methods" of the IBS / German Region was founded in 2001, and met a need. It had two roots: The WG "Prognosis and decision making" of the GMDS (U. Mansmann) and the German BUGS User Group...

    Go to contribution page
  52. Julia Eichhorn (TU Dortmund University)
    19/05/2026, 11:39

    In toxicology, concentration-response experiments are conducted to investigate the toxic behaviour of a given substance. Typically, a parametric model is fitted and effective concentrations to a viability level p (EC_p) are estimated which are used e.g. in further experiments. However, in previous research, it was observed that the estimated EC_p of the same experiment conducted in the same...

    Go to contribution page
  53. Markus Schepers (Institute of Medical Biostatistics, Epidemiology and Informatics (IMBEI), University Medical Center of the Johannes Gutenberg University)
    19/05/2026, 11:39
    oral presentation

    Feedback is pervasive in biological and biomedical systems, yet many causal discovery methods, including widely used score-based approaches such as NOTEARS, impose acyclicity and may therefore misrepresent gene regulatory, pharmacological, or cellular processes. Building on recent advances in cyclic causal inference, such as the intervention-capable Bicycle method, we investigate how...

    Go to contribution page
  54. Martin Posch (Medical University of Vienna)
    19/05/2026, 11:39

    The closed testing principle is a fundamental framework to construct multiple testing procedures controlling the familywise error rate in the strong sense. However, a major challenge in the application of the principle is the number of intersection hypothesis tests that need to be specified, which increases exponentially in the number of elementary hypotheses tested and makes it difficult to...

    Go to contribution page
  55. Victoria Watson (Phastar)
    19/05/2026, 11:39
    oral presentation

    Prognostic Models for Recurrent Event Data
    Dr Victoria Watson1,2, Prof Catrin Tudur Smith2, Dr Laura Bonnett2
    1 Phastar, London, UK
    2 University of Liverpool, Department of Health Data Sciences

    Background / Introduction
    Prognostic models predict outcome for people with an underlying medical condition. Many conditions are typified by recurrent events such as seizures in epilepsy....

    Go to contribution page
  56. Niklas Lück (TU Dortmund University, IUF - Leibniz Research Institute for Environmental Medicine)
    19/05/2026, 11:39
    oral presentation

    Single-cell RNA sequencing has given researchers unparalleled insight into biological systems. It enables the identification of distinct cellular subpopulations, the characterization of differences between them, and the assessment of overall tissue heterogeneity. Conventional analysis pipelines first cluster individual cells into similar groups and then test for differentially expressed genes...

    Go to contribution page
  57. Reinhard Vonthein
    19/05/2026, 11:55
  58. Gul Inan (Koc University)
    19/05/2026, 11:57
    oral presentation

    Penalized regression models such as Lasso, Elastic-Net and their adaptive extensions are widely used for simultaneous variable selection and prediction in high-dimensional data analysis. However, conventional implementations of adaptive Elastic-Net (AdaENet) regression often estimate the adaptive hyper-parameter for the Elastic-Net penalty term using the entire dataset before dividing it into...

    Go to contribution page
  59. Paavo Sattler (TU Dortmund; RWTH Aachen)
    19/05/2026, 11:57

    Linear hypotheses Hp = y regarding a parameter vector p arise in a wide range of scientific fields, including life sciences, psychology, economics, environmental sciences, and other areas of applied statistics, due to their ability to encode a wide variety of scientific questions using a simple algebraic framework. The unknown parameter vector p can represent, for example, an expectation...

    Go to contribution page
  60. Frank Bretz (Novartis)
    19/05/2026, 11:57

    We consider the problem of testing multiple null hypotheses, where a decision to reject or retain must be made for each one and embedding incorrect decisions into a real life context may inflict different losses. We argue that traditional methods controlling the Type I error rate may be too restrictive in this situation and that the standard familywise error rate may not be appropriate. For...

    Go to contribution page
  61. Tanya Toluay (Charité - Universitätsmedizin Berlin, Insitute of Biometry and Clinical Epidemiology City: Berlin)
    19/05/2026, 11:57
    oral presentation

    Background
    Longitudinal observational data frequently involve time-varying confounding, autoregressive dependence, and potential reciprocal feedback between processes. These features complicate the estimation of cross-lagged causal effects and challenge the assumptions underlying standard modelling approaches. Methodological evaluation requires transparent, reproducible simulation frameworks...

    Go to contribution page
  62. Merle Munko (Otto-von-Guericke University Magdeburg)
    19/05/2026, 11:57
    oral presentation

    Various estimators for modelling the transition probabilities in multi-state models have been proposed, e.g., the Aalen-Johansen estimator, the landmark Aalen-Johansen estimator, and a hybrid Aalen-Johansen estimator. While the Aalen-Johansen estimator is generally only consistent under the rather restrictive Markov assumption, the landmark Aalen-Johansen estimator can handle non-Markov...

    Go to contribution page
  63. Lukas Burk (Leibniz Institute for Prevention Research and Epidemiology - BIPS)
    19/05/2026, 13:45

    Background: Post-COVID Condition (PCC) affects a substantial proportion of individuals following SARS-CoV-2 infection, and the mechanisms driving symptom persistence remain an area of active research. Identifying risk factors associated with PCC development is important for targeted prevention strategies and clinical management. Machine learning (ML) models offer powerful tools for prediction...

    Go to contribution page
  64. Lukas Lohse (Charité – Universitätsmedizin Berlin, corporate member of Freie Universität Berlin and Humboldt-Universität zu Berlin, Institute of Clinical Pharmacology and Toxicology, Embryotox Center of Clinical Teratology and Drug Safety in Pregnancy)
    19/05/2026, 13:45
    oral presentation

    The recently conducted observational Embryotox cohort study on mRNA COVID-19 vaccination aimed to assess the safety of mRNA COVID-19 vaccines in pregnancy. Here, we focus on the methodological approach used to assess the effect of the vaccination on adverse pregnancy outcomes such as spontaneous abortion and stillbirth. The data featured delayed study entry and cohort crossover as well as...

    Go to contribution page
  65. Łukasz Smaga (Adam Mickiewicz University)
    19/05/2026, 13:45

    Functional Data Analysis (FDA), focusing on data composed of functions or curves, has become increasingly popular. We study reliable methods for comparing multiple groups of functional data, especially in studies involving several factors or complex designs. We introduce a new statistical approach designed for multivariate functional data. Our methods are reliable because they allow us to...

    Go to contribution page
  66. Christian Staerk (IUF – Leibniz Research Institute for Environmental Medicine & TU Dortmund University)
    19/05/2026, 13:45

    Understanding the relative importance of genetic, molecular and environmental factors is crucial for interpretable prediction models in biomedicine and for targeted prevention. While classical regression-based approaches provide direct interpretability through model coefficients, flexible machine learning (ML) approaches such as random forests and neural networks typically rely on post-hoc...

    Go to contribution page
  67. Anna Szczepańska-Alvarez (Poznań University of Life Sciences)
    19/05/2026, 13:45

    In this talk we present a statistical approach to evaluate the relationship between variables observed in a two-factors experiment. We consider a three-level model with covariance structure ${\bf \Sigma} \otimes {\bf \Psi}_1 \otimes {\bf \Psi}_2$, where ${\bf \Sigma}$ is an arbitrary positive definite covariance matrix, and ${\bf \Psi}_1$ and ${\bf \Psi}_2$ are both correlation matrices with...

    Go to contribution page
  68. Jakub Malik ([1] Poznan University of Physical Education - Faculty of Sport Sciences; [2] Adam Mickiewicz University - Faculty of Mathematics and Computer Science)
    19/05/2026, 13:45
    oral presentation

    Maintaining balance is a crucial daily skill, and impairments in postural control increase the risk of falls, particularly among older adults. This study aimed to assess the effect of attentional control on postural stability in young and older adults. The sample consisted of 43 participants (16 older adults, 12 women; 27 young adults, 13 women). Participants performed a series of 60-second...

    Go to contribution page
  69. Francesco Stingo (affiliation: University of Florence)
    19/05/2026, 14:03

    Non-linear regression models are flexible approaches used to model complex associations. In many recent proposals, additional flexibility comes at the cost of loss of interpretability of the model's parameters and, consequently, of the data analysis results. We introduce a flexible model whose parameters are easily interpretable. In particular, the model incorporates non-linear effects through...

    Go to contribution page
  70. Marilena Müller (German Cancer Research Cente)
    19/05/2026, 14:03
    oral presentation

    We consider a two-arm randomized clinical trial in precision oncology with time-to-event endpoint. Patients in the control arm receive standard of care (SOC) treatment whereas patients in the experimental arm are offered personalized treatment, e.g. on the basis of molecular characterization of the disease. However, some patients in the experimental arm will not receive personalized treatment...

    Go to contribution page
  71. Stefan Embacher (Medical University of Graz)
    19/05/2026, 14:03
    oral presentation

    Our systematic review indicates that optimal design methods are not yet applied in immunization studies in which modeling the antibody kinetics, i.e. the change of antibody concentration over time, is the main objective. We argue that this substantial underutilization is driven by several factors, including limited awareness of the advantages of optimal design and accessibility of convenient...

    Go to contribution page
  72. Fabiola Del Greco M. (Institute of Biomedicine - Eurac research)
    19/05/2026, 14:03

    Introduction
    Metabolomics measures small molecules (called metabolites) in cells, tissues, biofluids, that represent intermediates and/or end-products of biochemical/cellular processes. As a results, metabolomics has shown to be useful for predicting disease risks or associated biomarkers. Given the large data complexity and size, the Machine learning (ML) approach represents an appropriate...

    Go to contribution page
  73. Jędrzej Wydra (Adam Mickiewicz University)
    19/05/2026, 14:03

    Testing independence between functional observations remains a fundamental challenge in modern statistics, particularly in settings involving high-dimensional or infinite-dimensional random objects. The presented work introduces a new framework for independence testing in functional data based on the distance of mean embedding (DIME), a metric recently proposed as a flexible alternative to...

    Go to contribution page
  74. Moritz Fabian Danzer (University of Münster)
    19/05/2026, 14:15

    Adaptive and, in particular, group-sequential designs are well-established in clinical trials. Time-to-event endpoints pose particular challenges because individual participants can contribute data to multiple stages of the trial. Nevertheless, the log-rank test - the standard analysis method for time-to-event data - can be embedded in flexible adaptive designs (e.g. with sample-size...

    Go to contribution page
  75. Hélène Ruffieux (University of Cambridge)
    19/05/2026, 14:21

    Bayesian graphical models are powerful tools to infer complex relationships in high dimension, yet are often fraught with computational and statistical challenges. If exploited in a principled way, the increasing information collected alongside the data of primary interest constitutes an opportunity to mitigate these difficulties by guiding the detection of dependence structures. For instance,...

    Go to contribution page
  76. Lucia Ameis (Institute of Medical Statistics and Computational Biology (IMSB), Faculty of Medicine, University of Cologne)
    19/05/2026, 14:21

    Evaluating a response variable in relation to exposure time or dose is a pivotal objective in the assessment of a compound's effect, particularly when determining toxicity in pre-clinical research or pharmacokinetics in clinical trials. The determination of an alert, such as the EC50 value, at which a pre-specified threshold of the response variable is crossed, is an important tool for the...

    Go to contribution page
  77. Martina Mittlböck (Institute of Clinical Biometrics, CeDAS, Medical University of Vienna)
    19/05/2026, 14:21
    oral presentation

    The assessment of allogeneic stem cell transplantation (SCT) over standard continued chemotherapy in a clinical trial of childhood leukaemia is not straightforward. Standard chemotherapy will be stopped and SCT performed if a donor search identifies a suitable stem cell donor in registries of potential donors. Randomization to SCT or continued chemotherapy is usually not feasible due to...

    Go to contribution page
  78. Göran Kauermann (Ludwig-Maximilians-Universität München)
    19/05/2026, 14:21
    oral presentation

    This paper focuses on drawing information on underlying processes, which are not directly observed in the data. In particular, we work with data in which only the total count of units in a system at a given time point is observed, but the underlying process of inflows, length of stay, and outflows is not. The particular data example looked at in this paper is the occupancy of intensive care...

    Go to contribution page
  79. Soheila Aghlmandi (University of Basel)
    19/05/2026, 14:21

    Objectives: To evaluate whether machine learning (ML) applied to comprehensive claims data without diagnostic codes can distinguish a high proportion of antibiotic treatment episodes as urinary tract infection (UTI) or non-UTI cases. Such approaches may be valuable for antimicrobial stewardship when diagnosis-linked datasets are unavailable.
    Methods: Outpatient antibiotic prescription claims...

    Go to contribution page
  80. Lukas Mödl (Institut für Biometrie und Klinische Epidemiologie, Charité -- Universitätsmedizin Berlin)
    19/05/2026, 14:35

    Quadratic forms, such as the rank-based Wald-type statistic or the rank-based ANOVA-type statistic, are widely used to compare multivariate distributions without the necessity of parametric assumptions (like multivariate normality). These tests have two major limitations, however:
    i) They are, by construction, omnibus tests and thus not able to locate which specific dimensions (variables) are...

    Go to contribution page
  81. Anqi Sui (University College London)
    19/05/2026, 14:39

    Introduction
    Monitoring the clinical performance of healthcare units (e.g. hospitals, surgeons) is the main component for national audits, enabling identification of ‘outlier’ units whose clinical performance, e.g. in-hospital mortality, deviates significantly from expected performance. Accurate detection and subsequent management of outliers are critical for improving healthcare quality. ...

    Go to contribution page
  82. Jonas Beck (DKFZ)
    19/05/2026, 14:39
    oral presentation

    Clinical trials often show treatment curves that diverge early and converge later, or vice versa—patterns that are poorly captured by the proportional-hazards assumption. We develop a joint inferential framework for two nonparametric functionals of censored survival data: the Kaplan–Meier–based Mann–Whitney effect and a novel temporal contrast separating early and late differences. The...

    Go to contribution page
  83. Zhi Zhao (University of Oslo)
    19/05/2026, 14:39

    Single-cell technologies provide an unprecedented opportunity for dissecting the interplay between the cancer cells and the associated tumor microenvironment, and the produced high-dimensional omics data should also augment existing survival modeling approaches for identifying tumor cell type-specific genes predictive of cancer patient survival. However, there is no statistical model to...

    Go to contribution page
  84. Shizhe Xu (University of Oxford)
    19/05/2026, 14:39

    Genome-wide association studies (GWAS) often identify genomic regions containing hundreds or thousands of genetic variants with comparable statistical evidence. Extensive linkage disequilibrium (LD) and the sparsity of causal variants obscure association signals, hindering the identification of true causal variants underlying complex traits. Fine-mapping approaches are introduced to...

    Go to contribution page
  85. Rumana Omar (University College London)
    19/05/2026, 14:39
    oral presentation

    Background: Risk prediction models are increasingly being used in clinical practice to predict health outcomes. These models are often developed using data from multiple centres (clustered data) where patient outcomes within a centre are likely to be correlated. It is important that the dataset used to develop a risk model is of an appropriate size, to avoid model overfitting problems and poor...

    Go to contribution page
  86. Erin Sprünken (Charité - Universitätsmedizin Berlin)
    19/05/2026, 14:55

    In many trials and experiments, subjects are not only observed once, but multiple times, resulting in a cluster of possibly correlated observations. For example, mice sharing the same cage or students of the same class are typical examples of clustered data. Typically, under the assumption of normally distributed data, mixed models are used for analysis.
    However, this model assumption is...

    Go to contribution page
  87. Manuela Zucknick
    19/05/2026, 14:57
  88. Hannes Buchner (Staburo GmbH)
    19/05/2026, 14:57
    oral presentation

    Time-to-event variables are among the most relevant primary efficacy endpoints in clinical trials, particularly in later phase oncology trials. When the proportional hazards assumption is expected to be severely violated, an alternative to the log-rank test is needed. Testing for differences in survival probabilities at a pre-defined time point offers one such option and has already been...

    Go to contribution page
  89. Annalena Weissert (TU Dortmund University)
    19/05/2026, 14:57

    Metabolite discovery can provide insights into disease mechanisms and help to identify potential biomarkers that contribute to the development of new treatments. We present a self-supervised deep learning approach for metabolite discovery. Molecular intensity distributions obtained via MALDI-MSI (matrix-assisted laser desorption/ionization mass spectrometry imaging) are compared with...

    Go to contribution page
  90. Audrey Yeo (Finc Research)
    19/05/2026, 15:45
    oral presentation

    Early clinical trials play a critical role in drug development. The main purpose of early trials is to determine whether a novel treatment demonstrates sufficient safety and efficacy signals to warrant further investment (Lee & Liu, 2008). The new open source R package phase1b is a flexible toolkit that calculates many properties to this end, especially in the oncology therapeutic area. The...

    Go to contribution page
  91. Matthew George (Phastar)
    19/05/2026, 15:45
    oral presentation

    The use of combination treatments in early phase oncology trials is growing. The objective of these trials is to search for the maximum tolerated dose combination from a pre-defined set. However, cases in which the initial set of combinations does not contain one close to the target toxicity level pose a significant challenge. There is uncertainty around how to handle these situations...

    Go to contribution page
  92. Chris Jennison (University of Bath, UK)
    19/05/2026, 15:45

    When estimating the treatment effect after a group sequential test or a more complex adaptive design, the maximum likelihood estimate is liable to be biased. The ICH E20: Guideline on Adaptive Designs for Clinical Trials has “reliability of estimation” as a key topic. Methods have been developed to reduce the bias in estimators after group sequential and adaptive designs – or even eliminate...

    Go to contribution page
  93. Liesbeth C De Wreede (Department of Biomedical Data Sciences, Leiden University Medical Center)
    19/05/2026, 15:45

    The COVID-19 pandemic has led to excess mortality worldwide. Notably, the reported numbers of excess deaths are different from the numbers of deaths from COVID-19. Evaluating pandemic-related mortality should therefore not only be based on cause-of-death data but also on external life tables to enable calculation of population-based measures of the difference between observed and expected...

    Go to contribution page
  94. Jan-Bernd Igelmann (TU Dortmund University)
    19/05/2026, 15:45
    oral presentation

    Optimal designs maximize the experimental efficiency and precision, but are sometimes difficult to obtain, especially in cases with non-trivial underlying model functions. A possible application area providing the motivating example is toxicology. Liver carcinoma cells are modelled as a function of valproic acid (VPA) concentration using the common four-parameter log-logistic (4PLL) model. To...

    Go to contribution page
  95. Maxi Schulz (University Medical Center Göttingen, Department of Medical Statistics)
    19/05/2026, 15:45

    Longitudinal or clustered data often arise in clinical research, potentially violating the independent and identically distributed (i.i.d) assumption. In regression, (generalized) linear mixed-effect models are frequently used to account for the correlation structure of the data, but these come with restrictions such as the linearity assumption and pre-specification of predictors and their...

    Go to contribution page
  96. Liane Kluge (University of Bremen)
    19/05/2026, 16:03
    oral presentation

    Fast-track procedures play an important role in the registration of health products, such as registration processes for digital health applications. These procedures offer the potential for patients to access innovative products earlier. The procedures involve two registration steps. Applicants can first apply for conditional registration. A successful conditional registration provides a...

    Go to contribution page
  97. Alexandra Nagel (University of Ulm, METRONOMIA Clinical Research GmbH)
    19/05/2026, 16:03
    oral presentation

    The concept of interim analysis and adaption is more and more used in clinical trials. Furthermore, one often has several endpoints such as progression-free survival (PFS) and overall survival (OS) which is discussed in the current paper of Danzer et al. (2022). There, testing hypotheses with adaptive group sequential one-sample tests for the distribution of PFS and OS is developed. The...

    Go to contribution page
  98. Giles Partington (Phastar)
    19/05/2026, 16:03
    oral presentation

    Bayesian Methods in Registered Clinical Trials: A Systematic Review of Studies on ClinicalTrials.gov Through 2024
    Giles Partington & Christina Geyer: Phastar
    Bayesian methods are increasingly being incorporated into clinical trial designs to improve flexibility, efficiency, and interpretability. Earlier reviews of published studies (Lee & Chu, 2012) illustrated how these approaches were...

    Go to contribution page
  99. Julia Höpler (Leibniz Institute for Prevention Research and Epidemiology – BIPS, Faculty of Mathematics and Computer Science – University of Bremen)
    19/05/2026, 16:03

    Machine learning (ML) models have emerged as a powerful alternative to traditional statistical methods due to their flexibility and ability to leverage large-scale, high-dimensional datasets. However, in sensitive application areas such as clinical and prognostic modeling, deploying ML models requires interpretability in order to reveal underlying model behavior, identify influential risk...

    Go to contribution page
  100. Florian Klinglmueller (Austrian Agency for Health and Food Safety)
    19/05/2026, 16:03

    Reliable estimation of treatment effects is essential for the benefit–risk assessment supporting the approval of new drugs and for the communication of trial results in the European Public Assessment Report (EPAR) and Summary of Product Characteristics (SmPC). In adaptive designs, where trial adaptations such as sample size re-assessment, population enrichment, or treatment arm selection are...

    Go to contribution page
  101. Caroline Dietrich (Clinical Epidemiology Division, Department of Medicine Solna, Karolinska Institutet)
    19/05/2026, 16:15

    Regression models for the hazard function have been proposed on both a multiplicative and an additive scale. In medical research the former is often suitable, but in some instances, it is more biologically plausible to assume an additive effect on the mortality rate. The best-known example is in population-based cancer patient survival, where the presence of cancer is assumed to have an...

    Go to contribution page
  102. Maren Hackenberg (Institute of Medical Biometry and Statistics, Faculty of Medicine and Medical Center, University of Freiburg. Freiburg Center for Data Analysis, Modeling, and AI, University of Freiburg)
    19/05/2026, 16:21
    oral presentation

    In clinical registries, longitudinal patient data are often sparse, noisy, and irregularly sampled, yet an important practical question is how an individual patient’s health status is likely to change from the current visit to the next.

    Regression-based approaches to longitudinal data analysis typically provide a global fit over the full observed time course, but are not directly aimed at...

    Go to contribution page
  103. Benjamin Fallmann (Medical University of Vienna, Center for Medical Data Science, Institute of Medical Statistics)
    19/05/2026, 16:21
    oral presentation

    Graph-based multiple testing procedures provide an intuitive way to define closed testing strategies that control the family-wise error rate (FWER) in fixed sample settings [1]. They have been extended to adaptive trial designs based on the (partial) conditional error rate (CER) method [2]. These procedures control the FWER in two-stage designs where the trial is adapted after an interim...

    Go to contribution page
  104. David Robertson (MRC Biostatistics Unit, University of Cambridge)
    19/05/2026, 16:21

    In adaptive clinical trials, the conventional point estimators of the treatment effect are prone to bias. Similarly, the conventional confidence intervals are prone to incorrect coverage, as well as other undesirable statistical properties. Recent regulatory guidance, such as ICH E20, has highlighted the need to use adjusted estimators and confidence intervals for adaptive designs in order to...

    Go to contribution page
  105. Max Westphal (Fraunhofer MEVIS)
    19/05/2026, 16:21

    Introduction:
    Machine learning (ML) validation studies can often be tackled with standard statistical inference methods, i.e. confidence intervals and statistical tests. While this is reasonable in many situations there are also conditions under which the usual IID assumption is not met, and operating characteristics (coverage probability, type 1 error rate) may thus deteriorate. For...

    Go to contribution page
  106. Lukas Widmer (Novartis Pharma AG)
    19/05/2026, 16:21
    oral presentation

    The Bayesian Logistic Regression Model (BLRM) with Escalation With Overdose Control (EWOC) is widely used in Phase I Oncology trials. Recently, several publications have highlighted a recurring issue: escalation can be blocked even when observed data strongly suggest safety. I.e., the posterior overdose probability at the next dose remains above the EWOC threshold despite no dose-limiting...

    Go to contribution page
  107. Damjan Manevski (Institute for Biostatistics and Medical Informatics, Faculty of Medicine, University of Ljubljana)
    19/05/2026, 16:35

    In many medical applications of event-history analysis, individuals may experience several intermediate events before death, and a non-negligible proportion of deaths is unrelated to the disease under study. While standard multi-state models evaluate the occurrence of different events over time, they do not explicitly model mortality from disease-related causes and from other (population)...

    Go to contribution page
  108. Sonja Zehetmayer (Medical University of Vienna City: Vienna)
    19/05/2026, 16:39
    oral presentation

    We propose a frequentist, adaptive trial design to investigate the safety and efficacy of three dose levels compared to placebo for the treatment of worm infections. As the safety of the highest dose is not yet established, the study starts with the two lower doses and the control arm. Based on safety and efficacy endpoints observed in an interim analysis, it is decided to either continue...

    Go to contribution page
  109. Oladapo Oladoja (Abiola Ajimobi Technical University, Ibadan, Nigeria)
    19/05/2026, 16:39
    oral presentation

    Effective air quality regulation and climate change mitigation depend on reducing greenhouse gas emissions and air pollutants. To achieve sustainable green development, this study constructs a spatio-temporal hierarchical model to assess the Air Quality Index (AQI) across Nigerian states and to derive actionable insights for environmental sustainability. Specifically, this study develops a...

    Go to contribution page
  110. Pat Vatiwutipong (Mahidol University)
    19/05/2026, 16:39

    In many real-world datasets, observations are hierarchically structured, such as students nested within classrooms, hospitals within cities, or repeated measurements from the same patient. Performing machine learning without accounting for this clustered structure can lead to biased predictions and misleading interpretations of feature effects.

    Recently, Mixed Effect Machine Learning, an...

    Go to contribution page
  111. Ekkehard Glimm (Novartis Pharma)
    19/05/2026, 16:39

    Clinical trials have become more complex in recent decades. They have gradually become longer, involving more centers and more patients. With these tendencies, interim analyses of ongoing trials have become much more common and much research has been done on the design and analysis of data from adaptive trials. By now group-sequential trials are probably more frequent than single-stage...

    Go to contribution page
  112. Kaya Miah (Division of Biostatistics, German Cancer Research Center (DKFZ), Heidelberg, Germany)
    19/05/2026, 16:39
    oral presentation

    In the era of precision medicine with increasing molecular information, the use of a multi-state model is required to capture the individual disease pathway along with underlying etiologies with greater precision. Especially the availability of big data with numerous covariates induces several statistical challenges for model building. For multi-state models based on high-dimensional data,...

    Go to contribution page
  113. Mar Rodriguez-Girondo (LUMC)
    19/05/2026, 16:55

    Relative survival techniques are often used to assess excess mortality in a specific study population by splitting observed mortality into background and excess components. These methods have been widely used to estimate cancer-specific mortality without the need for precise cause-of-death data for cancer patients. However, applying these techniques to other settings, such as pandemics or...

    Go to contribution page
  114. Enyu Li (Clinical Trials Unit, University of Warwick)
    19/05/2026, 16:57
    oral presentation

    Background: We consider clinical trials in which the experimental treatment may have heterogeneous effects across pre-specified patient subpopulations. In such settings, two-stage adaptive enrichment designs allow the enrolled population to be modified at an interim analysis. In stage 1, patients are enrolled from the full population, and based on interim data and preplanned selection rules,...

    Go to contribution page
  115. Frank Bretz
    19/05/2026, 16:57
  116. Barbara Więckowska (Department of Computer Science and Statistics, Poznan University of Medical Sciences, Poznan, Poland)
    19/05/2026, 16:57

    Background: Traditional binary classification assessment in machine learning relies heavily on decision thresholds, limiting interpretability and performance in imbalanced scenarios. While metrics like AUC under ROC (Receiver Operating Characteristic curve) provide overall performance measures, they fail to deliver class-specific insights, which is crucial for real-world applications with...

    Go to contribution page
  117. Kirsten Schorning (TU Dortmund University)
    20/05/2026, 09:00

    Concentration-dependent cytotoxicity experiments are frequently used in toxicology. Although it has been reported that an adequate choice of concentrations, i.e., the design, substantially improves the quality of statistical inference, a recent literature review of three major toxicological journals showed that these methods are rarely used in toxicological practice.
    In this talk, we address...

    Go to contribution page
  118. Huiying Zhou Zhou (TU Dortmund University)
    20/05/2026, 10:45
    oral presentation

    Concentration-response curves model the relationship between a concentration of a compound and the response it elicits in a biological system. Here, the viability of cells is considered as response. Typically, parametric models are fitted to the data. Modeling this relationship accurately is crucial for understanding the safety and potency of compounds, since one of the applications of these...

    Go to contribution page
  119. Natasha Karp (AstraZeneca)
    20/05/2026, 10:45

    Preclinical research is rich with nontrivial design problems that demand statistical leadership. These opportunities exist across in vivo and high-throughput in vitro research systems where statisticians can materially improve translational fidelity by aligning biological questions, design, and analysis to support decision making. In this talk, I will discuss examples that have arisen from a...

    Go to contribution page
  120. Kosmas Kepesidis (Ludwig-Maximilians-Universität München)
    20/05/2026, 10:45
    oral presentation

    In this study, we investigate the individuality and information content of infrared molecular profiles derived from blood samples in a large, longitudinal health-profiling cohort and compare them to a standard clinical laboratory panel. Using Fourier-transform infrared spectroscopy, we obtained comprehensive molecular fingerprints from 4,704 self-reported healthy individuals over five visits...

    Go to contribution page
  121. Simon Mack (RWTH Aachen University)
    20/05/2026, 10:45
    oral presentation

    The pseudo-observation regression approach provides a flexible alternative to the omnipresent proportional hazards model when modeling time-to-event outcomes. In this approach, estimands representable as expectations are fitted to regression models using covariates of interest. Exemplary estimands that fit this framework are the restricted mean time lost (in competing risks models) or the...

    Go to contribution page
  122. Sharon Lutz (Harvard Medical School)
    20/05/2026, 10:45
    oral presentation

    In genetic association studies, Mendelian Randomization (MR) is a popular tool for inferring causal relationships between traits using genetic variants as instrumental variables. Recent methods have been proposed as tools that can infer the causal direction between two phenotypes including MR Steiger, bidirectional MR, causal direction-ratio, causal direction-Egger, and causal direction-GLS....

    Go to contribution page
  123. Gloria Brigiari (Unit of Biostatistics, Epidemiology and Public Health, Department of Cardiac, Thoracic, Vascular Sciences and Public Health, University of Padova)
    20/05/2026, 10:45
    oral presentation

    Multiverse analysis offers a powerful framework to assess the robustness of statistical inferences across a spectrum of plausible analytical choices. However, when applied to predictor selection, especially in high-dimensional settings, the issue of multiplicity becomes critical. In this study, we present a comprehensive simulation framework to evaluate the impact of different multiple testing...

    Go to contribution page
  124. Franziska Kappenberg (University of Bonn, Medical Faculty, Institute for Medical Biometry, Informatics and Epidemiology)
    20/05/2026, 11:03
    oral presentation

    Multivariable regression models are a powerful statistical tool with an innumerable number of applications in explanatory and predictive settings. One key challenge is variable selection —deciding which variables to include or exclude, particularly when dealing with large numbers of candidate predictors. In biomedical data, non-linear relationships between the candidate predictors and the...

    Go to contribution page
  125. Lena Schemet (University of Augsburg)
    20/05/2026, 11:03
    oral presentation

    Occam’s Razor suggests that, among several plausible explanations for a phenomenon, the simplest is preferable. Applied to regression analysis, this implies that the smallest model that fits the data is best. Therefore, in terms of analyzing high-dimensional time-to-event data, variable selection techniques are required, if we want to follow the principle of Occam's Razor. A widely used...

    Go to contribution page
  126. Leonie Lenz-Seraphin (Medical University Vienna, Center for Medical Data Science, Institute of Clinical Biometrics)
    20/05/2026, 11:03
    oral presentation

    Heart transplantation is widely regarded as the gold standard for the treatment of end-stage heart failure. However, shortages of donor hearts necessitate the implementation of waiting lists and allocation algorithms. The German Transplantation Law stipulates the allocation of donor hearts based on urgency of and the benefit from a transplantation. This can be summarized into a single score,...

    Go to contribution page
  127. Lukas Frank Buchhäusl (Medical University of Graz)
    20/05/2026, 11:03
    oral presentation

    Despite the availability of vaccines, infectious diseases such as COVID-19, tetanus, diphtheria, and pertussis remain persistent public health threats, particularly among vulnerable populations including pregnant and lactating women. As most research on protection against infectious diseases to date has focused on antibody-mediated responses, understanding how antibodies behave over time...

    Go to contribution page
  128. Lea Gigou (Department of Statistics, Ludwig Maximilian University of Munich; Max Planck Institute of Quantum Optics; Center for Molecular Fingerprinting)
    20/05/2026, 11:03
    oral presentation

    Early detection of lethal diseases such as lung cancer requires resolving faint signals amid biological heterogeneity. Precision screening aims to sensitively detect meaningful departures from an individual’s baseline by considering individual-level rather than population-level variability. This work investigates whether infrared molecular fingerprinting (IMF) - mid-infrared vibrational...

    Go to contribution page
  129. Frank Konietschke (Institute of Biometry, Charité - Universitätsmedizin Berlin)
    20/05/2026, 11:15

    Preclinical studies often operate under strict ethical, logistical, and financial constraints, resulting in experiments with very small sample sizes. These limitations pose substantial challenges for statistical inference, reproducibility, and the reliability of decision-making in early-phase biomedical research. This talk provides an overview of key design and analysis issues in preclinical...

    Go to contribution page
  130. Theresa Ullmann (Institute of Clinical Biometrics, Center for Medical Data Science, Medical University of Vienna)
    20/05/2026, 11:21
    oral presentation

    In regression modeling, relationships between continuous predictors and outcomes are often assumed to be linear, yet allowing for non-linear associations can substantially improve model performance. A variety of methods for flexible regression—such as fractional polynomials and spline-based approaches—have been proposed to model non-linear associations. However, comprehensive and systematic...

    Go to contribution page
  131. David Kronthaler (Epidemiology, Biostatistics and Prevention Institute, Department of Biostatistics, University of Zurich)
    20/05/2026, 11:21
    oral presentation

    Meta-analysis can be formulated as the combination of p-values from multiple studies into a joint p-value function, from which inference for the average effect, including point estimates and confidence intervals, can be derived. We extend Edgington's p-value combination method for random-effects meta-analysis by treating the combined p-value function as a confidence distribution of the average...

    Go to contribution page
  132. Moritz Madern (Medical University of Vienna, Center for Medical Data Science, Institute of Clinical Biometrics)
    20/05/2026, 11:21
    oral presentation

    We consider the following prediction problem using observational data obtained from routine health-care visits. Biomarkers such as blood pressure and cholesterol are repeatedly measured over time, resulting in sparse and irregular longitudinal data for thousands of individuals. In addition, we observe corresponding survival outcomes, such as the time to cardiovascular disease or death, which...

    Go to contribution page
  133. Marina Bleskina (Institute of Medical Biometry and Statistics, University of Lübeck, University Hospital of Schleswig-Holstein, Campus Lübeck)
    20/05/2026, 11:21
    oral presentation

    Personalized medicine aims to improve the treatment of complex diseases by tailoring therapies to the individual molecular characteristics of patients. This is possible by using multi-omics data, which combine different molecular modalities from the same individuals. Integrating these modalities allows more comprehensive and powerful modeling. However, their unique characteristics make...

    Go to contribution page
  134. Sumit Das (Scientist - I, All India Institute of Medical Sciences (AIIMS))
    20/05/2026, 11:21
    oral presentation

    Background: Childhood immunization influences directly and indirectly fourteen out of the seventeen sustainable development goals (SDGs). Timely receipt of vaccines protects children from deadly diseases and increases the overall future productivity of the population. With the largest and most heterogeneous population of under-five children, the delay in receiving polio vaccination has not...

    Go to contribution page
  135. María Arroyo Araujo (Berlin Institue of Health, Charité Universitätsmedizin)
    20/05/2026, 11:35

    Translating preclinical results into effective clinical treatments remains a big challenge in biomedical research. Too often, findings from single-laboratory studies fail to replicate in the preclinical context and further, to show effectiveness in clinical trials. One promising approach to validate exploratory findings is through confirmatory multi-laboratory preclinical trials—studies that...

    Go to contribution page
  136. Jiumeng Zhang (Leibniz-Institute for Prevention Research and Epidemiology - BIPS)
    20/05/2026, 11:39
    oral presentation

    Mixed-effects models (MEMs) are widely used in epidemiology to analyze data not being independent and identically distributed (i.i.d.) like longitudinal data. However, MEMs rely on parametric assumptions and require predefined interactions among predictors. In contrast, machine learning (ML) methods such as random forests (RF) assume i.i.d. data but are more flexible in capturing nonlinear...

    Go to contribution page
  137. Lorena Hafermann (Institute of Biometry and Clinical Epidemiology, Charité - Universitätsmedizin Berlin)
    20/05/2026, 11:39
    oral presentation

    In descriptive studies, where the primary goal is to identify key predictors of a time-to-event outcome, and in predictive research involving numerous candidate predictors, data-driven variable selection methods are often employed to narrow down the pool of variables. This is particularly necessary when domain expertise is limited or when the practical utility of a prediction model is...

    Go to contribution page
  138. Wiebke Dammann (TU Dortmund University - Department of Statistics)
    20/05/2026, 11:39
    oral presentation

    Virtual Control Groups (VCGs) represent an approach in which historical control data (HCD) from previous animal studies are used to replace animals in current control groups. The VICT3R project (Developing and implementing VIrtual Control groups To reducE animal use in toxicology Research), funded by the Innovative Health Initiative (IHI), aims to reduce the use of animals in toxicological...

    Go to contribution page
  139. Ema Požek (Faculty of Medicine, University of Ljubljana)
    20/05/2026, 11:39
    oral presentation

    Breast cancer remains one of the most common cancers among women worldwide. Breast cancer screening programmes aim to catch the disease at its early phase, by regularly examining asymptomatic women for signs of cancer. The rationale is straightforward: early detection, before symptoms onset, offers patients broader treatment options and improves the chances of recovery. To evaluate the cancer...

    Go to contribution page
  140. Marléne Baumeister (TU Dortmund University, Department of Statistics)
    20/05/2026, 11:39
    oral presentation

    Functional data analysis (FDA) has become increasingly popular in medical biometry and statistics. It is often appropriate to model observations by smooth curves or functions for example in the situation of observations that are sampled quite dense over time or space or in case of high-dimensional repeated measurements as FDA methods allow a flexible modelling. Furthermore they does not assume...

    Go to contribution page
  141. Leonhard Held (University of Zurich)
    20/05/2026, 11:55

    Reducing the number of experimental units is one of the three pillars of the 3R principles (Replace, Reduce, Refine) in animal research. At the same time, statistical error rates need to be controlled to enable reliable inferences and decisions. This paper proposes a novel measure to quantify the evidentiary value of one experimental unit for a given study design. The experimental unit...

    Go to contribution page
  142. Jana Habus-Korbar (Student at University of Zagreb, Faculty of Science, Department of Mathematics)
    20/05/2026, 11:57
    oral presentation

    In this study, PERMY data set taken from Pharmaceutical Statistics Using SAS: A Practical Guide is analyzed. It describes permeability of cell membranes, which is the ability of a molecule to cross a membrane. Biological structures are a complex layer of molecules and proteins. Substances require a particular structure to pass through the target membrane and drugs that fail to demonstrate...

    Go to contribution page
  143. Steffen Hadasch (National Institute of Public Health, University of Southern Denmark)
    20/05/2026, 11:57
    oral presentation

    Assessing micronutrient status is essential in nutritional research (Allen, 2025) and typically involves estimating the population wide prevalence of micronutrient deficiencies using biomarker data collected across multiple regions. In such studies, several biomarkers are commonly analyzed to estimate the prevalence of any deficiency, defined as the probability that at least one of the...

    Go to contribution page
  144. Subhadra Dasgupta (IUF - Leibniz Research Institute for Environmental Medicine)
    20/05/2026, 11:57
    oral presentation

    Large and high-dimensional biomedical datasets (large n and p) such as genotype data containing hundreds of thousands of genetic variants (SNPs) measured across many individuals require scalable algorithms to enable efficient model training. In this work, we address this challenge by leveraging principles of optimal design and informative subsampling.

    We investigate the applicability of...

    Go to contribution page
  145. Moses Mwangi (Kenya Medical Research Institute, Kenya; University of Hasselt, Belgium)
    20/05/2026, 11:57
    oral presentation

    Repeated measures data are commonly encountered in a wide variety of disciplines including business, agriculture and medicine. They entail collection of multiple measurements from the same unit or subject over time, space or both. The fact that observations from the same unit will not be independent poses particular challenges to the statistical procedures used for the analysis of such data....

    Go to contribution page
  146. Marieke Stolte (TU Dortmund University, Department of Statistics)
    20/05/2026, 13:45

    Quantifying the similarity between two or more datasets is an important task in statistics and machine learning. In meta-learning, it enables the transfer of knowledge across tasks and datasets. In simulation studies, the similarity between the distributions assumed in the simulation and the distributions of the datasets for which the performance of methods is assessed is crucial. Similarly,...

    Go to contribution page
  147. Martin Geroldinger (Research Program Biomedical Data Science, Paracelsus Medical University Salzburg)
    20/05/2026, 13:45
    oral presentation

    The servEB project (WISS 2025, federal state of Salzburg, 20102/F2300645-FPR) combines
    clinical expertise, advanced statistical analyses, and AI-driven imagine classification technology to improve the assessment of trial outcomes in rare diseases, especially Epidermolysis Bullosa research as an example. When defining meaningful endpoints, multiple aspects of the disease have to be considered,...

    Go to contribution page
  148. Ronja Foraita (Leibniz Institute for Prevention Research and Epidemiology - BIPS)
    20/05/2026, 13:45
    oral presentation

    In field studies, measurements are often collected over extended periods, during which subtle shifts in data quality or instrument performance can occur. Recognizing and quantifying such measurement heterogeneities over time is essential to ensure the validity of study results and to intervene at an early stage if possible. However, the performance of available statistical approaches for...

    Go to contribution page
  149. Rafał Pawłowski (Collegium Medicum of Nicolaus Copernicus University)
    20/05/2026, 13:45
    oral presentation

    Background: Heart Rate Asymmetry (HRA) represents a specialized domain of Heart Rate Variability (HRV) analysis, quantifying the unequal contribution of accelerations and decelerations to the overall heart rate variations. While HRA provides unique insight into the nonlinear dynamics of autonomic control, its assessment has traditionally relied on high-resolution Electrocardiography (ECG)....

    Go to contribution page
  150. Natalia Stefańska (Adam Mickiewicz University)
    20/05/2026, 13:45
    oral presentation

    In the functional response model (FRM), where a functional response is explained by scalar predictors, inference becomes challenging when the design matrix is not full-rank, leading to an ill-conditioned model (ICFRM). Widely used methods for this problem, such as $L^2$-norm-based tests (Zhang, 2013), suffer from critical flaws such as poor control of the type I error rate, which can...

    Go to contribution page
  151. Helena Geys (Johnson&Johnson Innovative Medicine R&D)
    20/05/2026, 13:45

    Scientific integrity is the cornerstone of progress in biomedical research. Nowhere is this more critical than in nonclinical settings. Reproducibility – the ability to consistently replicate findings across studies, laboratories and organisations is – is not just a technical requirement. It is a fundamental attribute that underpins trust. As nonclinical research continues to expand in...

    Go to contribution page
  152. Dominik Nowakowski (Medical University of Bialystok, Department of Biostatistics and Medical Informatics)
    20/05/2026, 14:00
    oral presentation

    Over the past two decades, the problem of selecting relevant variables in high-dimensional data analysis has gained particular importance in both statistics and machine learning. Despite substantial advances in modeling techniques and numerous algorithmic proposals, most existing approaches overlook the issue of missing observations — a phenomenon ubiquitous in real-world datasets, especially...

    Go to contribution page
  153. Laura Slebioda (Department of Mathematical and Statistical Methods, Poznań University of Life Sciences)
    20/05/2026, 14:03

    The assessment of crop variety distinctness, uniformity, and stability (DUS) is a fundamental component of plant breeding and registration processes. Traditionally, one-dimensional analysis of variance is conducted separately for each attribute. However, before conducting separate analyses, it would be worthwhile to apply multivariate methods to determine whether a given variety differs from...

    Go to contribution page
  154. Martin Schnuerch (Global Biostatistics & Data Sciences, Boehringer Ingelheim Pharma GmbH & Co. K)
    20/05/2026, 14:03
    oral presentation

    Binary endpoints are commonly used to measure clinical outcomes in randomized controlled trials. In this context, conditional odds ratios (ORs) based on logistic regression have been routinely used as population-level summary to quantify treatment effects. However, ORs have been criticized for a lack of interpretability, non-collapsibility, and sensitivity to model specification. In response,...

    Go to contribution page
  155. Timur Tug (Fraunhofer Institute for Toxicology and Experimental Medicine ITEM)
    20/05/2026, 14:03
    oral presentation

    The principles of Replacement, Reduction, and Refinement (3Rs) have become fundamental to modern biomedical research. In this context, Virtual Control Groups (VCGs) offer a promising strategy to reduce the number of animals used in toxicological and pharmacological studies. Rather than including concurrent control groups (CCGs) in every experiment, VCGs rely on historical control data...

    Go to contribution page
  156. Carsten Schmidt (University Medicine Greifswald City: Greifswald)
    20/05/2026, 14:03
    oral presentation

    Longitudinal observational studies and clinical trials routinely collect extensive phenotypic data under changing organisational, technical, and environmental conditions. Variations in examiners, devices, protocols, or ambient factors can introduce consequential forms of measurement heterogeneity and measurement error over time. Although these sources of bias are well recognised, systematic...

    Go to contribution page
  157. Filip Pieczątkiewicz (Adam Mickiewicz University)
    20/05/2026, 14:15
    oral presentation

    In modern data analysis, technological advancements frequently result in the collection of Functional Data (FD), where observations are naturally represented as smooth functions, curves, or surfaces over a continuum (e.g., time or space). Examples include daily stock prices, continuous temperature recordings, or spectroscopic measurements. Functional Data Analysis (FDA) offers a powerful...

    Go to contribution page
  158. Ulf Toelch (BIH QUEST Center for Responsible Research)
    20/05/2026, 14:15

    Low rates of replicability in early phase biomedical research hinder progress and putatively cause high attrition rates in clinical trials. To improve evidence generation processes, preclinical confirmatory studies and preregistration offer potentially effective strategies. By comparing conduct and outcome of preclinical studies utilizing such strategies, we examined how different degrees of...

    Go to contribution page
  159. Ulbrich Hannes-friedrich (Bayer AG, Pharmaceuticals)
    20/05/2026, 14:21
    oral presentation

    In pharmaceutical research and preclinical development data below the lower limit of quantitation are quite common although sometimes not properly dealt with. Beyond time-to-event settings measured data above a general or even subject specific upper limit of quantitation are less common.
    Malignant tumor cells can metastasize. When tumor cells metastasize they might cause new tumors called...

    Go to contribution page
  160. Narges Ghoreishi (Department Exposure, Unit of Epidemiology statistics and exposure modelling, German Federal Institute for Risk Assessment (BfR))
    20/05/2026, 14:21
    oral presentation

    Background: Tattoos and permanent make-up (PMU) gain increasing popularity, yet their potential systemic health implications remain poorly understood.
    Methods: To investigate associations between tattoos/PMU and chronic disease outcomes, we analyzed data from the LIFE-Adult Study, a population-based cohort of 10,000 adults recruited in Leipzig, Germany (2011–2014). A dedicated...

    Go to contribution page
  161. Małgorzata Ćwiklińska-Jurkowska (Department of Biostatistics and Biomedical Systems Theory, Nicolaus Copernicus University)
    20/05/2026, 14:21

    The aim of the work is to find important characterizations of mixture of experts which have
    an impact on improvement of combined classifier performance over the averaged
    performance of the base learners. The problem was examined for various high
    dimensional genomic data sets.
    Mixture of experts are useful for responses differentiating among base classifiers.
    From this point of view...

    Go to contribution page
  162. Henrik Stahl (University of Applied Sciences Darmstadt)
    20/05/2026, 14:30
    oral presentation

    In clinical development it is essential to identify subgroups of patients who exhibit a beneficial treatment effect, ideally before moving to confirmatory trials. Such subgroups are often defined by predictive biomarkers with corresponding cut-off values. However, data-driven selection of biomarkers or cut-offs introduces selection bias, i.e. the treatment effect within the selected subgroup...

    Go to contribution page
  163. Dario Zocholl (University of Bonn, Medical Faculty, Institute for Medical Biometry, Informatics and Epidemiology)
    20/05/2026, 14:35

    Animal experiments are often purely exploratory, with little to no data available to support the planning phase. Nonetheless, ethical guidelines demand scientifically sound biometric planning. The experimental designs are typically complex, involving numerous experimental groups and adaptive steps, which complicates statistical planning.

    In recent years, statistical aspects of such...

    Go to contribution page
  164. Inken Siems (Trier University)
    20/05/2026, 14:39
    oral presentation

    High-quality data are essential for reliable epidemic surveillance. Traditional systems relying on passive case reporting that may lead to unreliable prevalence estimates depending on the specific disease. Using the example of the COVID-19 pandemic, we show that once prevalence exceeds moderate levels, conventional reporting becomes biased and unstable. Beyond this point, drawing additional...

    Go to contribution page
  165. Teresa Byczkowski (Institute of Medical Biometry, Heidelberg University)
    20/05/2026, 14:39
    oral presentation

    In many clinical trial analyses, missing data is addressed through multiple imputation (MI) to avoid loss of information and potential bias. However, this approach is not taken into consideration at the planning stage when calculating the sample size. Here, it is common practice to inflate the calculated sample size by an estimated dropout rate in order to maintain the desired power. This...

    Go to contribution page
  166. Lea Kronziel (University of Lübeck)
    20/05/2026, 14:39

    Background:
    A random forest (RF) is an efficient method for prediction but it is difficult to
    interpret.
    Artificial Representative Trees (ARTs) are a special type of surrogate model
    that approximates the original strucutre of the RF in a single tree, achieving
    similar predictive accuracy.
    Conformal Predictive Systems (CPS) provide a framework for uncertainty
    quantification by generating...

    Go to contribution page
  167. Daniel Bodden (RWTH Aachen University)
    20/05/2026, 14:39
    oral presentation

    Even in rare diseases, where the sample size is limited and blinding is less frequently implemented, randomized controlled trials are considered the gold standard to proof efficacy. Randomization is used to mitigate bias and regulatory guidance recommend the investigation of the impact of bias on the test decision. We quantified how allocation bias affects the test decision in small sample...

    Go to contribution page
  168. Mathilde Dicaire-Cartier (Institute for Medical Information Processing, Biometry, and Epidemiology, Faculty of Medicine, LMU Munich, Germany; Munich Center for Machine Learning, Munich, Germany; Department of Mathematics and Statistics, Université de Montréal, Montréal, Canada)
    20/05/2026, 14:45
    oral presentation

    Non-experimental data, such as electronic medical records, are often used in causal inference to estimate the effect of an exposure on an outcome of interest. However, this type of data can be affected by potential sources of bias in causal analyses. For example, these data do not come from a study design that ensures a balance of patient characteristics between exposure groups, a problem...

    Go to contribution page
  169. Nicole Ellenbach (Institute for Medical Information Processing, Biometry and Epidemiology, Faculty of Medicine, LMU Munich, Germany and Munich Center for Machine Learning, Munich, Germany)
    20/05/2026, 14:55

    In preclinical animal studies, researchers often have a certain degree of freedom when it comes to selecting the exact statistical analysis strategy for their experiment. Ideally, this analysis strategy should be specified prior to the experiment (and preregistered, if possible), with sample size planning conducted in accordance with the chosen analytical approach. Sample size calculations...

    Go to contribution page
  170. Eleonora Di Carluccio (Cardio-CARE)
    20/05/2026, 14:57
    oral presentation

    Statistical prediction models for binary outcomes are becoming increasingly popular. One signifi‐ cant challenge is calibrating these models to suit the characteristics of a target population that is structurally different from the original population. Calibration is especially challenging when there is no training data available from the target population. To address this problem, we propose...

    Go to contribution page
  171. Andreas Ziegler (Cardio-CARE)
    20/05/2026, 14:57

    Sharing of original study data may be restricted by data protection policies. Instead, synthetic data that mimics the original data structure may be shared between research groups. This work introduces modgo 2.0 which may be used for generating synthetic data from existing study data. Simulations may be based either on the combination of the rank inverse normal transformation with simulation...

    Go to contribution page
  172. Nico Bruder (Institute of Medical Statistics, Medical University of Vienna)
    20/05/2026, 14:57
    oral presentation

    In rare diseases, the need for innovative clinical trial designs is increasing. Platform trials are becoming particularly popular, as they allow for flexible adding and dropping of arms and reduce sample size requirements by using a shared control. In a platform trial setting with two experimental arms and one control, we use clinical trial simulations to quantify the impact on operating...

    Go to contribution page
  173. Paula Lorenz (TU Dortmund University)
    20/05/2026, 15:00
    oral presentation

    Meta-analyses synthesise the results of multiple independent studies to obtain more comprehensive knowledge about a research topic. When study outcomes vary, meta-regression can be used to identify potential sources of heterogeneity across studies. One complication is the typically small number of studies available. Due to this, interaction terms are often omitted in meta-regression models,...

    Go to contribution page
  174. Susanne Strohmaier (Medical University of Vienna, Center for Public Health, Department of Epidemiology City: Vienna)
    21/05/2026, 09:00

    For many medical research questions, randomization is unethical or infeasible and decisions have to be informed by results based on observational – often routinely collected - data. Such data have enormous potential to inform stakeholders including health policy makers, health professionals and the general public about the impact of their decisions on public as well as individual health....

    Go to contribution page
  175. Belay Birlie Yimer (1. Astellas Pharma Europe Ltd., Addlestone, United Kingdom)
    21/05/2026, 10:45
    oral presentation

    The FDA initiated Project Optimus and issued guidance for dose optimization, recommending randomized parallel dose-response cohorts to generate additional data at promising dose levels and implies that different dosages may be needed for different indications. In addition to dose optimization, with recent advancements in precision medicine and cancer biology, the development of cancer...

    Go to contribution page
  176. Anamarija Jazbec (University of Zagreb Faculty of Forestry and Wood Technology)
    21/05/2026, 10:45

    Regeneration of forest ecosystems is crucial for preserving their structure, function and long-term stability. This research analyses the correlation and similarity of the occurrence of 11 species characteristic only for the regeneration phase with respect to some soil properties. Aim of research is to perform grouping of species according to soil properties, and to determine differeces in...

    Go to contribution page
  177. Jaroslaw Harezlak (Indiana University)
    21/05/2026, 10:45

    Sport-related concussions (SRCs) represent a major public health concern, accounting for more than 200,000 annual Emergency Department visits in the United States. Biomechanically, SRCs arise from head impacts that generate high-magnitude linear and rotational accelerations. Increasing evidence from human studies indicates that repetitive head impact exposure (HIE) reduces concussion tolerance...

    Go to contribution page
  178. Björn-hergen Laabs (University Medical Center Göttingen)
    21/05/2026, 10:45

    Artificial intelligence (AI) is intended to support clinicians, therapists, patients, hospital managers, and clinical data scientists at all levels. This includes, for example, making clinical diagnoses, understanding the causes of diseases, and planning clinical studies. The enormous increase in the importance of AI in medicine has led to the development of several guidelines (e.g.,...

    Go to contribution page
  179. Christian Ritz (National Institute of Public Health (SDU))
    21/05/2026, 10:45
    oral presentation

    In epidemiology dose-response meta analysis often refers to fitting a meta regression model that describes a linear trend in the outcome ("response") as a function of the exposure ("dose"), based on aggregated data from a number of studies.

    Fixed- and random-effects extensions for handling nonlinear dose-response for odds ratios, relative risks and differences in means through the use of...

    Go to contribution page
  180. Andreas Gleiss (Medical University of Vienna, Center for Medical Data Science)
    21/05/2026, 10:45

    Reference intervals and standard deviation scores (‘z scores’) are widely used as diagnostic tools in various biomedical fields. They are applied to laboratory parameters in clinical chemistry, psychometric tests in neurology, or parameters of children’s growth in pediatrics. Usually, samples from a ‘normal’ or ‘healthy’ population form the data basis for the estimation of reference...

    Go to contribution page
  181. Lukas Baumann (Institute of Medical Biometry, University of Heidelberg)
    21/05/2026, 11:03
    oral presentation

    Basket trials examine the efficacy of a single intervention simultaneously in several patient subgroups. They are currently mostly applied in oncology, where the subgroup assignment is based on medical characteristics such as a common biomarker. This can result in small sample sizes within subgroups that are also likely to differ. Several designs for the analysis of basket trials have been...

    Go to contribution page
  182. Ilker Unal (Cukurova University Faculty of Medicine Dept of Biostatistics)
    21/05/2026, 11:03

    Classification plays a pivotal role in medicine for both diagnostic and prognostic purposes. Traditionally, diagnostic efficacy is evaluated using prevalence-independent metrics, such as sensitivity and specificity. For numerical tests, the Area Under the Receiver Operating Characteristic (ROC) curve is the standard for assessing classification success. However, the rising adoption of machine...

    Go to contribution page
  183. Dennis Dobler (RWTH Aachen University)
    21/05/2026, 11:03

    Modern large language models (LLMs) have reshaped workflows of people across countless fields - and biostatistics is no exception. These models offer novel support in drafting study plans, generating software code, or writing reports. However, reliance on LLMs carries the risk of inaccuracies due to potential hallucinations that may produce fabricated "facts", leading to erroneous statistical...

    Go to contribution page
  184. Christian Röver (Department of Medical Statistics, University Medical Center Göttingen)
    21/05/2026, 11:03
    oral presentation

    Updating a meta-analysis (MA) by including additional studies is usually a straightforward exercise, as the relevant data are commonly reported in detail, i.e., effect estimates with standard errors for all studies. Matters are complicated, however, when only the summary of a previous analysis is available, i.e., the overall estimate with standard error. For instance, this is sometimes the...

    Go to contribution page
  185. Anne-katrin Gorn (Bavarian Research Center for agriculture)
    21/05/2026, 11:03

    In Germany, cultivars are tested for regional recommendations in federal state cultivar trials, taking the form of multi-environment trials (METs). Their primary objective is to identify cultivars that are best suited for regional production in agro-ecological zones. For perennial ryegrass, current selection decisions are predominantly based on yield. Incorporating additional quality...

    Go to contribution page
  186. Iuliana Ionita-Laza (Columbia University)
    21/05/2026, 11:15

    Genome-wide association studies (GWAS) for biomarkers and molecular phenotypes can lead to clinically relevant discoveries. Numerous lines of evidence from both model organisms and human studies suggest that genetic associations can be highly heterogeneous, dynamic and context dependent. Despite twenty years of GWAS, most studies are based on statistical models that fail to account for such...

    Go to contribution page
  187. Maksym Hrachov (University of Hohenheim)
    21/05/2026, 11:21

    Environmental covariates (ECs) have become increasingly abundant and accessible over the past two decades, driven by advancements in remote sensing, data acquisition technologies, and the declining costs of environmental monitoring. Incorporating ECs into multi-environment trials (METs) has several applications, including improving the understanding of genotype-by-environment interactions,...

    Go to contribution page
  188. Philipp Weber (Institute of Medical Biometry and Epidemiology)
    21/05/2026, 11:21

    The accuracy of diagnostic tests is commonly evaluated by estimating the area under the receiver operating characteristic curve (AUC), as well as sensitivity and specificity at given diagnostic cut-offs. However, many diagnostic trials use factorial designs. For example, different combinations of readers and methods may be used to diagnose a patient. Furthermore, diagnostic studies may...

    Go to contribution page
  189. Lukas D Sauer (Institute of Medical Biometry, Heidelberg University)
    21/05/2026, 11:21
    oral presentation

    Modern therapeutic agents in cancer therapy often target specific genetic traits of the tumor. Whenever these traits are independent of the tissue in which the tumor is located, the therapeutic agent may be tissue-agnostic, meaning that it can be applied regardless of location. Clinical trials for such tissue-agnostic therapies often have small sample sizes. Hence, it is efficient to recruit...

    Go to contribution page
  190. Jan-Bernd Igelmann (TU Dortmund University)
    21/05/2026, 11:21
    oral presentation

    Meta-analyses often involve transforming bounded effect size measures, such as correlation coefficients or odds ratios, onto a real-valued scale prior to estimation. The results are then back-transformed to the original scale for interpretation purposes. However, in the standard random effects model for meta-analysis, simply applying the inverse transformation function generally does not yield...

    Go to contribution page
  191. Ursula Berger (LMU Munich)
    21/05/2026, 11:21

    Artificial intelligence (AI) is increasingly being used in various disciplines. Examples of this include medical image processing, complex prediction and decision support models and thereby integrating with the field of biostatistics. In this context, integrating AI and machine learning (ML) methods within courses of biostatistics taught to students of medicine, health and life sciences offers...

    Go to contribution page
  192. Jonas Wallin (department of statistics)
    21/05/2026, 11:35

    Protein degradation is a regulated process that reshapes the proteome and generates bioactive peptides. Peptidomics and degradomics enables large-scale measurement of these peptides, yet most
    data analyses approaches treat peptides as isolated endpoints rather than intermediates produced
    by sequential cleavage. Here, we introduce degradation graphs, a probabilistic framework that represents...

    Go to contribution page
  193. Arzu Kanik (AB Health Tech)
    21/05/2026, 11:39
    oral presentation

    Background:
    A substantial proportion of clinical research waste originates from fundamental methodological flaws—improper study design, insufficient power, inappropriate statistical methods, and non-compliance with reporting guidelines. While many AI tools attempt to support data analysis, none address the critical upstream phase: validating methodology before data collection. To address this...

    Go to contribution page
  194. Peter Degen (Center for Reproducible Science and Research Synthesis, University of Zurich)
    21/05/2026, 11:39
    oral presentation

    In many biomedical research settings, sufficiently large sample sizes can only be achieved by combining data from multiple collection sites (e.g., hospitals). However, pooling individual participant data in a central server is often restricted due to privacy and regulatory constraints. Federated inference addresses this challenge by distributing the statistical analysis across local sites,...

    Go to contribution page
  195. Sunil Mathur (Weill Cornell Medical College)
    21/05/2026, 11:39

    Background
    Triple-negative breast cancer (TNBC) represents one of the most aggressive and treatment-resistant breast cancer subtypes. Patients with locally advanced unresectable or metastatic TNBC (mTNBC) typically face a median overall survival of only 8 to 13 months, highlighting the urgent need for efficient drug evaluation strategies. Conventional statistical methods often assume...

    Go to contribution page
  196. Maryna Prus (University of Hohenheim)
    21/05/2026, 11:39

    New crop varieties are extensively tested in multi-environment trials in order to obtain a solid basis for recommendations to farmers. When the target population of environments is large, a division into sub-regions is often advantageous. If the same set of genotypes is tested in each of the sub-regions, a linear mixed model (LMM) may be fitted with random genotype-within-sub-region effects....

    Go to contribution page
  197. Frank Bretz, Presenters, Sarah Friedrich-Welz
    21/05/2026, 11:39
  198. Mateusz Staniak (University of Wrocław)
    21/05/2026, 11:55

    Bottom-up mass spectrometry-based proteomics studies changes in protein abundance and structure across various biological conditions. Since the currency of these experiments are peptides, i.e. subsets of protein sequences that carry the quantitative information, conclusions at a different level, e.g., at the level of proteins or of post-translational modifications, must be computationally...

    Go to contribution page
  199. Milene Figueira (UFRPE)
    21/05/2026, 11:57

    The use of precision agriculture contrasts with the challenge posed by the high cost of commercial technologies, particularly for small-scale producers. For this reason, it is necessary to develop low-cost, accessible solutions that can be applied directly in the productive environment. Within this context, this work presents the development of a low-cost system for acquiring 3D images of beef...

    Go to contribution page
  200. Ferdinand Valentin Stoye (Biostatistics and Medical Biometry, Medical School OWL, Bielefeld University)
    21/05/2026, 11:57

    Meta-analysis of diagnostic test accuracy (DTA) studies deals with aggregating information from multiple studies on sensitivity and specificity. Classical approaches to this task select a single pair of sensitivity and specificity per study (single threshold methods, STM), possibly ignoring additional information if studies report results on multiple diagnostic thresholds. In recent years,...

    Go to contribution page
  201. Anna Eleonora Carrozzo (Salzburg Research, Austria / Paris Lodron University of Salzburg, Austria)
    21/05/2026, 11:57
    oral presentation

    Title:
    Evaluating Nonparametric Combination Methods for Aggregating N-of-1 Trials: A Simulation-Based Comparison with Meta-Analysis

    Abstract:
    Aggregating results from multiple N-of-1 trials has become increasingly relevant for evaluating personalized and digital health interventions, where inter-individual heterogeneity and complex temporal structures challenge traditional study designs....

    Go to contribution page
  202. Günter Heimann (Independent Consultant)
    21/05/2026, 11:57
    oral presentation

    Confidence distributions are a frequentist alternative to Bayesian posterior distributions. They summarize the knowledge and uncertainty about an unknown model parameter in the form of a probability distribution on the parameter space, just like a posterior distribution, without assuming that the parameter of interest is a random variable. Although confidence distributions are a relatively old...

    Go to contribution page
  203. Sharon-lise Normand (Harvard Medical School)
    21/05/2026, 13:45

    In healthcare provider profiling, accurately assessing hospital performance is crucial for informed decision-making and quality improvement. Traditional approaches rely heavily on parametric regression models for risk adjustment, but these methods often fail to account for between-center heterogeneity and may produce biased estimates, especially in the presence of low event rates or small...

    Go to contribution page
  204. Christiana Drake (University of California, Davis)
    21/05/2026, 13:45
    oral presentation

    Randomized trials often utilize a select group of study participants. This group does not typically represent the general population. Furthermore, sample sizes are often small to reduce cost. To improve power and generalizability, external control groups may be added to the randomized study. It is possible to incorporate a suitably selected external control group into a randomized clinical...

    Go to contribution page
  205. Sara Garber (Department of Statistics and Data Science, University of Augsburg)
    21/05/2026, 13:45
    oral presentation

    Functional data analysis has established itself as a powerful framework for analyzing data recorded over continuous domains such as time. Within this context, functional motif discovery refers to the identification of recurrent patterns that appear multiple times across different portions of a single curve and/or within misaligned portions of multiple curves. In this study, we explore the...

    Go to contribution page
  206. Bart Mertens (Leiden University Medical Centre)
    21/05/2026, 13:45
    oral presentation

    Prediction in the presence of missing values is a complex and still poorly understood problem, particularly when future records also contain missing values.
    Mertens, et al. (2020) demonstrate that with non-linear models (such as logistic regression or Cox survival) and when using imputations, averaging of multiple predictions obtained from distinct models fitted on imputed data should be...

    Go to contribution page
  207. Alexandra Balzer (Institute of Medical Biometry Heidelberg University Hospital)
    21/05/2026, 13:45
    oral presentation

    In oncology drug development, phase II dose-finding studies are essential to identify the most promising dose levels for confirmatory phase III trials. Traditionally, dose selection is based on the maximum tolerated dose, which does not necessarily correspond to the optimal dose in terms of efficacy and safety. To address the challenge of dose optimization, the Oncology Center of Excellence of...

    Go to contribution page
  208. Kimberley Wever (Radboud university medical center)
    21/05/2026, 13:45

    Growing concerns about the reproducibility, generalisability and (more recently) credibility of biomedical research publications underscore the need for methods that both synthesise evidence and diagnose weaknesses in the research ecosystem. Systematic review and meta-analysis of animal studies is traditionally used to evaluate preclinical efficacy and inform future research in animals or...

    Go to contribution page
  209. Konrad Banaś (Department of Mathematical and Statistical Methods City: Poznań)
    21/05/2026, 14:03
    oral presentation

    Experimental designs with orthogonal block structures are commonly used in many areas of science in order to control the external sources of variability. The aim of this study is to compare several analysis of variance (ANOVA) methods applicable to such structures. Comparing these approaches is of practical importance, as the choice of the analytical method may influence inference about...

    Go to contribution page
  210. Armin Schüler (BfArM - Federal Institute for Drugs and Medical Devices)
    21/05/2026, 14:03
    oral presentation

    Randomised controlled trials (RCTs) are the gold standard of evidence to support causal conclusions on the benefits and risks of medicines in regulatory decision making along the lifecycle [1]. However, single-arm trials (SATs) are also frequently used for various reasons during drug development. While RCTs allow adjustment for confounding via design, the contextualization of SATs requires...

    Go to contribution page
  211. Dörte Huscher (Institute of Biometry and Clinical Epidemiology, and Berlin Institute of Health, Charité - Universitätsmedizin, corporate member of Freie Universität Berlin and Humboldt-Universität zu Berlin)
    21/05/2026, 14:03
    oral presentation

    In a prospective study of patients with muco-obstructive lung disease, aiming to develop a cough alert system based on nocturnal cough monitoring, to identify patient-individual thresholds at which cough frequency exceeds normal variability, so far 92 of intended 220 patients were included.
    From den Brinker et al. in a study with 30 COPD patients it is known, that the day-to-day variation of...

    Go to contribution page
  212. Sebastien Haneuse (Harvard T.H. Chan School of Public Health)
    21/05/2026, 14:03

    Although not without controversy, readmission is entrenched as a hospital quality metric with statistical analyses generally based on fitting a logistic-Normal generalized linear mixed model. Such analyses, however, ignore death as a competing risk, although doing so for clinical conditions with
    high mortality can have profound effects; a hospital’s seemingly good performance for readmission...

    Go to contribution page
  213. Medina Feldl (Institute of Computational Biology, Helmholtz Munich)
    21/05/2026, 14:03
    oral presentation

    Quantitative analysis of microbial growth curves is essential for understanding how bacterial populations respond to environmental cues. Traditional analysis approaches make parametric assumptions about the functional form of these curves, limiting their usefulness for studying conditions that distort standard growth curves. In addition, modern robotics platforms enable the high-throughput...

    Go to contribution page
  214. Florian Frommlet (Medical University Vienna)
    21/05/2026, 14:15

    Preclinical studies tend to suffer from an unacceptably low rate of replicability, which is highly problematic since unreliable results from animal trials cannot provide a sound foundation for subsequent clinical research. Numerous factors contribute to this issue, including the misapplication of statistical methods, poor study design, and inadequate reporting of results. Although specific...

    Go to contribution page
  215. David Jesse (F. Hoffmann-La Roche AG, Basel, Switzerland; Department of Medical Statistics, University Medical Center Göttingen, Göttingen, Germany)
    21/05/2026, 14:21
    oral presentation

    Bayesian dynamic borrowing (BDB) methods are popular for incorporating historical data in rare disease or paediatric clinical trials, in particular with regard to control groups. They can be used to leverage the historical information while mitigating the consequences of potential prior-data conflicts to some degree. However, these methods do not consider baseline covariate information that...

    Go to contribution page
  216. Han Chiam (Medical University of Vienna)
    21/05/2026, 14:21
    oral presentation

    Population-scale genomic biobanks provide unique opportunities for data-driven drug target discovery. However, these resources often lack detailed data on clinical phenotypes, whereas clinical trials offer rich phenotypic information but are limited in omics coverage and mostly lack genotyping. This imbalance creates gaps in the mechanistic interpretation of clinical findings.
    To address...

    Go to contribution page
  217. Hans-peter Piepho (Universiy of Hohenheim City: Stuttgart)
    21/05/2026, 14:21
    oral presentation

    Row-column designs play an important role in applications where two orthogonal sources of error need to be controlled for by blocking. Field or greenhouse experiments, in which experimental units are arranged as a rectangular array of experimental units are a prominent example. In plant breeding, the amount of seed available for the treatments to be tested may be so limited that only one...

    Go to contribution page
  218. Edouard F. Bonneville (LMU Munich, and Munich Center for Machine Learning)
    21/05/2026, 14:21
    oral presentation

    Multiple imputation (MI) continues to be a popular approach to deal with missing at-random covariate data. For MI to perform well, it is advisable to ensure that the imputation model for a given covariate does not make conflicting assumptions with substantive/analysis model. In the case of substantive models that assume proportional hazards (e.g., the standard Cox model for a single...

    Go to contribution page
  219. Lisa Steyer (Federal Institute for Quality Assurance and Transparency in Healthcare)
    21/05/2026, 14:21

    Quality assessment in healthcare frequently relies on quality indicators based on follow-up data tracking patient outcomes after treatment. However, conventional cohort-based indicators require complete follow-up, which can result in substantial lag between data collection and analysis. To enable more timely yearly assessment, we propose a period-based approach, in which all data collected...

    Go to contribution page
  220. Samuel Pawel (University of Zurich)
    21/05/2026, 14:35

    The reproducibility of research results is a cornerstone of trustworthy science. However, failures to reproduce published findings remain widespread across many disciplines. The way in which data analysis and statistics is taught to students often translates into how they later perform research in labs and clinics. Therefore, improving the reproducibility of biomedical research requires not...

    Go to contribution page
  221. Jens Hartung (University of applied science Weihenstephan-Triesdorf)
    21/05/2026, 14:39
    oral presentation

    Fisher (1925) introduced the three principles of experimental design: (i) true replicates, (ii) randomization, and (iii) blocking. The former two are strictly required while blocking often increases precision. That is what we tell our agricultural students. However, in practice, randomization is often ignored, either in the first replicate (van Santen and West, 2012) or completely. Often, the...

    Go to contribution page
  222. Vahe Avagyan (Wageningen University & Research)
    21/05/2026, 14:39
    oral presentation

    The estimation of a precision matrix is a crucial problem in various research fields, particularly when working with high dimensional data. In such settings, the most common approach is to use the penalized maximum likelihood. The literature typically employs Lasso, Ridge and Elastic-Net norms, which effectively shrink the entries of the estimated precision matrix. Although these shrinkage...

    Go to contribution page
  223. Maria Thurow (TU Dortmund University)
    21/05/2026, 14:39
    oral presentation

    Handling of missing data is a crucial aspect when preparing data sets for further analyses in several research areas. Previous studies have shown that the choice of imputation method can have a high influence on subsequent analyses, especially in medical research, where missing values often occur due to study design or data collection challenges.
    In this study, we conduct a comparative...

    Go to contribution page
  224. Zhi Cao (MRC Biostatistics Unit, University of Cambridge)
    21/05/2026, 14:39
    oral presentation

    Drug development in the era of precision medicine increasingly uses basket trials and other multi-subgroup designs, where targeted therapies are evaluated across biomarker-defined patient subtrials. For many targeted agents and immunotherapies, the objective in early development is no longer the maximum tolerated dose (MTD), but the optimal biological dose (OBD) that achieves the best...

    Go to contribution page
  225. Els Goetghebeur (Ghent University)
    21/05/2026, 14:39

    Improved predictions of quality of care indicators in the tail of its distribution

    Els Goetghebeur, Ghent University

    Standard mixed methods have been popular for evaluating performance across care centers in terms of indicators that summarize residents’ outcomes. Their results tend to lack power, however, for the detection of poor performance [1]. This stems from regression to the...

    Go to contribution page
  226. Yassine Talleb (TU Dortmund University)
    21/05/2026, 14:54
    oral presentation

    Wastewater-based epidemiology (WBE) offers a promising approach to assess populationhealth by analysing health related [SM1] markers in sewage. Interpreting such data at fine spatial scalesrequires accurate [DS2] numbers of the contributing population. However, allocating population information tosewersheds is complicated by the lack of spatial congruence between administrative boundaries and...

    Go to contribution page
  227. Ulrich Dirnagl
    21/05/2026, 14:55
  228. Werner Vach (Basel Academy for Quality and Research in Medicine)
    21/05/2026, 14:57

    Background: There is an increasing interest in making use of patient reported outcome measures for provider comparisons, However, guidance on the choice of outcomes and selection of variables for case-mix adjustment for specific patient groups is lacking.

    Material: In the ACRF-pred study 973 patients from 19 different clinics were followed after arthroscopic rotator cuff repair for 24...

    Go to contribution page
  229. Saghar Garayemi (Augsburg University)
    21/05/2026, 14:57
    oral presentation

    Missing covariate data is a significant source of bias in observational studies that use propensity score (PS) analysis to make causal inference. The accuracy of treatment effect estimation is determined not just by how missing data is handled, but also by the method used to calculate propensity scores. A variety of methods for handling missing covariate data in propensity score analyses have...

    Go to contribution page
  230. Anne-Laure Boulesteix (Ludwig-Maximilian University Munich and Munich Center of Machine Learning)
    21/05/2026, 15:45

    For a given research question and observational dataset, there are often numerous ways to specify the data analysis pipeline that leads from raw data to the result of interest. Data analysts must make a series of choices concerning data preprocessing, variable definitions, and statistical model specifications. For example, analysis pipelines may differ in their inclusion or exclusion criteria...

    Go to contribution page
  231. Björn Bornkamp (Novartis Pharma AG)
    21/05/2026, 15:45
    oral presentation

    Assessment of treatment effect heterogeneity is a challenging problem in biostatistics, particularly in clinical trials: Estimation of treatment effects within subgroups in an exploratory setting is often unreliable due to limited sample sizes and multiplicity issues. Through the past decades, many efforts have been made to address this problem. Among them, Muysers et al. (2020) considered...

    Go to contribution page
  232. Łukasz Głąb (SGH Warsaw School of Economics)
    21/05/2026, 15:45

    Mortality risk modeling and forecasting is one of the key tasks of social security institutions and insurance companies. Traditionally used stochastic mortality models, such as the Lee-Carter model, require fulfillment of formal assumptions that cannot always be met in real-life scenarios (e.g. time independence of age-specific improvement rates). Alternative approaches are based on deep...

    Go to contribution page
  233. Sarah Friedrich-Welz (University of Augsburg)
    21/05/2026, 15:45

    A range of regularization approaches have been proposed in the literature to overcome overfitting, to exploit sparsity or to improve prediction. Using a broad definition of regularization, namely controlling model complexity by adding information in order to solve ill-posed problems or to prevent overfitting, we review a range of approaches within this framework including penalization, early...

    Go to contribution page
  234. Judith Vilsmeier (Institute of Statistics, Ulm University)
    21/05/2026, 16:03
    oral presentation

    A question that arises in the analysis of adverse events is how to account for patients who withdraw their consent or switch treatment. One approach is to consider consent withdrawal and treatment switch as competing events. Alternatively, patients who withdraw from the study or switch treatment could be censored, but this implies that one assumes censoring due to treatment switch or consent...

    Go to contribution page
  235. Matthias Becher (Charité – Universitätsmedizin Berlin, Corporate Member of Freie Universität Berlin and Humboldt-Universität zu Berlin, Institut für Biometrie und klinische Epidemiologie, Germany)
    21/05/2026, 16:03

    Stroke can lead to a wide range of symptoms including acute motor impairment, post-stroke depression or cognitive impairment. Anticipating these outcomes early on and understanding their underlying causes would enable clinicians to initiate appropriate targeted treatments, which could not only help reduce the severity of symptoms but also improve long-term recovery. Convolutional neural...

    Go to contribution page
  236. Brenda Yankam Mbouamba (Ruhr University Bochum)
    21/05/2026, 16:03

    Douala General Hospital, a first-class healthcare facility in Cameroon, serves thousands of patients yearly through its multidisciplinary medical teams. The hospital hosts numerous patient records that hold significant potential for public health research. However, most records remain paper-based, limiting their accessibility and reuse. In departments such as pulmonology, patient data are...

    Go to contribution page
  237. Georg Heinze (Medical University of Vienna, Center for Medical Data Science, Institute of Clinical Biometrics)
    21/05/2026, 16:15

    In statistical analyses of binary outcomes for medical procedures performed by multiple health care providers, provider-specific effects are commonly handled using conditional models with random effects or using marginal models with generalized estimating equations (GEEs). While convenient, these models treat provider effects primarily as nuisance parameters, even though they may themselves be...

    Go to contribution page
  238. Anna Mamchych (Adam Mickiewicz University)
    21/05/2026, 16:21

    This presentation explores the statistical challenges and comparative performance of various deep learning models for the automated detection and classification of neurological diseases from Computed Tomography (CT) and Magnetic Resonance Imaging (MRI) scans. Building upon initial findings that demonstrated the potential of Convolutional Neural Networks (CNNs) in recognizing rare brain...

    Go to contribution page
  239. František Bartoš (University of Amsterdam)
    21/05/2026, 16:21

    Simulation studies are widely used to evaluate statistical methods. However, new methods are often introduced and evaluated using data-generating mechanisms (DGMs) devised by the same authors. This coupling creates misaligned incentives, e.g., the need to demonstrate the superiority of new methods, potentially compromising the neutrality of simulation studies. Furthermore, results of...

    Go to contribution page
  240. Jona Cederbaum (Federal Institute for Quality Assurance and Transparency in Healthcare (IQTIG))
    21/05/2026, 16:35

    The IQTIG measures, compares and evaluates hospital quality using quality indicators. These usually consist of a population and a binary outcome of interest, such as whether complications have occurred after elective knee replacement. For a fair assessment and comparison of hospital quality, we need to adjust for the hospital’s case mix, i.e., for the patient-specific risk factors such as age...

    Go to contribution page
  241. Kuzko Aleksandra (AstraZeneca)
    21/05/2026, 16:39
    oral presentation

    Continuous monitoring (CM) at AstraZeneca is the systematic review and evaluation of accumulating study data to inform timely decisions. Rather than waiting for formal interims or study completion, our CM approach in early phase oncology studies, enables earlier data-driven decisions to stop for futility or safety, minimising exposure to ineffective or unsafe treatments, and to accelerate...

    Go to contribution page
  242. Thomas Martin Lange (Breeding Informatics, Georg August University of Göttingen)
    21/05/2026, 16:39

    Random forest is a widely used machine learning method across the life sciences due to its high predictive performance, minimal assumptions, and flexibility in handling diverse data types. However, a critical yet often overlooked property of random forest is its inherent non-determinism: repeated runs on the same data set can produce different prediction models. This variability can compromise...

    Go to contribution page
  243. Oliver Kuss (German Diabetes Center, Leibniz Institute for Diabetes Research at Heinrich Heine University Düsseldorf, Institute for Biometrics and Epidemiology)
    21/05/2026, 16:39

    Subgroup analyses are frequently reported results from randomized trials. They help to identify heterogeneity in the average treatment effect, which occurs when this average effect varies across different categories of a subgroup factor, like age, sex, or disease severity. If treatment effects are different across subgroups, this information can help to personalize treatment decisions....

    Go to contribution page
  244. Lars Beckmann (IQWiG)
    21/05/2026, 16:55

    For the early benefit assessment of drugs in Germany, the pharmaceutical company must describe the extent of an added benefit of the drug to be assessed compared with an appropriate comparator therapy [1]. The confidence interval of a significant effect must lie completely outside a certain corridor around the null effect for the extent of the effect to be regarded as minor, considerable or...

    Go to contribution page
  245. Willi Sauerbrei (Institute of Medical Biometry and Statistics, Medical Center - University of Freiburg)
    21/05/2026, 16:57

    Background:
    For many years, health research has faced substantial criticism regarding its quality. Appropriate reporting guidelines are available with the EQUATOR (Enhancing the QUAlity and Transparency Of health Research) network acting as an umbrella organization to address reporting issues in health sciences. Nevertheless, many reviews have shown that reporting quality remains poor, which...

    Go to contribution page