<p>Counterfactual explanations identify minimal input changes needed to alter a machine learning model’s prediction, offering actionable insights in tasks like churn analysis. However, existing methods often produce counterfactuals that vary in quality, coherence, and plausibility, limiting their practical value. We propose an ensemble evaluation framework that integrates multiple generation techniques and ranks their outputs using a tunable scoring function balancing multiple relevant metrics. Our approach addresses two key deployment scenarios: (i) in-house churn analysis, where decision-makers can interactively adjust scoring weights for tailored, user-driven explanations; and (ii) outsourced churn prediction, where counterfactuals must be generated on synthetic data to preserve privacy while remaining representative of real cases. Experiments on benchmark churn datasets demonstrate that our ensemble approach improves the consistency, interpretability, and utility of counterfactuals across both real and synthetic settings, supporting more reliable and privacy-aware decision-making.</p>

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Counterfactual ensembles for interpretable churn prediction: from real-world to privacy-preserving synthetic data

  • Samuele Tonati,
  • Marzio Di Vece,
  • Fosca Giannotti,
  • Roberto Pellungrini

摘要

Counterfactual explanations identify minimal input changes needed to alter a machine learning model’s prediction, offering actionable insights in tasks like churn analysis. However, existing methods often produce counterfactuals that vary in quality, coherence, and plausibility, limiting their practical value. We propose an ensemble evaluation framework that integrates multiple generation techniques and ranks their outputs using a tunable scoring function balancing multiple relevant metrics. Our approach addresses two key deployment scenarios: (i) in-house churn analysis, where decision-makers can interactively adjust scoring weights for tailored, user-driven explanations; and (ii) outsourced churn prediction, where counterfactuals must be generated on synthetic data to preserve privacy while remaining representative of real cases. Experiments on benchmark churn datasets demonstrate that our ensemble approach improves the consistency, interpretability, and utility of counterfactuals across both real and synthetic settings, supporting more reliable and privacy-aware decision-making.