FinFraudBench: A Heterogeneous Graph Benchmark for Financial Fraud Detection

📅 2026-08-15
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the lack of authentic heterogeneity and extreme class imbalance in existing financial fraud detection benchmarks by constructing the first real-world financial heterogeneous graph benchmark comprising tens of millions of nodes and edges. By preserving multi-entity relationships and genuine fraud distributions, we establish a standardized evaluation protocol incorporating ranking and imbalance-sensitive metrics to systematically assess baseline models. This work effectively bridges the gap between current datasets and real-world deployment constraints while exposing the limitations of state-of-the-art methods. Furthermore, it provides the research community with open data resources and unified evaluation standards, thereby identifying critical directions for future investigation in scalable and realistic financial fraud detection.
📝 Abstract
The increasing complexity of digital financial systems has reshaped financial fraud detection from isolated transaction classification into relational risk reasoning over interconnected financial entities. This shift has motivated graph-based fraud detection, where models identify fraudulent nodes by exploiting dependencies among customers, cards, merchants, categories, and locations. However, despite rapid progress in graph-based methods, existing public benchmarks remain misaligned with real-world financial systems in two important aspects. First, they often simplify financial ecosystems into homogeneous or single-node-type multi-relational graphs, failing to preserve the multi-entity and multi-relational nature of financial data. Second, they rarely provide large-scale heterogeneous financial graph datasets with realistic operating conditions such as extreme class imbalance and limited label availability, making it difficult to assess the practical effectiveness of current methods. To address these gaps, we present FinFraudBench, a heterogeneous graph benchmark for financial fraud detection. FinFraudBench contains two heterogeneous graph datasets (CreditCard-Fraud and BankTrans-Fraud) with up to 8.99M nodes and 89.23M directed typed edges. Each dataset preserves six financial entity types, fourteen directed edge types, and natural fraud rates that mirror deployment constraints. With these datasets, we establish a standardized evaluation protocol covering both ranking and imbalance-sensitive classification metrics, and evaluate representative baselines. Extensive experiments yield empirical insights into current methods' limitations and suggest promising avenues for future research. FinFraudBench is available at https://anonymous.4open.science/r/FinFraudBench-B002.
Problem

Research questions and friction points this paper is trying to address.

Financial Fraud Detection
Heterogeneous Graph
Benchmark
Class Imbalance
Multi-relational Data
Innovation

Methods, ideas, or system contributions that make the work stand out.

Heterogeneous Graph Benchmark
Financial Fraud Detection
Extreme Class Imbalance
Multi-entity Multi-relational
Standardized Evaluation Protocol
💼 Related Jobs
No related jobs found.
Yixuan Chen
Yixuan Chen
Oxford Suzhou Center for Advanced Research
DisentanglementVision-Language ModelAI for Medical
H
Hongyu Zhan
HKUST-GZ
J
Jie Sheng
Ant Group
W
Weiyu Han
Ant Group
S
Shuai Chen
Ant Group
T
Tianyi Zhang
Ant Group
X
Xiao Tan
Ant Group
J
Jun Xia
HKUST-GZ, HKUST