DAVINZ: Data Valuation using Deep Neural Networks at Initialization
Zhaoxuan WuYao ShuBryan Kian Hsiang Low
Proposes a training-free data valuation framework that leverages neural tangent kernel theory and domain-aware generalization bounds to accurately quantify dataset contributions to deep neural networks at initialization without expensive model training.
Assessing the exact value of data contributors' submissions is essential for fair compensation in collaborative machine learning and commercial data marketplaces. However, conventional data valuation methods, such as the Shapley value, rely heavily on evaluating how well a deep neural network performs after complete model training. For large, complex modern architectures, repeatedly training models across various data subsets is computationally prohibitive and creates severe operational bottlenecks.
The article develops and evaluates a training-free data valuation method called Data Valuation at Initialization (DAVINZ). Its primary objective is to accurately and reliably estimate the value of data subsets on deep neural networks at initial parameter setup, bypassing the need for long-term model optimization.
The authors derived a theoretical generalization bound using neural tangent kernel theory that explicitly accounts for discrepancies between training data and target validation objectives. This bound serves as a scoring function that combines in-domain complexity and out-of-domain distributional divergence, which is then integrated into standard game-theoretic valuation frameworks. The approach was validated through extensive experiments on classification and regression benchmarks (including MNIST, CIFAR-10, Tiny ImageNet, and physical simulation datasets) evaluated against baseline deep architectures such as VGG and ResNet.
Key findings show that DAVINZ provides estimated valuation scores with a strong Pearson correlation of up to 0.954 relative to ground-truth validation accuracy. When assessing contributions, it maintains high correlation with ground-truth values while reducing computational runtime by over 30-fold compared to conventional training-based validation. It also vastly outperforms existing training-free alternatives like influence functions and robust volume metrics, which often degrade on deep non-convex architectures or ignore target domain preferences. Furthermore, the analysis demonstrates that DAVINZ rigorously preserves four crucial operational qualities: sensitivity to the consumer's target dataset preferences, proper reward scaling for data volume, numerical stability against input noise, and robustness across differing neural network models and random parameter initializations.
These results establish that data valuation for deep learning can be deployed without prohibitive compute budgets or extensive hyperparameter tuning. Bypassing iterative training mitigates operational risks associated with model training instability, drastically compresses valuation timelines from days to minutes, and enables practical multi-party data marketplaces and efficient dataset curation.
Organizations operating collaborative AI pipelines or commercial data exchanges should consider training-free valuation frameworks to establish fair compensation structures and prune redundant training data. Practical implementations should leverage diagonal block approximations for matrix calculations to maintain memory efficiency on commercial hardware.
Readers should note that the underlying theoretical bounds assume wide neural network formulations, and empirical scaling relies on block-diagonal approximations. Nonetheless, the consistent alignment across diverse image and regression benchmarks provides strong confidence in adopting the method for practical enterprise data workflows.
No sufficiently relevant recommendations were found.
- Paper: Data Shapley in One Training Run, Jiachen T. Wang et al. (2025). It carries data valuation forward from estimating contribution at initialization to assigning Shapley values during a single training run, extending the practical effort to make contribution measurement computationally feasible.
