Volume 8 • Issue 2 • PP: 19–27 • 2026
Metaheuristic-Optimized Conditional Diffusion Networks for Industrial Sensor Data Synthesis: An Applied Computational Intelligence Benchmark
Open Access & Copyright
© 2026 The Author(s). Published by ASPG. This article is licensed under the Creative Commons Attribution 4.0 International License (CC BY 4.0).
Abstract
Synthetic industrial telemetry can alleviate scarcity and confidentiality constraints, but its value depends on more than distributional similarity. Generated sequences must support downstream engineering analysis, limit disclosure risk, and satisfy operational relations. This paper presents a reproducible three-axis benchmark and a differential-evolution calibration layer for conditional denoising diffusion. Six generators–bootstrap resampling, a shrinkage-Gaussian model, Fourier surrogates, a conditional variational autoencoder, a conditional diffusion model, and its calibrated counterpart–are evaluated on chronologically partitioned manufacturing telemetry. The protocol covers 18,000 generated 80-minute windows and combines train-synthetic/test-real alarm classification, marginal and temporal fidelity measures, record-proximity and membership-inference tests, and five engineering-rule audits. Differential evolution selects four post-generation controls using validation data only. Relative to the uncalibrated diffusion model, calibration reduces the aggregate physical-violation rate by 94.7% and the Wasserstein error by 29.2%, while increasing mean downstream ROC–AUC by 0.009. The improvement is accompanied by a 0.073 increase in membership-inference AUC and a small deterioration in autocorrelation error. Bootstrap resampling provides the strongest mean predictive utility but exactly reproduces 77.9% of its outputs; the conditional variational autoencoder attains the lowest Wasserstein error (0.046) with a 0.020% physical-violation rate. No generator dominates utility, privacy, and plausibility simultaneously. Synthetic industrial data should therefore be selected through deploymentspecific acceptance regions rather than a single realism score.
Keywords
References
[1] V. Buggineni, C. Chen, and J. Camelio, “Enhancing manufacturing operations with synthetic data: A systematic framework for data generation, accuracy, and utility,” Frontiers in Manufacturing Technology, vol. 4, p. 1320166, 2024.
[2] F. K. Dankar and M. Ibrahim, “Fake it till you make it: Guidelines for effective synthetic data generation,” Applied Sciences, vol. 11, no. 5, p. 2158, 2021.
[3] M. Goyal and Q. H. Mahmoud, “A systematic review of synthetic data generation techniques using generative AI,” Electronics, vol. 13, no. 17, p. 3509, 2024.
[4] A. Figueira and B. Vaz, “Survey on synthetic data generation, evaluation methods and GANs,” Mathematics, vol. 10, no. 15, p. 2733, 2022.
[5] D. Atzeni, R. Ramjattan, R. Figliè, G. Baldi, and D. Mazzei, “Data-driven insights through industrial retrofitting: An anonymized dataset with machine learning use cases,” Sensors, vol. 23, no. 13, p. 6078, 2023.
[6] J. Ho, A. N. Jain, and P. Abbeel, “Denoising diffusion probabilistic models,” in Advances in Neural Information Processing Systems, vol. 33, 2020, pp. 6840–6851.
[7] Y. Song, J. Sohl-Dickstein, D. P. Kingma, A. Kumar, S. Ermon, and B. Poole, “Score-based generative modeling throughstochastic differential equations,” in International Conference on Learning Representations, 2021.
[8] Y. Tashiro, J. Song, Y. Song, and S. Ermon, “CSDI: Conditional score-based diffusion models for probabilistic time series imputation,” in Advances in Neural Information Processing Systems, vol. 34, 2021, pp. 24 804–24 816.
[9] X. Yuan and Y. Qiao, “Diffusion-TS: Interpretable diffusion for general time series generation,” in International Conference on Learning Representations, 2024.
[10] L. Ren, H. Wang, and Y. Laili, “Diff-MTS: Temporalaugmented conditional diffusion-based AIGC for industrial time series toward the large model era,” IEEE Transactions on Cybernetics, vol. 54, no. 12, pp. 7187–7197, 2024.
[11] L. Lin, Z. Li, R. Li, X. Li, and J. Gao, “Diffusion models for time-series applications: A survey,” Frontiers of Information Technology & Electronic Engineering, vol. 25, no. 1, pp. 19– 41, 2024.
[12] L. P. Mora-de León, D. Solís-Martín, J. Galán-Páez, and J. Borrego-Díaz, “Text-conditioned diffusion-based synthetic data generation for turbine engine sensor analysis and RUL estimation,” Machines, vol. 13, no. 5, p. 374, 2025.
[13] B. Van Breugel, Z. Qian, and M. van der Schaar, “Synthetic data, real errors: How (not) to publish and use synthetic data,” in Proceedings of the 40th International Conference on Machine Learning, ser. Proceedings of Machine Learning Research, vol. 202, 2023, pp. 34 793–34 808.
[14] T. Stadler, B. Oprisanu, and C. Troncoso, “Synthetic data– anonymisation groundhog day,” in 31st USENIX Security Symposium, 2022, pp. 1451–1468.
[15] H. Murtaza, M. Ahmed, N. F. Khan, G. Murtaza, S. Zafar, and A. Bano, “Synthetic data generation: State of the art in health care domain,” Computer Science Review, vol. 48, p. 100546, 2023.
[16] K.-M. Kim and J. W. Kwak, “PVS-GEN: Systematic approach for universal synthetic data generation involving parameterization, verification, and segmentation,” Sensors, vol. 24, no. 1, p. 266, 2024.
[17] S. Jeon and J. T. Seo, “A synthetic time-series generation using a variational recurrent autoencoder with an attention mechanism in an industrial control system,” Sensors, vol. 24, no. 1, p. 128, 2024.
[18] G. E. Karniadakis, I. G. Kevrekidis, L. Lu, P. Perdikaris, S. Wang, and L. Yang, “Physics-informed machine learning,” Nature Reviews Physics, vol. 3, no. 6, pp. 422–440, 2021.
[19] J. Willard, X. Jia, S. Xu, M. Steinbach, and V. Kumar, “Integrating scientific knowledge with machine learning for engineering and environmental systems,” ACM Computing Surveys, vol. 55, no. 4, pp. 1–37, 2022.
[20] A. Kotelnikov, D. Baranchuk, I. Rubachev, and A. Babenko, “TabDDPM: Modelling tabular data with diffusion models,” in Proceedings of the 40th International Conference on Machine Learning, ser. Proceedings of Machine Learning Research, vol. 202, 2023, pp. 17 564–17 579.
[21] A. Alaa, B. Van Breugel, E. S. Saveliev, and M. van der Schaar, “How faithful is your synthetic data? sample-level metrics for evaluating and auditing generative models,” in Proceedings of the 39th International Conference on Machine Learning, ser. Proceedings of Machine Learning Research, vol. 162, 2022, pp. 290–306, .
[22] N. Carlini, J. Hayes, M. Nasr, M. Jagielski, V. Sehwag, F. Tramèr, B. Balle, D. Ippolito, and E. Wallace, “Extracting training data from diffusion models,” in 32nd USENIX Security Symposium, 2023, pp. 5253–5270.
[23] H. Hu and J. Pang, “Loss and likelihood based membership inference of diffusion models,” in Information Security, ser. Lecture Notes in Computer Science. Springer, 2023, vol. 14411, pp. 121–141.
[24] Bilal, M. Pant, H. Zaheer, L. Garcia-Hernandez, and A. Abraham, “Differential evolution: A review of more than two decades of research,” Engineering Applications of Artificial intelligence, vol. 90, p. 103479, 2020.
[25] J. Drefs, E. Guiraud, and J. Lücke, “Evolutionary variational optimization of generative models,” Journal of Machine Learning Research, vol. 23, no. 21, pp. 1–51, 2022.
[26] F. Karl, T. Pielok, J. Moosbauer, F. Pfisterer, S. Coors, M. Binder, L. Schneider, J. Thomas, J. Richter, M. Lang, E. C. Garrido-Merchán, J. Branke, and B. Bischl, “Multiobjective hyperparameter optimization in machine learning–an overview,” ACM Transactions on Evolutionary Learning and Optimization, vol. 3, no. 4, pp. 1–50, 2023.
[27] R. Shwartz-Ziv and A. Armon, “Tabular data: Deep learning is not all you need,” Information Fusion, vol. 81, pp. 84–90, 2022.
[28] M. Meiser and I. Zinnikus, “A survey on the use of synthetic data for enhancing key aspects of trustworthy AI in the energy domain: Challenges and opportunities,” Energies, vol. 17, no. 9, p. 1992, 2024.
Cite This Article
Choose your preferred format
Publisher's Note
The statements, opinions, and data presented in this article are solely those of the author(s) and do not necessarily represent those of ASPG, the journal, or its editors. ASPG and the editors disclaim responsibility for any harm arising from the use of any ideas, methods, instructions, or products described in this article, to the fullest extent permitted by applicable law.