paper-with-me

홈 › Papers

Automated Batch Distillation Process Simulation for a Large Hybrid Dataset for Deep Anomaly Detection

2026-04-10 · Jennifer Werner, Justus Arweiler, Indra Jungjohann, Jochen Schmid, Fabian Jirasek, Hans Hasse, Michael Bortz arxiv

Anomaly detection (AD) in chemical processes based on deep learning offers significant opportunities but requires large, diverse, and well-annotated training datasets that are rarely available from industrial operations. In a recent work, we introduced a large, fully annotated experimental dataset for batch distillation under normal and anomalous operating conditions. In the present study, we augment this dataset with a corresponding simulation dataset, creating a novel hybrid dataset. The simulation data is generated in an automated workflow with a novel Python-based process simulator that employs a tailored index-reduction strategy for the underlying differential-algebraic equations. Leveraging the rich metadata and structured anomaly annotations of the experimental database, experimental records are automatically translated into simulation scenarios. After calibration to a single reference experiment, the dynamics of the other experiments are well predicted. This enabled the fully automated, consistent generation of time-series data for a large number of experimental runs, covering both normal operation and a wide range of actuator- and control-related anomalies. The resulting hybrid dataset is released openly. From a process simulation perspective, this work demonstrates the automated, consistent simulation of large-scale experimental campaigns, using batch distillation as an example. From a data-driven AD perspective, the hybrid dataset provides a unique basis for simulation-to-experiment style transfer, the generation of pseudo-experimental data, and future research on deep AD methods in chemical process monitoring.

📄 PDF Abstract BibTeX arXiv:2604.09166

Code (0)

등록된 구현이 없습니다.

Tasks

Anomaly DetectionStyle Transfer

Similar Papers 제목 키워드 기반

Faster Inference of Flow-Based Generative Models via Improved Data-Noise Coupling

2026-03-16 · Aram Davtyan, Leello Tadesse Dadi, Volkan Cevher, Paolo Favaro arxiv

Conditional Flow Matching (CFM), a simulation-free method for training continuous normalizing flows, provides an efficient alternative to diffusion models for key tasks like image and video generation. The performance of…

Video Generation

Are Large-scale Soft Labels Necessary for Large-scale Dataset Distillation?

2024-10-21 · Lingao Xiao, Yang He

In ImageNet-condensation, the storage for auxiliary soft labels exceeds that of the condensed dataset by over 30 times. However, are large-scale soft labels necessary for large-scale dataset distillation? In this paper, …

Dataset DistillationDiversity

Controlling the Quality of Distillation in Response-Based Network Compression

2021-12-19 · Vibhas Vats, David Crandall

The performance of a distillation-based compressed network is governed by the quality of distillation. The reason for the suboptimal distillation of a large network (teacher) to a smaller network (student) is largely att…

Knowledge Distillation

Marginal Advantage Accumulation for Memory-Driven Agent Self-Evolution

2026-06-18 · Mingyu Yang, Keye Zheng, Congchao Cheng, Yujie Liu 외 arxiv

In batch-style trace distillation, the same memory operation may receive contradictory feedback across different batches. Existing methods lack a cross-batch, operation-level evidence accumulation mechanism, making it im…

Online Density-Based Clustering for Real-Time Narrative Evolution Monitorin

2026-01-28 · Ostap Vykhopen, Viktoria Skorik, Maksym Tereshchenko, Veronika Solopova arxiv

Automated narrative intelligence systems for social media monitoring face significant scalability challenges when relying on batch clustering methods to process continuous data streams. We investigate replacing offline H…

Computational EfficiencyOnline Clustering