Towards Unlocking Insights from Logbooks Using AI
Electronic logbooks contain valuable information about activities and events concerning their associated particle accelerator facilities. However, the highly technical nature of logbook entries can hinder their usability and automation. As natural language processing (NLP) continues advancing, it offers opportunities to address various challenges that logbooks present. This work explores jointly testing a tailored Retrieval Augmented Generation (RAG) model for enhancing the usability of particle accelerator logbooks at institutes like DESY, BESSY, Fermilab, BNL, SLAC, LBNL, and CERN. The RAG model uses a corpus built on logbook contributions and aims to unlock insights from these logbooks by leveraging retrieval over facility datasets, including discussion about potential multimodal sources. Our goals are to increase the FAIR-ness (findability, accessibility, interoperability, and reusability) of logbooks by exploiting their information content to streamline everyday use, enable macro-analysis for root cause analysis, and facilitate problem-solving automation.
Code (0)
등록된 구현이 없습니다.
Tasks
RAGRetrievalRetrieval-augmented GenerationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Transfer Learning Methods for Domain Adaptation in Technical Logbook Datasets
Event identification in technical logbooks poses challenges given the limited logbook data available in specific technical domains, the large set of possible classes, and logbook entries typically being in short form and…
Domain AdaptationTransfer LearningUnlocking Insights into Business Trajectories with Transformer-based Spatio-temporal Data Analysis
The world of business is constantly evolving and staying ahead of the curve requires a deep understanding of market trends and performance. This article addresses this requirement by modeling business trajectories using …
ArticlesTextual genre based approach to use WordNet in language-for-specific-purpose classroom as dictionary
When teaching language for specific purposes (LSP) linguistic resources are needed to help students understand and write specialised texts. As building a lexical resource is costly, we explore the use of wordnets to repr…
A maximum likelihood estimate of natural mortality for brown tiger prawn (Penaeus esculentus) in Moreton Bay (Australia)
The delay difference model was implemented to fit 21 years of brown tiger prawn (Penaeus esculentus) catch in Moreton Bay by maximum likelihood to assess the status of this stock. Monte Carlo simulations testing of the s…
NLP Tools for Predictive Maintenance Records in MaintNet
Processing maintenance logbook records is an important step in the development of predictive maintenance systems. Logbooks often include free text fields with domain specific terms, abbreviations, and non-standard spelli…
ClusteringPOSPOS Tagging