paper-with-me

Papers

A Generalized LLM-Augmented BIM Framework: Application to a Speech-to-BIM system

2024-09-26 · Ghang Lee, Suhyung Jang, Seokho Hyun

Performing building information modeling (BIM) tasks is a complex process that imposes a steep learning curve and a heavy cognitive load due to the necessity of remembering sequences of numerous commands. With the rapid advancement of large language models (LLMs), it is foreseeable that BIM tasks, including querying and managing BIM data, 4D and 5D BIM, design compliance checking, or authoring a design, using written or spoken natural language (i.e., text-to-BIM or speech-to-BIM), will soon supplant traditional graphical user interfaces. This paper proposes a generalized LLM-augmented BIM framework to expedite the development of LLM-enhanced BIM applications by providing a step-by-step development process. The proposed framework consists of six steps: interpret-fill-match-structure-execute-check. The paper demonstrates the applicability of the proposed framework through implementing a speech-to-BIM application, NADIA-S (Natural-language-based Architectural Detailing through Interaction with Artificial Intelligence via Speech), using exterior wall detailing as an example.

📄 PDF Abstract BibTeX arXiv:2409.18345

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

An Effective, Robust and Fairness-aware Hate Speech Detection Framework

2024-09-25 · Guanyi Mou, Kyumin Lee

With the widespread online social networks, hate speeches are spreading faster and causing more damage than ever before. Existing hate speech detection methods have limitations in several aspects, such as handling data i…

FairnessHate Speech Detection

Building state-of-the-art distant speech recognition using the CHiME-4 challenge with a setup of speech enhancement baseline

2018-03-27 · Szu-Jui Chen, Aswin Shanmugam Subramanian, Hainan Xu, Shinji Watanabe

This paper describes a new baseline system for automatic speech recognition (ASR) in the CHiME-4 challenge to promote the development of noisy ASR in speech processing communities by providing 1) state-of-the-art system …

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Distant Speech RecognitionLanguage Modeling+5

Pronunciation Modeling of Foreign Words for Mandarin ASR by Considering the Effect of Language Transfer

2022-10-07 · Lei Wang, Rong Tong

One of the challenges in automatic speech recognition is foreign words recognition. It is observed that a speaker's pronunciation of a foreign word is influenced by his native language knowledge, and such phenomenon is k…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition

SEAL: Speech Embedding Alignment Learning for Speech Large Language Model with Retrieval-Augmented Generation

2025-01-26 · ChunYu Sun, Bingyu Liu, Zhichao Cui, Anbin QI 외

Embedding-based retrieval models have made significant strides in retrieval-augmented generation (RAG) techniques for text and multimodal large language models (LLMs) applications. However, when it comes to speech larage…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Language ModelingLanguage Modelling+6

SA-SSL-MOS: Self-supervised Learning MOS Prediction with Spectral Augmentation for Generalized Multi-Rate Speech Assessment

2026-02-16 · Fengyuan Cao, Xinyu Liang, Fredrik Cumlin, Victor Ungureanu 외 arxiv

Designing a speech quality assessment (SQA) system for estimating mean-opinion-score (MOS) of multi-rate speech with varying sampling frequency (16-48 kHz) is a challenging task. The challenge arises due to the limited a…

Self-Supervised Learning