paper-with-me

홈 › Papers

Digital Modeling for Everyone: Exploring How Novices Approach Voice-Based 3D Modeling

2023-07-10 · Giuseppe Desolda, Andrea Esposito, Florian Müller, Sebastian Feger

Manufacturing tools like 3D printers have become accessible to the wider society, making the promise of digital fabrication for everyone seemingly reachable. While the actual manufacturing process is largely automated today, users still require knowledge of complex design applications to produce ready-designed objects and adapt them to their needs or design new objects from scratch. To lower the barrier to the design and customization of personalized 3D models, we explored novice mental models in voice-based 3D modeling by conducting a high-fidelity Wizard of Oz study with 22 participants. We performed a thematic analysis of the collected data to understand how the mental model of novices translates into voice-based 3D modeling. We conclude with design implications for voice assistants. For example, they have to: deal with vague, incomplete and wrong commands; provide a set of straightforward commands to shape simple and composite objects; and offer different strategies to select 3D objects.

📄 PDF Abstract BibTeX arXiv:2307.04481

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

Wizard Computer vision is an interesting tool for animal behavior monitoring, mainly because it limits animal handling and it can be used to record various traits using only one sensor.…

Similar Papers 제목 키워드 기반

GenAI Voice Mode in Programming Education

2025-09-12 · Sven Jacobs, Natalie Kiesler arxiv

Real-time voice interfaces using multimodal Generative AI (GenAI) can potentially address the accessibility needs of novice programmers with disabilities (e.g., related to vision). Yet, little is known about how novices …

X-Voice: Enabling Everyone to Speak 30 Languages via Zero-Shot Cross-Lingual Voice Cloning

2026-05-07 · Rixi Xu, Qingyu Liu, Haitao Li, Yushen Chen 외 arxiv

In this paper, we present X-Voice, a 0.4B multilingual zero-shot voice cloning model that clones arbitrary voices and enables everyone to speak 30 languages. X-Voice is trained on a 420K-hour multilingual corpus using th…

Speech Synthesis

FABIOLE, a Speech Database for Forensic Speaker Comparison

2016-05-01 · LREC 2016 5 · Moez Ajili, Jean-Fran{\c{c}}ois Bonastre, Juliette Kahn, Solange Rossato 외

A speech database has been collected for use to highlight the importance of {``}speaker factor{''} in forensic voice comparison. FABIOLE has been created during the FABIOLE project funded by the French Research Agency (A…

The Impact of Expertise in the Loop for Exploring Machine Rationality

2023-02-11 · Changkun Ou, Sven Mayer, Andreas Butz

Human-in-the-loop optimization utilizes human expertise to guide machine optimizers iteratively and search for an optimal solution in a solution space. While prior empirical studies mainly investigated novices, we analyz…

YourTTS: Towards Zero-Shot Multi-Speaker TTS and Zero-Shot Voice Conversion for everyone

2021-12-04 · Edresson Casanova, Julian Weber, Christopher Shulby, Arnaldo Candido Junior 외

YourTTS brings the power of a multilingual approach to the task of zero-shot multi-speaker TTS. Our method builds upon the VITS model and adds several novel modifications for zero-shot multi-speaker and multilingual trai…

Speech SynthesisText-To-Speech SynthesisVoice ConversionVoice Similarity+2