paper-with-me

홈 › Papers

Thinkink: 2D Spatial Ink-native Interaction with LLMs

2026-07-23 · Mohammad Hasan Payandeh, Daniel Vogel, Jian Zhao arxiv

People often use handwritten notes and sketches to externalize ideas for ideation. To integrate large language models (LLMs) into this practice, we propose Thinkink. Prompts can be handwritten text or drawn sketches with LLM-generated responses visualized as ink-like text and sketches spatially integrated into a shared canvas. A semantic tree streamlines ink interpretation, and a lightweight UI provides explicit control using a state machine. The tool was designed using a three-stage process. A formative study (N=12) examined current practices with conventional and digital inking methods. The results informed a technical probe for a diagnostic study (N=6) identifying usability and human-LLM interaction challenges. This motivated the design of Thinkink, with a final study (N=10) examining how people incorporate it into their ideation practices. We contribute design implications and a tool for ink-native LLM interaction where the user and LLM write and draw in a shared 2D canvas.

📄 PDF Abstract BibTeX arXiv:2607.21468

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Conversations in Space: Non-Linear LLM Interaction in Everyday Use

2026-05-15 · Rifat Mehreen Amin, Alperen Adatepe, Daniela Fernandes, Daniel Buschek 외 arxiv

As LLM conversations grow, their histories capture alternative directions, decisions, and evolving lines of thought that can be difficult to navigate through chat alone. We investigate an interaction concept that represe…

Convolutions for Spatial Interaction Modeling

2021-04-15 · CVPR 2022 1 · Zhaoen Su, Chao Wang, David Bradley, Carlos Vallespi-Gonzalez 외

In many different fields interactions between objects play a critical role in determining their behavior. Graph neural networks (GNNs) have emerged as a powerful tool for modeling interactions, although often at the cost…

Autonomous Vehicles

SpaceMind++: Toward Allocentric Cognitive Maps for Spatially Grounded Video MLLMs

2026-05-10 · Bo Gu, Zhikang Zhang, Zizhuang Wei, Zhenyuan Chen 외 arxiv

Recent multimodal large language models (MLLMs) have made remarkable progress in visual understanding and language-based reasoning, yet they lack a persistent world-centered representation for spatially consistent reason…

Learning person-object interactions for action recognition in still images

2011-12-01 · NeurIPS 2011 12 · Vincent Delaitre, Josef Sivic, Ivan Laptev

We investigate a discriminatively trained model of person-object interactions for recognizing common human actions in still images. We build on the locally order-less spatial pyramid bag-of-features model, which was show…

Action RecognitionAction Recognition In Still ImagesObjectTemporal Action Localization

Probing Synergistic High-Order Interaction in Infrared and Visible Image Fusion

2024-01-01 · CVPR 2024 1 · Naishan Zheng, Man Zhou, Jie Huang, JunMing Hou 외

Infrared and visible image fusion aims to generate a fused image by integrating and distinguishing complementary information from multiple sources. While the cross-attention mechanism with global spatial interactions…

Infrared And Visible Image Fusion