任务(项目管理)
计算机科学
接口(物质)
人机交互
航程(航空)
用户界面
封面(代数)
语言模型
认知心理学
自然语言处理
心理学
程序设计语言
经济
并行计算
复合材料
管理
材料科学
气泡
最大气泡压力法
工程类
机械工程
作者
Felix Gröner,Erin K. Chiou
标识
DOI:10.1177/10711813241260399
摘要
Large Language Models (LLMs) with their novel conversational interaction format could create incorrectly calibrated expectations about their capabilities. The present study investigates human expectations toward a generic LLM’s capabilities and limitations. Participants of an online study were shown a series of prompts that cover a wide range of tasks and asked to assess the likelihood of the LLM being able to help with those tasks. The result is a catalog of people’s general expectations of LLM capabilities across various task domains. Depending on the actual capabilities of a specific system, this could inform developers of potential over- or under-reliance on this technology due to these misconceptions. To explore a potential way of correcting misconceptions we also attempted to manipulate their expectations with three different interface designs. In most of the tested task domains, such as computation and text processing, however, these seem to be insufficient to overpower people’s initial expectations.
科研通智能强力驱动
Strongly Powered by AbleSci AI