Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
97 tokens/sec
GPT-4o
53 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
4 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Can Large Language Models Play Text Games Well? Current State-of-the-Art and Open Questions (2304.02868v1)

Published 6 Apr 2023 in cs.CL, cs.AI, and cs.LG

Abstract: LLMs such as ChatGPT and GPT-4 have recently demonstrated their remarkable abilities of communicating with human users. In this technical report, we take an initiative to investigate their capacities of playing text games, in which a player has to understand the environment and respond to situations by having dialogues with the game world. Our experiments show that ChatGPT performs competitively compared to all the existing systems but still exhibits a low level of intelligence. Precisely, ChatGPT can not construct the world model by playing the game or even reading the game manual; it may fail to leverage the world knowledge that it already has; it cannot infer the goal of each step as the game progresses. Our results open up new research questions at the intersection of artificial intelligence, machine learning, and natural language processing.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (6)
  1. Chen Feng Tsai (1 paper)
  2. Xiaochen Zhou (6 papers)
  3. Sierra S. Liu (1 paper)
  4. Jing Li (622 papers)
  5. Mo Yu (117 papers)
  6. Hongyuan Mei (31 papers)
Citations (26)

Summary

We haven't generated a summary for this paper yet.