Skip to main navigation Skip to search Skip to main content

Pet-Bench: Benchmarking the Abilities of Large Language Models as E-Pets in Social Network Services

  • Hongcheng Guo
  • , Zheyong Xie
  • , Shaosheng Cao*
  • , Boyang Wang
  • , Weiting Liu
  • , Zheyu Ye
  • , Zhoujun Li
  • , Zuozhu Liu
  • , Wei Lu
  • *Corresponding author for this work
  • Fudan University
  • Xiaohongshu
  • Beihang University
  • Zhejiang University
  • Singapore University of Technology and Design

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

As interest in using Large Language Models for interactive and emotionally rich experiences grows, virtual pet companionship emerges as a novel yet underexplored application. Existing approaches focus on basic pet role-playing interactions without systematically benchmarking LLMs for comprehensive companionship. In this paper, we introduce PET-BENCH, a dedicated benchmark that evaluates LLMs across both self-interaction and human-interaction dimensions. Unlike prior work, PET-BENCH emphasizes self-evolution and developmental behaviors alongside interactive engagement, offering a more realistic reflection of pet companionship. It features diverse tasks such as intelligent scheduling, memory-based dialogues, and psychological conversations, with over 7,500 interaction instances designed to simulate pet behaviors. Evaluation of 28 LLMs reveals significant performance variations linked to model size and inherent capabilities, underscoring the need for specialized optimization in this domain. PET-BENCH serves as a foundational resource for benchmarking pet-related LLM abilities and advancing emotionally immersive human-pet interactions.

Original languageEnglish
Title of host publicationCIKM 2025 - Proceedings of the 34th ACM International Conference on Information and Knowledge Management
PublisherAssociation for Computing Machinery, Inc
Pages6402-6407
Number of pages6
ISBN (Electronic)9798400720406
DOIs
StatePublished - 10 Nov 2025
Event34th ACM International Conference on Information and Knowledge Management, CIKM 2025 - Seoul, Korea, Republic of
Duration: 10 Nov 202514 Nov 2025

Publication series

NameCIKM 2025 - Proceedings of the 34th ACM International Conference on Information and Knowledge Management

Conference

Conference34th ACM International Conference on Information and Knowledge Management, CIKM 2025
Country/TerritoryKorea, Republic of
CitySeoul
Period10/11/2514/11/25

Keywords

  • benchmark
  • e-pet
  • emotional support
  • large language model

Fingerprint

Dive into the research topics of 'Pet-Bench: Benchmarking the Abilities of Large Language Models as E-Pets in Social Network Services'. Together they form a unique fingerprint.

Cite this