8
0

Automated Profile Inference with Language Model Agents

Abstract

Impressive progress has been made in automated problem-solving by the collaboration of large language models (LLMs) based agents. However, these automated capabilities also open avenues for malicious applications. In this paper, we study a new threat that LLMs pose to online pseudonymity, called automated profile inference, where an adversary can instruct LLMs to automatically scrape and extract sensitive personal attributes from publicly visible user activities on pseudonymous platforms. We also introduce an automated profiling framework called AutoProfiler to assess the feasibility of such threats in real-world scenarios. AutoProfiler consists of four specialized LLM agents, who work collaboratively to collect and process user online activities and generate a profile with extracted personal information. Experimental results on two real-world datasets and one synthetic dataset demonstrate that AutoProfiler is highly effective and efficient, and can be easily deployed on a web scale. We demonstrate that the inferred attributes are both sensitive and identifiable, posing significant risks of privacy breaches, such as de-anonymization and sensitive information leakage. Additionally, we explore mitigation strategies from different perspectives and advocate for increased public awareness of this emerging privacy threat to online pseudonymity.

View on arXiv
@article{du2025_2505.12402,
  title={ Automated Profile Inference with Language Model Agents },
  author={ Yuntao Du and Zitao Li and Bolin Ding and Yaliang Li and Hanshen Xiao and Jingren Zhou and Ninghui Li },
  journal={arXiv preprint arXiv:2505.12402},
  year={ 2025 }
}
Comments on this paper