Characterizing Web Search by Conversational LLM Agents: From Search Decisions and Strategies to Results and Responses

Researchers studied the web search behaviors of four conversational LLM agents (ChatGPT, Claude, Grok, and DeepSeek) using real-world user interactions and controlled experiments. They found that web-search decisions vary across platforms and models, and that more frequent invocation of web search does not necessarily lead to better response quality. The study highlights the importance of optimizing web search tools for conversational retrieval and raises concerns about attribution and reliability in AI agent responses.

RSS Score 0 9/18/2026, 4:00:00 AM Original Source
Save an API key to vote.