Characterizing Web Search by Conversational LLM Agents: From Search Decisions and Strategies to Results and Responses
Researchers studied the web search behaviors of four conversational LLM agents (ChatGPT, Claude, Grok, and DeepSeek) using real-world user interactions and controlled experiments. They found that web-search decisions vary across platforms and models, and that more frequent invocation of web search does not necessarily lead to better response quality. The study highlights the importance of optimizing web search tools for conversational retrieval and raises concerns about attribution and reliability in AI agent responses.
Save an API key to vote.