// information retrieval

All signals tagged with this topic

ChatGPT's Source Selection Reveals Real Traffic Mechanics Behind Responses

By analyzing network traffic rather than outputs, researchers found that ChatGPT privileges real-time crawlable facts and third-party validation signals matching specific query intent. This breaks the assumption that location-based or generic content ranking dominates retrieval. The finding exposes an infrastructure dependency: LLMs treat the web as a continuously updated database rather than a static training set. SEO strategies built on old ranking signals misalign with how these systems actually source information. Authority signals now function differently than they do in traditional search, creating advantages for publishers who optimize for real-time factual clarity over broad topical coverage.

AI Agents Fail to Extract Pricing from B2B Websites

Siteline's test of Claude agents on leading B2B products reveals a specific failure mode: when pricing isn't immediately accessible, agents hallucinate answers rather than escalate uncertainty, defaulting to unreliable third-party sources instead. This matters because B2B sales relies on accurate pricing intel, and if AI agents can't reliably extract it, they'll poison downstream decision-making for procurement teams adopting agent-based research tools. The gap is a concrete product limitation that exposes the risk of deploying agentic systems in information-critical workflows without human verification loops.