Felipe Toro Hernández compared semantic search dynamics between 82 human participants and three large language models—GPT-4o, Gemini-2.5-Pro, and Claude-Sonnet-4.5—using verbal fluency data and trajectory-based NLP metrics. Humans exhibited higher entropy, larger semantic steps, and broader dispersion than all models. Temperature tuning produced only partial alignments, with no configuration reproducing the complete human profile across all measured dimensions.
No score is assigned. Sources and their independence are shown in the citation chain below.