This is why I like AI for evaluations. It’s probably not 100% because with the ESS and an improved crossover it’s best guess, but AIs guess is more educated than mine.
Many times I’ve asked AI to compare two models of speaker that I own/owned. In approximately 2/3rds of these queries, it suggested a preference for the speaker that I determined was inferior in real-world experience. For the most part, LLMs only make predictions based on info they gather from the web, and they often miss critical details in exchange for a speedy answer. For example, AI will tend to conclude a larger diameter woofer can move more air simply for the fact it’s larger, without considering each woofer’s excursion capabilities, sensitivity etc. They also have a tendency to treat subjective forum opinions as fact. This is especially the case with Gemini and Chat GPT LLMs. Both of those particular AIs are practically useless for choosing HiFi gear unless you merely want to compare basic specs. Of course, when you call them out on their idiosyncrasies or contradictions, they tend to backpedal and have an excuse for their misleading suggestions. It’s scary how human-like they are in that regard.

