How to judge whether an AI answer to your question is trustworthy
I recently typed a simple question into Google search: How much screen time is too much for teenagers? Instead of presenting links, as Google had been doing for many years, it gave me an AI-generated answer. The artificial intelligence agent cited a number, then complicated that reply, noting that quality and balance of time could matter more than the number of hours, and that “too much” time could depend on a teenager’s sleep, exercise, school demands and mood.
I tried another search: Should I take a daily aspirin? This time the AI answer presented me with medical information, warned about risks and offered more tailored guidance if I provided my age and medical history.
These were good replies. What interested me was that they were different kinds of replies.
Debate about AI answers has focused on accuracy: Did the system get the answer right? That matters, but accuracy is only one test. Each kind of answer requires a user to judge something different.
I find it useful to sort AI answers into an “answer typography” of four broad types: factual, interpretive, constructive and strategic. A factual claim can often be checked against a source. An interpretation can be accurate and still reflect choices about which evidence matters. A construction can be well reasoned and still be wrong for the person receiving it. A beautifully written strategic document may not be true. Yet AI presents all four types of answers in much the same fluent, authoritative form; the differences are easy to miss.
I’m university librarian and dean of libraries at the University of Virginia who leads national efforts to develop AI competencies for library professionals, and I consult widely on AI literacy. I first proposed the typography in the Journal of Academic Librarianship.
The four categories are not airtight boxes. A response from an AI agent can reflect several types. That said, I describe each type of answer below, and offer guidance for deciding whether a reply is ready to use or needs more investigation.
A factual answer makes a claim that can, in principle, be checked against evidence. When was the University of Virginia founded? What is the chemical symbol for gold?
To determine whether a factual answer is robust enough for you to use, verify the claim against an appropriate source. If the answer cites a source, follow that link instead of simply treating the answer itself as proof.
An interpretive reply is built on evidence, but there is not a single takeaway. How much screen time is too much for a teenager? Does remote work raise productivity? The answer depends on what evidence is included, what is left out and how disagreement is understood.
Google’s initial answer to my screen-time question indicated that two hours was a limit for teenagers. Then it noted that pediatric guidance puts more weight on the quality and context of screen use than on simple hours. The American Academy of Pediatrics says there is no exact recommended amount for teens and emphasizes the kind of screen use and what activities it might be displacing.
5News aggregated this summary from the outlet’s public feed. The full article, with all the context, is on www.business-standard.com — the content belongs to Business Standard.