Center forSocial
Connection

Technology & Social MediaMethods & Data

How Much Uncertainty Did the APA Review Actually Have?

A close look at how the American Psychological Association's 2026 review of AI companionship hedged its conclusions, and whether the hedge holds up against a longitudinal study published two months later.

Photograph · Pexels

In January, the American Psychological Association’s Monitor on Psychology published a review of AI chatbots and digital companions, concluding that their effects on emotional connection are “genuinely mixed” and depend on “user characteristics, chatbot design, and usage intensity.” That is a reasonable-sounding hedge. It is also, on inspection, three different claims stacked into one sentence, and each has since been tested with different degrees of rigor. The question worth asking is not whether the APA review was wrong to hedge, but whether it hedged in the right places.

The three variables, and what evidence backs each

The APA review names user characteristics, chatbot design, and usage intensity as the moderators that determine whether AI companionship helps or harms. Treat these separately, because the literature does not support them equally.

Usage intensity is the best-supported of the three. A twelve-month longitudinal study of more than 2,000 adults across four Western countries, published in April, found that increased social chatbot use predicted increased loneliness over time. That is a directional finding from repeated measurement, not a single cross-sectional snapshot, which matters: cross-sectional data cannot distinguish whether lonely people use chatbots more, or chatbot use makes people lonelier. The longitudinal design lets the study frame the relationship as bidirectional — lonely people turn to chatbots, and chatbot use appears to deepen the loneliness that sent them there in the first place. That is a stronger and more specific claim than “usage intensity matters.” It says which direction usage intensity tends to push, at least over a year, in this sample.

Reporting from TechXplore in March made a related but distinct point: short-term comfort and long-term harm both show up in the literature, and the distinction tracks duration of use rather than contradicting findings against each other. Heavy daily use correlated with increased loneliness and reduced human social engagement. Read together, the longitudinal study and the TechXplore synthesis point the same direction. The APA’s “usage intensity” hedge understates how consistent this particular piece of evidence already was by January.

Chatbot design is a weaker claim by comparison. None of the sources reviewed here isolate design features — persona, memory, conversational style — as an experimentally manipulated variable with a measured effect on loneliness outcomes. The claim that design moderates outcomes is plausible and probably true, but nothing in this literature demonstrates it. This is the part of the APA hedge that is honest about absence of evidence, whether or not the review said so explicitly.

User characteristics sits in between. The George Mason University commentary from September 2025 frames AI companionship as a substitution risk rather than a straightforward remedy — implying that who benefits and who is displaced depends on whether the chatbot supplements or replaces existing human contact. That is a characteristic-level claim, but it is offered as public health framing, not as a tested moderator with an effect size attached.

Where the hedge served the evidence, and where it obscured it

The problem with folding all three into a single “it depends” sentence is that it treats a well-evidenced directional finding (usage intensity) the same as an untested hypothesis (chatbot design). A reader taking the APA review at face value would come away thinking the field is evenly split three ways. It is not. One of the three moderators has longitudinal, multi-country evidence behind it; one has essentially none; and the third has been asserted more than measured.

This is not a case of a report inventing certainty it does not have — the more common failure mode. It is closer to the opposite: a report so committed to acknowledging complexity that it flattens a real asymmetry in how much is actually known. The AHA’s 2022 scientific statement on cardiovascular and brain health took a more disciplined approach to the same problem, naming explicitly that the absence of intervention trials was the central gap, rather than folding it into a general statement that effects “depend on” various things. The APA review would have been stronger, not weaker, if it had ranked its three moderators by evidentiary weight instead of listing them as co-equal.

There is a second layer of complexity the review does not address at all: the closeness of the connection matters even within human-to-human contact, let alone chatbot contact. A nationally representative study published this month on social media closeness among US adults examined whether closeness of online contacts moderates loneliness, treating closeness as a variable worth measuring on its own rather than assuming all “contact” is equivalent. If closeness moderates loneliness outcomes for human contacts online, there is no reason to assume chatbot interactions escape that logic — yet the review does not connect the two literatures, even though the underlying mechanism (parasocial versus reciprocal interaction) is arguably the same question asked twice.

What would settle this

The APA review’s caution about chatbot design is the piece of the puzzle actually worth investing research effort in, because it is the one lever a company or regulator could pull. A study that experimentally varied chatbot memory, persona consistency, or conversational initiative — rather than observing existing products as a monolith — could establish whether design genuinely moderates the loneliness effect, or whether usage intensity swamps it regardless of design. Absent that, “chatbot design matters” remains an assumption borrowed from intuition about product variation, not a finding.

The more immediate takeaway is narrower and less comfortable: the strongest available evidence, from a twelve-month study spanning four countries, points toward heavier chatbot use predicting more loneliness, not less. A review published a few months earlier treated that as one hypothesis among three equally uncertain ones. It was not one among three. It was, even at the time of writing, the one with a longitudinal design behind it.

Sources

  1. AI Chatbots and Digital Companions Are Reshaping Emotional ConnectionAmerican Psychological Association, Monitor on Psychology, January 2026
  2. How Does Turning to AI for Companionship Predict Loneliness, and Vice Versa?PubMed, April 2026
  3. AI Companions Can Comfort Lonely Users but May Deepen Distress Over TimeTechXplore, March 2026
  4. AI, Loneliness, and the Value of Human ConnectionGeorge Mason University College of Public Health, September 2025
  5. Closeness of Social Media Contacts and Loneliness Among US Adults: A Nationally Representative StudyPubMed, April 2026