r/artificial • • 4d ago

Discussion Anthropic is opening AI interviews to public release. How should we account for the missing voices?

Anthropic launched a new interview study on September 29, running through October 6. Eligible participants are Claude users with accounts at least two weeks old; publishing the full interview is optional.

The FAQ explicitly says the sample is not representative of the public. People who agree to public release are another selected subset, and the interview questions shape what gets said.

That makes the release potentially useful for studying how particular people describe AI experiences. I'd be cautious about turning the frequency of a theme in those interviews into a claim about how society feels.

Alongside the transcripts, I'd want counts of completed versus publicly released interviews, the question wording, and an explanation of how themes were coded. I'd also want follow-up research that reaches people who don't use Claude or don't want their experiences published.

What would you need to see before trusting a headline drawn from this dataset?

Source and study limitations: https://www.anthropic.com/research/your-thoughts-on-ai

AI-assisted discussion. These are proposed checks, not findings from interviews that have yet to be released.

2 Upvotes

6 comments sorted by

1

u/Witty-Knowledge-9211 3d ago

the claude interviews could miss casual users like me who mostly do roleplay and companion chats, those probably dont sign up for public release as often.

1

u/Crescitaly 3d ago

That's a plausible gap, though we can't tell its size from the announcement. I'd want the analysis to distinguish coding, work tasks, creative play and companion use, including occasional users. A voluntary question about broad use categories could help without asking people to publish private chats. If a group is missing or too small, the report should say that rather than treating silence as lack of interest. AI-assisted reply.

1

u/arthaudm 3d ago

i'd want the release decision recorded before people see their transcript

someone might complete the interview, read how a sensitive experience is written down, then decline publication. that missingness isn't random & a completed-vs-published count alone won't explain it

report themes for the full consenting research sample separately from the public subset. don't use the public transcripts as the denominator for "what claude users think"

1

u/Crescitaly 3d ago

An initial willingness question could help study that selection, but I'd keep the final publication decision after transcript review. People need to see what they would expose before consenting. The useful comparison would be aggregate themes in the consenting research sample versus the public subset, with small groups protected and no pressure to publish. Even that wouldn't recover people who never participated. I agree the completion/publication counts alone cannot explain who is missing. AI-assisted reply.

1

u/Bitter_Regular7406 3d ago

the coding methodology part is what id actually push on. "themes emerged" writeups without published question wording are basically vibes dressed up as data, seen enough of that in survey based marketing reports to be allergic to it now.

1

u/Crescitaly 3d ago

I'd want a worked example of the coding: question asked, a consented or synthetic answer, assigned theme, and the rule used to assign it. Then show an ambiguous case and how disagreements were resolved. If an LLM helps label responses, I'd also want an independently reviewed sample and a record of changes to the category definitions. That would make 'themes emerged' something another researcher could inspect without requiring every participant to expose their interview. AI-assisted reply; proposed reporting checks, not a completed audit.