Skip to content
Original study

Reddit appeared in 17 of 528 observed OpenAI API responses with citations

Reddit was cited in a minority of responses in the Bee LLM Observatory sample. Giving each prompt equal weight raises the rate to 4.1%, without changing that reading.

5 min read

A minority presence within a defined sample

Reddit appeared as a cited source in 17 of 528 observed responses with citations from the OpenAI API channel, a rate of 3.22%. That is the main finding from this Bee LLM Observatory sample. The percentage describes responses that already contained at least one valid citation; it does not tell us how often all AI answers cite Reddit.

The limits belong alongside the finding. These 528 responses cover 130 distinct prompts, with repetitions and unequal weights. They are not a representative survey of AI use. OpenAI API also identifies the collection channel: this result cannot be carried over to the consumer ChatGPT application.

Results for the sample described in this article
Engine / channelResponses with citationsResponses meeting the criterionPercentage
OpenAI API528173.22%

What counts as a Reddit appearance

The measure checks whether a response includes at least one citation with a host name equal to reddit.com or one of its subdomains. This provides a specific way to observe the platform among cited sources. Each response counts only once toward the indicator, even if it contains several valid Reddit links.

So 17 is not a count of cited links, posts or communities. Nor is 3.22% Reddit’s share of all citations. A response can cite Reddit alongside other websites and still contribute exactly one case to the numerator. The statistic measures presence within responses, rather than the volume of links to the platform.

Equal prompt weights produce a somewhat higher rate

The primary calculation gives each observed response the same weight. Prompts repeated more often can therefore have more influence on the result. An additional sensitivity check gives each distinct prompt equal weight. Under that calculation, the rate is 4.1%, compared with the primary response-weighted figure of 3.22%.

The shift shows that weighting affects the number, while leaving the basic interpretation intact: Reddit appears in a minority of responses with citations in this segment. The 4.1% figure does not come from a second collection of responses and is not a representative estimate for all users. It also does not establish that unequal account weights have been removed.

A citation shows presence, not influence

The finding is useful as a record of Reddit’s presence among cited sources, without turning that observation into a claim about its general importance to AI. These aggregates cannot establish whether a link was central to an answer, supplied a secondary detail or pointed to content the response viewed favorably.

They also do not measure the accuracy of the answers or the quality of the linked posts. A recorded citation is insufficient to reconstruct how a response was produced. Reading this statistic as a measure of trust, endorsement or Reddit’s influence would require additional evidence that is not included here.

The observation week is separate from publication

This article is published on October 1, 2026. The observations cover seven complete local days, beginning September 23, inclusive, and ending at the start of September 30, exclusive, in the Europe/Madrid time zone. Publication does not extend that window or make the finding a measurement of activity on publication day.

Across collection channels, the citation-bearing sample contains 1,648 responses, 164 distinct prompts, 30 projects and 27 accounts. Those totals describe the wider sample and cannot all be assigned to OpenAI API. That segment contributes the 528 responses and 130 prompts behind the main finding. The available aggregate cannot explain Reddit’s presence by subject matter.

How the denominator was built

Collection included completed executions, excluded branded prompts and removed accounts with “demo” in their email address. It retained the latest response per execution and stored model, provided the text was nonempty and not an error. Citation measurement used links recorded as cited through a known method, rather than links merely detected in the text.

Across the full sample, there were 1,979 valid responses from identified channels. Missing citation telemetry excluded 18, leaving 1,961 measured responses, of which 1,648 had valid citations. Empty or missing citation records do not enter the headline denominator. Keeping these categories separate avoids confusing missing citation information with a response that can be assessed for this indicator.

What the evidence can support

Published segments must include at least 100 responses, 10 prompts, three projects and three accounts; positive cases must span at least three accounts. These thresholds provide some breadth for aggregate reporting, but they do not correct prompt selection, repetitions or differences in collection methods.

The supplied evidence reports only one channel segment for Reddit. It cannot support engine rankings, market-share estimates or treating absent or suppressed channels as zeros. Any cross-channel comparison would need to acknowledge different prompts and collection methods. There is also no demonstrated time trend, and no causal basis for promising that publishing on Reddit will earn citations.

Frequently asked questions

Does 3.22% refer to all OpenAI responses?

No. It refers to 528 observed responses from the OpenAI API channel that contained at least one valid citation. It represents neither all its responses nor the ChatGPT application.

Do several Reddit links make a response count more?

No. Each response counts once if it cites reddit.com or one of its subdomains, regardless of how many links it includes.

Why report the 4.1% figure as well?

It checks the result when each distinct prompt receives equal weight. The rate rises from the primary 3.22%, but Reddit remains a minority presence in this segment.