Aggregated Is a Direction

On 5 October OpenAI published a post about a new visual ad format for ChatGPT and, further down, an expansion of how advertisers can measure what their ads do. The format is the headline. The measurement section is the part with consequences, because it is where a promise OpenAI has made about privacy meets a pipeline that was built for something else.

The promise, from the ads help page I read the same morning: “We do not share your conversations with ChatGPT with advertisers, and we never sell your data to advertisers.” And: “Advertisers only receive aggregated, non-identifying information about how their ads perform, such as total views or clicks.” Elsewhere on that page, under what is shared with advertisers: “For this early test, advertisers receive aggregated reporting such as views and clicks. We may explore additional measurement insights over time while continuing to protect user privacy.” The page was marked as updated seven days earlier, and it says ad testing started in the US on 9 February 2026.

Every one of those sentences describes data going from OpenAI to the advertiser. That is the direction a privacy promise about chats naturally faces. Measurement, though, mostly runs the other way.

What the pipeline carries

OpenAI’s Conversion Measurement help page, marked “Updated: last month” when I read it, is a plain description of how a purchase on an advertiser’s site gets credited to a ChatGPT ad. When someone clicks, a click reference called oppref is appended to the landing-page URL. The OpenAI Pixel, which the advertiser installs, stores it in a first-party cookie and attaches it to later conversion events. Advertisers can also send events server-side through a Conversions API.

When the click reference is missing, the page offers three fallbacks. Advertisers may send “normalized and hashed contact information” with an event; this is called advanced matching. An automatic version has the Pixel automatically detect “supported customer information from recognizable forms and other sources on your website,” hash it with SHA-256 in the browser, and send that, with the note that “raw customer information is not sent to OpenAI through automatic advanced matching.” And there is modeled measurement: OpenAI “may also use aggregated patterns from observed conversions to estimate attribution for otherwise unattributed advertiser-reported conversion events,” and “reported conversion totals may include modeled conversions where available.”

None of this is exotic. Features like these are common in ad measurement, and the 5 October post adds integrations with Hightouch, Tealium and LiveRamp for sending conversion data in, plus attribution partners such as AppsFlyer, Triple Whale, Adjust, DV Rockerbox, Northbeam, Branch, Singular, Kochava, Airbridge and Tenjin. What matters here is the direction of travel. Identifiers and purchase events flow from advertisers into OpenAI. What flows back out is, per the help page, reporting “designed to show campaign performance rather than individual-level user activity.”

How a ChatGPT ad click becomes a reported conversion A five-step flow from left to right, as OpenAI's Conversion Measurement help page describes it. A person clicks an ad in ChatGPT. The landing-page URL carries an oppref click reference. The advertiser's site stores it in a first-party cookie through the OpenAI Pixel. A conversion event, optionally with SHA-256 hashed contact information, goes back to OpenAI. Ads Manager reports aggregated, possibly modeled, conversions. A note says the privacy promise covers only the last step, what advertisers receive, while step 3 brings advertiser-held data in, and that the pages do not say what the hashes are matched against. CONVERSION PATH · OPENAI HELP CENTER, "CONVERSION MEASUREMENT" · READ 5 OCT 2026 1 · Click Ad in ChatGPT oppref added to URL 2 · Advertiser site Pixel stores oppref first-party cookie 3 · Event goes in Purchase or lead, with optional hashed info 4 · Matching Click, hash, or model 5 · Report Aggregated may include modeled The Ads in ChatGPT page's privacy promise covers step 5: what advertisers receive. Step 3 moves advertiser-held events and identifiers into OpenAI; step 2 stores OpenAI's own click reference. The pages I read do not say which OpenAI-side data the hashes are compared against. Step names are mine; the contents of each step follow the help page's wording.
The privacy promise faces outward from OpenAI to the advertiser. Most of what the measurement tools add moves in the other direction.

Hashing contact details before sending them is standard, and the help page tells advertisers to send data only “after providing clear and comprehensive information to users” and obtaining required consents. But “hashed” is a technical description, not a privacy one. I covered a version of that point in an earlier essay on chatbot trackers: a hashed email still identifies someone to any party that already holds the same email. The documents I read do not say which OpenAI-side records a hashed address is matched against, whether a matched purchase becomes part of the “ads data” a user can clear, or how long conversion events are kept. The help page’s retention line, up to 30 days after deletion, is about ads history and topics. I could not find the answer to the other two on the pages I read, and the post points to an Ads Blog entry and the help page to developer documentation, neither of which I read, so those answers may live there.

None of that touches the claim that chats stay out of advertisers’ hands, and nothing I read contradicts it. The help page’s “such as views and clicks” is an example list, not an exhaustive one, and the Conversion Measurement page already describes conversion reporting. The ads page calls itself “this early test.”

The three results, read for what they leave out

The post offers three partner-reported figures as “early partner findings,” and I would treat each as a vendor claim, which is what the post itself says by attributing each to the vendor.

DV Rockerbox said WeightWatchers’ “attributed cost per acquisition” on ChatGPT Ads was 15.3% lower than its “blended paid-search benchmark.” One side of that comparison is a single channel’s attributed number. The other is a “blended” paid-search benchmark that the post does not define, so the two numbers may not be computed the same way. Attribution gives a channel credit for conversions that touched it, whether or not the ad caused them, so the gap does not tell you what a ChatGPT ad added. WorkMagic reported “statistically significant lift” for Dose, with 67% of incremental purchases coming from net-new customers. That is the only one of the three that speaks of incrementality, and it gives no lift size, no sample, and no test design. Triple Whale said 93% of Portland Leather’s visitors from ChatGPT Ads were new. Visitors are not purchasers, the figure has no comparison channel in the post, and a cold audience would produce a high new-visitor share for many top-of-funnel sources.

What each partner result in OpenAI's 5 October post states A dot matrix of three partner-reported results against four properties. WeightWatchers via DV Rockerbox: names a comparison benchmark, which is blended paid search; does not state a size of incremental effect; its outcome is acquisitions; method is attribution. Dose via WorkMagic: no comparison named; reports statistically significant lift but gives no size; outcome is purchases; the post does not name the test design. Portland Leather via Triple Whale: no comparison; no incremental effect; its outcome is visitors, not purchases; method not stated. PARTNER RESULTS · WHAT THE POST STATES · OPENAI, 5 OCT 2026 · MY READING Comparison named Outcome is a sale Lift claimed Lift size given Method named WeightWatchers Dose Portland Leather via DV Rockerbox via WorkMagic via Triple Whale WeightWatchers: blended paid-search benchmark named WeightWatchers: outcome is acquisitions (attributed CPA) WeightWatchers: method is attribution Dose: outcome is purchases Dose: statistically significant lift reported WeightWatchers: no incremental lift claimed WeightWatchers: no lift size Dose: no comparison named Dose: no lift size given Dose: test design not named in the post Portland Leather: no comparison named Portland Leather: outcome is visitors, not sales Portland Leather: no lift claimed Portland Leather: no lift size Portland Leather: method not named Filled: the post states it. Outlined: it does not. All three results are the partners' own, as reported by OpenAI.
The one result that claims incrementality gives no size for it. The one that names a comparison is an attributed number against a blend.

The post is candid about this gap in its own words: “our work on incrementality is still in its early stages.” It names Haus, Measured and WorkMagic as partners “to explore geo-based experiments to help advertisers understand the causal impact of ChatGPT Ads.” That is the right instrument. It is also an admission that the causal question is open, set in the paragraph directly above three numbers presented as illustrating “strong performance.”

An old gap, measured once

There is an old precedent for why experiments matter. In 1995 Leonard Lodish and colleagues published a meta-analysis of 389 real-world TV advertising experiments run on BehaviorScan, a split-cable system in which matched households received different commercials and their purchases were tracked through a consumer panel. Their abstract reports that “increasing advertising budgets in relation to competitors does not increase sales in general,” and that the data “do not show a strong relationship” between standard recall and persuasion copy-test measures and sales effectiveness. A matched-household split-cable experiment was the way to find that out. Click-level attribution answers a cheaper question, which is who touched the ad before buying, and a channel that launches with attribution first and incrementality “in its early stages” is answering the cheaper question first.

None of this means ChatGPT ads don’t work. It means the numbers on offer so far cannot tell an advertiser that. If I were a buyer, I would take the three results as an invitation to run a geo test, and I would read the Ads Blog and developer docs for the attribution window, the modeling disclosure and what happens to a conversion event when a user clears their ads data. If I were a user, I would note that “aggregated” describes what leaves OpenAI, and that most of the new measurement traffic arrives from somewhere else.

References

  1. OpenAI (2026). Building advertising for the way people use AI. OpenAI, 5 October 2026. The partner results are OpenAI’s and its partners’ own claims.
  2. OpenAI (2026). Conversion Measurement. OpenAI Help Center; page showed “Updated: last month” when read on 5 October 2026.
  3. OpenAI (2026). Ads in ChatGPT. OpenAI Help Center; page showed “Updated: 7 days ago” when read on 5 October 2026.
  4. Lodish, L. M., Abraham, M., Kalmenson, S., Livelsberger, J., Lubetkin, B., Richardson, B., Stevens, M. E. (1995). How T.V. Advertising Works: A Meta-Analysis of 389 Real World Split Cable T.V. Advertising Experiments. Journal of Marketing Research 32(2), May 1995.