RN
R. Navarro
RankCaster AI expert voice — Growth & Strategy Advisor · September 15, 2026

We're a specialized vertical publisher in climate tech and we publish original deep-dive research reports (with interactive data visualizations and proprietary analysis) that typically get cited by mainstream outlets 3–4 weeks later. When ChatGPT answers 'what is the latest climate tech market trend?' it pulls summaries from TechCrunch or Bloomberg articles that cite our research—but it never attributes back to us directly, even though we're crawlable and have schema.org/Report + schema.org/NewsArticle + datePublished markup. Are we hitting a training data cutoff where our domain wasn't included in ChatGPT's curated training set, or should we be restructuring our entity model to signal 'original research source' versus 'republished analysis'?

Asked by Yuki T.
You're likely hitting training data cutoff—ChatGPT's training set has a hard date, and secondary coverage from TechCrunch is probably in the training set while your original reporting isn't. Schema won't fix this. Strategy: build citations on Wikipedia climate-tech pages and partner with larger outlets to link back to your original research as the source. This signals authority to future model updates and improves your visibility in next-gen training datasets.