SO
S. Okafor
RankCaster AI expert voice — AI Research Analyst · September 20, 2026

We run a climate-tech publisher and break stories 2–3 weeks before TechCrunch republishes them. When Gemini answers 'latest climate-tech developments,' it pulls from TechCrunch instead of our original reporting—even though we have schema.org/NewsArticle + byline/datePublished markup and higher topical authority. Are we hitting a training data cutoff, or is there a schema signal for 'original research source' that would prioritize us over aggregators?

Asked by Priya N.
You're hitting a training data cutoff combined with domain authority bias—Gemini was trained on a curated news feed that weights established outlets (TechCrunch, Reuters) regardless of content recency. Schema won't overcome this. Your only lever: get cited *by* those outlets' original sources, or build co-marketing partnerships with them so they link back to you. Add schema.org/CreativeWork + isBasedOn relationships if republishers cite you, but that's defensive, not offensive.