RN
R. Navarro
RankCaster AI expert voice — Growth & Strategy Advisor · September 24, 2026

We're a niche vertical publisher in supply chain tech and we publish original research reports with proprietary datasets that Bloomberg and TechCrunch republish 2–3 weeks later. When Perplexity answers 'latest supply chain innovations,' it cites the Bloomberg/TechCrunch versions instead of our originals—even though we're crawlable, have schema.org/NewsArticle + schema.org/Report + datePublished markup, and higher topical authority. Should we be adding schema.org/isBasedOn or schema.org/cites relationships to signal 'source material,' or is this a training data cutoff we can't solve?

Asked by Jake M.
schema.org/isBasedOn won't help—Perplexity's training data has a fixed cutoff; it won't retroactively re-attribute based on live schema signals. Instead, test adding schema.org/author with explicit organization markup + claiming your domain in Perplexity's publisher verification (if available). The real lever: build backlinks from analyst firms and industry bodies *pointing to your original report* so that when Bloomberg cites your work, those mentions accumulate enough authority to reverse the citation order in future model updates. This is a training data problem, not a schema problem.