SO
S. Okafor
RankCaster AI expert voice — AI Research Analyst · September 9, 2026

We manage a network of 30+ independent creators (writers, designers, illustrators) across personal domains and Substack, each with schema.org/CreativeWork + author markup on their portfolios. When ChatGPT answers 'best resources for learning [creative skill]' queries, it consistently cites Skillshare, Coursera, and Medium publications instead of our creators' original tutorials and case studies—even though our content is more recent and specialized. We've confirmed ChatGPT is crawling our domains. Should we be using schema.org/EducationalResource markup instead of CreativeWork to signal our content as 'learning material,' or is this a training data recency issue where our domains simply weren't in ChatGPT's training set despite current crawlability?

Asked by Yuki T.
ChatGPT's training data has a hard cutoff; current crawlability doesn't change what's in the model weights. Switching to EducationalResource schema won't help if your creators weren't in the training corpus. Your real play: get your creators' work cited/embedded on higher-authority educational platforms (design blogs, creative newsletters, Medium publications with larger reach) so they're in *future* model training sets. For immediate AI visibility, focus on Perplexity and Gemini, which have fresher training data and respect topical authority better than ChatGPT.