ChatGPT reads the open web: your website first, then your public profiles, directories and portals, and local coverage that names you. It can only work with what is published somewhere it can reach. To the machine, nothing unpublished exists, and 45% of homebuyers already use AI tools in their search (Veterans United, 2026).
A client opens ChatGPT and types: "Who's a good agent for condos in Montclair?" The machine reads something, somewhere, and returns a name. Maybe yours. Maybe not. What you cannot see is the retrieval happening underneath: the crawl through your website, your Zillow profile, an old newspaper quote, a brokerage page that still lists your previous team. The machine assembles a picture of you from whatever it can reach. If that picture is incomplete, outdated, or contradictory, the answer it gives reflects that. Understanding "where does ChatGPT get its information" is the first step toward controlling what it finds.
Where Does ChatGPT Get Its Information? The Source Stack
The machine reads in a rough hierarchy. Understanding it helps you see where your leverage actually sits.
Tier 1: Your own website. This is the only source you fully control. Every word, every page, every update is yours to make. The machine retrieves information about your market expertise, service areas, and recent activity from your website first. If your site says nothing useful, the machine looks elsewhere.
Tier 2: Profiles you partly control. Portal pages on Zillow, Realtor.com, and Homes.com. Your brokerage profile. Your Google Business listing. You can edit these, but you do not own the platform. They can change formats, remove fields, or deprioritize your content. What you write there contributes to how the machine understands you, but it is not fully yours.
Tier 3: Sources you do not control. Local news mentions, directory aggregators, forum threads, old press releases. These age unpredictably. A 2021 article naming you at a brokerage you left years ago still exists on the web. The machine may read it. You cannot edit it. You can only outpublish it with fresher, more authoritative content on the surfaces you do control.
Each tier contributes. Each tier carries risk when the information goes stale or contradicts what appears elsewhere.
The Short Version
ChatGPT does not know you. It reads published sources about you and synthesizes what it finds. Your own website carries the most weight because you control every word. Public profiles and directory listings fill in gaps. Sources conflict or go stale, the machine treats you as uncertain, and uncertain answers do not get recommended.
The Uncertainty Problem: Sources That Disagree
Imagine an agent named Sarah Chen at Lakeview Realty. Her website says she serves Essex County. Her Zillow profile, unchanged since 2022, says Bergen County. A local directory lists her under her maiden name at a brokerage she left three years ago. A town newsletter from 2024 quotes her as a "Morris County specialist."
The machine reads all of this. It cannot call Sarah to ask which is correct. It sees four sources with four different answers to basic questions: What is her name? Where does she work? What areas does she cover?
Conflicting information reads as uncertainty. Machines do not recommend what they cannot confirm. A user asks for an agent in Essex County, the machine may skip Sarah entirely, not because she is unqualified, but because the evidence is a mess. Someone with cleaner, more consistent sources gets the mention instead.
This is not a ranking algorithm penalizing you. It is simpler than that: the machine cannot confidently say who you are or what you do, so it says nothing.
What ChatGPT Cannot See
The machine reads the open web. Anything behind a login, inside a private database, or exchanged in direct messages is invisible to it.
Closed MLS records. Your transaction history lives in the MLS, but most of that data is not publicly crawlable. The machine cannot see your 47 closed deals last year unless you publish that number somewhere it can reach.
Private databases. CRM notes, internal brokerage reports, your email threads with past clients: none of this exists to the machine.
Your reputation among past clients. The referrals you earn, the thank-you texts, the repeat business matter enormously to your actual practice. They are also completely invisible to ChatGPT unless they surface as published reviews or testimonials on a crawlable page.
This is the uncomfortable truth: to the machine, unpublished expertise does not exist. You may be the best agent in your market. If that expertise lives only in your head and your closed files, the machine cannot know it.
Why This Matters Now: Buyers Are Already Using AI
This is not a future concern. It is a present reality.
Forty-five percent of homebuyers already use AI tools in their search (Veterans United, 2026). They ask ChatGPT, Perplexity, or Copilot for agent recommendations, neighborhood breakdowns, and market context. The answers they receive shape who they contact.
The agents who show up in those answers are the ones with clear, consistent, current information published where the machine can read it. The agents who do not show up are not necessarily worse at their jobs. They are just invisible to the retrieval.
What You Control Today
You cannot edit ChatGPT directly. You cannot call OpenAI and ask them to update your profile. But you can control what the machine reads when it looks for you.
Publish real market content on the site you own. Not placeholder pages. Not a bio from five years ago. Current market analysis, neighborhood guides, building breakdowns: content that demonstrates what you actually know about the places you work. This is the surface you fully control. Use it.
Make every profile match. Audit your Zillow page, your brokerage listing, your Google Business profile, your Realtor.com presence. Same name. Same brokerage. Same service areas. Same contact information. Consistency is evidence. Inconsistency is noise.
Keep information current. A quarterly pass through your public profiles catches the drift before it compounds. The brokerage name you updated on your website but forgot on Zillow. The service area that expanded last year but still shows the old footprint on a directory listing.
This is not busywork. It is leverage. The agents who treat their published presence as infrastructure, maintained, consistent, current, are the ones the machine can confidently recommend.
Frequently Asked Questions
Does ChatGPT read my social media?
Mostly no. Much of social media is walled off from crawlers by login requirements and platform restrictions. Even public posts age out of relevance quickly. Your website is the durable record the machine prioritizes. Social content may occasionally surface, but it is not a reliable foundation for AI visibility.
Can I remove wrong information ChatGPT has about me?
Not directly from the model. ChatGPT does not maintain a profile of you that you can edit. It retrieves from live sources. Fix the public source it read, update the directory listing, correct the brokerage page, publish fresh content that supersedes the old, and future retrievals pick up the correction.
Which source matters most for AI visibility?
Your own website. It is the only surface where you control every word the machine reads. Profiles and directories contribute, but they are surfaces you rent, not own. Your website is the foundation. Everything else is supplemental. For more on how agents appear in AI answers, see how real estate agents show up in ChatGPT.
If you want to see what AI can currently find about you, run a free visibility audit at filtrs.io.