We ran the same property-risk test on 6 AIs — Claude, Qwen, Llama, DeepSeek. Four rival labs.
Every single one confidently made things up: wrong energy ratings, invented earthquake zones, guessed radon levels. 😬
Ungrounded, their accuracy sat around ~0.4. That's not an assistant — that's a very articulate guess.
Then we fed them Secrets.RealEstate's (https://secrets.realestate/mission) knowledge-graph grounding. Accuracy jumped to ~1.0. Every model. Every time.
The lesson isn't "AI is bad." It's this: without verified facts underneath, even frontier models will confidently mislead you.
Grounding isn't a nice-to-have. It's the line between an answer and a guess.