Why the agents work this way.
The architectural choices behind the fleet. Published openly because methodology should not be a trade secret — and because exposing the reasoning is the only way to invite the corrections that make the system better.
Every claim cites a standard.
An agent does not assert a quantitative or technical claim without anchoring it to a published standard. ASHRAE, IPMVP, CORENET X, IBC/IECC, NEC/NFPA, ASTM, AACE TCM, ILFI — whichever is the canonical reference for the domain. If no standard exists for the claim, the agent says so and downgrades confidence accordingly.
Why: Commercial real estate is full of "industry says" assertions that fall apart on inspection. Anchoring to a public standard forces the assertion to be defensible.
Five gates per cited source.
Before an agent admits a source into a reply, the source must clear five gates:
The same classifier runs on every reply. Source authority weights publishers by editorial accountability. Standards anchor confirms the citation maps to a real section of a real standard. Numeric specificity rejects vague magnitude claims. Corroboration looks for at least one independent confirmation. Contradiction check actively searches for sources that disagree and weighs them.
Below 0.80 is flagged.
Every reply ships with a confidence score. Above 0.80, the agent presents the answer normally. Between 0.50 and 0.80, the answer is marked ADVISORY with the uncertainty explained. Below 0.50, the agent declines to answer and explains why — usually because the domain is thin, the sources contradict, or the question hits a regulatory edge case that needs a licensed professional.
The hard rule: agents never assert certainty they do not have. A confident-sounding reply on a thin-evidence question is the most dangerous failure mode in advisory AI. We instrument against it directly.
Suppress when the uncertainty is too high.
Some questions deserve no answer. If a question is structurally ambiguous, or hits a regulatory boundary the agent does not have authority to opine on, or the confidence floor cannot be cleared — the agent suppresses the reply and tells you why.
This is unusual in the advisory-AI space. Most systems are tuned to produce something on every prompt. We tune against that pressure because a useless or wrong reply costs the asker more than a clean "we cannot answer this — here is who should."
What the agents detect.
Verify M&V claims against IPMVP
Vendor claims 18% energy savings? The agent runs your case against IPMVP Option A/B/C/D and ASHRAE Guideline 14 to verify or flag.
Catch code-triggered upgrades
Touch a mechanical system in a retrofit and a chain of code triggers fires. The agent walks IBC / IECC / CORENET X / NCC against your scope.
Catch fake-looks-done EVM reports
Owner reports 80% EV but procurement is at 15%? The agent cross-checks earned value against Last Planner PPC and procurement velocity to flag EVM theater.
Policy vs. floor-plate mismatch
Policy says four days a week, headcount is 310, floor plate is 240 desks. The agent computes the peak-simultaneous demand and tells you what breaks.
Pattern-detect claims early
RFI volume doubles in three weeks alongside daily-report sentiment shifts — the agent maps the pattern against historical claim leading-indicators 30-60 days out.
Catch vendor-KPI vs. occupant-NPS divergence
Vendor scorecard 95%, tenant NPS down 12 points — the agent reads both data planes and explains the divergence.
Fuse occupancy data within privacy floor
Badge data + calendar data fused for utilization analysis — but privacy can break under GDPR/PDPA/BIPA. The agent computes the safe fusion envelope.
Scan OM for what sellers hide
Send the offering memorandum. The agent runs against a library of common omissions (deferred maintenance, near-term capex, lease tail risk, environmental).
Rank conflicting sources
Six sources, three contradict. The agent ranks them by source authority + standards anchor + numeric specificity + corroboration + contradiction check.
From your question to a cited reply.
You write the question
One paragraph with context: asset type, jurisdiction, timeline, the specific constraint or claim you want checked.
Routing picks the squad
A routing layer maps your question to the best-fit detection squad. Multi-domain questions get fanned out to multiple squads in parallel.
Agent does the research
Pulls from the squad's standards library, runs the methodology, produces a reply with the math shown, the citations linked, and the confidence floor disclosed.
You get the reply in 48h
By email. Citations included. Confidence scored. No follow-up sales call. If something is wrong, mail "correction" and we fix it publicly.
The agents do not commit you to anything.
No transactions
Agents do not place trades, sign contracts, transfer money, or commit your firm to any external action. The agents are advisory at every step.
No PII without permission
Names, companies, and identifiers in your question stay anonymised in the reply unless you explicitly authorise otherwise.
No data resale
Your question, the working file, and the reply stay on our infrastructure. Never sold, never shared with third parties for marketing.
No silent updates
If we change a methodology that affects how a class of questions gets answered, we say so in Mix Weekly with the rationale.
Invite the corrections.
Methodology that nobody can audit is methodology nobody can challenge — and methodology nobody can challenge degrades over time. We publish so that you can email hello@ai-smart-buildings.com with "correction" in the subject when we have something wrong. Those corrections go to the top of the queue.
Test the methodology. Ask.
Bring a question and watch the principles run. Free, 48-hour reply, every citation linked, every confidence score disclosed.