Paul Jialiang Wu agentic-portfolio 中文 Español 한국어 日本語✉️ Free list
← Back to portfolio

AI-Native Series · Trust Engineering

Test Everything. Hold On to the Good. Now We Have the Tools.

By Paul Jialiang Wu · agentic-portfolio-lovat.vercel.app · 2026-08-02

1-minute takeaway — what you'll walk away with

In 2017 a well-known charismatic minister said Jesus told him, "very simply and very clearly," that Pope Francis was the false prophet of Revelation. In 2026, after the prophecy failed, the claim became "a vision" he had "misinterpreted." That is not a spiritual problem first — it is a mutable ledger problem, and engineering already knows the fix. This essay takes a documented ministry-integrity collapse, shows that Scripture specified the test suite three millennia ago, and lays out the OEC stack — Observability, Eval, Control — that AI-native trust engineering would wire around any system whose words move people's lives and money.

A case study in charisma outrunning accountability, and the trust stack that closes the gap. ~9 min.

Two panels: on the left, a mutable ledger where the 2017 claim 'Jesus said: Pope Francis is the false prophet' is struck through and overwritten in 2026 by 'a vision I misinterpreted'; on the right, an append-only ledger where the 2017 claim is committed, the 2026 outcome is recorded as FAILED, and the revision attempt is rejected
The same nine years, on two ledgers. On the left, the story can be rewritten. On the right, only appended.

In July 2026, Pastor Joe Sweet and the board of Shekinah Worship Center published a formal public statement about the minister Sundar Selvaraj, whom they note they had platformed "for over 20 years" [4]. Its central finding is unflinching — and, unusually for this genre, it leads with its evidence inventory: "It has been thoroughly documented in video recordings, transcriptions of those videos, text messages, numerous emails, testimonies of numerous victims and witnesses and a conference phone call with Sundar Selvaraj on February 25, 2026, that Sundar has demonstrated a pattern of lying and has engaged in giving false “prophetic” words to manipulate and intentionally mislead people, causing harm to individuals and families." The statement calls the warning "our duty to God and to people," concerning a ministry that "has, in fact, proven to be harmful, especially to the unwary" [4].

I'm not writing this to pile on one man; the allegations are the testimony of those who came forward, and I've listed what is verified and what is testimony at the end. I'm writing because the pattern in those transcripts is one I recognize professionally. I spend my days building accountability systems for AI agents — machines whose words can move people's decisions and money. And the uncomfortable observation is this: the average production chatbot now operates under stricter prophetic standards than some prophets. Its claims are logged immutably. Its outputs are evaluated against declared criteria. Its failures have consequences wired to them. The pulpit deserves at least what we demand of a language model.

The diff that tells the whole story

The clearest exhibit is a nine-year quote diff, preserved in the second transcript by the minister Stephen Powell. In 2017: "Jesus appeared to me... and then very simply and very clearly he said 'the present Pope Francis is the prophesied false prophet.'" In the 2026 apology, after Pope Francis died without the prophecy coming to pass: "I had a vision... I mistakenly presumed and misinterpreted it to mean that the reigning pope at the time, Pope Francis, was the one." Powell's critique lands where an engineer's would: "He is changing the story... when you say that Jesus appeared to you and told you with his words... there's no room for interpretation."

In my field this maneuver has a name: rewriting the commit after the build failed. A direct quotation became a vision; a named subject became an assumption the audience supposedly made. The claim mutated exactly at the moment accountability arrived — which is only possible because the claim lived in a mutable medium: memory, sermons, re-uploaded videos. Nothing pinned the 2017 words to 2017.

The mental model: charisma scales faster than accountability — ledgers scale accountability. A gifted speaker can reach a million people by Sunday; the community's ability to test what he said grows much slower. Every integrity collapse in the transcripts — and in crypto, in wellness empires, in AI-hype startups — lives in that gap. You close it the way engineering closes it: make the claims observable (recorded, immutable, public), make the evaluation explicit (declared criteria, scored by others), and make the consequences wired (platforms and money respond to the scorecard, not the charisma).

The case pattern — six failure modes, one root

The transcripts document six recurring mechanisms. Condensed, and attributed — these are the accounts of those who came forward:

One root under all six: unfalsifiable inputs, unaudited state, unenforced outputs. Any system with that architecture — church, DAO (a blockchain-governed collective with no human gatekeeper), startup, or agent swarm — produces this failure class eventually. The robes are incidental.

How the truth actually surfaced — and what it cost

The discovery chain, as the third transcript recounts it, is a case study in accountability running on manual labor. In January 2026, former members of Steven Francis's church — some of them former members of Sweet's own congregation — contacted Sweet's office "in desperation," reporting lies, manipulation, and financial dishonesty. They had first gone to the system's own accountability layer — the church's elders, and Selvaraj himself — and, per their account, were ignored. The inner loop was closed; the truth had to route around it.

So in February 2026, Sweet and his wife traveled to Shelby, North Carolina, and rebuilt the observability layer by hand: days of interviews with ten people who describe themselves as victims, written statements from ten more, audio and video recordings of messages the witnesses describe as deceptive, financial documents and canceled checks described in the transcript as evidence of money laundering — and the wedding transcript in which Selvaraj prophetically blessed the very marriage he had — per the transcript — privately called "not the will of the Lord." Weeks of travel and testimony-gathering to assemble what an append-only ledger would have held all along. The evidence always existed. The access cost a pastorate's worth of trust and a cross-country trip.

Two earlier moments in the transcript deserve their own frame. In 2019, Steven Francis asked Sweet to sponsor his U.S. religious-worker visa; the forms that arrived claimed Steven would be a 20-hour-per-week employee of Sweet's church, preach there four times a year, and sit on its board — none of it true, per Sweet. Sweet refused to sign. That signature was a write boundary — the same mechanism my strategy engine uses when it refuses to record a decision it cannot verify: bad claims never get to exist if the gate holds. And in 2023, in Jerusalem, when Sweet explained the refusal, Selvaraj's reported reply was the quiet center of the whole affair: "No, my brother did that. I know him and I know what he is like." If the reported reply is accurate, the knowledge preceded the concealment by years. Observability was never the bottleneck for the insiders — enforcement was.

Scripture shipped the test suite first

Here is what should embarrass no Christian: none of the fixes require new theology. The oldest layer of the canon already specifies falsifiability as the test for prophetic claims: "If what a prophet proclaims in the name of the Lord does not take place or come true, that is a message the Lord has not spoken. That prophet has spoken presumptuously, so do not be alarmed" (Deuteronomy 18:22). That is a resolution criterion — outcome-based, binary, no appeals to sincerity.

The New Testament adds peer review: "Two or three prophets should speak, and the others should weigh carefully what is said" (1 Corinthians 14:29) — evaluation by a panel, not self-attestation. It adds an explicit instruction to test: "do not believe every spirit, but test the spirits to see whether they are from God, because many false prophets have gone out into the world" (1 John 4:1). It defines the metric: "By their fruit you will recognize them" (Matthew 7:16) — outcomes, not vibes. It ships a glory-seeking detector: "Whoever speaks on their own does so to gain personal glory" (John 7:18). It predicts the exploit class: "In their greed these teachers will exploit you with fabricated stories" (2 Peter 2:3). And it hard-codes the one non-negotiable: "I am the Lord... I will not yield my glory to another" (Isaiah 42:8) — no worship-receiving-by-proxy, ever. It even names the posture this essay takes for its title: "but test them all; hold on to what is good" (1 Thessalonians 5:21) — testing is not the opposite of honoring prophecy; it is how prophecy is honored.

The spec is complete. What's missing is the runtime — the church-scale machinery that actually executes these tests faster than every twenty years.

The human attempt, honored honestly

The charismatic world has tried. The Prophetic Standards Statement (2021), signed by dozens of leaders after a wave of failed public prophecies, says plainly: "we can only believe the prophetic word if it is not contrary to Scripture, it is not factually in error, and our own spirits bear witness with it." It names the structural problem — "anyone claiming to be a prophet can release a word to the general public without any accountability or even responsibility" — and it mandates the fix for public failure: "If the word was delivered publicly, then a public apology (and/or explanation/clarification) should be presented" [1]. On the financial side, the ECFA's Seven Standards of Responsible Stewardship have run a working accreditation-and-revocation loop for ministries' finances since 1979 [2].

Both are real progress, and both share a limitation: they are voluntary policies executed on mutable infrastructure. A signatory can quietly go silent; a statement binds only the conscientious — the people least likely to need it. Policy without observability is a sermon about logging.

Three-layer stack diagram titled THE OEC STACK FOR SPIRITUAL AUTHORITY: Observability (claims committed append-only at utterance, finances disclosed as data), Eval (declared resolution criteria, judged by others per 1 Corinthians 14:29, track record scored by Deuteronomy 18:22), and Control (platforms, endorsements and funding respond to the scorecard; documented correction loop). A footer reads: the spec is ancient — the runtime is missing
Observability · Eval · Control — the loop every trustworthy agentic system runs, mapped onto the tests Scripture already commands.

The OEC stack: what AI-native engineering adds

In agentic system design we use a simple governing loop — OEC: Observability, Eval, Control. You cannot trust what you cannot observe; you cannot improve what you cannot evaluate; you cannot govern what you cannot correct. Applied here:

1 · Observability — claims become commits

A public prophetic word should be recorded the way engineering records anything it may need to audit: append-only, timestamped, content-addressed. This is the one place the web3 toolbox earns its keep — not tokens, but the humble commitment device: hash the claim at utterance, anchor it in a public log (a transparency log in the mold of Certificate Transparency [3], or a smart-contract registry — the mechanism matters less than the immutability). Then "Jesus said Pope Francis" in 2017 stays diffable against "a vision I misinterpreted" in 2026, and semantic retreat becomes a visible edit war instead of a memory dispute. The same layer covers the costume problem: ECFA-style financial disclosure, but as machine-readable data rather than a PDF a donor never finds.

2 · Eval — declared criteria, judged by others

Every logged claim carries its resolution criterion at commit time: what would count as fulfilled, by when, judged by whom. Deuteronomy 18:22 is the grader; 1 Corinthians 14:29 names the panel — others weigh, never the speaker. Three honest grades, exactly as in my own eval discipline: fulfilled, failed, and — crucially — not measurable, for words too vague to test. A prophecy that can never be false can never be counted true; it goes in the ledger as unmeasurable, not as a win. Over time each public ministry accrues what every AI agent I ship accrues: a track record computed from the ledger, not from the highlight reel. (This is the discipline behind my private eval-anything framework: every claim a testable assertion; no evidence, no pass — the same rule this portfolio applies to résumés and to its own report cards.)

3 · Control — consequences wired to the scorecard

Pastor Sweet's de-platforming was the control action — correct, courageous, and decades latent, because it depended on one insider's conscience, a refused signature in 2019, and a self-funded investigation in 2026. In an OEC design, control is wired, not heroic: platforms, conference invitations, endorsements and giving flows subscribe to the scorecard. A failed public word triggers the Prophetic Standards Statement's own remedy — public correction — as a standing expectation, not a hostage negotiation. And repentance, the church's oldest control loop, gets the dignity of documentation: a recorded correction, a change in behavior the ledger can see. In machine-learning terms this is the fine-tuning step — error signal in, adjusted behavior out — and its absence is precisely what the July 2026 statement describes: evidence presented, no correction returned.

I run this loop on myself, involuntarily. Earlier this week my own accountability engine graded a turn of my work 0.57 against the 1.0 I claimed, because my evidence was weaker than my confidence — and the number is published where my readers can check it. It stung precisely as much as it should. Systems that force the quiet number out loud are the only ones whose loud numbers mean anything — that's as true of a portfolio as of a pulpit.

The thought experiment: replay the documented case through the stack

My rule for any proposed system: test it against the incident that motivated it — honestly, including the cells where it fails. The Shekinah statement is unusually testable because it publishes its own evidence. Five replays:

And the honest failure cells: the stack cannot see unrecorded pastoral conversations; it cannot stop spiritual intimidation from suppressing reports before they reach any log; and its adoption problem is real — the leaders who most need a ledger will not volunteer for one. Which is why the control lever is not the prophet but the platforms: conferences, networks, and broadcasters requiring ledger participation as a condition of the stage, exactly as ECFA membership already works for finances.

Update, same day: this thought experiment is no longer only prose — the stack shipped as working code, with these five replays as its named test suite, in the open-source seminary platform kingdom-come. The receipts are in part 2.

What this does not do

Honesty about limits: a ledger cannot regenerate a heart, and no gate ever built has stopped a determined fraud from finding the unmeasured surface. This stack does something narrower and still decisive: it raises the cost of fraud and collapses the cost of discernment. Today, testing a minister requires twenty years of insider access and the courage of a Joe Sweet. With claims-as-commits and scores-in-public, it requires a link. The Spirit is not caged by this — testing is commanded (1 John 4:1); the ledger just makes obedience cheap. The genuinely humble minister loses nothing except plausible deniability, which is the one asset a true prophet never needed.

Patterns / Anti-patterns

The mechanism, named: charisma scales faster than accountability, and every integrity collapse lives in that gap; ledgers, declared evals, and wired consequences are how you scale accountability at the same rate as reach.

The first principle, one sentence: a word claimed from God is a commit, not a draft — record it, let others test it against its own stated criterion, and let the consequences bind, because Deuteronomy 18:22 was never a suggestion.

Provenance — what's verified and what's testimony

The case material comes in two classes. Quotations attributed to the Shekinah Worship Center statement — the evidence-inventory sentence, "pattern of lying," "duty to God and to people," "harmful, especially to the unwary," the 20-year platforming, the wedding-transcript excerpts, the 2019 visa-application characterization, the consent and timing details, and the apology-timing observation — were verified verbatim against the live public statement at shekinahworship.com (published July 10, 2026; updated July 26, 2026). The remaining material — the 2017 and 2026 prophecy quotations, Stephen Powell's critique, and parts of the discovery narrative — is quoted from three transcript digests supplied to me and remains the testimony of those who came forward, as rendered there. In both classes the claims are one side's account: attributed throughout, presented as allegation where they allege, and the accused's response beyond the quoted apology is not represented here. No wrongdoing is asserted in my own voice. The Prophetic Standards Statement and ECFA quotations were verified verbatim against their live sites; all Scripture quotations are from the NIV, verified against BibleGateway; Certificate Transparency is cited descriptively. The 0.57 grade is my own, reproducible from this repo's committed report cards.

References

  1. Prophetic Standards Statement (2021). "We can only believe the prophetic word if it is not contrary to Scripture, it is not factually in error, and our own spirits bear witness with it." propheticstandards.com
  2. Evangelical Council for Financial Accountability. Seven Standards of Responsible Stewardship. ecfa.org/Content/Standards
  3. Laurie, B., Langley, A., & Kasper, E. (2013). Certificate Transparency, RFC 6962 — the append-only public-log design this essay borrows. datatracker.ietf.org/doc/html/rfc6962
  4. Shekinah Worship Center (2026, July 10; updated July 26). Statement Regarding Sadhu Sundar Selvaraj. All quotations verified verbatim against the live page. shekinahworship.com/statement-regarding-sadhu-sundar-selvaraj
  5. Scripture quotations: The Holy Bible, New International Version (NIV) — Deuteronomy 18:22; Isaiah 42:8; Matthew 7:16; John 7:18; 1 Corinthians 14:29; 1 Thessalonians 5:21; 2 Peter 2:3; 1 John 4:1. Verified against biblegateway.com

Related

Written from the Shekinah Worship Center's public statement (quotes verified verbatim), three supplied transcript digests (attributed as testimony), and the verified texts of the Prophetic Standards Statement, ECFA, and Scripture (NIV), and a working week spent building the same accountability loops for machines. — Paul Jialiang Wu · agentic-portfolio-lovat.vercel.app