Start with channel cost.
A text message should require stronger evidence than a digest entry or review queue item.
A proactive text is one of the most expensive actions a personal AI agent can take. Before it sends, make it score the evidence, explain the threshold, and preserve a correction path.
The agent should know when a confident summary still does not deserve a text. Evidence scoring makes that decision inspectable.
A text message should require stronger evidence than a digest entry or review queue item.
Show the top evidence, the threshold, and the correction path in language the user can judge.
Use with a text message AI assistant when the score clears the highest threshold.
Send qualified but non-urgent work into computer-use cache workflows.
Require stronger receipts when an AI agent builds websites or edits public pages.
Capture the message, source, timestamp, sender, related calendar entry, prior receipt, and any user correction that applies.
Rank source quality, freshness, deadline proximity, sender authority, reversibility, and known user tolerance from low to high.
Text requires the highest score. Review queue requires medium score. Digest accepts lower scores. Ignore stale or weak signals.
Every proactive text should include or link to a receipt explaining why the agent chose SMS instead of batching.
When the user marks a text as unnecessary, early, late, duplicate, or missed, update the system prompt with what went wrong and what should happen instead.
| Factor | Question | Text threshold |
|---|---|---|
| Source quality | Is the signal first-party, direct, and fresh? | High |
| Urgency | Will waiting materially harm the user? | High |
| Reversibility | Can the action be undone if the agent is wrong? | Medium to high |
| User preference | Has the user asked for this category to interrupt? | High |
| Correction history | Has this pattern caused false positives before? | Must be reviewed |
Before sending a proactive text, score the evidence by source quality, urgency, freshness, reversibility, user preference, and correction history. If the evidence does not clear the SMS threshold, route the item to review or digest. Always provide a receipt and correction path.
Use a high threshold and tune it from user corrections. SMS should be reserved for strong, fresh, user-relevant evidence.
Route to review, digest, or ignore depending on the score and user preference.
Show a readable receipt with the top reasons. The exact numeric score can stay internal unless the user asks.
Start with a text message AI assistant, then reuse the same receipts for browser and page-building workflows.
Supers-style personal agents become more trustworthy when proactive messages carry evidence receipts, correction labels, and channel thresholds.