Developer finds AI email assistant rates every reply 0.85 confidence, including spam
A developer building InboxSync, a RAG-based email reply suggestion tool, discovered that the system's confidence score was returning a fixed value of 0.85 for every response regardless of context. Testing revealed the AI confidently drafted replies to spam, out-of-office messages, and even a formal GDPR data deletion request — situations where no reply should be sent at all. In the GDPR case, the system produced an authoritative-sounding legal commitment with unfilled placeholders, which if sent could constitute a binding legal acknowledgment. The root cause was that the confidence metric was never properly wired to any decision logic, making it a meaningless static number rather than a genuine quality signal. The findings highlight a broader risk in AI pipelines: a system that cannot recognize when it is out of its depth may cause real harm by acting confidently in domains it has no training data for.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in