Octomind adds image and voice support, refuses queries when model lacks vision

Developer tool Octomind has shipped a new update introducing image attachments and voice input across browser, Telegram, Slack, and WhatsApp sessions. Unlike most AI tools that silently strip unsupported images and respond anyway, Octomind explicitly refuses such requests upfront and names a compatible model instead. Voice input includes safeguards like a 15-second silence cutoff, a two-minute hard stop, and a minimum three-word threshold before any query is sent. All inputs — screenshots, voice notes, or typed messages — share a single persistent session across devices, with no re-uploading needed after model switches. Images are available on all plans at no extra cost, while voice features are limited to paid plans at $0.010 per minute to listen and $0.025 per minute to speak.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in