Voice & Chat

A persistent conversation with an agent that can see your workspace — by voice or keyboard, same thread.

user Version 1 Updated 2026-07-16

What problem this solves

Sometimes the question isn't "show me the board", it's "what did I leave unfinished on Pointr?" or "add 'test the export on real data' to FieldLog's backlog" — while you're walking. Chat answers those: the agent reads your projects, roadmaps, and activity through Mind's tools mid-conversation, and can write roadmap items when you ask.

Voice and text are the same conversation. Speak a question in the car, read the answer later at your desk, type the follow-up — one retained thread.

Before it works: two switches

"mic needs HTTPS — open TOM at its https:// address" means you're on http://100.88.227.62:5176. Same app, wrong door. A real "microphone permission denied" means you (or iOS) declined the permission prompt — check Settings → Safari (or the site permissions) and retry.

Talking: tap-to-talk, not open mic

The mic is push-to-start, not always-listening:

tap 🎤 → listening (mic turns live, transcript area appears) speak… → pause briefly, and it finalizes + sends on its own or tap ✓ → finalize + send now assistant speaks → replies stream in and are read aloud, sentence by sentence tap 🎤 mid-speech → barge-in: speech stops instantly, mic is yours again

Replies are spoken with the device's own voices and stream as they're generated — you'll hear the first sentence while the rest is still being written. The speaker icon (top right of a conversation) mutes speech entirely; the stop button cancels a reply you don't want to wait through.

If the live transcript says it's unavailable and asks you to tap ✓ when done, the streaming transcriber is offline and TOM fell back to whole-utterance mode. Everything still works — you just don't see words appear as you speak, and you finish with ✓ (or a pause).

What the agent can actually do

Mid-turn, the agent calls Mind's tools. You'll see small tool notes appear in the thread as it works. Out of the box it can:

It cannot run shell commands, queue autonomous work, or change settings from chat — the allowed toolset is a Mind setting (chat.tool_allowlist), deliberately conservative by default.

When you mention a project and the agent loads its context, the conversation binds to that project — you'll see its name appear under the conversation title, and later questions assume that project until you steer elsewhere.

Conversations persist

Each thread keeps its history (titled automatically from your first message) and survives restarts on both ends. Old threads stay useful — the agent carries a rolling summary, so a conversation you left three days ago still knows what you were doing. Start a New conversation when the topic genuinely changes; reuse an existing one when it's the same piece of work.

Gotchas

Where to find this in the app