Watch your speech live, see its tags, step inside the image you captured
Speech is now visible end to end on the Dictation side: the live bar watches the speech and its context, the transformed panel shows the rectangles you drew and the images with their tags, and hovering a captured imag…
Speech is now visible end to end on the Dictation side: the live bar watches the speech and its context, the transformed panel shows the rectangles you drew and the images with their tags, and hovering a captured image steps inside the content you annotated earlier. The recording bar and the Dictation panel were renewed in the same release.
From the live bar to four panels; from the image queue into annotated content
You watch the speech from the live bar; the context shows which app it was started in, and hovering the flow indicator opens the four panels. The panel holds the transformed form of your speech: rectangles you draw while speaking appear as tags in the transcript, images as references with the moment they were saved. You read the content in three views: Speech is your plain words, Events is what goes to the agent, Context is the human form. Capture actions open inside the conversation; the most used is Capture Screen — select an area with Shift and it lands tagged into your speech. Captured images drop into the queue: the app they were captured in, the form they will reach the agent in and their size sit on one row; press the form badge to send the image as text or as a picture, and when size cannot be counted as tokens it falls back to kilobytes. Hovering the row opens the final step: the content you annotated earlier stands in front of you, picture inside picture, with its rectangles and numbers.
This card is two nested visuals: hovering the queue row steps into the preview of previously annotated Cockpit content.
The recording bar tracks context live; speeches paste into a single session
As you speak and capture, the context list grows; the system actively tracks the context and updates itself while you talk. The red waveform shows your speech is being heard; the small indicator says it is being transcribed in the background. While an old speech waits unpasted you can record a new one: they wait separately, and pasting merges them into a single session.
The Dictation panel: tag, relate, listen, search
Your conversations list in the panel with their durations and tags; tag with ⌘D or right click. Conversations merged by pasting count as related and their bond is recorded. Read your own conversation on the right; click anywhere in the text and the audio starts playing from that moment, light grey regions are where you stayed silent. For Whisper users our own algorithm cleans mistranscribed subtitle-like patterns; the panel names what was hidden and one click brings it back. Your total speech record sits in the corner of the panel; the list updates as you type a search, and typing guide then pressing Tab keeps every guide-tagged conversation in the panel.