Studio · macOS live workspace

Call Copilot

Call Copilot listens to system audio, your microphone, or both, builds a live transcript, and uses the configured model to surface concise answers, useful contributions, diagrams, and structured meeting notes.

It stays beside the conversation; it does not join it. Call Copilot captures audio already playing on your Mac or reaching its microphone. It is not a Zoom, Meet, or Teams bot and does not enter meetings as a participant.

Live transcript

Interleave the far side and your mic with revisable speaker labels.

Useful help

Answers, key facts, speakable lines, insights, and architecture diagrams.

Living notes

Maintain summary, decisions, actions, open questions, and room read.

Explicit handoff

Copy Markdown, open notes in Review, or sync a linked Interview.

Availability: native audio capture is currently macOS-only. Call Copilot needs Speech Recognition permission and, depending on the selected sources, Screen & System Audio Recording and/or Microphone permission.

What it is#

The Studio has a transcript on the left and two right-side views: Live suggestions and Notes. One session hub keeps the transcript, notes, speaker attribution, source state, and current suggestions alive while the overlay is open. You can also type a direct question that uses the transcript gathered so far.

Open it from the Studio picker or ask Sage to open Call Copilot. There is currently no dedicated Call Copilot slash command.

Capture & permissions#

System audioCaptures what the Mac is playing—the far side of a call—through ScreenCaptureKit. macOS treats this as Screen & System Audio Recording access.
MicrophoneCaptures your side of the call or the room through the selected Mac input device. It requires Microphone access.
BothInterleaves both sources. Initial labels are “Them” and “You”; later attribution may split the far side into additional speakers.

Apple Speech Recognition creates partial and final transcript segments. Recognition is on-device when the current locale supports it; otherwise Apple’s dictation service may handle recognition. If system-audio permission was just granted, macOS can require a full app quit and reopen before capture starts.

Choose a mode#

The modes change prompting, not capture. You can switch during a session; the same transcript remains available.

Transcript & speakers#

Partial speech appears live and settles into timestamped segments. Source chips keep microphone and system audio distinguishable. The notes model can later re-attribute far-side segments to speaker IDs and proposed names; those labels update in place.

Speaker attribution is an inference. It can merge people, split one person, or assign the wrong name. Treat the transcript and names as working notes, not a legal record or a verbatim source of truth.

Use Copy to put the current transcript on the clipboard as Markdown. Call Copilot does not advertise an audio-recording archive; its working artifact is transcript text and derived notes.

Live help#

The fast model loop reviews finalized transcript chunks and returns nothing when there is no useful contribution. When it does respond, the Studio can render:

Type into Ask Call Copilot when you need something specific. The question is answered against the current transcript; it is not sent into the meeting automatically.

Notes & room read#

A slower model loop maintains a topic, summary, decisions, action items, open questions, and a list of voices. It may also produce a compact read of mood, energy, and tension. That room read is interpretive and should be checked against your own judgment.

Call Copilot also exposes Voice Mode for “Sage…” utterances. Notes addressed to Sage wait in a local batch and can be filed together; other addressed requests are acted on through the assistant path. Ordinary call audio is not treated as an instruction merely because someone says something imperative.

Handoffs#

Each handoff is a user action. Call Copilot does not silently publish notes to a meeting platform or send them to attendees.

Privacy & limits#

Get permission to transcribe. Recording and transcription laws and workplace rules vary. You are responsible for notifying participants and obtaining consent where required.

When on-device speech recognition is supported, audio stays on the Mac for transcription and the resulting transcript text is used by the configured model for answers, notes, room read, and attribution. When the locale lacks on-device recognition, Apple’s dictation service may process recognition. Model-route privacy still applies to transcript text.