# ChatGPT 桌面版應用程式推出 ChatGPT Voice 讓使用者能透過即時語音控制電腦

> 📖 本站完整內容索引（documentation index）：[llms.txt](/llms.txt)

> 原作者：OpenAI (@OpenAI) · 策展與摘要：EasyVibeCoding · 平台：X (Twitter) · 熱度：🔥🔥🔥 · 日期：2026-07-24

> 原始來源：https://x.com/OpenAI/status/2080378182469857576

## 中文摘要

ChatGPT 桌面版應用程式推出 ChatGPT Voice 讓使用者能透過即時語音控制電腦。

**功能與運作機制**
OpenAI 於 2026 年 7 月 24 日宣布將 ChatGPT Voice 導入桌面版應用程式，由 GPT-Live 技術驅動，具備同時說話、聆聽與協調應用程式內工作的能力。
- 支援透過語音直接控制電腦，並指揮在 ChatGPT Work 或 Codex 中運作的多個 Agent。
- 支援 macOS 與 Windows 系統，向 Plus、Pro、Business、Edu 及 Enterprise 方案全球推出。
- 使用者亦可透過 iOS 應用程式搭配配對的遠端存取功能，在 Codex 中使用 ChatGPT Voice，Android 支援則將陸續推出。

**實際應用與體驗**
開發者 Tibo（@thsottiaux）表示使用者現在能擺脫鍵盤進行日常工作。使用者 Guinness Chen（@guinnesschen）分享其為理想的規劃模式，能從模糊的想法開始，透過 AI 提出犀利的提問協助釐清需求、邊對話邊繪製架構圖，最後啟動 Codex 任務進行建置。

<video src="https://pub-75d4fe1e4e80421b9ecb1245a7ae0d1a.r2.dev/curated/1784874664571-0rphiszx.mp4" poster="https://pub-75d4fe1e4e80421b9ecb1245a7ae0d1a.r2.dev/curated/0fa39ed718af110c.jpg" controls playsinline preload="metadata" style="max-width:100%;height:auto;display:block;margin:1rem 0"></video>
> 一位工程師透過 OpenAI 桌面版應用程式的 Voice 功能進行開發。

**桌面協作與畫面展示**
在實際應用展示中，工程師可透過桌面版應用的 Voice 功能進行協同開發，介面呈現 macOS 桌面 App 風格的對話記錄、專案看板與檔案列表，並包含如 `launch-preview`、`audio-phase-inputs`、`feature-flags` 等目錄與檔案。介面亦展示了音訊處理邏輯與架構圖比較的任務設置，包含「Current behavior」與「Proposed behavior」的流程對比，並能將變更後的架構指派給任務執行。本站先前策展過〈[OpenAI 推出 GPT-Live 即時語音對話](/curated/2415)〉，當時展示了語音對話支援，而本次更新則進一步將該能力延伸至桌面版與多 Agent 協同調度。

<video src="https://pub-75d4fe1e4e80421b9ecb1245a7ae0d1a.r2.dev/curated/1784874693069-9o0j539j.mp4" poster="https://pub-75d4fe1e4e80421b9ecb1245a7ae0d1a.r2.dev/curated/0d820ec5406866eb.jpg" controls playsinline preload="metadata" style="max-width:100%;height:auto;display:block;margin:1rem 0"></video>
> 某 AI 程式介面透過對話呈現 Voice-leading Engine 變更的架構圖比較與任務設置。

## 媒體內容

**一位工程師透過 OpenAI 桌面版應用程式的 Voice 功能進行開發。**

**影片中的 Prompt 與操作**

操作步驟：

1. （00:01）啟動投影並播放日常播放清單
2. （00:04）透過桌面 App 語音助理撰寫部落格貼文草稿
3. （00:39）呼叫 Codex 檢查 feature flags 設定
4. （00:45）建立新執行緒尋找 bug 根源並建立 pull request
5. （00:58）透過語音設定專案目標並標記團隊審查

**逐字稿**

- `00:00` 嘿，chat，把畫面投射出來。（Hey, chat, put yourself on the projector.）
- `00:03` 播放我的播放清單。（Play my day list.）
- `00:04` 請稍候。（One moment.）
- `00:05` 所以我們得為下週的語音功能發布準備一些發表素材。（So we got to prepare some launch materials for the voice launch next week.）
- `00:10` 你可以幫我們草擬一些部落格文章嗎？（Could you help us draft some blog posts?）
- `00:13` 沒問題。（Absolutely.）
- `00:14` 一個簡單的初步版本。（A simple first pass.）
- `00:16` 在桌面版應用程式中體驗語音功能。（Meet voice in the desktop app.）
- `00:18` 語音能讓使用者更輕鬆地自然推進工作。（Voice makes it easier to move work forward naturally.）
- `00:22` 使用你的電腦而不中斷心流。（Use your computer without breaking flow.）
- `00:24` 我們把它發布吧。（Let's ship it.）
- `00:26` 我對於一些可以做的最後潤飾有一個功能點子。（I got this feature idea for, like, some last-minute polish we can do.）
- `00:30` 我要稍微碎念一下，然後我要你問我一些問題並提出反對意見。（I'm going to ramble a little bit, and then I want you to ask me some questions and push back on me.）
- `00:33` 看起來光是在 shader 裡面就能做到很多事，因為它已經有音訊和 phase 輸入了。（It looks like we can get pretty far just inside the shader, since it already has the audio and phase inputs.）
- `00:39` Codex，你能去檢查功能旗標並確保它們都正確設定了嗎？（Codex, could you go check the feature flags and make sure they're all configured properly?）
- `00:42` 等等，意見回饋管道裡還有一個新的錯誤回報。（Wait, there's also a new bug report in the feedback channel.）
- `00:45` 你能建立一個新的執行緒、發起 pull request 並找出根本原因嗎？（Can you spin up a new thread, make a pull request, find the root cause?）
- `00:48` 讓我看一下。（Let me take a look.）
- `00:49` 啊，找到了。（Ah, found it.）
- `00:51` 就是那個功能旗標。（It was that feature flag.）
- `00:53` 完全設定錯誤。（Totally misconfigured.）
- `00:54` 難免啦。（Happens.）
- `00:54` 我們發布的所有準備工作都搞定了，對吧？（We've got everything set for launch, right?）
- `00:56` 對，我想我們準備好了。（Yeah, I think we're ready to go.）
- `00:57` 你懂的。（You get the idea.）
- `00:58` 只要設定一個目標，然後標記團隊稍後審查。（Just set a goal, and then tag the team for review later.）
- `01:00` 我會用手機跟你保持聯繫。（I'll check in with you from my phone.）
- `01:02` 知道了。（Got it.）
- `01:02` 我會為此建立一個後續目標。（I'll create a follow-up goal for that.）
- `01:04` 然後你知道的，就是別出錯。（And then just, you know, make no mistakes.）

**某 AI 程式介面透過對話呈現 Voice-leading Engine 變更的架構圖比較與任務設置。**

**影片中的 Prompt 與操作**

Prompt（00:01）：

```
可以秀一下你想做的變更的圖表嗎？
```

原文：Can you show me a diagram of the changes you want to make?

Prompt（00:33）：

```
你可以開始一個任務來做這件事嗎？
```

原文：Could you start a task to go do this?

操作步驟：

1. （00:42）開啟工作佇列與任務指派

**逐字稿**

- `00:00` 你可以給我一張你想要進行變更的圖表嗎？（Can you show me a diagram of the changes you want to make?）
- `00:04` 可以，我會在螢幕上放一個簡單的圖表。（Yeah, I'll put a simple diagram on screen.）
- `00:08` 在螢幕上，左側，按鍵輸入提供給 UI，這有時會讓求解器延遲，導致錯誤的前一個和弦溜進來。（On screen, left side, keys feed the UI, which sometimes lags the solver, so the wrong previous chord slips in.）
- `00:17` 右側，按鍵提供和弦意圖。純求解器使用最後彈奏的和弦配置。（Right side, keys give chord intent. A pure solver uses the last played voicing.）
- `00:22` 它輸出下一個和弦配置、彈奏它、將其儲存為最後彈奏的和弦，並更新顯示。（It outputs the next voicing, plays it, saves it as last played, and updates the display.）
- `00:27` 讓音訊引擎成為真實來源可以防止過時的比較。（Making the sound engine the source of truth prevents stale comparisons.）
- `00:32` 對，這看起來很合理。你可以建立一個任務來執行這個嗎？（Yeah, this seems legit. Could you start a task to go do this?）
- `00:36` 沒問題，我來設定。（Sure, I'll set that up.）
- `00:44` 我已經指派了一個包含求解器重構的任務，記住最後彈奏的和弦以及目標迴歸覆蓋率。（I've assigned a task with the solver rewrite, remembering the last played chord and targeted regression coverage.）

## 標籤

功能更新, Voice, 新產品, macOS, Windows, iOS, OpenAI, ChatGPT, Codex
