Footer

    Download on the App StoreGet it on Google Play

    關於

    • 認識 VoiceTube
    • 學習服務介紹
    • 加入我們
    • 常見問題
    • 熱門搜尋主題
    • 企業英文培訓
    • 社群推廣分潤計畫

    服務總覽

    • 口說挑戰
    • 單字單句本
    • Hero 智能學習
    • Tutor 真人家教
    • Vclass 名師課程
    • Campus 教育版
    • 字典查詢
    • 匯入影片並生成字幕
    • 部落格

    精選頻道

    影片分級

    • A1 初級
    • A2 初級
    • B1 中級
    • B2 中高級
    • C1 高級
    • C2 高級

    隱私權˙條款˙
    ©2026 VoiceTube Corporation. All rights reserved

    multimodal

    US

    ・

    UK

    B1 中級
    adj.形容詞多峰
    The brain is multi-modal, it has so many interacting components

    影片字幕

    Intel 英特爾 AI 晶片發佈會,10 分鐘濃縮精華帶你看!(Intel's Lunar Lake AI Chip Event: Everything Revealed in 10 Minutes)

    09:46Intel 英特爾 AI 晶片發佈會,10 分鐘濃縮精華帶你看!(Intel's Lunar Lake AI Chip Event: Everything Revealed in 10 Minutes)
    • So you can see a typical LLM, we're getting the text answer here, standard, but it's a multimodal LLM.

      是以,你可以看到一個典型的 LLM,我們在這裡得到的是標準的文本答案,但這是一個多模式 LLM。

    • We're getting the text answer here, standard, but it's a multimodal LLM.

      我也不怎麼樣。

    B2 中高級

    超強大 AI 個人助手登場!Google「I/O大會」重點 10 分鐘內一次看! (Google I/O '24 in under 10 minutes)

    09:58超強大 AI 個人助手登場!Google「I/O大會」重點 10 分鐘內一次看! (Google I/O '24 in under 10 minutes)
    • Unlocking knowledge across formats is why we built Gemini to be multimodal from the ground up.

      跨格式解鎖知識是我們從頭開始將 Gemini 打造為多模式的原因。

    • is why we built Gemini to be multimodal from the ground up.

      這就是為什麼我們從一開始就將雙子座設計為多模式的原因。

    B1 中級

    9招讀書技巧大公開!告別NG讀書法! (9 BEST Study Strategies Ranked | Stop Studying Wrong)

    11:129招讀書技巧大公開!告別NG讀書法! (9 BEST Study Strategies Ranked | Stop Studying Wrong)
    • You can also create multimodal study resources yourself.

      你也可以自己製作 multimodal 的學習資源。

    • You can also create multimodal study resources yourself.

      第七項是 interleaving。

    B1 中級

    AI 什麼時候會取代放射科醫師?🤖 (When Will Artificial Intelligence Replace Radiologists? 🤖)

    15:49AI 什麼時候會取代放射科醫師?🤖 (When Will Artificial Intelligence Replace Radiologists? 🤖)
    • These models are increasingly multimodal and sometimes called vision language models, or VLMs, because they can now process images and videos alongside text.

      這些模型越來越多模態,有時也被稱為視覺語言模型(VLMs),因為它們現在可以同時處理文字、影像和影片。

    • These models are increasingly multimodal and sometimes called vision language models or VLMs because they can now process images and videos alongside text.

      一項針對兒科影像的研究顯示,這些模型僅正確診斷了 27.8% 的病例,另一項研究的準確度則低於 50%。

    B1 中級

    Gemini 3:智慧新紀元登場! (A new era of intelligence with Gemini 3)

    01:57Gemini 3:智慧新紀元登場! (A new era of intelligence with Gemini 3)
    • Gemini has been multimodal since the beginning.

      Gemini 從一開始就是 multimodality。

    • Gemini has been multimodal since the beginning.

      Gemini 從一開始就是 multimodality。

    B1 中級

    2026年必備的 Google Gemini 3.0 Pro 功能全解析! (Every Google Gemini 3.0 Pro Feature You'll NEED in 2026)

    13:352026年必備的 Google Gemini 3.0 Pro 功能全解析! (Every Google Gemini 3.0 Pro Feature You'll NEED in 2026)
    • It's multimodal, which means it doesn't just read text;

      現在你可以把它想像成驅動 Google 所有 AI 工具的大腦。

    • It's multimodal, which means it doesn't just read text, it can understand and generate text, images, audio, video, PDFs and even entire code bases all at once.

      它是多模態的,這表示它不只讀文字,還能同時理解和生成文字、圖片、音訊、影片、PDF 甚至完整的程式碼庫。

    B1 中級

    30分鐘精通 Claude Code 程式碼! (Mastering Claude Code in 30 minutes)

    28:0730分鐘精通 Claude Code 程式碼! (Mastering Claude Code in 30 minutes)
    • You mentioned giving an image to Cloud Code, which made me wonder if there's some sort of multimodal functionality that I'm not aware of.

      你提到要為雲代碼提供影像,這讓我想到是否有某種我不知道的多模式功能。

    • Yeah, so Cloud Code is fully multimodal.

      是的,所以雲代碼是完全多模式的。

    A2 初級

    Google 的 Tulsee Doshi:Gemini 3、Deep Think、Nano Banana Pro 中的 Layer-Based 編輯 // AI Inside 110 (Google's Tulsee Doshi: Gemini 3, Deep Think, Layer-Based Editing in Nano Banana Pro // AI Inside 110)

    33:55Google 的 Tulsee Doshi:Gemini 3、Deep Think、Nano Banana Pro 中的 Layer-Based 編輯 // AI Inside 110 (Google's Tulsee Doshi: Gemini 3, Deep Think, Layer-Based Editing in Nano Banana Pro // AI Inside 110)
    • And I'm finding, especially because Gemini has stretched such a strong multimodal model, it's actually been huge to be able to take a picture of my dishwasher and be like, hey, one, tell me more about this brand.

      我覺得你的觀點其實非常精準,關於如何利用 Gemini 來處理那些過去需要找別人幫忙,或者你還沒想到的使用情境。

    • We just moved into a new house and I'm finding like, especially because Gemini is such a strong multimodal model, it's actually been huge to be able to take a picture of my dishwasher and be like, hey, one, tell me more about this brand.

      我們剛搬進新家,我發現,特別是因為 Gemini 是個非常強大的多模態模型,它其實在幫我拍下洗碗機的照片,然後問它「嘿,一、告訴我更多關於這個品牌」這方面幫了大忙。

    A2 初級

    Sydney 創意夫妻的 60坪超美小宅 📚🛍️ (圖書館兼商店) (Inside a Creative Sydney Couple’s Small Home, Public Library and Shop, 60sqm/646sqft)

    19:42Sydney 創意夫妻的 60坪超美小宅 📚🛍️ (圖書館兼商店) (Inside a Creative Sydney Couple’s Small Home, Public Library and Shop, 60sqm/646sqft)
    • We really wanted a space where we could spend all of our time in the house and we could evolve the rooms and be a bit more multimodal.

      我們真的很想要一個空間,在這裡我們可以把所有的時間都花在房子裡,我們可以發展房間,讓房間變得更加多姿多彩。

    • and we could evolve the rooms and be a bit more multimodal.

      這裡的露臺是一排微型露臺的一部分,寬約三米,在雪梨算是很小的了。

    B1 中級

    如何讓 Gemini 3.0 Pro 的使用技巧超越 99% 的人! (How to Use Gemini 3.0 Pro Better than 99% of People)

    09:00如何讓 Gemini 3.0 Pro 的使用技巧超越 99% 的人! (How to Use Gemini 3.0 Pro Better than 99% of People)
    • It's multimodal, which means it doesn't just understand text, it can process images, videos, audio, PDFs, and

      它是 multimodal,這表示它不只懂 text,還能同時處理 images、videos、audio、PDFs,還有

    • It's multimodal, which means it doesn't just understand text; it can process images, videos, audio, PDFs, and

      它是 multimodal,這表示它不只懂 text,還能同時處理 images、videos、audio、PDFs,還有

    B1 中級