large language model

原始詞典資料1 筆記錄

這個字我們還沒有寫成教學內容。以下是外部詞典記錄的用法,原文照列,方便你先看懂它的意思。

large language modelnoun

  • 1

    A type of neural network specializing in language, typically including billions of parameters.

    • GPT-3 belongs to a category of deep learning known as a large language model, a complex neural net that has been trained on a titanic data set of text: in GPT-3’s case, roughly 700 gigabytes of data drawn from across the web, including Wikipedia, supplemented with a large collection of text from digitized books. GPT-3 is the most celebrated of the large language models, and the most publicly available, but Google, Meta (formerly known as Facebook) and DeepMind have all developed their own L.L.M.s in recent years.
    • GPT-2 is a large language model with 1.5 billion parameters, trained on a dataset of 8 million web pages scraped from the internet.

資料來源:英文維基詞典 English WiktionaryCC BY-SA 4.0)。 原文照列,未經改寫。