LLM (Large Language Model)

/ɛl ɛl ɛm/

A neural network with billions of parameters trained on massive text datasets to understand and generate human language.

In italiano: LLM (Large Language Model)

LLMs like GPT-4, Claude, and LLaMA are trained on diverse text from the internet. They demonstrate emergent abilities like reasoning, few-shot learning, and task generalization. Scale is key - larger models show better performance.

Examples

  • GPT-4 with 1.76 trillion parameters
  • LLaMA 2 for open-source applications
  • Claude for long-context understanding