The chapter begins with the introduction of basic modules in modern neural networks. Then, we provide details about transformers, which are state-of-the-art neural network architectures and popular backbones for foundation models. Finally, we summarize major components in large language models, including next-token prediction, decoding, alignment, and parameter-efficient finetuning.

错误:搜索内容不能为空,请输入英文关键词
错误:关键词超出字数限制,请精简
高级检索

Neural Networks

  • Pin-Yu Chen,
  • Sijia Liu

摘要

The chapter begins with the introduction of basic modules in modern neural networks. Then, we provide details about transformers, which are state-of-the-art neural network architectures and popular backbones for foundation models. Finally, we summarize major components in large language models, including next-token prediction, decoding, alignment, and parameter-efficient finetuning.