← Researchへ戻る

Keywords

LLMキーワード一覧

キーワードをクリックすると、LLM技術知識体系の該当ページを開きます。

1-6|基礎・学習・推論

Transformerp.5encoder-onlyp.5encoder-decoderp.5decoder-onlyp.5causal maskp.5cross-attentionp.5QKVp.5scaled dot-product attentionp.5Multi-Head Attentionp.5MHAp.5MQAp.5GQAp.5KV cachep.5FlashAttentionp.5FlashAttention-3p.5FlashAttention-4p.5FFNp.5SwiGLUp.5pre-normp.5post-normp.5RoPEp.5ALiBip.5QK-Normp.5logit soft-cappingp.5誤差逆伝播p.12forward propagationp.12backpropagationp.12勾配降下法p.12SGDp.12Adamp.12AdamWp.12学習率スケジュールp.12warmupp.12cosine decayp.12weight decayp.12cross entropyp.12label smoothingp.12gradient clippingp.12activation checkpointingp.12BF16p.12FP16p.12FP32p.12loss scalingp.12tokenizerp.20BPEp.20WordPiecep.20Unigramp.20SentencePiecep.20byte fallbackp.20vocabularyp.20special tokensp.20BOS / EOSp.20chat templatep.20detokenizationp.20multilingual tokenizerp.20scaling lawsp.25Chinchilla則p.25compute-optimalp.25pretraining datap.25data mixturep.25データ重複除去p.25data contaminationp.25Common Crawlp.25FineWebp.25RefinedWebp.25Dolmap.25Nemotron-CCp.25synthetic datap.25SFTp.32instruction tuningp.32preference datap.32reward modelp.32RLHFp.32PPOp.32DPOp.32IPOp.32KTOp.32ORPOp.32RLAIFp.32RLVRp.32GRPOp.32DAPOp.32rejection samplingp.32knowledge distillationp.32judge modelp.32推論モデルp.39Chain-of-Thoughtp.39hidden CoTp.39test-time computep.39推論時計算p.39inference-time scalingp.39thinking budgetp.39reasoning effortp.39best-of-Np.39self-consistencyp.39tree searchp.39verifierp.39process rewardp.39outcome rewardp.39budget forcingp.39DeepSeek-R1p.39

7-12|アーキテクチャ・評価・安全性

13-18|効率化・分散学習・推論システム

19-23|モデル・企業・OSSエコシステム

24-27|RAG・エージェント・コーディング

28-31|プロンプト・LLMOps・実務設計