Large language models: Section catalog — 大規模言語モデル:セクションカタログ

From Systems analysis Wiki
Jump to navigation Jump to search

Systems Analysis Wikiにおける、大規模言語モデル (large language model, LLM)をテーマとしたシステムアプローチ関連の記事カタログです。

サイト:systems-analysis.ru

Large Language Models (LLM) - 大規模言語モデル (LLM)

  • 大規模言語モデル
  • LLMの理論的基礎
  • LLMアーキテクチャ
  • Transformerアーキテクチャ
  • エンコーダー
  • デコーダー
  • エンコーダーオンリー
  • デコーダーオンリー
  • エンコーダー・デコーダー
  • トークン化
  • トークン
  • 埋め込み
  • コンテキストウィンドウ
  • 大規模言語モデルの学習
  • 事前学習
  • ファインチューニング
  • インコンテキスト学習
  • Top-p
  • Top-k
  • 温度 (LLM)
  • LLMのハルシネーションと不正確な回答
  • データ歪みとバイアス
  • コンテキスト忘却
  • 生成におけるバイアス
  • Mixture-of-Experts (MoE)
  • LLMのエラー削減
  • LLM利用コストの最適化
  • オープンウェイトモデルとクローズドウェイトモデル
  • Constitutional AI
  • 説明可能なAI
  • RLHF
  • Direct Preference Optimization
  • Low‑Rank Adaptation (LoRA)
  • PEFT
  • ベクトルデータベース
  • マルチモーダルLLM
  • ジェイルブレイク
  • FlashAttention
  • FlashAttention-2
  • FlashAttention-3
  • 停止シーケンス
  • 合成データ生成
  • マルチモーダル推論
  • 確率的オウム

大規模言語モデル (LLM) カタログ

  • T5
  • LaMDA
  • PaLM
  • BERT
  • Chinchilla
  • Huawei PanGu
  • IBM Granite
  • BLOOM
  • Mixtral
  • DBRX
  • GPT
  • Claude
  • Gemma
  • Gemini
  • LLaMA
  • Mistral
  • DeepSeek
  • Grok
  • Qwen
  • Phi
  • Jais
  • Jamba
  • Cohere
  • Falcon
  • Perplexity
  • YandexGPT
  • Huggingface
  • OpenAIの大規模言語モデル
  • Googleの大規模言語モデル
  • 大規模言語モデル:モデルカタログ

Prompt Engineering (LLM) - プロンプトエンジニアリング (LLM)

  • プロンプト
  • プロンプトエンジニアリング
  • プロンプトとコンテキスト
  • プロンプトエンジニアリングの基本テクニック
  • Retrieval‑Augmented Generation (RAG)
  • Chain-of-Thoughtプロンプティング
  • Few-shotとZero-shot
  • ロールプロンプティング
  • Tree of Thoughts
  • Self-refineプロンプティング
  • Self-consistencyプロンプティング
  • メタプロンプティング
  • マルチエージェントプロンプティング
  • プロンプト圧縮
  • Program of Thoughtsプロンプティング
  • Generated Knowledge Prompting
  • Multimodal CoT Prompting
  • Graph-of-Thoughts
  • Chain-of-Verification
  • Toolformer
  • Least-to-Mostプロンプティング
  • Automatic Prompt Engineer (APE)
  • ReActプロンプティング
  • Function Calling
  • RAGパターン
  • GraphRAG
  • MM-RAG (Multimodal RAG)
  • Hypothetical Document Expansion
  • ハイブリッド検索
  • パッケージングとコンテキスト処理
  • プロンプトエンジニアリング:セクションカタログ

AI Agents (LLM) - AIエージェント (LLM)

評価とメトリック比較 (LLM)

  • LLMの評価
  • LLMの品質メトリクス
  • パープレキシティ
  • BLEU
  • ROUGE
  • BERTScore
  • METEOR
  • MAUVE
  • LLM-as-a-Judge

Benchmarks and Datasets (LLM) - ベンチマークとデータセット (LLM)

  • LLMベンチマーク
  • MMLU benchmark
  • HellaSwag benchmark
  • HumanEval benchmark
  • TruthfulQA benchmark
  • MT-Bench benchmark
  • GLUE benchmark
  • SuperGLUE
  • Humanity's Last Exam
  • GSM8K (Grade School Math 8K)
  • WinoGrande benchmark
  • AgentHarm
  • SafetyBench
  • SWE-bench
  • BIG-bench
  • MATH benchmark
  • FLORES‑200
  • RealToxicityPrompts
  • PromptRobust
  • BOLD
  • BBQ
  • LMArena
  • モデルのELOレーティング