Large language models: Section catalog — 大規模言語モデル:セクションカタログ
Jump to navigation
Jump to search
Systems Analysis Wikiにおける、大規模言語モデル (large language model, LLM)をテーマとしたシステムアプローチ関連の記事カタログです。
Large Language Models (LLM) - 大規模言語モデル (LLM)
- 大規模言語モデル
- LLMの理論的基礎
- LLMアーキテクチャ
- Transformerアーキテクチャ
- エンコーダー
- デコーダー
- エンコーダーオンリー
- デコーダーオンリー
- エンコーダー・デコーダー
- トークン化
- トークン
- 埋め込み
- コンテキストウィンドウ
- 大規模言語モデルの学習
- 事前学習
- ファインチューニング
- インコンテキスト学習
- Top-p
- Top-k
- 温度 (LLM)
- LLMのハルシネーションと不正確な回答
- データ歪みとバイアス
- コンテキスト忘却
- 生成におけるバイアス
- Mixture-of-Experts (MoE)
- LLMのエラー削減
- LLM利用コストの最適化
- オープンウェイトモデルとクローズドウェイトモデル
- Constitutional AI
- 説明可能なAI
- RLHF
- Direct Preference Optimization
- Low‑Rank Adaptation (LoRA)
- PEFT
- ベクトルデータベース
- マルチモーダルLLM
- ジェイルブレイク
- FlashAttention
- FlashAttention-2
- FlashAttention-3
- 停止シーケンス
- 合成データ生成
- マルチモーダル推論
- 確率的オウム
大規模言語モデル (LLM) カタログ
- T5
- LaMDA
- PaLM
- BERT
- Chinchilla
- Huawei PanGu
- IBM Granite
- BLOOM
- Mixtral
- DBRX
- GPT
- Claude
- Gemma
- Gemini
- LLaMA
- Mistral
- DeepSeek
- Grok
- Qwen
- Phi
- Jais
- Jamba
- Cohere
- Falcon
- Perplexity
- YandexGPT
- Huggingface
- OpenAIの大規模言語モデル
- Googleの大規模言語モデル
- 大規模言語モデル:モデルカタログ
Prompt Engineering (LLM) - プロンプトエンジニアリング (LLM)
- プロンプト
- プロンプトエンジニアリング
- プロンプトとコンテキスト
- プロンプトエンジニアリングの基本テクニック
- Retrieval‑Augmented Generation (RAG)
- Chain-of-Thoughtプロンプティング
- Few-shotとZero-shot
- ロールプロンプティング
- Tree of Thoughts
- Self-refineプロンプティング
- Self-consistencyプロンプティング
- メタプロンプティング
- マルチエージェントプロンプティング
- プロンプト圧縮
- Program of Thoughtsプロンプティング
- Generated Knowledge Prompting
- Multimodal CoT Prompting
- Graph-of-Thoughts
- Chain-of-Verification
- Toolformer
- Least-to-Mostプロンプティング
- Automatic Prompt Engineer (APE)
- ReActプロンプティング
- Function Calling
- RAGパターン
- GraphRAG
- MM-RAG (Multimodal RAG)
- Hypothetical Document Expansion
- ハイブリッド検索
- パッケージングとコンテキスト処理
- プロンプトエンジニアリング:セクションカタログ
AI Agents (LLM) - AIエージェント (LLM)
- AIエージェント
- エージェントワークフロー
- マルチエージェントフレームワーク
- LangChain
- AutoGPT
- マルチエージェントディベート
評価とメトリック比較 (LLM)
- LLMの評価
- LLMの品質メトリクス
- パープレキシティ
- BLEU
- ROUGE
- BERTScore
- METEOR
- MAUVE
- LLM-as-a-Judge
Benchmarks and Datasets (LLM) - ベンチマークとデータセット (LLM)
- LLMベンチマーク
- MMLU benchmark
- HellaSwag benchmark
- HumanEval benchmark
- TruthfulQA benchmark
- MT-Bench benchmark
- GLUE benchmark
- SuperGLUE
- Humanity's Last Exam
- GSM8K (Grade School Math 8K)
- WinoGrande benchmark
- AgentHarm
- SafetyBench
- SWE-bench
- BIG-bench
- MATH benchmark
- FLORES‑200
- RealToxicityPrompts
- PromptRobust
- BOLD
- BBQ
- LMArena
- モデルのELOレーティング