---
title: "Phi (Microsoft) (ID)"
source: "https://systems-analysis.info/int/Phi_(Microsoft)_(ID)"
wiki: "systems-analysis.info/int"
article: "Phi_(Microsoft)_(ID)"
language: "id"
categories:
  - "Category:Indonesian"
  - "Category:Large language models"
  - "Category:LLM families"
  - "Category:Machine learning"
revision_id: 5606
wiki_created_at: 2026-09-06T23:51:38Z
wiki_modified_at: 2026-09-06T23:51:38Z
downloaded_at: 2026-09-07T23:09:25Z
---

# Phi (Microsoft) (ID)

**Phi** — adalah keluarga model bahasa kecil (Small Language Models, SLM) yang dikembangkan oleh Microsoft Research. Model-model ini merepresentasikan pergeseran paradigma dalam pengembangan AI dan menunjukkan bahwa model yang kompak dan efisien secara komputasi dapat mencapai performa yang sebanding dengan sistem yang jauh lebih besar. Berbeda dengan pendekatan tradisional yang didasarkan pada penskalaan jumlah parameter, filosofi Phi berfokus pada kualitas data pelatihan dan metode pelatihan yang inovatif<sup>[\[1\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-1)</sup>.

Model Phi dioptimalkan untuk tugas-tugas yang membutuhkan penalaran logis mendalam, seperti pemrograman, matematika, dan analisis teks. Berkat ukurannya yang kecil, model ini sangat cocok untuk penerapan pada perangkat lokal (on-device AI), termasuk smartphone dan laptop, yang membuka peluang baru bagi demokratisasi AI<sup>[\[2\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-2)</sup>.

## Filosofi: «Buku Teks Adalah Segalanya yang Anda Butuhkan»

Hipotesis sentral yang mendasari proyek Phi adalah bahwa untuk melatih model berkinerja tinggi, kualitas data lebih penting daripada volumenya. Gagasan ini pertama kali dirumuskan dalam makalah penelitian **«Textbooks Are All You Need»**<sup>[\[3\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-3)</sup>. Alih-alih dilatih pada triliunan token dari web yang tidak tersaring, model Phi dilatih pada dataset yang dipilih dengan cermat dan dihasilkan secara sintetis, yang kualitasnya menyerupai buku teks.

Prinsip-prinsip utama pendekatan ini:

- **Data «berkualitas buku teks»:** Korpus pelatihan terdiri dari materi yang bersih, koheren secara logis, dan bersifat penjelasan, terinspirasi dari buku anak-anak.
- **Data sintetis:** Sebagian besar data dihasilkan menggunakan model besar (misalnya, GPT-4). Sebagai contoh, untuk pelatihan **Phi-4** dibuat 400 miliar token konten sintetis berkualitas tinggi melalui lebih dari 50 pipeline pengguna<sup>[\[4\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-4)[\[5\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-5)</sup>.
- **Pelatihan iteratif:** Proses pembuatan data dan pelatihan model berlangsung secara iteratif, yang memungkinkan peningkatan kualitas data maupun model itu sendiri secara berkelanjutan.

Pendekatan ini memungkinkan model Phi mengembangkan kemampuan penalaran yang mendalam, bukan sekadar menghafal pola statistik.

## Evolusi Model Phi

- **Phi-1 (1,3 miliar parameter):** Model pertama yang diperkenalkan pada Juni 2023, difokuskan pada pemrograman dalam bahasa Python. Model ini menunjukkan performa unggul pada benchmark HumanEval dan MBPP, membuktikan efektivitas pendekatan berbasis data berkualitas<sup>[\[6\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-6)</sup>.

<!-- -->

- **Phi-2 (2,7 miliar parameter):** Dirilis pada Desember 2023, Phi-2 memperluas kemampuannya ke pemahaman bahasa umum dengan tetap mempertahankan arsitektur yang kompak. Model ini menunjukkan bahwa SLM dapat mencapai performa yang sebanding dengan model yang puluhan kali lebih besar<sup>[\[7\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-7)</sup>.

<!-- -->

- **Phi-3 (3,8 - 14 miliar parameter):** Keluarga yang diperkenalkan pada April 2024 menjadi terobosan di bidang AI mobile. **Phi-3-mini** (3,8 miliar) mampu berjalan di smartphone, mencapai performa yang sebanding dengan Mixtral 8x7B dan GPT-3.5<sup>[\[8\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-8)</sup>. Keluarga ini juga mencakup versi **Phi-3-small** (7 miliar) dan **Phi-3-medium** (14 miliar).

<!-- -->

- **Phi-3.5 (3,8 - 6,6 miliar parameter aktif):** Diumumkan pada 2024, keluarga ini mencakup tiga model utama:
  - **Phi-3.5-mini-instruct:** Versi yang dioptimalkan dengan dukungan multibahasa yang ditingkatkan.
  - **Phi-3.5-MoE-instruct:** Model berbasis arsitektur Mixture-of-Experts dengan 16 ahli dan 6,6 miliar parameter aktif.
  - **Phi-3.5-Vision-instruct:** Model multimodal untuk pemrosesan teks dan gambar<sup>[\[9\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-9)</sup>.

<!-- -->

- **Phi-4 (14 miliar parameter):** Model yang dispesialisasikan untuk penalaran matematis yang kompleks. Model ini menunjukkan performa yang sebanding dengan Gemini-1.5-Flash dan GPT-4o-mini, dengan ukuran yang jauh lebih kecil. **Phi-4-reasoning** melampaui DeepSeek-R1-Distill-Llama-70B<sup>[\[10\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-10)</sup>.
- **Phi-4-Multimodal (5,6 miliar parameter):** Model multimodal penuh pertama dalam keluarga ini, yang mampu memproses teks, gambar, dan audio secara bersamaan. Model ini menggunakan pendekatan inovatif **Mixture-of-LoRAs** untuk pemrosesan berbagai modalitas secara efisien tanpa interferensi satu sama lain<sup>[\[11\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-11)</sup>.

## Arsitektur dan Karakteristik Teknis

- **Arsitektur:** Model Phi menggunakan arsitektur transformer standar bertipe decoder-only dengan optimasi utama seperti **Grouped Query Attention** dan **Flash Attention** untuk meningkatkan efisiensi<sup>[\[12\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-12)</sup>.
- **Penerapan lokal:** Model dioptimalkan untuk berjalan pada perangkat dengan sumber daya terbatas. Misalnya, **Phi-3-mini** hanya membutuhkan 1,8 GB memori dengan kuantisasi 4-bit dan dapat berjalan di iPhone 14<sup>[\[13\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-13)</sup>.
- **Dukungan framework:** Model Phi tersedia melalui **Microsoft Azure AI Model Catalog**, **Hugging Face**, **Ollama**, dan **NVIDIA NIM microservices**, yang memastikan integrasi luas dan aksesibilitas bagi para pengembang<sup>[\[14\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-14)</sup>.

## Performa dan Benchmark

| Model            | Parameter | MMLU | MT-Bench | HumanEval       |
|------------------|-----------|------|----------|-----------------|
| **Phi-3-mini**   | 3.8B      | 69%  | 8.38     | \-              |
| **Phi-3-small**  | 7B        | 75%  | 8.7      | \-              |
| **Phi-3-medium** | 14B       | 78%  | 8.9      | \-              |
| **Phi-4**        | 14B       | \-   | \-       | Melampaui GPT-4 |

Performa komparatif model Phi pada benchmark utama

**Phi-4** menunjukkan hasil yang luar biasa dalam tugas matematika, termasuk American Mathematics Competitions (AMC), dengan performa yang sebanding dengan Gemini-1.5-Flash<sup>[\[15\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-15)</sup>. **Phi-3.5-Vision** multimodal melampaui pesaing berukuran serupa, mencapai 57,0% pada benchmark BLINK<sup>[\[16\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-16)</sup>.

## Aplikasi Khusus

Model Phi menunjukkan efisiensi tinggi di bidang-bidang khusus:

- **Kedokteran:** Penelitian menunjukkan korelasi moderat antara jawaban Phi-3 dengan penilaian para ahli dalam teks medis dan olahraga<sup>[\[17\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-17)</sup>.
- **Deteksi ujaran kebencian:** Model HateTinyLLM berbasis Phi-2 mencapai akurasi lebih dari 80% dalam tugas ini menggunakan LoRA fine-tuning<sup>[\[18\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-18)</sup>.
- **Strategi permainan:** Model SC-Phi2 menunjukkan kemampuan dalam memprediksi strategi dalam permainan StarCraft II<sup>[\[19\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-19)</sup>.

## AI yang Bertanggung Jawab dan Keamanan

Keluarga Phi dikembangkan sesuai dengan standar **Microsoft Responsible AI**, yang mencakup prinsip-prinsip akuntabilitas, transparansi, keadilan, dan keamanan. Model-model ini menjalani evaluasi keamanan multidimensi, termasuk **Supervised Fine-Tuning (SFT)** dan **Direct Preference Optimization (DPO)**, serta pengujian dalam berbagai bahasa dan kategori risiko<sup>[\[20\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-20)</sup>.

## Keterbatasan

Meskipun menunjukkan hasil yang mengesankan, model Phi dapat kalah dari model besar yang terspesialisasi dalam beberapa tugas kompleks. Misalnya, **Phi-4** menunjukkan hasil yang baik dalam penalaran bertipe chain-of-thought, tetapi dibatasi oleh ketiadaan kemampuan pemanggilan fungsi (function calling)<sup>[\[21\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-21)</sup>. Selain itu, meskipun **Phi-3.5** mendukung lebih dari 20 bahasa, performanya dapat bervariasi, dan penelitian menunjukkan adanya ketidakakuratan dalam jawaban pada bahasa selain bahasa Inggris<sup>[\[22\]](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_note-22)</sup>.

## Daftar Pustaka

- Gunasekar, S.; et al. (2023). *Textbooks Are All You Need*. arXiv:2306.11644.
- Gunasekar, S.; et al. (2023). *Textbooks Are All You Need II: phi‑1.5 Technical Report*. arXiv:2309.05463.
- Dao, T.; et al. (2022). *FlashAttention: Fast and Memory‑Efficient Exact Attention with IO‑Awareness*. arXiv:2205.14135.
- Zheng, S.; et al. (2023). *GQA: Training Generalized Multi‑Query Transformer Models for Faster Decoding*. arXiv:2305.13245.
- Feng, W.; et al. (2024). *Mixture‑of‑LoRAs: An Efficient Multitask Tuning for Large Language Models*. arXiv:2403.03432.
- Wu, X.; et al. (2024). *Mixture of LoRA Experts*. arXiv:2404.13628.
- Microsoft Research (2024). *Phi‑3 Technical Report*. arXiv:2404.14219.
- Abdin, M.; et al. (2024). *Phi‑4 Technical Report*. arXiv:2412.08905.
- Microsoft Research (2025). *Phi‑4‑reasoning Technical Report*. PDF.
- Microsoft Research (2025). *Phi‑4‑Multimodal: Mixture‑of‑Modality‑LoRAs*. arXiv:2503.01743.

## Catatan

1.  <span id="cite_note-1">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-1) «The Phi-3 small language models with big potential». *Microsoft Source Features*. <a href="https://news.microsoft.com/source/features/ai/the-phi-3-small-language-models-with-big-potential/" class="external autonumber" rel="nofollow">[1]</a></span>
2.  <span id="cite_note-2">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-2) «Microsoft's Phi-3: Revolutionising AI with efficient and accessible small language models». *Landing.Jobs Blog*. <a href="https://landing.jobs/blog/microsofts-phi-3-revolutionising-ai-with-efficient-and-accessible-small-language-models/" class="external autonumber" rel="nofollow">[2]</a></span>
3.  <span id="cite_note-3">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-3) «Textbooks Are All You Need». *Microsoft Research*. <a href="https://www.microsoft.com/en-us/research/publication/textbooks-are-all-you-need/" class="external autonumber" rel="nofollow">[3]</a></span>
4.  <span id="cite_note-4">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-4) «Introducing Phi-4: Microsoft’s newest Small Language Model, specializing in complex reasoning». *Microsoft Tech Community*. <a href="https://techcommunity.microsoft.com/t5/ai-at-microsoft/introducing-phi-4-microsoft-s-newest-small-language-model/ba-p/4357090" class="external autonumber" rel="nofollow">[4]</a></span>
5.  <span id="cite_note-5">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-5) «Exploring Phi-4: A Deep Dive into Microsoft's Latest Language Model». *OpenCV Blog*. <a href="https://opencv.org/blog/phi-4/" class="external autonumber" rel="nofollow">[5]</a></span>
6.  <span id="cite_note-6">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-6) «Unlocking the Power of Small Language Models (SLMs): The Evolution of Phi». *LinkedIn*. <a href="https://www.linkedin.com/pulse/unlocking-power-small-language-models-slms-evolution-vijay-ru4xc" class="external autonumber" rel="nofollow">[6]</a></span>
7.  <span id="cite_note-7">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-7) «Новая ИИ-модель Phi-2 от Microsoft училась по учебникам». *TechInsider*. <a href="https://www.techinsider.ru/news/news-1625165-novaya-ii-model-phi-2-ot-microsoft-uchilas-po-uchebnikam/" class="external autonumber" rel="nofollow">[7]</a></span>
8.  <span id="cite_note-8">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-8) «Phi-3 Technical Report». *arXiv*. <a href="https://arxiv.org/abs/2404.14219" class="external autonumber" rel="nofollow">[8]</a></span>
9.  <span id="cite_note-9">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-9) «Discover the new multi-lingual, high-quality Phi-3.5 SLMs». *Microsoft Tech Community*. <a href="https://techcommunity.microsoft.com/blog/azure-ai-services-blog/discover-the-new-multi-lingual-high-quality-phi-3-5-slms/4225280" class="external autonumber" rel="nofollow">[9]</a></span>
10. <span id="cite_note-10">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-10) «Phi-4 Technical Report». *arXiv*. <a href="https://arxiv.org/abs/2412.08905" class="external autonumber" rel="nofollow">[10]</a></span>
11. <span id="cite_note-11">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-11) «Mixture-of-Modality-LoRAs: A Low-Rank Approach to Natively Multimodal Foundation Models». *arXiv*. <a href="https://arxiv.org/abs/2503.01743" class="external autonumber" rel="nofollow">[11]</a></span>
12. <span id="cite_note-12">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-12) «Phi-3: A Tutorial on Microsoft's Small Language Models (SLMs)». *DataCamp*. <a href="https://www.datacamp.com/tutorial/phi-3-tutorial" class="external autonumber" rel="nofollow">[12]</a></span>
13. <span id="cite_note-13">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-13) «Unlocking the Power of Small Language Models (SLMs): The Evolution of Phi». *LinkedIn*. <a href="https://www.linkedin.com/pulse/unlocking-power-small-language-models-slms-evolution-vijay-ru4xc" class="external autonumber" rel="nofollow">[13]</a></span>
14. <span id="cite_note-14">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-14) «Microsoft Phi». *Microsoft Azure*. <a href="https://azure.microsoft.com/en-us/products/phi" class="external autonumber" rel="nofollow">[14]</a></span>
15. <span id="cite_note-15">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-15) «Exploring Phi-4: A Deep Dive into Microsoft's Latest Language Model». *OpenCV Blog*. <a href="https://opencv.org/blog/phi-4/" class="external autonumber" rel="nofollow">[15]</a></span>
16. <span id="cite_note-16">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-16) «Phi-3.5-vision-instruct». *Hugging Face*. <a href="https://huggingface.co/microsoft/Phi-3.5-vision-instruct" class="external autonumber" rel="nofollow">[16]</a></span>
17. <span id="cite_note-17">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-17) «Small But Mighty: Exploring the Capabilities of Small Language Models in Medical and Sport-Specific Applications». *arXiv*. <a href="https://arxiv.org/abs/2504.08764" class="external autonumber" rel="nofollow">[17]</a></span>
18. <span id="cite_note-18">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-18) «HateTinyLLMs: A Small Language Model for Hate Speech Detection». *arXiv*. <a href="https://arxiv.org/abs/2405.01577" class="external autonumber" rel="nofollow">[18]</a></span>
19. <span id="cite_note-19">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-19) «SC-Phi2: A Specialized Small Language Model for StarCraft II». *MDPI*. <a href="https://www.mdpi.com/2673-2688/5/4/115" class="external autonumber" rel="nofollow">[19]</a></span>
20. <span id="cite_note-20">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-20) «Microsoft’s Phi-3.5: a responsible, small language model». *Skymod*. <a href="https://skymod.tech/microsofts-phi-3-5-small-language-models/" class="external autonumber" rel="nofollow">[20]</a></span>
21. <span id="cite_note-21">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-21) «Phi-4: A New Era of Small Language Models». *Meta-quantum.today*. <a href="https://meta-quantum.today/?p=3386" class="external autonumber" rel="nofollow">[21]</a></span>
22. <span id="cite_note-22">[↑](https://systems-analysis.info/int/Phi_(Microsoft)_(ID)#cite_ref-22) «A Multi-faceted Analysis of Language-specific Bias in Large Language Models». *U.S. Securities and Exchange Commission*. <a href="https://www.sec.gov/Archives/edgar/data/789019/000119312524242883/d858775ddef14a.htm" class="external autonumber" rel="nofollow">[22]</a></span>
