Microsoft's efficient 3.8B parameter language model — state-of-the-art performance in a compact package, designed for on-device and resource-constrained scenarios.
Runs efficiently on CPU, GPU, and even mobile devices. Optimized for bf16 with just ~7.6 GB memory footprint.
Competitive with models 2-3x its size on benchmarks like MMLU, HellaSwag, and GSM8K.
Full fine-tuning and LoRA/QLoRA support. Compatible with Hugging Face Transformers and TRL.
Trained on multilingual data with strong performance across English, French, German, Spanish, and more.
Instruction-tuned variant with a custom chat template for conversational use cases.
Fully open-source with a permissive MIT license — use it freely in commercial and research projects.
Try Phi-3 Mini right now — powered by Hugging Face Chat with serverless inference.
Open HF ChatWrite functions, debug code, explain complex algorithms, and generate programming snippets.
Draft emails, blog posts, social media content, and creative writing with contextual understanding.
Explain concepts, answer questions, and serve as a tutoring assistant across various subjects.
Analyze data, generate reports, summarize documents, and extract insights from text.