Small language model from Microsoft supporting a context of up to 128k tokens, with a 4k version as well. It has been retired and Phi-4-mini-instruct succeeds it.
Also called Phi-3-mini-128k-instruct.
Read more: Microsoft Learn
In the Ultra Transcenders books
Each book explains Phi-3-mini in context, with comparison tables and the common traps.
Terms in this definition
- SLM
Smaller than an LLM, with roughly under 10 billion parameters but a similar kind of architecture. Running it costs less and is quicker, at the price of narrower capability.
Related terms
- Context window
How many tokens, counting input and output together, a model can handle in one request; GPT-4.1-mini manages roughly 1M, for instance, against 128k for the now-retired Phi-3-mini.