Phi-4

Microsoft small reasoning model. Runs locally on 8 GB VRAM via Ollama.

Type: local

Context: 16,384 tokens

Use cases: coding, reasoning

Run in Ollama

Limits checked against official docs on 2026-08-23. Rates change without notice. Not affiliated with providers unless noted. Verify before production use.