Phi-4
Microsoft small reasoning model. Runs locally on 8 GB VRAM via Ollama.
Type: local
Context: 16,384 tokens
Use cases: coding, reasoning
Limits checked against official docs on 2026-08-23. Rates change without notice. Not affiliated with providers unless noted. Verify before production use.