Data points: 57
Model compare
The readout for GPT-OSS and Phi 4 Multimodal Instruct, before the detailed comparison sheet.
Weighted outcome: GPT-OSS. Benchmark capability categories carry 80%, while price, API performance, and availability carry 20%.
Decision read
GPT-OSS
GPT-OSS has the higher weighted result; Model A / B score 80 to 20.
Evidence depth
57 data points
Includes 0 benchmark rows, 0 audit samples, and 8 provider examples.
Selection signal
Start with GPT-OSS
The charts below split 8 high-signal samples across speed, scores, and audit health.
Switch either side of this report to compare another model with the same LMSpeed data pipeline.
Select a different model to open a new comparison URL.
This report only uses LMSpeed data for GPT-OSS and Phi 4 Multimodal Instruct: pricing, speed aggregates, third-party benchmark scores, and shared provider samples.
| Model compare | GPT-OSS | Phi 4 Multimodal Instruct |
|---|---|---|
| Overall leader | Leading | Contender |
| Weighted overall score | 80.0 pts | 20.0 pts |
| Benchmark category leads | 0 categories | 0 categories |
| Operational advantages | Cheapest input price, Average speed, Free providers, Provider coverage | First-token latency |
| Context window | 131.1K tokens | No data |
| Max output | 131.1K tokens | No data |
| Modalities | Input Text Output Text | No data |
| Features |
The overall result weights benchmark capability categories at 80% and price, API speed/latency, and availability at 20%. Recent test volume does not affect the winner, and missing benchmark categories are excluded.
| Model compare | GPT-OSS | Phi 4 Multimodal Instruct |
|---|---|---|
| Developer | No data | No data |
| Released | Aug 2025 | No data |
| Parameters | 120B | No data |
| Tokenizer | GPT | No data |
| Knowledge cutoff | 2024-06-30 | No data |
| OpenRouter ID | openai/gpt-oss-120b:free | No data |
| References | No data | No data |
This report only uses LMSpeed data for GPT-OSS and Phi 4 Multimodal Instruct: pricing, speed aggregates, third-party benchmark scores, and shared provider samples.
GPT-OSS
GPT-OSS has these operational advantages: Cheapest input price, Average speed, Free providers, Provider coverage.
Phi 4 Multimodal Instruct
Phi 4 Multimodal Instruct has these operational advantages: First-token latency.
Third-party benchmark profile synced into LMSpeed; only metrics available for both models are shown.
Compare benchmark category scores on a 0-100 scale. Select a category to inspect the gap.
Avg. score
GPT-OSS
-
Avg. score
Phi 4 Multimodal Instruct
35.9
No data
Phi 4 Multimodal Instruct
Phi 4 Multimodal Instruct
Metric-level scores with benchmark source, rank depth, confidence, error, and evaluation date where available.
No shared professional benchmark scores are available yet.
Latest completed audits from shared providers, with four safety and integrity score groups plus report links.
| Provider | GPT-OSS | Phi 4 Multimodal Instruct |
|---|---|---|
| No completed audits are available from shared providers yet. | ||
Speed aggregates and input/output pricing share each provider row for real API selection and migration cost checks.
| Provider | GPT-OSS | Phi 4 Multimodal Instruct |
|---|---|---|
120 tests | GPT-OSS speed / latency 153 tok/s / 1011ms input / output No data | Phi 4 Multimodal Instruct speed / latency 83 tok/s / 393ms input / output No data |
35 tests | GPT-OSS speed / latency 1061 tok/s / 857ms input / output No data | Phi 4 Multimodal Instruct speed / latency N/A / N/A input / output No data |
35 tests | GPT-OSS speed / latency 347 tok/s / 2670ms input / output No data | Phi 4 Multimodal Instruct speed / latency N/A / N/A input / output No data |
20 tests | GPT-OSS speed / latency 164 tok/s / 1247ms input / output No data | Phi 4 Multimodal Instruct speed / latency N/A / N/A input / output No data |
10 tests | GPT-OSS speed / latency 473 tok/s / 827ms input / output No data | Phi 4 Multimodal Instruct speed / latency N/A / N/A input / output No data |
GPT-OSS gpt-oss-120b speed / latency No data input / output $0.150/M/$0.600/M | Phi 4 Multimodal Instruct microsoft/phi-4-multimodal-instruct speed / latency No data input / output $0/M | |
GPT-OSS gpt-oss-120b speed / latency No data input / output $0.150/M/$0.600/M | Phi 4 Multimodal Instruct microsoft/phi-4-multimodal-instruct speed / latency No data input / output $0/M | |
GPT-OSS openai/gpt-oss-20b speed / latency No data input / output $0.290/M/$1.40/M | Phi 4 Multimodal Instruct microsoft/phi-4-multimodal-instruct speed / latency No data input / output $0/M |
Weighted outcome: GPT-OSS. Benchmark capability categories carry 80%, while price, API performance, and availability carry 20%.
Text inputText outputTool callingReasoning |
| None listed |
No data
Phi 4 Multimodal Instruct
No data
No data
No data