Model comparison

Grok Build 0.1 vs Phi-3.5-MoE

Grok Build 0.1 has enough public results to be ranked (#216); Phi-3.5-MoE does not yet, so treat this comparison as directional.

Last verified . 0 shared benchmarks.

Grok Build 0.1 xAI

36.4

Rank #216 Reported

Phi-3.5-MoE Microsoft

—

Unranked

Summary

  • Phi-3.5-MoE has downloadable open weights; the other is API-only.

Side by side

Grok Build 0.1 and Phi-3.5-MoE specifications
Grok Build 0.1Phi-3.5-MoE
ProviderxAIMicrosoft
Noometry Index36.4—
Released2026-04-162024-08-17
WeightsProprietaryOpen
Context window256K—
Max output256K—
Input $ / M tokens$1—
Output $ / M tokens$2—
Results tracked33

Sponsored placements are available on pages like this one. Advertise on Noometry

Category by category

Coding Not comparable

Grok Build 0.1: 43.1 (#91), Phi-3.5-MoE: —

Coding benchmarks
BenchmarkGrok Build 0.1Phi-3.5-MoE
SciCode50.2%—

Agentic & Tool Use Not comparable

Grok Build 0.1: 22.7 (#129), Phi-3.5-MoE: —

Agentic & Tool Use benchmarks
BenchmarkGrok Build 0.1Phi-3.5-MoE
GBAEval2.4%—

Reasoning Not comparable

Grok Build 0.1: 32.2 (#77), Phi-3.5-MoE: —

Reasoning benchmarks
BenchmarkGrok Build 0.1Phi-3.5-MoE
CritPt9.1%—
PIQA—88.6%

Math Not comparable

Grok Build 0.1: —, Phi-3.5-MoE: —

Math benchmarks
BenchmarkGrok Build 0.1Phi-3.5-MoE
GSM8K—88.7%

Knowledge Not comparable

Grok Build 0.1: —, Phi-3.5-MoE: —

Knowledge benchmarks
BenchmarkGrok Build 0.1Phi-3.5-MoE
BoolQ—84.6%

Frequently asked questions

Is Grok Build 0.1 better than Phi-3.5-MoE?

Grok Build 0.1 has enough public results to be ranked (#216); Phi-3.5-MoE does not yet, so treat this comparison as directional.

How many benchmarks do Grok Build 0.1 and Phi-3.5-MoE share?

0 benchmarks have published results for both models. Grok Build 0.1 has 3 scored results on Noometry and Phi-3.5-MoE has 3.

Related comparisons

Go deeper