Grok Build vs. Muse Code
Vergleich von Grok Build und Muse Code im Artificial Analysis Coding Agent Index, einschließlich Benchmark-Werten, Kosten, Ausführungszeit und Tokenverbrauch.
Einzelheiten zu unserer Methodik finden Sie auf unserer Methodikseite.
Weitere Vergleiche entdecken
vs
Wichtigste Ergebnisse
Artificial Analysis Coding Agent Index v1.4 · Higher is better
Not publicly available
Vergleich
Direkter Vergleich von Grok Build und Muse Code.
Programmieragenten-Vergleich
Metrik | Analyse | ||
|---|---|---|---|
Agent-Harness | Grok Build | Muse Code | |
Repräsentatives Modell | Grok 4.5 (high) | Muse Spark 1.3 (max) | |
Coding Agent Index | 64 | 68 | Muse Code hat einen höheren Coding Agent Index als Grok Build |
DeepSWE | 60% | 68% | Muse Code hat einen höheren DeepSWE-Wert als Grok Build |
Terminal-Bench v2.1 | 84% | 84% | Grok Build hat einen höheren Terminal-Bench v2.1-Wert als Muse Code |
SWE-Atlas-QnA | 48% | 52% | Muse Code hat einen höheren SWE-Atlas-QnA-Wert als Grok Build |
Kosten pro Aufgabe | $2.44 | $0.00 | Vergleich nicht verfügbar |
Zeit pro Aufgabe | 15.5m | 24.6m | Grok Build hat eine geringere Zeit pro Aufgabe als Muse Code |
Runden pro Aufgabe | 60.7 | 128.5 | Grok Build hat weniger Runden pro Aufgabe als Muse Code |
Token-Nutzung pro Aufgabe | 3.6M | 14.2M | Grok Build hat eine geringere Token-Nutzung pro Aufgabe als Muse Code |
Cache-Trefferquote | 92% | 95% | Muse Code hat eine höhere Cache-Trefferquote als Grok Build |
Modellvarianten
Bewertete Modellvarianten für Grok Build und Muse Code.
Modellvarianten
64 | 60% | 84% | 48% | $2.44 | 15.5m | 3.6M | ||
68 | 68% | 84% | 52% | $0.00 | 24.6m | 14.2M | ||
64 | 67% | 82% | 44% | $1.72 | 12.8m | 14.8M | ||
62 | 58% | 82% | 45% | $2.07 | 40.8m | 20M |
Leistung
Leistung im Artificial Analysis Coding Agent Index.
Artificial Analysis Coding Agent Index
Artificial Analysis Coding Agent Index v1.4 incorporates 3 benchmarks: DeepSWE, Terminal-Bench v2.1, and SWE-Atlas-QnA · Higher is better
Not publicly available
Since benchmarking, we have observed a higher rate of content safety filtering on this endpoint.
Token-Nutzung
Token-Verbrauch im Artificial Analysis Coding Agent Index.
Token-Nutzung pro Aufgabe
Durchschnittliche Eingabe-, Cache- und Ausgabe-Tokens pro Aufgabe
Prompt cache hit rates can vary significantly by provider routing, which can materially change effective cost.
Artificial Analysis Coding Agent Index vs. Gesamt-Tokens
Artificial Analysis Coding Agent Index vs. durchschnittliche Gesamt-Tokens pro Aufgabe
Most attractive quadrant
Kosten
Pay-per-Token-API-Kosten im Artificial Analysis Coding Agent Index, basierend auf den aktuellen Preisen pro Token.
Kosten pro Aufgabe
Average pay-per-token API cost per task (USD) · Lower is better
Artificial Analysis Coding Agent Index vs. Kosten pro Aufgabe
Artificial Analysis Coding Agent Index vs. durchschnittliche Pay-per-Token-API-Kosten pro Aufgabe (USD)
Most attractive quadrant
Ausführungszeit
Aktive Agent-Laufzeit im Artificial Analysis Coding Agent Index.
Zeit pro Aufgabe
Average agent wall time per task · Lower is better
Not publicly available
Artificial Analysis Coding Agent Index vs. Ausführungszeit
Artificial Analysis Coding Agent Index vs. durchschnittliche Agent-Laufzeit pro Aufgabe
Most attractive quadrant