Grok Build vs. Muse Code
Vergleich von Grok Build und Muse Code im Artificial Analysis Coding Agent Index, einschließlich Benchmark-Werten, Kosten, Ausführungszeit und Tokenverbrauch.
Einzelheiten zu unserer Methodik finden Sie auf unserer Methodikseite.
Weitere Vergleiche entdecken
vs
Wichtigste Ergebnisse
Vergleich
Direkter Vergleich von Grok Build und Muse Code.
Programmieragenten-Vergleich
Metrik | Analyse | ||
|---|---|---|---|
Agent-Harness | Grok Build | Muse Code | |
Repräsentatives Modell | Grok 4.7 (xhigh) | Muse Spark 1.3 (max) | |
Coding Agent Index | 56 | 54 | Grok Build hat einen höheren Coding Agent Index als Muse Code |
DeepSWE v1.1 | 73% | 72% | Grok Build hat einen höheren DeepSWE v1.1-Wert als Muse Code |
Terminal-Bench 4.0 | 33% | 32% | Grok Build hat einen höheren Terminal-Bench 4.0-Wert als Muse Code |
SWE-Atlas-QnA | 63% | 59% | Grok Build hat einen höheren SWE-Atlas-QnA-Wert als Muse Code |
Kosten pro Aufgabe | $8.82 | $3.98 | Muse Code hat geringere Kosten pro Aufgabe als Grok Build |
Zeit pro Aufgabe | 39.2m | 18.4m | Muse Code hat eine geringere Zeit pro Aufgabe als Grok Build |
Runden pro Aufgabe | 162.6 | 185.2 | Grok Build hat weniger Runden pro Aufgabe als Muse Code |
Token-Nutzung pro Aufgabe | 14.3M | 16.7M | Grok Build hat eine geringere Token-Nutzung pro Aufgabe als Muse Code |
Cache-Trefferquote | 94% | 97% | Muse Code hat eine höhere Cache-Trefferquote als Grok Build |
Modellvarianten
Bewertete Modellvarianten für Grok Build und Muse Code.
Modellvarianten
56 | 73% | 33% | 63% | $8.82 | 39.2m | 14.3M | ||
47 | 65% | 18% | 58% | $3.57 | 19.5m | 5.5M | ||
54 | 72% | 32% | 59% | $3.98 | 18.4m | 16.7M | ||
48 | 73% | 17% | 55% | $3.47 | 18.3m | 16.2M |
Leistung
Leistung im Artificial Analysis Coding Agent Index.
Artificial Analysis Coding Agent Index
Artificial Analysis Coding Agent Index v1.5 incorporates 3 benchmarks: DeepSWE v1.1, Terminal-Bench 4.0, and SWE-Atlas-QnA · Higher is better
Color by
Artificial Analysis Coding Agent Index vs. Kosten pro Aufgabe
Artificial Analysis Coding Agent Index vs. durchschnittliche Pay-per-Token-API-Kosten pro Aufgabe (USD)
Color by
Most attractive quadrant
Pareto line
Token-Nutzung
Token-Verbrauch im Artificial Analysis Coding Agent Index.
Token-Nutzung pro Aufgabe
Durchschnittliche Eingabe-, Cache- und Ausgabe-Tokens pro Aufgabe
Prompt cache hit rates can vary significantly by provider routing, which can materially change effective cost.
Artificial Analysis Coding Agent Index vs. Gesamt-Tokens
Artificial Analysis Coding Agent Index vs. durchschnittliche Gesamt-Tokens pro Aufgabe
Color by
Most attractive quadrant
Kosten
Pay-per-Token-API-Kosten im Artificial Analysis Coding Agent Index, basierend auf den aktuellen Preisen pro Token.
Kosten pro Aufgabe
Average pay-per-token API cost per task (USD) · Lower is better
Color by
Ausführungszeit
Aktive Agent-Laufzeit im Artificial Analysis Coding Agent Index.
Zeit pro Aufgabe
Average agent wall time per task · Lower is better
Color by
Artificial Analysis Coding Agent Index vs. Ausführungszeit
Artificial Analysis Coding Agent Index vs. durchschnittliche Agent-Laufzeit pro Aufgabe
Color by
Most attractive quadrant