MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 4681-4700 of 6443

👍0
Research the history of Sardis excavations comprehensively

👍0
hi how r u

👍0
Regarding your question, Bain and McKinsey do not really do ...

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
你好你好你好

👍0
Create code for 3d world visually stunning of a indian city

👍0
You are an administrative operations lead in a government de...

👍0
CS Higher education and the Ne Professor
CS and Ne

👍0
تعرف ايه عن Dify.ai

👍0
compare the token sizes of Claude Opus 5 and GPT-5.6 Sol

👍0
game
game

👍0
Research EXTENSIVELY and reason hard on the following statem...

👍0
Ancient India, Lumbini at sunrise. Queen Maya lovingly holdi...

👍0
name every song in linkin park's hybrid theory tracklist

👍0
Write the full strategy criteria and checklist for a low-vol...

👍0
tier2 test

👍0
test_amacli

👍0
1. detect_language(text, allowed=()) throws a runtime IndexE...

👍0
You are an expert in computational quantum mechanics and sci...

👍0
Выбираю LLM для код агента чтобы ее захостить на H100 (польз...