MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 4681-4700 of 9736
👍0
Creativity test
Let's see how creative can they get with a simple prompt

👍0
Biskupnica
👍0
Estimate the 85th percentile household-networth of the follo...

👍0
5.5pro

👍0
Where does your favourite LLM rank currently?
👍0
Random
Random

👍0
Дарова
👍0
workly
finding the worker

👍0
You are an administrative operations lead in a government de...

👍0
Compare all reasoning effort levels in GPT-5.6 Luna.

👍0
Viral LLM Coding Challenges
👍0
hej
👍0
resume

👍0
Основываясь на предоставленных данных из файлов, я провел де...

👍0
Aldrich Ocampo Nuevo aldrichnuevo369@gmail.com

👍0
Estimate the level of recoverable easily extractable commerc...

👍0
ki trading
strict prompt following

👍0
A robe takes 2 bolts of blue fiber and half that much white ...

👍0
P22.5 — INTERNAL SHADOW / WEB RADAR / INTER-RESULT WORK / AG...

👍0
egy nema refluxosnak lehet halalos vagy extra veszelyes egy ...