MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 5841-5860 of 8611

👍0
recipe descramble

👍0
compare between claude, chat gpt, gemini
for software QA usa...
👍0
deadliest insects
list of deadliest insects in the world

👍0
Selam

👍0
Test

👍0
You are being consulted as a world-class expert to stress-te...

👍0
I'm scoring several AI models on one fixed battery.
Work wit...

👍0
enhance this prompt into .md plan
code me an EpicBot local ...
👍0
TeaVal

👍0
test

👍0
Is P=NP?

👍0
deliver profound, useful improvements in novelty and impact:...

👍0
Simple test

👍0
Absolutely. If we're building this, let's build something yo...

👍0
You are an administrative operations lead in a government de...

👍0
Delegation

👍0
how are you
👍0
You are an administrative operations lead in a government de...

👍0
You are an administrative operations lead in a government de...

👍0
Qwen v Gemma logic battle
testing Qwen and Gemma side by side