MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 8481-8500 of 8776

👍0
Súper mario

👍0
You are an administrative operations lead in a government de...

👍0
create a vido of a dog eating bhelpuri

👍0
Kırkkilit bitkisini çayının günün hangi saatinde tüketmek da...

👍0
Tìm nền tảng tương tự arena.ai và zenmux.ai

👍0
deliver profound upgrades to enhance novelty, usefulness to ...

👍0
saq
👍0
Finance

👍0
You are a creative coder and an expert in computer graphics,...

👍0
模拟我大师赛

👍0
You are a Senior Business Development Strategist with expert...
👍0
111

👍0
You are the administrative services manager responsible for ...

👍0
Как да си отворим фирма ЕООД

👍0
In 2015, he said, a study completed in cooperation with the ...

👍0
Выступи в роли прагматичного, жесткого карьерного стратега и...

👍0
OscarGolf-3529

👍0
Generate a fully synthetic, production-ready pre-training da...

👍0
In 2015, he said, a study completed in cooperation with the ...

👍0
Runner system-prompt efficiency — improvement task
How good models are at improving existing systems.