MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2501-2520 of 6712

👍0
make recipe for pizza

👍0
C# Master
Create a better C# master if possible

👍0
Bewerbung

👍0
234
234

👍0
Trem

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
TIER 2

👍0
Hola quiero saber porque se dreba tanto la batería de mi poc...

👍0
fix this code:
# employee_data is a list of dicts: each ha...

👍0
Can we build a python script to convert PDF charts into exce...

👍0
Solve this problem step-by-step. Show your work for each ste...

👍0
content creation prompt

👍0
ai chat app Send back the complete code with all the fixes. ...

👍0
testing
just testing something

👍0
сделай мне презентацию по следующей информации
Результативн...

👍0
Tóm tắt chính xác tài liệu RP2A03 / CPU NES, Cycle Reference...

👍0
Game modding deep research challenge

👍0
Design an AI Chip

👍0
Make a roadmap with the best resources to learn (Philosophy ...

👍0
Reinforcement Learning Agent Eval
A set of 10 RL questions spanning topics like RL agent training in Minecraft, Rocket League, etc.