MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 3241-3260 of 8697

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
TIER 2

👍0
Hola quiero saber porque se dreba tanto la batería de mi poc...

👍0
fix this code:
# employee_data is a list of dicts: each ha...

👍0
Can we build a python script to convert PDF charts into exce...

👍0
Solve this problem step-by-step. Show your work for each ste...

👍0
Roll out spend-meter.sh to 3 hosts via rsync; one host has a...

👍0
content creation prompt

👍0
ai chat app Send back the complete code with all the fixes. ...

👍0
testing
just testing something

👍0
сделай мне презентацию по следующей информации
Результативн...

👍0
Tóm tắt chính xác tài liệu RP2A03 / CPU NES, Cycle Reference...
👍0
test

👍0
Game modding deep research challenge

👍0
Design an AI Chip

👍0
Make a roadmap with the best resources to learn (Philosophy ...
👍0
I want to redesign the UI/UX of my existing Telegram bot Adm...

👍0
Reinforcement Learning Agent Eval
A set of 10 RL questions spanning topics like RL agent training in Minecraft, Rocket League, etc.
👍0
You are a retail general manager at a bridal store. You need...

👍0
https://artificialanalysis.ai/evaluations/artificial-analysi...