MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 5741-5760 of 6217

👍0
what is the best free ai for research

👍0
You are an administrative operations lead in a government de...

👍0
Jsi expertní vědecký, medicínský a Evidence-Based Medicine a...

👍0
Make a roadmap with the best resources to learn PHILOSOPHY o...

👍0
Project Blackbox

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
botanicula

👍0
I am planning a trip to Hangzhou province, and I would like ...

👍0
# CLAUDE.md
Behavioral guidelines to reduce common LLM codi...

👍0
Berätta om barnboken "Pippi och spöket på Jönköpings central...

👍0
Code Eval

👍0
What's initContainer in kubernetes

👍0
Minecraftを完全に再現

👍0
teste
teste

👍0
Gym business operation
To compare how different models approach a general business operation/analysis problem and how well it responds in Chinese

👍0
i wish to see cost over time for gpt-5-series

👍0
teljes szinasztriat vagy mi az erted az osszes orb egyutt al...

👍0
What's 9.9-9.11?

👍0
Testing

👍0
test
test