MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 5761-5780 of 6242

👍0
Bir sepette 100 yumurta 6 çıktı kaç kaldı bu bir bilmecedir ...

👍0
یک گیم داییناسور گوگل

👍0
AI Agent Expressif

👍0
what is the best free ai for research

👍0
You are an administrative operations lead in a government de...

👍0
Jsi expertní vědecký, medicínský a Evidence-Based Medicine a...

👍0
Make a roadmap with the best resources to learn PHILOSOPHY o...

👍0
Project Blackbox

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
botanicula

👍0
I am planning a trip to Hangzhou province, and I would like ...

👍0
# CLAUDE.md
Behavioral guidelines to reduce common LLM codi...

👍0
Berätta om barnboken "Pippi och spöket på Jönköpings central...

👍0
Code Eval

👍0
What's initContainer in kubernetes

👍0
Minecraftを完全に再現

👍0
teste
teste

👍0
Gym business operation
To compare how different models approach a general business operation/analysis problem and how well it responds in Chinese

👍0
i wish to see cost over time for gpt-5-series

👍0
teljes szinasztriat vagy mi az erted az osszes orb egyutt al...