MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2601-2620 of 3008

👍0
Statement 1 | A permutation that is a product of m even perm...

👍0
https://ziglang.org/documentation/0.16.0/ ezt fogok le crawo...

👍0
Cyclic subgroup of Z_24 generated by 18 has order:

👍0
Sales Headline

👍0
Create a p5.js animation that is a cool interactive thing ba...

👍0
سایت

👍0
Consulting Slides

👍0
Poem

👍0
Actúa como un Ingeniero de Software Senior experto en optimi...

👍0
The cyclic subgroup of Z_24 generated by 18 has order
A) 4
...

👍0
Necip Fazıl Kısakürek'in çile adlı eserinden esinlen erek g...

👍0
Ответ Grok 4.3 (high):
1. Совершенствование системы управлен...

👍0
Just another Local VS Cloud

👍0
Fındıkta bulunan Fito kimyasallar sağlığa faydaları ve mikta...

👍0
==================================================
QISSAH — ...

👍0
Arithmetic
What complexity of arithmetic is it safe to trust LLMs to do without a code interpreter? The purpose of this eval is to understand what level is completely safe, and what level you should instruct and LLM to use a

👍0
sprites

👍0
I want to write an article on dopamine and that how cheap do...

👍0
Research, think, plan, and let me know if it's possible to b...

👍0
Ferry