MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2201-2220 of 2529

👍0
The cyclic subgroup of Z_24 generated by 18 has order
A) 4
...

👍0
Necip Fazıl Kısakürek'in çile adlı eserinden esinlen erek g...

👍0
Ответ Grok 4.3 (high):
1. Совершенствование системы управлен...

👍0
Just another Local VS Cloud

👍0
Fındıkta bulunan Fito kimyasallar sağlığa faydaları ve mikta...

👍0
==================================================
QISSAH — ...

👍0
Arithmetic
What complexity of arithmetic is it safe to trust LLMs to do without a code interpreter? The purpose of this eval is to understand what level is completely safe, and what level you should instruct and LLM to use a

👍0
sprites

👍0
Research, think, plan, and let me know if it's possible to b...

👍0
Ferry

👍0
zgdocs

👍0
Best coder

👍0
You are a professional experienced genius intelligent creati...

👍0
What's the most liked video featuring or starring a person w...

👍0
Сделай объёмный список наиболее древнейших письменных источн...

👍0
Render a rotating kebab in JavaScript.

👍0
pygame test
a good test to see how coding works

👍0
Create a pro level complex mobile friendly web based profess...

👍0
hi whats news

👍0
Сам придумай и создай какой нибудь сложный и интересный вебс...