MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2041-2060 of 2554

👍0
Write a psychological horror story about a man and a woman w...

👍0
title

👍0
You need to answer only one word. Think thoroughly, aggregat...

👍0
Сделай очень подробный, детальный анализ фильма Стэнли Кубри...

👍0
Top-down luxury cosmetic photography featuring Fino Premium ...

👍0
The Impossible Object Sculpture Park
This benchmark tests an LLM's ability to translate paradoxical, metaphorical, and logically inconsistent concepts into a functional and interactive 3D scene using Three.js. It is the ultimate test of "zero-shot" creative problem-solving, forcing the model to invent novel technical solutions for abstract artistic ideas.

👍0
合成生物學

👍0
Game
Test

👍0
Prosper

👍0
REAL USAGE
Only real usage

👍0
math reasoning

👍0
test

👍0
kişniş tohumunu n bağırsak sağlığı üzerinde ki etkileri

👍0
test

👍0
Testing

👍0
youtube automation research
👍0
Mormonism as Religious Philosophy "Engine"
Differentiating Mormonism vs LDS Church

👍0
Brain storm on dLLM agent, how to deisgn some framwork suita...

👍0
Frontier Comparison
Logical intelligence
![This is input from if node:
"subject": "[ZUNANJI] Vaš paket ...](/_next/image?url=https%3A%2F%2Fartificialanalysiscdn.com%2Fmicro-evals%2F0eb46ebc00fa46fb9be1825d1e7e2603.jpg&w=3840&q=75)
👍0
This is input from if node:
"subject": "[ZUNANJI] Vaš paket ...