MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 161-180 of 7809

👍1
f35 simulation

👍1
Coding Harness
To build a good AI Agent Harness

👍1
hello

👍1
A robe takes 2 bolts of blue fiber and half that much white ...

👍1
退伍老兵薛起受昔日战友杨树林邀请,只身前往探望。二十年前的山地遭遇战中,杨树林为掩护薛起撤退,替他挡下致命子弹壮烈牺牲,...

👍1
Among
Among code system

👍1
Info graph
Use svg

👍1
https://blog.google/intl/pl-pl/nowosci-produktowe/sztuczna-i...

👍1
Build a stepper showcase with progress indication, step vali...

👍1
agent-orchestrator-mcp
agent-orchestrator-mcp

👍1
https://artificialanalysis.ai/video/model-families/gemini-om...

👍1
dose etomidate for adult RSI BW 60kg and what ER doctor need...

👍1
app

👍1
Complete the following Python function:
```python
def fib4(...

👍1
面饼泡面

👍1
Universal dragon Aslam

👍1
hello

👍1
What is gradation

👍1
tính số sách

👍1
Models on Round 1
Models on Round 1