MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 5281-5300 of 9326

👍0
Act as a senior full-stack software engineer, UI/UX designer...

👍0
geenrate images realted to 2050 in technology
👍0
Estimate the 85th percentile household-networth of the follo...

👍0
from __future__ import annotations
import os
from pathlib i...
👍0
Great
👍0
Complete the following Python function:
```python
from typi...
👍0
Create a highly detailed and hyper-realistic 3D model of a m...

👍0
Comparacion Modelos Busqueda Piso
👍0
Chem test

👍0
nailrahil19@gmail.com
👍0
Hazme un texto breve y muy filosófico

👍0
hih

👍0
wrsrr
helyes válasz: Sasha Banks

👍0
PROMPT 1 - P30.1a - CUMULATIVE EXPOSURE / RECHECK / PROTECTI...

👍0
ChatGPT

👍0
Summarize all the previous points in a few sentence:
A komme...

👍0
Imagine there is a building called Bronte tower whose height...
👍0
Spreadsheet Engine: Spec Fidelity + Performance (sheet.py)
A hard, original agentic coding task. Models must implement a spreadsheet engine from a precise written spec: a recursive-descent formula parser, integer arithmetic with truncating division, lazy IF, row-major error ordering in range aggregates, static cycle detection that ignores IF laziness, and UNDO. The spec also lists six performance families (deep chains, 100k-cell ranges with point updates, cycle toggling, wide rectangles, 10k-character nested formulas) that naive solutions fail. Outputs are deterministic stdin/stdout, so each submission can be scored objectively by running it against the public example and hidden tests.
👍0
hey

👍0
Every day, Wendi feeds each of her chickens three cups of mi...