MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 3281-3300 of 7648

👍0
Silly token in/out benchmark (short)
How fast does each model read a long prompt and print out a long answer. Measures raw time-to-solution for a predictable, fixed output. Short variant.

👍0
analyze my resume - Sahil Bhatt + Toronto, ON # sahil.bhatt@...

👍0
Complete the following Python function:
```python
from typi...

👍0
chem ds

👍0
"""Host a downloaded JAIDE checkpoint as a Modal web service...

👍0
Make an interactive education app UI like Duolingo, Brillian...

👍0
strarberry有几个r
strarberry有几个r

👍0
Estimate the 85th percentile household net worth of the foll...

👍0
Which parameter has the greatest impact on how quickly a tas...

👍0
Experiments
123

👍0
How to learn mathematics well

👍0
ptvgff

👍0
Consider a universe consisting of a vast network of nodes (r...

👍0
Best Winning Project developing, writing and producing model

👍0
Output of ('hhh'*0) in python and why is the output is given

👍0
zelda breath of the wild
Three.js game offline capabilities

👍0
Pipe Puzzle
classic game; by mnf

👍0
Find the order of the factor group (Z_4 x Z_12)/(<2> x <2>)
...

👍0
In 2015, he said, a study completed in cooperation with the ...

👍0
JVM CORE VHDL