MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2461-2480 of 8662

👍0
Merhaba

👍0
OMDQS
lkl

👍0
laknans

👍0
jsglass

👍0
Aaba

👍0
Complete the following Python function:
```python
from typi...

👍0
算f(2)和f(3): f(x) = -x for x<0; f(x) = f(x - f(x-1))/2 for x≥...

👍0
LMC COUNTEREXAMPLE C-01
Nezávislý audit; bez Web Search a p...

👍0
test
👍0
P35-v1.1a
PROMPT 1 - GOVERNANCE / SELECTION-AWARE EVIDENCE /...
👍0
Write a 300+ word summary of the wikipedia page "https://en....
👍0
I want to perform an exhaustive, line-by-line static analysi...

👍0
Shanon

👍0
B.M.5 — FULL-METHOD RED-TEAM AUDIT PORTABLE v20 + WEB CROSS-...

👍0
خودتون رو معرفی کنید

👍0
wres
👍0
Execute tasks as a pure computational synthesis and mathemat...

👍0
HTML WebDev Challenge
A simple web development challenge that tests an LLM's ability to create HTML web pages given a prompt. The eval is 8 prompts. The LLM is ONLY allowed to build with HTML, Javascript, and TailwindCSS. It may pull Javascript libraries from a CDN (like Jsdelivr or Cloudflare), but only if it is SURE they exist and are needed for the build. The first four prompts are very specific, and then the rest give the model more freedom.

👍0
idk
👍0
Hi!