MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 4421-4440 of 9330

👍0
Write a blog post about the history of the internet and how ...

👍0
hiii

👍0
you want to transcribe a financial news livestream, so who a...

👍0
Same topology, but: external access ALSO works sometimes fro...

👍0
Please provide the names of 5 famous moms in JSON format. Pl...
👍0
Scene 1 – Village Morning: A beautiful Indian village at sun...

👍0
Academic Text Logic and Language Quality Test
This test evaluates how well different models revise academic text. It focuses on logical completeness, theoretical reasoning, linguistic precision, coherence, and readability. Models must preserve the original meaning and citations while correcting weak reasoning, conceptual ambiguity, and redundant expression.

👍0
Building ORGANIZATION

👍0
Hibakfi

👍0
SEL GUARD TEST
👍0
Analyst stock
👍0
Helo

👍0
Riddle

👍0
hi hello

👍0
Welchen Sinn hat der Schlossberg in Sternenfels?

👍0
models jinki elo ratings highest ho ya prelaunched models ji...

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
tent
👍0
Generate a 5sec. Video of a housewife in 1950s

👍0
how to make a full stack web development project with ai hel...