MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2661-2680 of 5605

👍0
6y6
65y5

👍0
Using the reference image of [Alex - locked character descri...

👍0
test2

👍0
Warum sind Bananen krumm?

👍0
hi can you generate a video

👍0
Написать такой код вба для эксель 2013 который должен:
- У ...

👍0
Что нужно для канонизации святого в православной церкви?

👍0
what is this doing, please analyse following code snippet
...

👍0
Let me clarify Ok suppose 3.8 gpa at math cs Econ hard cours...

👍0
Write a function that returns the perimeter of a square give...

👍0
Without the connotation that an acutal french person would h...

👍0
what should I do? I am current selling premium portfolio web...

👍0
Please provide the names of 5 famous moms in JSON format. Pl...

👍0
Academic Text Logic and Language Quality Test
This test evaluates how well different models revise academic text. It focuses on logical completeness, theoretical reasoning, linguistic precision, coherence, and readability. Models must preserve the original meaning and citations while correcting weak reasoning, conceptual ambiguity, and redundant expression.

👍0
Building ORGANIZATION

👍0
Hibakfi

👍0
Riddle

👍0
hi hello

👍0
Welchen Sinn hat der Schlossberg in Sternenfels?

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...