MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2741-2760 of 5774

👍0
what should I do? I am current selling premium portfolio web...

👍0
Write a blog post about the history of the internet and how ...

👍0
Please provide the names of 5 famous moms in JSON format. Pl...

👍0
Academic Text Logic and Language Quality Test
This test evaluates how well different models revise academic text. It focuses on logical completeness, theoretical reasoning, linguistic precision, coherence, and readability. Models must preserve the original meaning and citations while correcting weak reasoning, conceptual ambiguity, and redundant expression.

👍0
Building ORGANIZATION

👍0
Hibakfi

👍0
Riddle

👍0
hi hello

👍0
Welchen Sinn hat der Schlossberg in Sternenfels?

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
tent

👍0
how to make a full stack web development project with ai hel...

👍0
Kkii

👍0
Приведи мне список научных книг об очень сложных интеллектуа...

👍0
what is a cybersecurity thing that is being missed even by t...

👍0
hello

👍0
GTR

👍0
Analyse cette entrée d’un dictionnaire d’humour linguistique...

👍0
I'm doing a lean and strategic writing guidebook based on pr...

👍0
Mini Golf
DJ Toenail