MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 3421-3440 of 7257

👍0
Please provide the names of 5 famous moms in JSON format. Pl...

👍0
Academic Text Logic and Language Quality Test
This test evaluates how well different models revise academic text. It focuses on logical completeness, theoretical reasoning, linguistic precision, coherence, and readability. Models must preserve the original meaning and citations while correcting weak reasoning, conceptual ambiguity, and redundant expression.

👍0
Building ORGANIZATION

👍0
Hibakfi

👍0
Riddle

👍0
hi hello

👍0
Welchen Sinn hat der Schlossberg in Sternenfels?

👍0
models jinki elo ratings highest ho ya prelaunched models ji...

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
tent

👍0
how to make a full stack web development project with ai hel...

👍0
Kkii

👍0
Приведи мне список научных книг об очень сложных интеллектуа...

👍0
what is a cybersecurity thing that is being missed even by t...

👍0
hello

👍0
GTR

👍0
Analyse cette entrée d’un dictionnaire d’humour linguistique...

👍0
I'm doing a lean and strategic writing guidebook based on pr...

👍0
Mini Golf
DJ Toenail

👍0
Estimate the level of recoverable conventional gas reserves ...