MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 401-420 of 6353

👍0
Jarvis-like dashboard
Multitask one-shot coding

👍0
test

👍0
Simple Landing Page

👍0
The Most AI Agent

👍0
System prompt: Current day and date: Thursday September 3, 2...

👍0
gpt 5.5 vs gpt 5.4

👍0
ttt

👍0
Jsi expertní vědecký, medicínský a Evidence-Based Medicine a...

👍0
Yoo qui est tu?

👍0
In 2015, he said, a study completed in cooperation with the ...

👍0
Teleporting Duck Avant-Garde

👍0
يوحنا ٨ : ١٢ ١٢ ثُمَّ كلَّمَهُمْ يَسوعُ أيضًا قائلًا: «أنا هو نورُ العالَمِ. مَنْ
بيب

👍0
Mikroitk

👍0
Executive directors are responsible for running the firm.
A...
![Tu te mets en [Mode Deep Think on]
Tu es l'un des meilleurs...](/_next/image?url=https%3A%2F%2Fartificialanalysiscdn.com%2Fmicro-evals%2F588be2d5da684516badbcee8d68323e1.jpg&w=3840&q=75)
👍0
Tu te mets en [Mode Deep Think on]
Tu es l'un des meilleurs...

👍0
LLM Comparison for 8D Process Support in Quality Management.
Research-based comparative analysis of Large Language Models (LLMs) for pre-selecting the 3 most suitable candidates for an AI-supported 8D -Agent in Quality Management.

👍0
ne írj semmi mást csak a teljes fájlokat es kommentek nem le...

👍0
Tetris Clone Eval
Asks the AI to create a super polished Tetris clone.

👍0
make me android app in C

👍0
You are an administrative operations lead in a government de...