MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 1121-1140 of 9017

👍0
Universal AI Real-World Capability Benchmark
A comprehensive evaluation of AI models across real-world problem solving, instruction following, factual knowledge, coding, debugging, reasoning, software architecture, security, QA, product thinking, communication, and failure analysis. Designed to compare models on practical usefulness, accuracy, reliability, and ability to follow complex constraints.

👍0
Pick a random number 1-1000

👍0
The morning temperature in a city is 41°F. If a sunny, mild ...

👍0
perfect points and counterpoints for Group Discussion., for ...
👍0
You’ve been hired as an In Ear Monitor (IEM) Tech for a tour...

👍0
我之前是一位香港證監會持牌人的負責人員,持有RA1,4號負責人員牌照,但現在已失業,我在招聘廣告見到有一個要求就是要幫公...

👍0
Work

👍0
SCB SCB-forensics-T2 MK3

👍0
Hangi Yapay Zeka modelisin

👍0
You are a professional experienced genius intelligent creati...

👍0
Apigee vs Kong

👍0
Create code for 3d world visually stunning of a indian city

👍0
היי
👍0
New

👍0
hi!

👍0
error handling in go gin

👍0
Task: zero Downtime Database SQL Migration
👍0
I am a 23-year-old male, MCA graduate, currently in an exten...

👍0
trial

👍0
jopta