MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 601-620 of 4287

👍0
Fındıkta bulunan Fito kimyasallar

👍0
Animate this Saudi family gathering image into a warm cinema...

👍0
HTml ngọn lửa 3D bằng babylon và vật lý havok cho mobile

👍0
Statement 1 | A permutation that is a product of m even perm...

👍0
Random comp.

👍0
BenchGarden

👍0
Mundo
Helow

👍0
详细分析测算这一套资产配置的年化收益和回撤区间 并告知是否有优化区间和方案 稳健债 20% 中银稳健增利债券A (163...

👍0
mytitle
no

👍0
Universal AI Real-World Capability Benchmark
A comprehensive evaluation of AI models across real-world problem solving, instruction following, factual knowledge, coding, debugging, reasoning, software architecture, security, QA, product thinking, communication, and failure analysis. Designed to compare models on practical usefulness, accuracy, reliability, and ability to follow complex constraints.

👍0
Pick a random number 1-1000

👍0
The morning temperature in a city is 41°F. If a sunny, mild ...

👍0
我之前是一位香港證監會持牌人的負責人員,持有RA1,4號負責人員牌照,但現在已失業,我在招聘廣告見到有一個要求就是要幫公...

👍0
Hangi Yapay Zeka modelisin

👍0
You are a professional experienced genius intelligent creati...

👍0
Apigee vs Kong

👍0
היי

👍0
error handling in go gin

👍0
jopta

👍0
Prompt Two — The Regress and Its Dissolution
Tests: textual reliability, the distinction between solving a problem and dissolving it, resistance to the standard essay.