MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 5721-5740 of 10627

👍0
CAS

👍0
The cyclic subgroup of Z_24 generated by 18 has order
A) 4
...

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...
👍0
大模型评测
大模型评测问题咨询

👍0
STBN X 3.0
aa_ANMBo•••• 22 de julio de 2026
👍0
free
👍0
you go your way i go mine mit jelent

👍0
new

👍0
test
first try

👍0
写python代码制作gif,1比1复刻claude icon的呼吸

👍0
You are an administrative operations lead in a government de...

👍0
Novel
👍0
Ez agent nek keszul shema hogy hasznalja az exa ai t
from e...

👍0
title

👍0
Roblox SVG
Testing how models make Roblox-related SVGs

👍0
Viết 1 file .c và .h:
-Dùng cho 1 chess engine pattern-match...

👍0
生成真实比例的太阳系3d-2

👍0
DBD
👍0
GPT-6 Astra VS Deepseek v4 Flash Latest
👍0
Project AEGIS: Frontier LLM Multi-Constraint & Epistemic Stress Test
Zero-shot stress test evaluating model performance across multi-layered combinatorial constraints, distractor rejection, epistemic independence, and negative instruction following.