MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 5721-5740 of 6208

👍0
better for coding agent and better for debugging agent

👍0
You are an expert qa engineer, experiend in mobile app autoa...

👍0
Jsi expertní vědecký, medicínský a Evidence-Based Medicine a...

👍0
Yon
👍0
MinecraftQA
A mixed set of 20 Minecraft-related questions, ranging from trivia to complex logic

👍0
claude vs codex

👍0
3D studio
AI generated 3D studio

👍0
// =========================================================...

👍0
Мне нужно выбрать самую лучшую тему для написания дипломной ...

👍0
Bir sepette 100 yumurta 6 çıktı kaç kaldı bu bir bilmecedir ...

👍0
یک گیم داییناسور گوگل

👍0
AI Agent Expressif

👍0
what is the best free ai for research

👍0
You are an administrative operations lead in a government de...

👍0
Jsi expertní vědecký, medicínský a Evidence-Based Medicine a...

👍0
Make a roadmap with the best resources to learn PHILOSOPHY o...

👍0
Project Blackbox

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
botanicula

👍0
I am planning a trip to Hangzhou province, and I would like ...