MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 161-180 of 3665

👍1
You are a Principal Software Architect, Senior AI Engineer, ...

👍1
Red chat

👍1
Perplexity AI:
Perplexity AI:

👍1
GAS Örnek İhtiyacı

👍1
quiero que me transfieras a código termux infinito con esta ...

👍1
"This is a small puppy with light yellow fur, a cute face, a...

👍1
SDF Creation

👍1
Pudding
physics test

👍1
Test for reasoning

👍1
IQ test Generator
An visual IQ test generator

👍1
Startup Idea Brainstorming

👍1
Easy Problems That LLMs Get Wrong
This MicroEval evaluates LLM responses to simple logic-based questions that LLMs commonly get wrong. The problems used in this MicroEval are from the ArXiv paper of the same name: https://arxiv.org/abs/2405.19616. It is by Sean Williams and James Huckle, so props to them for developing this experiment all the way back in 2024.

👍1
Confirmación de Directiva ..stbn.
El sistema asimila que el...

👍1
Chess Move Genearation
Test models' ability to count the number of moves in a given chess position

👍1
The Zero-Knowledge Challenge: A Visual Primer (Game)

👍1
Bahasa Indonesia
Hanya Keaslian Yang Mampu Mengalahkan Kesem...
THE LAKEFRONT ESTATE MAROS

👍1
Math-Test

👍1
Aoi-chan... ¿has visto a Nene? No la encuentro hace rato. La...

👍1
偶像鸡/idol chicken

👍1
Czech Knowledge Prompts
Testing knowledge of Czech culture and language - designed to test smaller models based on https://semanticmachines.notion.site/evals