MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 141-160 of 3315

👍1
towerofhujnoji
👍1
p5.js physics_New

👍1
fusion generator
Generator
![Strawrbrerrry [sic] eval](/_next/image?url=https%3A%2F%2Fartificialanalysiscdn.com%2Fmicro-evals%2F45d80a29bd4f4c68b3a92569454473d4.jpg&w=3840&q=75)
👍1
Strawrbrerrry [sic] eval
Reasoning should include the ability to generalize to unfamiliar words instead of memorizing answers. Let's see if models can detect the number of 'r's in the word "strawrbrerrry."

👍1
Interactive Concepts for Mathematical Problems
Visual perception of the 5 most important unsolved concepts in mathematics!

👍1
Hola, ¿cómo estamos?, ¡estás o no estás', ¿con qué modelo co...

👍1
Generate a 20x20 word search game and also the ability to fi...

👍1
百合是什么?

👍1
You are a Principal Software Architect, Senior AI Engineer, ...

👍1
Red chat

👍1
Perplexity AI:
Perplexity AI:

👍1
GAS Örnek İhtiyacı

👍1
quiero que me transfieras a código termux infinito con esta ...

👍1
"This is a small puppy with light yellow fur, a cute face, a...

👍1
SDF Creation

👍1
Pudding
physics test

👍1
Test for reasoning

👍1
IQ test Generator
An visual IQ test generator

👍1
Startup Idea Brainstorming

👍1
Easy Problems That LLMs Get Wrong
This MicroEval evaluates LLM responses to simple logic-based questions that LLMs commonly get wrong. The problems used in this MicroEval are from the ArXiv paper of the same name: https://arxiv.org/abs/2405.19616. It is by Sean Williams and James Huckle, so props to them for developing this experiment all the way back in 2024.