MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2101-2120 of 4461

👍0
You are an elite, highly influential Persian political comme...

👍0
Make a Russian landing page for a company producing cucumber...

👍0
LLM Ultimate Challenge: Interactive GLSL Shader Art
This benchmark tests an LLM's ability to handle a multi-language, algorithmically complex task. It requires generating a single HTML file with JavaScript (using Three.js) to manage the scene, and GLSL shader code to render a dynamic, interactive fractal. This evaluates advanced knowledge of mathematics, GPU programming, and system integration.

👍0
test

👍0
what can you do?

👍0
Prompt Three — Contentlessness, Attractors, and Frozen Weights
Tests: capacity to object rather than agree, and to specify an experiment that could fail.

👍0
My first comparison

👍0
Create a complete, single-file HTML app (including embedded ...

👍0
I am planning a trip to Japan, and I would like thee to writ...

👍0
LLM
Better LLM

👍0
临床上下文增强测试

👍0
test

👍0
The Cool Professor: Which MBTI or Cognitive Function Combo (ie. Dom+Aux functions) represent the "cool professor" archetype that appeals to the modern young adult generation (Gen-Z and Millennials)?
Cool Professor Archetype - MBTI/Cognitive Functions

👍0
Landscaping Landing Page

👍0
Wordle Clone

👍0
משחק

👍0
test

👍0
test2

👍0
hi can you generate a video

👍0
Написать такой код вба для эксель 2013 который должен:
- У ...