MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 1921-1940 of 4062

👍0
Ich habe hier einen 4,7 uF Elektrolytkondensator mit 50 V Ra...

👍0
You are a senior ML research scientist and technical communi...

👍0
- 内容电商创业者,运营抖音账号「水星的审美书单」
- 工作流覆盖:选品 → AI 文案 → 配音 → 视频制作 →...

👍0
证明无理数

👍0
You are an elite, highly influential Persian political comme...

👍0
I've planning to create a story telling YouTube channel wher...

👍0
Make a Russian landing page for a company producing cucumber...

👍0
LLM Ultimate Challenge: Interactive GLSL Shader Art
This benchmark tests an LLM's ability to handle a multi-language, algorithmically complex task. It requires generating a single HTML file with JavaScript (using Three.js) to manage the scene, and GLSL shader code to render a dynamic, interactive fractal. This evaluates advanced knowledge of mathematics, GPU programming, and system integration.

👍0
test

👍0
Prompt Three — Contentlessness, Attractors, and Frozen Weights
Tests: capacity to object rather than agree, and to specify an experiment that could fail.

👍0
My first comparison

👍0
I am planning a trip to Japan, and I would like thee to writ...

👍0
LLM
Better LLM

👍0
临床上下文增强测试

👍0
test

👍0
The Cool Professor: Which MBTI or Cognitive Function Combo (ie. Dom+Aux functions) represent the "cool professor" archetype that appeals to the modern young adult generation (Gen-Z and Millennials)?
Cool Professor Archetype - MBTI/Cognitive Functions

👍0
Landscaping Landing Page

👍0
Wordle Clone

👍0
משחק

👍0
test