MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2641-2660 of 5613

👍0
Janet’s ducks lay 16 eggs per day. She eats three for breakf...

👍0
- 内容电商创业者,运营抖音账号「水星的审美书单」
- 工作流覆盖:选品 → AI 文案 → 配音 → 视频制作 →...

👍0
Non conventional introduction of personal finance

👍0
证明无理数

👍0
You are an elite, highly influential Persian political comme...

👍0
best way to make and display a Relative Rotation graphh for ...

👍0
Make a Russian landing page for a company producing cucumber...

👍0
i am preparing for jee mains and advanced and scored 70 perc...

👍0
LLM Ultimate Challenge: Interactive GLSL Shader Art
This benchmark tests an LLM's ability to handle a multi-language, algorithmically complex task. It requires generating a single HTML file with JavaScript (using Three.js) to manage the scene, and GLSL shader code to render a dynamic, interactive fractal. This evaluates advanced knowledge of mathematics, GPU programming, and system integration.

👍0
你作为资深的财务专家,现在需要对企业内的办公系统内增加财务系统的预算管理,你会如何设计

👍0
test

👍0
what can you do?

👍0
Prompt Three — Contentlessness, Attractors, and Frozen Weights
Tests: capacity to object rather than agree, and to specify an experiment that could fail.

👍0
My first comparison

👍0
Tell me about gandhiji

👍0
Write a detailed, emotional, snarky, sarcastic sci-fi story.

👍0
Create a complete, single-file HTML app (including embedded ...

👍0
I am planning a trip to Japan, and I would like thee to writ...

👍0
LLM
Better LLM

👍0
临床上下文增强测试