MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 3061-3080 of 6588

👍0
### General Instructions
- Planning: Step-back and take a b...

👍0
Ich habe hier einen 4,7 uF Elektrolytkondensator mit 50 V Ra...

👍0
ZERIKO
This app was made for the riders

👍0
hiii kitty

👍0
You are a senior ML research scientist and technical communi...

👍0
Janet’s ducks lay 16 eggs per day. She eats three for breakf...

👍0
- 内容电商创业者,运营抖音账号「水星的审美书单」
- 工作流覆盖:选品 → AI 文案 → 配音 → 视频制作 →...

👍0
Non conventional introduction of personal finance

👍0
证明无理数

👍0
You are an elite, highly influential Persian political comme...

👍0
best way to make and display a Relative Rotation graphh for ...

👍0
Make a Russian landing page for a company producing cucumber...

👍0
Google seo
Google seo 的价值与各大模型的反应时间与回答问题的专业程度

👍0
i am preparing for jee mains and advanced and scored 70 perc...

👍0
LLM Ultimate Challenge: Interactive GLSL Shader Art
This benchmark tests an LLM's ability to handle a multi-language, algorithmically complex task. It requires generating a single HTML file with JavaScript (using Three.js) to manage the scene, and GLSL shader code to render a dynamic, interactive fractal. This evaluates advanced knowledge of mathematics, GPU programming, and system integration.

👍0
你作为资深的财务专家,现在需要对企业内的办公系统内增加财务系统的预算管理,你会如何设计

👍0
Buyers and Purchasing Agents
You are the senior category bu...

👍0
test

👍0
what can you do?

👍0
Prompt Three — Contentlessness, Attractors, and Frozen Weights
Tests: capacity to object rather than agree, and to specify an experiment that could fail.