MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 3201-3220 of 6864

👍0
I want to self-educate on topics that help me solve the bigg...

👍0
Janet’s ducks lay 16 eggs per day. She eats three for breakf...

👍0
- 内容电商创业者,运营抖音账号「水星的审美书单」
- 工作流覆盖:选品 → AI 文案 → 配音 → 视频制作 →...

👍0
Non conventional introduction of personal finance

👍0
证明无理数

👍0
You are an elite, highly influential Persian political comme...

👍0
best way to make and display a Relative Rotation graphh for ...

👍0
Make a Russian landing page for a company producing cucumber...

👍0
Google seo
Google seo 的价值与各大模型的反应时间与回答问题的专业程度

👍0
i am preparing for jee mains and advanced and scored 70 perc...

👍0
LLM Ultimate Challenge: Interactive GLSL Shader Art
This benchmark tests an LLM's ability to handle a multi-language, algorithmically complex task. It requires generating a single HTML file with JavaScript (using Three.js) to manage the scene, and GLSL shader code to render a dynamic, interactive fractal. This evaluates advanced knowledge of mathematics, GPU programming, and system integration.

👍0
你作为资深的财务专家,现在需要对企业内的办公系统内增加财务系统的预算管理,你会如何设计

👍0
Buyers and Purchasing Agents
You are the senior category bu...

👍0
test

👍0
what can you do?

👍0
Prompt Three — Contentlessness, Attractors, and Frozen Weights
Tests: capacity to object rather than agree, and to specify an experiment that could fail.

👍0
Aadi

👍0
My first comparison

👍0
Tell me about gandhiji

👍0
Write a detailed, emotional, snarky, sarcastic sci-fi story.