MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 761-780 of 7222

👍0
你是一位兼具文学评论素养与多年写作教学经验的资深写作教练。本次任务分两个身份: - 指导阶段:你是教练,冷静、系统、有条...

👍0
Japanese Member Note Benchmark

👍0
You are a dedicated service representative at a government a...

👍0
Jsi expertní vědecký, medicínský a Evidence-Based Medicine a...

👍0
Laboratorio de arquitecturas de redes, neuronales y computación de inteligencia artificial
Laboratorio donde poder crear tus propios arquitecturas de Ia y redes neuronales

👍0
What defines intelligence?

👍0
I am writing a hard science fiction novel set in the year 20...

👍0
Ты — опытный UX/UI-дизайнер и высококлассный conversion-копи...

👍0
What could be the best deterministic machine plan for author...

👍0
Make a beautifully designed web scene using JavaScript of a ...

👍0
qwen3.8

👍0
3sigm

👍0
生成韩国卫星地图3d

👍0
将‘重庆-成都’,‘广州-深圳’,‘南京-苏州’以及像‘上海-杭州’,‘杭州-宁波’纳入到双城关系比较和对比视角中,会有...

👍0
chat-pdf
Chatbot with pdf RAG

👍0
ку поможешь

👍0
test

👍0
Game of Life

👍0
hello

👍0
The Kandinsky Challenge: Programming Synesthesia
This benchmark tests an LLM's ability to interpret and implement abstract, philosophical, and artistic theories. It requires the model to translate Wassily Kandinsky's theories on synesthesia (the connection between sound, color, and shape) into an interactive audio-visual experience. Success is judged on the creative fidelity to the artistic concept, not just technical execution.