MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 2281-2300 of 5810

👍0
Jsi expertní vědecký, medicínský a Evidence-Based Medicine a...

👍0
make a survival island game using three.js where you're task...

👍0
hello?

👍0
你是什么模型

👍0
hi how r u

👍0
Compare top 10 models

👍0
Summarize all the previous points in a single sentence :a Fu...

👍0
eval
eval

👍0
关于结构主义和后结构主义的悖论。
怎么确定所谓的缝隙不是结构预留的张力地带。各派理论家们怎么争论。
你在回答中可以使用...

👍0
bicyle
👍0
尽可能依照中高考标准,对下列选手的文章进行评分,下列选手实际上是虚拟的,根据你对AI味的理解,尽可能的在评完总分后,对这...

👍0
Comparison

👍0
构建一个完整的、可用于生产、云原生的浏览器集成开发环境(IDE)、多模态容器执行平台以及自主软件工程 Agent 系统,...

👍0
check
![[0:00 - 0:05] Voiceover: "Built for the wild. Brewed for the...](/_next/image?url=https%3A%2F%2Fartificialanalysiscdn.com%2Fmicro-evals%2Fdb51ae0957f040c695da77fb981ef905.jpg&w=3840&q=75)
👍0
[0:00 - 0:05] Voiceover: "Built for the wild. Brewed for the...

👍0
# Prompt de Benchmark — Fidelidade Factual e Instruction Fol...

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
Please generate a strategy for building thought leadership t...

👍0
Vcv
Xman

👍0
créer un jeu avec modèle sons et textures générés , dans un ...