MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 4141-4160 of 6995

👍0
generate me an snake game in html

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
分析casioGD-B500手表的优缺点

👍0
caa
ca

👍0
I am planning a trip to Japan, and I would like thee to writ...

👍0
teting
lets see what happen

👍0
hitman codename 47 oyunun karakter modelleri tasarlanırken, ...

👍0
what is span margin

👍0
You are an administrative operations lead in a government de...

👍0
veste couleur

👍0
Gsgs
Stst
![[SYSTEM] Convert a natural-language user idea + target aspec...](/_next/image?url=https%3A%2F%2Fartificialanalysiscdn.com%2Fmicro-evals%2F81e2375a0ebb4729b2d1b27869f1bb82.jpg&w=3840&q=75)
👍0
[SYSTEM] Convert a natural-language user idea + target aspec...

👍0
Create a fun arcade racer in JS. The game should be set in a...

👍0
roma

👍0
how to win friends and influence people

👍0
kimi max3

👍0
# 全面实施提示:50亿参数混合稀疏MoE语言模型训练系统
构建一个完整的生产级训练系统,用于训练一个50亿参数混合稀...

👍0
EXTERNAL METHODOLOGY-ARCHITECTURE AUDIT — ITERATION 23
You a...

👍0
Book Segmentation
Reasoning ability to see if the model can reason which sentence/text chunk is spoken by which character

👍0
HumanEval Question 32
HumanEval Question 32