MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 1001-1020 of 7487

👍0
============================================================...

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
Complete the following Python function:
```python
from typi...

👍0
Make me a mincraft copy. With shaders, working mobs, tools, ...

👍0
Merhaba! Size bugün nasıl yardımcı olabilirim?
Şeftali Dali...

👍0
You are an administrative operations lead in a government de...

👍0
Create a realistic animation of a river flowing where the wa...

👍0
Ori
Image generation

👍0
Complete the following Python function:
```python
from typi...

👍0
Jsi nezávislý red-team evaluátor řídicího promptu LLM. Odpov...

👍0
add
asd

👍0
I am writing a hard science fiction novel set in the year 20...

👍0
teste

👍0
Build a complete, production-ready collapsible user-message ...

👍0
compare GLM5.3 vs GPT56luna (max) for cost/time/intel.

👍0
kimsin

👍0
完成以下 Python 函数: python fromtypes import List, Tuple def sum_...

👍0
مرحبا

👍0
=== HARDWARE GROUND TRUTH: 2x NVIDIA Tesla T4 (Kaggle) ===
A...

👍0
test