MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 1301-1320 of 4475

👍0
llj

👍0
Controleer de onderstaande lijst met AI-termen en bijbehoren...

👍0
The cyclic subgroup of Z_24 generated by 18 has order
A) 4
...

👍0
живой глубокий

👍0
Here is a fully completed, production-ready version of your ...
![# [Topic Name]
## 🧠 Core Intuition (2-3 sentences)
What i...](/_next/image?url=https%3A%2F%2Fartificialanalysiscdn.com%2Fmicro-evals%2Fc99f81f7c4314927827e9a0f03e34f19.jpg&w=3840&q=75)
👍0
# [Topic Name]
## 🧠 Core Intuition (2-3 sentences)
What i...

👍0
Quais as melhores IAs para criar, construir e desenvolver ar...

👍0
You're my chief-of-staff-level assistant. I'm a director at ...

👍0
Benchmarks between most popular LLM

👍0
Shrodinger equation
Quantum Physics

👍0
laknans

👍0
jsglass

👍0
Complete the following Python function:
```python
from typi...

👍0
test

👍0
Shanon

👍0
خودتون رو معرفی کنید

👍0
wres

👍0
HTML WebDev Challenge
A simple web development challenge that tests an LLM's ability to create HTML web pages given a prompt. The eval is 8 prompts. The LLM is ONLY allowed to build with HTML, Javascript, and TailwindCSS. It may pull Javascript libraries from a CDN (like Jsdelivr or Cloudflare), but only if it is SURE they exist and are needed for the build. The first four prompts are very specific, and then the rest give the model more freedom.

👍0
idk

👍0
What is the best build in mass builder?