MicroEvals
Run your prompts across multiple models to compare their performance.
Public evaluations
Showing 381-400 of 3014

👍0
Budgeting

👍0
Japanese Member Note Benchmark

👍0
Laboratorio de arquitecturas de redes, neuronales y computación de inteligencia artificial
Laboratorio donde poder crear tus propios arquitecturas de Ia y redes neuronales

👍0
What defines intelligence?

👍0
qwen3.8

👍0
chat-pdf
Chatbot with pdf RAG

👍0
Game of Life

👍0
The Kandinsky Challenge: Programming Synesthesia
This benchmark tests an LLM's ability to interpret and implement abstract, philosophical, and artistic theories. It requires the model to translate Wassily Kandinsky's theories on synesthesia (the connection between sound, color, and shape) into an interactive audio-visual experience. Success is judged on the creative fidelity to the artistic concept, not just technical execution.

👍0
Analyse cette entrée d’un dictionnaire d’humour linguistique...

👍0
ai aiai

👍0
Genérame una cancha de baloncesto reglamentaria FIBA en svg
👍0
What if this scene include in Ace Combat 7 2018 E3 trailer(N...

👍0
Hello!

👍0
test
to test how it works

👍0
SeahorseBench
The infamous Mandela Effect prompt

👍0
Covariance Matrix explanation in 4 levels

👍0
Fizikteki gecikmeli seçim kuantum silgi deneyinin gerçekçi g...

👍0
Create photorelatic world of man at horse explaining abou ma...

👍0
Taze fasulye yemeğini düdüklü tencerede pişirsek Fito kimyas...

👍0
Solve this with the constraint that you cannot use the obvio...