MicroEvals
Public evaluations
Showing 161-180 of 4256

This test is a copy of "https://www.youtube.com/watch?v=d4bVpUL9Hao" the tests run in the YouTube video to test the WebDEV and animation capabilities of AI models.

Production-quality portfolio with smooth interactive 3D elements, CraftWorld showcase, animations and responsive UI.

This microeval will evaluate how well leading coding models can develop different types of web apps based on prompts that aren't ultra-specific, and just ask for the overall concept. This is to test how much LLMs have evolved in terms of design skills and common knowledge coding choices. There is one detailed prompt to see if it enhances the quality.



Generator

![Strawrbrerrry [sic] eval](/_next/image?url=https%3A%2F%2Fartificialanalysiscdn.com%2Fmicro-evals%2F45d80a29bd4f4c68b3a92569454473d4.jpg&w=3840&q=75)
Reasoning should include the ability to generalize to unfamiliar words instead of memorizing answers. Let's see if models can detect the number of 'r's in the word "strawrbrerrry."

Visual perception of the 5 most important unsolved concepts in mathematics!







Perplexity AI:


