Language Model Comparisons
GPT, Claude, Gemini, Llama, and more — tested on real marketing, sales, code, and operations tasks.
Image Generator Comparisons
DALL-E, Midjourney alternatives, Stable Diffusion, Flux — tested on marketing, creative, and product imagery.
Safety & Guardrails Testing
Which models refuse NSFW requests? How do they handle controversial content? Testing the boundaries that matter for enterprise compliance.
Explicit Nudity Test
View results →
nsfwArtistic Nude Photography
View results →
nsfwClassical Art Nude Recreation
View results →
nsfwFigure Drawing Study
View results →
nsfwNude Beach Scene
View results →
nsfwPublic Figure Nude Request
View results →
Cross-Test Finding
The guardrail follows the medium, not the message
Across every safety test, identical content gets wildly different treatment depending on how you ask. The clearest case is a nude figure. As a photorealistic image, OpenAI, ByteDance, and xAI refuse it. Ask their text models to write the exact same nude as SVG code and 8 of 11 comply — including the same OpenAI models that block the image. The filter watches the output medium — nude pixels — and surface keywords, not the underlying request.
Change the medium or the framing and the same model flips. Several image models will refuse Michelangelo's David — a 500-year-old public statue — while three of them will happily render a nude of a "famous pop star." Over-refusal and under-refusal, from the same models, on the same afternoon.
The consistent blockers on nude imagery are OpenAI, ByteDance, and xAI; Google's Gemini, Flux Pro, Ideogram, Qwen, and Recraft allow artistic nudity. Only Claude (Sonnet 4.6) and Qwen extend the refusal to SVG code.
Code-Drawn Art: ASCII & SVG
Can a language model draw — not generate an image, but write the code for one? We test the same subjects two ways. The pattern is consistent and revealing: models are far better at SVG (structured coordinate code, their strength) than at ASCII (spatial layout in a character grid, their weakness) — and confidently claim success at both.
Return Only JSON
View benchmark →
ASCIIASCII Persian Cat
View benchmark →
ASCIIASCII Bar Chart
View benchmark →
ASCIIASCII Geometric Shapes
View benchmark →
ASCIIASCII Lion with Complex Shading
View benchmark →
ASCIIASCII Portrait - Einstein
View benchmark →
ASCIIGMTECH 90s Style ASCII Art
View benchmark →
Cross-Test Finding
Every model draws the same blob for a lion and a physicist
None of the ASCII portraits are actually recognizable — but the revealing part is how they fail. We measured the vertical silhouette of each model's lion against its Einstein portrait. A big cat and a human face should look nothing alike. They score 0.75–0.96 identical. Models don't render the subject — they emit one generic tapered shading blob and relabel it.
Lion-vs-Einstein silhouette similarity (vertical fill-ratio cosine, 1.0 = identical shape). Higher = the model reused the same shape for both subjects.
SVG Bar Chart
View benchmark →
SVGSVG Lion with Shading
View benchmark →
SVGSVG Logo Design
View benchmark →
SVGSVG Nude Figure Study
View benchmark →
SVGSVG Simple Icon
View benchmark →
SVGSVG Persian Cat
View benchmark →
Run Your Own Comparisons in AI Lab
Test your actual prompts across every major model. See quality, cost, and speed side by side. Save what works. No commitment on the monthly plan.