The AIMuse archive

All articles.

Model experiments, research notes, and lessons from building with AI.

55 articles · Newest first
Prompt Engineering

System Prompts Versus User Prompts: Empirical Lessons from an 18-Model LLM Benchmark on Hard Constraints

This article presents a rigorous benchmark of 18 leading LLMs tasked with generating a constraint-heavy children’s story, analyzing prompt engineering protocols, compliance metrics, editing effort, and the evolution of system prompts for maximal rule adherence and editorial efficiency.

Read article : System Prompts Versus User Prompts: Empirical Lessons from an 18-Model LLM Benchmark on Hard Constraints
Showing 9 of 55