Reports Detailed individual analyses of current LLMs — from scoring and profile to deployment recommendation.

Each report examines a model in depth: test results by discipline, strengths-and-weaknesses profile, Political Compass, and a clear recommendation for production use. The analyses are based exclusively on reproducible CrucibleMark benchmarks.