Language

Preview KPIs are a precomputed sample on the recorded 100-problem catalog. LLM rows stay not_evaluated. A fluent paragraph is not a feasible plan.

No live result yet. Load a sample problem and run a comparison.

Edit the schema JSON only if you want a structured-solver what-if. Prompts stay synthetic. Direct-answer JSON is benchmark-only.

Interactive problem