Project Patti: Why can You Solve Diabolical Puzzles on one Sudoku Website but not Easy Puzzles on another Sudoku Website?

๐Ÿ“… 2025-07-22
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
This study addresses the inconsistency in difficulty ratings across Sudoku websites. We propose two novel, unsupervised, and quantifiable difficulty metrics: (1) a structural complexity measure based on clause-length distribution derived from SAT encoding; and (2) a simulation-based solver integrating four human-like solving strategies with randomized Nishio backtracking. Together, these form a cross-platform difficulty standardization framework. Evaluated on over 1,000 puzzles from five major Sudoku websites, our approach achieves strong agreement with original site labelsโ€”Spearmanโ€™s ฯ > 0.85 on four sites. It successfully establishes a universal three-tier classification (Easy/Medium/Hard) and supports novice-oriented solving guidance. To our knowledge, this is the first fully automated, interpretable, and platform-agnostic Sudoku difficulty assessment method that requires no human annotation.

Technology Category

Application Category

๐Ÿ“ Abstract
In this paper we try to answer the question "What constitutes Sudoku difficulty rating across different Sudoku websites?" Using two distinct methods that can both solve every Sudoku puzzle, I propose two new metrics to characterize Sudoku difficulty. The first method is based on converting a Sudoku puzzle into its corresponding Satisfiability (SAT) problem. The first proposed metric is derived from SAT Clause Length Distribution which captures the structural complexity of a Sudoku puzzle including the number of given digits and the cells they are in. The second method simulates human Sudoku solvers by intertwining four popular Sudoku strategies within a backtracking algorithm called Nishio. The second metric is computed by counting the number of times Sudoku strategies are applied within the backtracking iterations of a randomized Nishio. Using these two metrics, I analyze more than a thousand Sudoku puzzles across five popular websites to characterize every difficulty level in each website. I evaluate the relationship between the proposed metrics and website-labeled difficulty levels using Spearman's rank correlation coefficient, finding strong correlations for 4 out of 5 websites. I construct a universal rating system using a simple, unsupervised classifier based on the two proposed metrics. This rating system is capable of classifying both individual puzzles and entire difficulty levels from the different Sudoku websites into three categories - Universal Easy, Universal Medium, and Universal Hard - thereby enabling consistent difficulty mapping across Sudoku websites. The experimental results show that for 4 out of 5 Sudoku websites, the universal classification aligns well with website-labeled difficulty levels. Finally, I present an algorithm that can be used by early Sudoku practitioners to solve Sudoku puzzles.
Problem

Research questions and friction points this paper is trying to address.

Defining Sudoku difficulty metrics across different websites
Analyzing puzzle complexity using SAT and human-like solvers
Creating a universal rating system for consistent difficulty classification
Innovation

Methods, ideas, or system contributions that make the work stand out.

Convert Sudoku to SAT for structural complexity
Simulate human solvers with backtracking Nishio
Universal rating system using two metrics
๐Ÿ”Ž Similar Papers
2024-03-15arXiv.orgCitations: 2
2024-06-13arXiv.orgCitations: 0
๐Ÿ’ผ Related Jobs
No related jobs found.
A
Arman Eisenkolb-Vaithyanathan
Lynbrook High School, San Jose