REVIEW 4 major objections 6 minor 16 references
Learning Fricke signs from Maass form Coefficients
T0 review · 4 major / 6 minor · reviewed 2026-08-10 · deepseek-v4-flash
Pith's one-line read This paper shows that supervised machine learning on the first 1,000 Fourier coefficients predicts the Fricke sign of a Maass form with 94–96% accuracy, and that predictions for forms with unknown signs agree with a heuristic algorithm…
desk verdict A useful empirical paper whose core claim is probably right, but whose headline numbers are sloppy and whose transfer to unknown-sign forms is only heuristically supported. read the letter →
The pith
A machine-rendered reading of the paper's core claim, the machinery that carries it, and where it could break.
The reading
What carries the argument
The central object is the normalized feature vector $$D = \{((-1)^{\$\sigma$(f)} a_n)_{n=1}^{1000} : f \in \mathcal{L}\},$$ in which each coefficient is multiplied by the parity sign so that even and odd forms are aligned by root number, together with the factorization of the Fricke sign into local factors $w_N = \prod_{p\mid N} w_p$. The machinery that carries the argument is Linear Discriminant Analysis (LDA), which fits a linear decision boundary under the assumption that the two sign classes share a covariance structure; the paper checks this assumption with a standard equal-covariance test. A second piece of machinery is the averaging operation that produces murmuration plots, which both motivates LDA and validates predictions by comparing average coefficients of predicted-sign forms with known-sign forms. The paper also uses a neural network with the spectral parameter $R$ appended, and a heuristic algorithm that guesses Fricke signs by solving approximate overdetermined linear systems; agreement with that heuristic provides the external check on the unknown-sign predictions.
What would settle it
Take a random sample of the 15,423 forms with unknown Fricke sign, compute their signs rigorously at higher precision, and compare with the LDA predictions; if the error rate on that sample is near 50% rather than near 5%, the transfer claim is false.
Extended reading notes
Core claim
On the paper's own terms, the discovery is that the Fricke sign of a Maass form is learnable from the first 1,000 Fourier coefficients, with accuracy far above the 86% obtainable from prime-indexed coefficients alone, and robust to masking the coefficients whose indices share a factor with the level—the obvious place where the sign is encoded. The authors argue the classifier is not merely reading off $a_p$ for primes dividing the level, because setting those coefficients to zero leaves accuracy nearly unchanged for the best feature set and because level-1 forms, where no coefficient directly encodes the sign, are classified successfully. Instead, the full coefficient vector carries extra predictive signal: indices with one or two prime factors give 95.3% accuracy, so the multiplicative structure itself appears informative. This connects the predictive task to murmurations: the same sign-conditioned averages that oscillate in the plots are what the linear classifier exploits.
Load-bearing premise
The classifier's transfer to the 15,423 unknown-sign forms assumes that forms whose signs are unknown because of computational difficulty are statistically similar, on the coefficient features used, to forms whose signs are known.
Editorial extensions
If this is right
- If the central claim holds, Fricke signs for the 15,423 unknown forms can be assigned probable values using only stored coefficients, matching independent heuristics on the subset where the heuristic is confident.
- The accuracy advantage of full coefficient vectors over prime-indexed vectors (96.1% versus 86.2%) implies the predictive information is not contained only in $a_p$ for primes dividing the level, nor in primes generally; composite-index coefficients add real signal.
- Because LDA works without hyperparameter tuning, the sign is nearly linearly separable in coefficient space after parity normalization, suggesting a simple statistical description rather than a deep structural one.
- Because the model transfers across levels without being given the level, the learned boundary is a function of coefficient patterns rather than of the level or analytic conductor alone.
- The connection to murmurations suggests sign-conditioned coefficient averages are a stable phenomenon, not an artifact of a particular classifier.
Reading between the lines
- Because the root number is the product of parity and Fricke sign, a classifier with this accuracy also yields a root-number estimator; this could be used to prioritize candidates for rigorous certification.
- The finding that full coefficient vectors beat prime-indexed ones suggests a testable hypothesis: the signal lives in the multiplicative semigroup structure of indices, so features derived from divisor counts should retain most of the accuracy.
- Since unknown signs become more frequent at higher level (as the paper's own level-by-level plot shows), a level-stratified retraining experiment would be the cleanest way to test whether the transfer assumption holds.
- The same averaging-plus-classifier pipeline could be tried on other expensive invariants of automorphic forms, such as symmetry type or eigenvalue location, wherever a labeled subset exists.
Signed reviews
Editorial analysis
A structured set of objections, weighed in public.
Referee Report
Summary. The paper applies supervised machine learning, specifically Linear Discriminant Analysis (LDA) and feed-forward neural networks, to the first 1000 Fourier coefficients of Maass newforms from the LMFDB in order to predict the Fricke sign. On the 19,993 forms with rigorously known sign, LDA is reported to reach about 96% validation accuracy. The trained model is then applied to the 15,423 forms with unknown Fricke sign, and the resulting predictions are checked in two ways: by comparing averaged coefficient patterns ('murmurations') for predicted signs with those for known signs, and by comparing with heuristic Hejhal guesses on a subset of 4,595 forms, where agreement is about 95%. The paper also addresses the concern that the Fricke sign is directly encoded in coefficients at primes dividing the level by repeating experiments with a modified feature in which such coefficients are zeroed.
Significance. If the transfer claim holds, the paper provides a useful empirical data-scientific tool for guessing missing Fricke signs in the LMFDB and evidence that coefficient vectors contain recoverable information beyond the direct local-factor readout. The paper is commendably explicit about the direct-encoding confound and tests a zeroed feature variant, and it compares its predictions with an independent heuristic algorithm. The main weaknesses are the lack of uncertainty quantification and, more importantly, the absence of evidence that the 4,595-form Hejhal-confirmed subset is representative of the full unknown-sign set, which is load-bearing for the claim that predictions on all unknown-sign forms are reasonable.
major comments (4)
- [Abstract and Section 2.3/Table 2.2] The abstract states 96% (resp. 94%) accuracy for even (resp. odd) parity, but Section 2.3 and Table 2.2 report 94.9% for even forms and 96.3% for odd forms. Since these are the headline accuracy figures, the abstract should be corrected to match the table.
- [Sections 2.3, 2.7, Table 2.4, Figure 2.9] The application to the 15,423 unknown-sign forms is a central output of the paper, but the only direct validation is on the 4,595 forms for which the heuristic Hejhal algorithm converged. Figure 2.9 shows that the unknown fraction increases with level, and the convergence of a heuristic root-finding algorithm is likely easier at lower levels and for larger eigenvalue gaps. The paper does not report the level or spectral-parameter distribution of the 4,595 Hejhal-confirmed forms, nor accuracy stratified by level. Without such evidence, the claim that predictions on the full unknown-sign set are 'reasonable' is not established.
- [Section 2.4, Eq. (2.4)] The sentence 'Similarly, when gcd(n,N)>1, we have a_n = 0' is inconsistent with Eq. (2.4), which gives the nonzero value a_p = -w_p/sqrt(p) for each prime p dividing N. The intended statement is presumably that entries not rigorously computed are set to zero in the database. This matters because the a'_n experiment is the main evidence that the classifier learns beyond the direct encoding, so the exact zeroing rule (which indices are zeroed, and whether coefficients with mixed prime factors are handled correctly) must be stated precisely.
- [Tables 2.2-2.6 and Section 2.3] All accuracy figures are single-split point estimates with no confidence intervals, standard errors, or repeated-seed variation. The reported differences between feature sets, such as 0.9612 for a_n versus 0.9456 for a'_n, could be within sampling noise. Please provide bootstrap intervals or results over several random splits, especially for the comparison that supports the 'learning something more' claim.
minor comments (6)
- [Figures 2.1 and 2.2] The text says these figures provide 'clear evidence' of separation, but no quantitative measure is given; consider adding a simple statistic such as the L2 distance between the averaged coefficient sequences or the area between the curves.
- [Section 2.3] The training sizes 7772 and 5023 for even and odd forms are not tied to the described 80-20 splits; please clarify how these counts are obtained from Table 2.1.
- [Section 2.3] The use of Box's M test to 'satisfy' equal covariance is not rigorous: rejecting equality for only 33 of 1000 features is not the same as establishing equality, and the test is sensitive to sample size. A direct comparison of covariance matrices would be more appropriate.
- [Sections 2.3 and 2.6] The phrase 'without any hyperparameter tuning' appears in Section 2.3, but the neural network experiments in Section 2.6 use Adam with a learning rate of 1e-3 and 4e4 iterations; the claim should be restricted to the LDA experiments.
- [Section 2.7 and Table 2.5] There is a typo in the caption of Table 2.5: 'Maaass forms' should be 'Maass forms', and the space in 'F ricke sign' in the Section 2.3 heading should be removed.
- [General] No code, data, or reproducibility statement is included. Since the experiments are computational, please provide a link to a repository with the exact data-processing steps, random seeds, and model configurations, or at least specify the seed and split procedure.
Circularity Check
Acknowledged partial tautology in raw-coefficient LDA accuracy; central prediction claim remains independently supported.
-
self definitional
[Section 2.1, Eq. (2.4); Section 2.3 validation accuracy; Section 2.4 a'_n experiment]
"Given complete information about the coefficients, the Fricke sign is easily computable; on Γ0(N ) with N squarefree, the coefficient aN encodes the Fricke sign. ... If p divides the level N then we have (2.4) a_p = −w_p /√p. ... we trained the LDA using all available training data (12795 observations) and recorded 96 .1% accuracy on the validation data."
Under Eq. (2.4), the Fricke sign of a labeled form is directly recoverable from the feature vector a_n, because a_p = −w_p/√p for every p|N and w_N is the product of the local signs w_p. An LDA trained on the raw (a_n) vector can therefore reach high validation accuracy by reading the label out of the input features, so the 96.1% figure is partly true by construction. The paper explicitly anticipates this objection, defines a'_n = 0 when gcd(n,N)>1 (matching the unknown-sign data), and reports 94.56% accuracy, so the transfer-to-unknown-sign claim is not carried by the tautological features. This is a partial, acknowledged circularity in one headline number, not a load-bearing derivation step.
full rationale
The only genuine circularity in the paper is the raw-coefficient LDA accuracy: for p|N, a_p = −w_p/√p, so the sign is literally embedded in the input for those indices. The paper identifies this in Section 2.4, removes the direct encoding via a'_n, and obtains essentially unchanged accuracy (94.56% vs. 96.12%), which is real evidence that the classifier is learning more than the deterministic readout. The application to the 15,423 unknown-sign forms is supported by this a'_n experiment and by a separate comparison with Hejhal's heuristic on 4,595 forms, an external though heuristic benchmark. The murmuration checks on predicted labels are weaker, since a classifier trained on the same features will tend to reproduce separation in those features, but the paper presents them only as reasonableness checks, not as the load-bearing derivation. Prior-work citations are contextual and not used to prove the ML result. Therefore the central claim retains independent content, and the partial tautology is explicitly acknowledged; a score of 2 reflects that minor, non-load-bearing circularity rather than a serious defect.
Assumptions & free parameters
free parameters (3)
- LDA discriminant weights and threshold =
1000-dimensional coefficient vector plus threshold, values not reported
- Neural network weights and biases =
Architecture weights and biases, values not reported
- Per-feature normalization statistics for NN inputs =
Per-feature means and variances from the dataset
assumptions (5)
- domain assumption Selberg eigenvalue conjecture for the dataset: spectral parameter R satisfies lambda = 1/4 + R^2 with R real and nonnegative
- standard math Hecke multiplicativity and Atkin-Lehner local sign relations, including a_p = -w_p / sqrt(p) for p dividing the level
- domain assumption The LMFDB rigorous computations and labeled Fricke signs for the 19,993 known-sign forms are correct
- ad hoc to paper The 4,595 heuristic Hejhal outputs labeled 'probably correct' are accurate enough to serve as comparison ground truth
- domain assumption LDA's equal-covariance assumption holds for the feature distributions
Cite this review
Pith. "Pith review of Learning Fricke signs from Maass form Coefficients." pith.science (2026). https://pith.science/paper/KOZCKVTE
@misc{pith2026250102105,
author = {Pith},
title = {Pith review of: Learning Fricke signs from Maass form Coefficients},
year = {2026},
howpublished = {\url{https://pith.science/paper/KOZCKVTE}},
note = {Machine review of arXiv:2501.02105}
}
read the original abstract
In this paper, we conduct a data-scientific investigation of Maass forms. We find that averaging the Fourier coefficients of Maass forms with the same Fricke sign reveals patterns analogous to the recently discovered "murmuration" phenomenon, and that these patterns become more pronounced when parity is incorporated as an additional feature. Approximately 43% of the forms in our dataset have an unknown Fricke sign. For the remaining forms, we employ Linear Discriminant Analysis (LDA) to machine learn their Fricke sign, achieving 96% (resp. 94%) accuracy for forms with even (resp. odd) parity. We apply the trained LDA model to forms with unknown Fricke signs to make predictions. The average values based on the predicted Fricke signs are computed and compared to those for forms with known signs to verify the reasonableness of the predictions. Additionally, a subset of these predictions is evaluated against heuristic guesses provided by Hejhal's algorithm, showing a match approximately 95% of the time. We also use neural networks to obtain results comparable to those from the LDA model.
Figures
Figures from the paper (11 more)
Reference graph
Works this paper leans on
-
[1]
A. R. Booker, M. Lee, D. Lowry-Duda, A. Seymour-Howell, N. Zubrilina Murmurations of Maass forms , arXiv:2409.00765
-
[2]
A. R. Booker, A. Str\"ombergsson, A. Venkatesh. Effective computation of Maass cusp forms , International Mathematics Research Notices, 2006
work page 2006
-
[3]
Certification of Maass cusp forms of arbitrary level and character
K. Child Certification of Maass cusp forms of arbitrary level and character , arXiv:2204.11761
-
[4]
Friedlander, and Henryk Iwaniec
Duke, William, John B. Friedlander, and Henryk Iwaniec. The subconvexity problem for Artin L –functions, Inventiones Mathematicae 149 (2002): 489-577
work page 2002
- [5]
-
[6]
M. Kazalicki, D. Vlah Ranks of elliptic curves and deep neural networks , Research in Number Theory, 9(3), 2023
work page 2023
- [7]
-
[8]
He, K.-H
Y.-H. He, K.-H. Lee, and T. Oliver, Machine-learning the Sato--Tate conjecture , J. Symb. Comput. 111 (2022), 61--72
2022
Show all 16 references
-
[9]
, Machine-learning Number Fields , Mathematics, Computation and Geometry of Data 2(1) (2022), 49--66
2022
-
[10]
, Machine learning invariants of arithmetic curves , J. Symb. Comput. 115 , (2023), 478--491
2023
-
[11]
He, K.-H
Y.-H. He, K.-H. Lee, T. Oliver, and A. Pozdnyakov, Murmurations of elliptic curves , Experimental Mathematics (2024), 1--13
2024
-
[12]
D. A. Hejhal, On eigenfunctions of the Laplacian for Hecke triangle groups , in Emerging applications of number theory (Minneapolis, MN, 1996) , 291--315, IMA Vol. Math. Appl., 109
1996
-
[13]
Lowry-Duda and A
D. Lowry-Duda and A. Seymour-Howell, A Rigorous Implementation of Hejhal's Algorithm , forthcoming
-
[14]
Lowry-Duda, Heuristic Hejhal repository , Available at https://github.com/davidlowryduda/heuristic_hejhal, 2024
D. Lowry-Duda, Heuristic Hejhal repository , Available at https://github.com/davidlowryduda/heuristic_hejhal, 2024
2024
-
[15]
The LMFDB Collaboration, The L-functions and modular forms database , Available at https://www.lmfdb.org , 2024, [Online; accessed 27 November 2024]
2024
-
[16]
Seymour-Howell, A rigorous computation of Maass cusp forms of squarefree level , Res
A. Seymour-Howell, A rigorous computation of Maass cusp forms of squarefree level , Res. Number Theory 8 (2022). https://doi.org/10.1007/s40993-022-00393-y
2022 doi
Reviewed August 10, 2026 · model on record in the stance chip above.
Discussion (0). Continue with ORCID to comment.