You are currently browsing the tag archive for the ‘Paul Erdos’ tag.
I’ve just uploaded to the arXiv my paper “Local Bernstein theory, and lower bounds for Lebesgue constants“. This paper was initially motivated by a problem of Erdős} on Lagrange interpolation, but in the course of solving that problem, I ended up modifying some very classical arguments of Bernstein and his contemporaries (Boas, Duffin, Schaeffer, Riesz, etc.) to obtain “local” versions of these classical “Bernstein-type inequalities” that may be of independent interest.
Bernstein proved many estimates concerning the derivatives of polynomials, trigonometric polynomials, and entire functions of exponential type, but perhaps his most famous inequality in this direction is:
Lemma 1 (Bernstein’s inequality for trigonometric polynomials) Letbe a trigonometric polynomial of degree at most
, with
for all
. Then
for all
.
Similar inequalities concerning norms of derivatives of Littlewood-Paley components of functions are now ubiquitious in the modern theory of nonlinear dispersive PDE (where they are also called Bernstein estimates), but this will not be the focus of this current post.
A trigonometric polynomial of degree
is of exponential type
in the sense that
for complex
. Bernstein in fact proved a more general result:
Lemma 2 (Bernstein’s inequality for functions of exponential type) Letbe an entire function of exponential type at most
, with
for all
. Then
for all
.
There are several proofs of this lemma – see for instance this survey of Queffélec and Zarouf. In the case that is real-valued on
, there is a nice proof by Duffin and Schaeffer, which we sketch as follows. Suppose we normalize
, and adjust
by a suitable damping factor so that
actually decays slower than
as
. Then, for any
and
, one can use Rouche’s theorem to show that the function
has the same number of zeroes as
in a suitable large rectangle; but on the other hand one can use the intermediate value theorem to show that
has at least as many zeroes than
in the same rectangle. Among other things, this prevents double zeroes from occuring, which turns out to give the desired claim
after some routine calculations (in fact one obtains the stronger bound
for all real
).
The first main result of the paper is to obtain localized versions of Lemma 2 (as well as some related estimates). Roughly speaking, these estimates assert that if is holomorphic on a wide thin rectangle passing through the real axis, is bounded by
on the intersection of the real axis with this rectangle, and is “locally of exponential type” in the sense that it is bounded by
on the upper and lower edges of this rectangle (and obeys some very mild growth conditions on the remaining sides of this rectangle), then
can be bounded by
plus small errors on the real line, with some additional estimates away from the real line also available. The proof proceeds by a modification of the Duffin–Schaeffer argument, together with the two-constant theorem of Nevanlinna (and some standard estimates of harmonic measures on rectangles) to deal with the effect of the localization. (As a side note, this latter argument was provided to me by ChatGPT, as I was not previously aware of the Nevanlinna two-constant theorem.)
Once one localizes this “Bernstein theory”, it becomes suitable for the analysis of (real-rooted, monic) polynomials of a high degree
, which are not bounded globally on
(and grow polynomially rather than exponentially at infinity), but which can exhibit “local exponential type” behavior on various intervals, particularly in regions where the logarithmic potential
This becomes relevant in the theory of Lagrange interpolation. Recall that if are real numbers and
is a polynomial of degree less than
then one has the interpolation formula
If one chooses the interpolation points poorly, then the Lebesgue constant can be extremely large. However, if one selects these points to be the roots of the aforementioned monic Chebyshev polynomials, then it is known that
for all fixed intervals
in
. In the case
, it was shown by Erdős} that this is the best possible value of the Lebesgue constant up to
errors for interpolation on
, thus
In terms of the monic polynomial , these two estimates can be written as
Problem 3 Letbe a trigonometric polynomial of degree
with
roots
in
.
It is easy to check that the lower bounds of and
are sharp by considering the case when
is a sinusoid
.
The bound (3) is immediate from Bernstein’s inequality (Lemma 1). By applying a local version of this inequality, I was able to get a weak version of the claim (1) in which was replaced with
; see this early version of the paper, which was developed through conversations with Nat Sothanaphan and Aron Bhalla. By combining this argument with ideas from the older work of Erdős}, I was able to establish (1).
The bound (2) took me longer to establish, and involved a non-trivial amount of playing around with AI tools, the story of which I would like to share here. I had discovered the toy problem (4), but initially was not able to establish this inequality; AlphaEvolve seemed to confirm it numerically (with sinusoids appearing to be the extremizer), but did not offer direct clues on how to prove this rigorously. At some point I realized that the left-hand side factorized into the expressions and
, and tried to bound these expressions separately. Perturbing around a sinusoid
, I was able to show that the
norm
was a local minimum as long as one only perturbed by lower order Fourier modes, keeping the frequency
coefficients unchanged. Guessing that this local minimum was actually a global minimum, this led me to conjecture the general lower bound
A basic problem in sieve theory is to understand what happens when we start with the integers (or some subinterval of the integers) and remove some congruence classes
for various moduli
. Here we shall concern ourselves with the simple setting where we are sieving the entire integers rather than an interval, and are only removing a finite number of congruence classes
. In this case, the set of integers that remain after the sieving is periodic with period
, so one work without loss of generality in the cyclic group
. One can then ask: what is the density of the sieved set
In this blog post I would like to note one simple fact, due to Rogers, that one can say about this problem:
Theorem 1 (Rogers’ theorem) For fixed, the density of the sieved set is maximized when all the
vanish. Thus,
Example 2 If one sieves out,
, and
, then only
remains, giving a density of
. On the other hand, if one sieves out
,
, and
, then the remaining elements are
and
, giving the larger density of
.
This theorem is somewhat obscure: its only appearance in print is in pages 242-244 of this 1966 text of Halberstam and Roth, where the authors write in a footnote that the result is “unpublished; communicated to the authors by Professor Rogers”. I have only been able to find it cited in three places in the literature: in this 1996 paper of Lewis, in this 2007 paper of Filaseta, Ford, Konyagin, Pomerance, and Yu (where they credit Tenenbaum for bringing the reference to their attention), and is also briefly mentioned in this 2008 paper of Ford. As far as I can tell, the result is not available online, which could explain why it is rarely cited (and also not known to AI tools). This became relevant recently with regards to Erdös problem 281, posed by Erdös and Graham in 1980, which was solved recently by Neel Somani through an AI query by an elegant ergodic theory argument. However, shortly after this solution was located, it was discovered by KoishiChan that Rogers’ theorem reduced this problem immediately to a very old result of Davenport and Erdös from 1936. Apparently, Rogers’ theorem was so obscure that even Erdös was unaware of it when posing the problem!
Modern readers may see some similarities between Rogers’ theorem and various rearrangement or monotonicity inequalites, suggesting that the result may be proven by some sort of “symmetrization” or “compression” method. This is indeed the case, and is basically Rogers’ original proof. We can modernize a bit as follows. Firstly, we can abstract into a finite cyclic abelian group
, with residue classes now becoming cosets of various subgroups of
. We can take complements and restate Rogers’ theorem as follows:
Theorem 3 (Rogers’ theorem, again) Letbe cosets of a finite cyclic abelian group
. Then
Example 4 Take,
,
, and
. Then the cosets
,
, and
cover the residues
, with a cardinality of
; but the subgroups
cover the residues
, having the smaller cardinality of
.
Intuitively: “sliding” the cosets together reduces the total amount of space that these cosets occupy. As pointed out in comments, the requirement of cyclicity is crucial; four lines in a finite affine plane already suffice to be a counterexample otherwise.
By factoring the cyclic group into -groups, Rogers’ theorem is an immediate consequence of two observations:
Theorem 5 (Rogers’ theorem for cyclic groups of prime order) Rogers’ theorem holds whenfor some prime power
.
Theorem 6 (Rogers’ theorem preserved under products) If Rogers’ theorem holds for two finite abelian groupsof coprime orders, then it also holds for the product
.
The case of cyclic groups of prime order is trivial, because the subgroups of are totally ordered. In this case
is simply the largest of the
, which has the same size as
and thus has lesser or equal cardinality to
.
The preservation of Rogers’ theorem under products is also routine to verify. By the coprime orders of and standard group theoretic arguments (e.g., Goursat’s lemma, the Schur–Zassenhaus theorem, or the classification of finite abelian groups), one can see that any subgroup
of
splits as a direct product
of subgroups of
respectively, so the cosets
also split as
Thomas Bloom’s erdosproblems.com site hosts nearly a thousand questions that originated, or were communicated by, Paul Erdős, as well as the current status of these questions (about a third of which are currently solved). The site is now a couple years old, and has been steadily adding features, the most recent of which has been a discussion forum for each individual question. For instance, a discussion I had with Stijn Cambie and Vjeko Kovac on one of these problems recently led to it being solved (and even formalized in Lean!).
A significantly older site is the On-line Encyclopedia of Integer Sequences (OEIS), which records hundreds of thousands of integer sequences that have some mathematician has encountered at some point. It is a highly useful resource, enabling researchers to discover relevant literature for a given problem so long as they can calculate enough of some integer sequence that is “canonically” attached to that problem that they can search for it in the OEIS.
A large fraction of problems in the Erdos problem webpage involve (either explicitly or implicitly) some sort of integer sequence – typically the largest or smallest size of some
-dependent structure (such as a graph of
vertices, or a subset of
) that obeys a certain property. In some cases, the sequence is already in the OEIS, and is noted in the Erdos problem web page. But in a large number of cases, the sequence either has not yet been entered into the OEIS, or it does appear but has not yet been noted on the Erdos web page.
Thomas Bloom and I are therefore proposing a crowdsourced project to systematically compute the hundreds of sequences associated to the Erdos problems and cross-check them against the OEIS. We have created a github repository to coordinate this process; as a by-product, this repository will also be tracking other relevant statistics about the Erdos problem website, such as the current status of formalizing the statements of these problems in the Formal Conjectures Repository.
The main feature of our repository is a large table recording the current status of each Erdos problem. For instance, Erdos problem #3 is currently listed as open, and additionally has the status of linkage with the OEIS listed as “possible”. This means that there are one or more sequences attached to this problem which *might* already be in the OEIS, or would be suitable for submission to the OEIS. Specifically, if one reads the commentary for that problem, one finds mention of the functions for
, defined as the size of the largest subset of
without a
-term progression. It is likely that several of the sequences
,
, etc. are in the OEIS, but it is a matter of locating them, either by searching for key words, or by calculating the first few values of these sequences and then looking for a match. (EDIT: a contributor has noted that the first foursequences appear as A003002, A003003, A003004, and A003005 in the OEIS, and the table has been updated accordingly.)
We have set things up so that new contributions (such as the addition of an OEIS number to the table) can be made by a Github pull request, specifically to modify this YAML file. Alternatively, one can create a Github issue for such changes, or simply leave a comment either on the appropriate Erdos problem forum page, or here on this blog.
Many of the sequences do not require advanced mathematical training to compute, and so we hope that this will be a good “citizen mathematics” project that can bring in the broader math-adjacent community to contribute to research-level mathematics problems, by providing experimental data, and potentially locating relevant references or connections that would otherwise be overlooked. This may also be a use case for AI assistance in mathematics through generating code to calculate the sequences in question, although of course one should always stay mindful of potential bugs or hallucinations in any AI-generated code, and find ways to independently verify the output. (But if the AI-generated sequence leads to a match with an existing sequence in the OEIS that is clearly relevant to the problem, then the task has been successfully accomplished, and no AI output needs to be directly incorporated into the database in such cases.)
This is an experimental project, and we may need to adjust the workflow as the project progresses, but we hope that it will be successful and lead to further progress on some fraction of these problems. The comment section of this blog can be used as a general discussion forum for the project, while the github issue page and the erdosproblems.com forum pages can be used for more specialized discussions of specific problems.
First things first: due to an abrupt suspension of NSF funding to my home university of UCLA, the Institute of Pure and Applied Mathematics (which had been preliminarily approved for a five-year NSF grant to run the institute) is currently fundraising to ensure continuity of operations during the suspension, with a goal of raising $500,000. Donations can be made at this page. As incoming Director of Special Projects at IPAM, I am grateful for the support (both moral and financial) that we have already received in the last few days, but we are still short of our fundraising goal.
Back to math. Ayla Gafni and I have just uploaded to the arXiv the paper “Rough numbers between consecutive primes“. In this paper we resolve a question of Erdös concerning rough numbers between consecutive gaps, and with the assistance of modern sieve theory calculations, we in fact obtain quite precise asymptotics for the problem. (As a side note, this research was supported by my personal NSF grant which is also currently suspended; I am grateful to recent donations to my own research fund which have helped me complete this research.)
Define a prime gap to be an interval between consecutive primes. We say that a prime gap contains a rough number if there is an integer
whose least prime factor is at least the length
of the gap. For instance, the prime gap
contains the rough number
, but the prime gap
does not (all integers between
and
have a prime factor less than
). The first few
for which the
prime gap contains a rough number are
Erdös initially thought that all but finitely many prime gaps should contain a rough number, but changed his mind, as per the following quote:
…I am now sure that this is not true and I “almost” have a counterexample. Pillai and Szekeres observed that for every , a set of
consecutive integers always contains one which is relatively prime to the others. This is false for
, the smallest counterexample being
. Consider now the two arithmetic progressions
and
. There certainly will be infinitely many values of
for which the progressions simultaneously represent primes; this follows at once from hypothesis H of Schinzel, but cannot at present be proved. These primes are consecutive and give the required counterexample. I expect that this situation is rather exceptional and that the integers
for which there is no
satisfying
and
have density
.
In fact Erdös’s observation can be made simpler: any pair of cousin primes for
(of which
is the first example) will produce a prime gap that does not contain any rough numbers.
The latter question of Erdös is listed as problem #682 on Thomas Bloom’s Erdös problems website. In this paper we answer Erdös’s question, and in fact give a rather precise bound for the number of counterexamples:
Theorem 1 (Erdos #682) For, let
be the number of prime gaps
with
that do not contain a rough number. Then
Assuming the Dickson–Hardy–Littlewood prime tuples conjecture, we can improve this to
for some (explicitly describable) constant
.
In fact we believe that , although the formula we have to compute
converges very slowly. This is (weakly) supported by numerical evidence:
While many questions about prime gaps remain open, the theory of rough numbers is much better understood, thanks to modern sieve theoretic tools such as the fundamental lemma of sieve theory. The main idea is to frame the problem in terms of counting the number of rough numbers in short intervals , where
ranges in some dyadic interval
and
is a much smaller quantity, such as
for some
. Here, one has to tweak the definition of “rough” to mean “no prime factors less than
” for some intermediate
(e.g.,
for some
turns out to be a reasonable choice). These problems are very analogous to the extremely well studied problem of counting primes in short intervals, but one can make more progress without needing powerful conjectures such as the Hardy–Littlewood prime tuples conjecture. In particular, because of the fundamental lemma of sieve theory, one can compute the mean and variance (i.e., the first two moments) of such counts to high accuracy, using in particular some calculations on the mean values of singular series that go back at least to the work of Montgomery from 1970. This second moment analysis turns out to be enough (after optimizing all the parameters) to answer Erdös’s problem with a weaker bound
Vjeko Kovac and I have just uploaded to the arXiv our paper “On several irrationality problems for Ahmes series“. This paper resolves (or at least makes partial progress on) some open questions of Erdős and others on the irrationality of Ahmes series, which are infinite series of the form for some increasing sequence
of natural numbers. Of course, since most real numbers are irrational, one expects such series to “generically” be irrational, and we make this intuition precise (in both a probabilistic sense and a Baire category sense) in our paper. However, it is often difficult to establish the irrationality of any specific series. For example, it is already a non-trivial result of Erdős that the series
is irrational, while the irrationality of
(equivalent to Erdős problem #69) remains open, although very recently Pratt established this conditionally on the Hardy–Littlewood prime tuples conjecture. Finally, the irrationality of
(Erdős problem #68) is completely open.
On the other hand, it has long been known that if the sequence grows faster than
for any
, then the Ahmes series is necessarily irrational, basically because the fractional parts of
can be arbitrarily small positive quantities, which is inconsistent with
being rational. This growth rate is sharp, as can be seen by iterating the identity
to obtain a rational Ahmes series of growth rate
for any fixed
.
In our paper we show that if grows somewhat slower than the above sequences in the sense that
, for instance if
for a fixed
, then one can find a comparable sequence
for which
is rational. This partially addresses Erdős problem #263, which asked if the sequence
had this property, and whether any sequence of exponential or slower growth (but with
convergent) had this property. Unfortunately we barely miss a full solution of both parts of the problem, since the condition
we need just fails to cover the case
, and also does not quite hold for all sequences going to infinity at an exponential or slower rate.
We also show the following variant; if has exponential growth in the sense that
with
convergent, then there exists nearby natural numbers
such that
is rational. This answers the first part of Erdős problem #264 which asked about the case
, although the second part (which asks about
) is slightly out of reach of our methods. Indeed, we show that the exponential growth hypothesis is best possible in the sense a random sequence
that grows faster than exponentially will not have this property, this result does not address any specific superexponential sequence such as
, although it does apply to some sequence
of the shape
.
Our methods can also handle higher dimensional variants in which multiple series are simultaneously set to be rational. Perhaps the most striking result is this: we can find an increasing sequence of natural numbers with the property that
is rational for every rational
(excluding the cases
to avoid division by zero)! This answers (in the negative) a question of Stolarsky Erdős problem #266, and also reproves Erdős problem #265 (and in the latter case one can even make
grow double exponentially fast).
Our methods are elementary and avoid any number-theoretic considerations, relying primarily on the countable dense nature of the rationals and an iterative approximation technique. The first observation is that the task of representing a given number as an Ahmes series
with each
lying in some interval
(with the
disjoint, and going to infinity fast enough to ensure convergence of the series), is possible if and only if the infinite sumset
Proposition 1 (Iterative approximation) Letbe a Banach space, let
be sets with each
contained in the ball of radius
around the origin for some
with
convergent, so that the infinite sumset
is well-defined. Suppose that one has some convergent series
in
, and sets
converging in norm to zero, such that
for all
. Then the infinite sumset
contains
.
Informally, the condition (2) asserts that occupies all of
“at the scale
“.
Proof: Let . Our task is to express
as a series
with
. From (2) we may write
In one dimension, sets of the form are dense enough that the condition (2) can be satisfied in a large number of situations, leading to most of our one-dimensional results. In higher dimension, the sets
lie on curves in a high-dimensional space, and so do not directly obey usable inclusions of the form (2); however, for suitable choices of intervals
, one can take some finite sums
which will become dense enough to obtain usable inclusions of the form (2) once
reaches the dimension of the ambient space, basically thanks to the inverse function theorem (and the non-vanishing curvatures of the curve in question). For the Stolarsky problem, which is an infinite-dimensional problem, it turns out that one can modify this approach by letting
grow slowly to infinity with
.
I’ve just uploaded to the arXiv my paper “Planar point sets with forbidden -point patterns and few distinct distance“. This (very) short paper was a byproduct of my recent explorations of the Erdös problem website in recent months, with a vague emerging plan to locate a suitable problem that might be suitable for some combination of a crowdsourced “Polymath” style project and/or a test case for emerging AI tools. The question below was one potential candidate; however, upon reviewing the literature on the problem, I noticed that the existing techniques only needed one additional tweak to fully resolve the problem. So I ended up writing this note instead to close off the problem.
I’ve arranged this post so that this additional trick is postponed to below the fold, so that the reader can, if desired, try to guess for themselves what the final missing ingredient needed to solve the problem was. Here is the problem (Erdös problem #135), which was asked multiple times by Erdös over more than two decades (and who even offered a small prize for the solution on one of these occasions):
Problem 1 (Erdös #135) Letbe a set of
points such that any four points in the set determine at least five distinct distances. Must
determine
many distances?
This is a cousin of the significantly more famous Erdös distinct distances problem (Erdös problem #89), which asks what is the minimum number of distances determined by a set of
points in the plane, without the restriction on four-point configurations. The example of a square grid
(assuming for sake of argument that
is a perfect square), together with some standard analytic number theory calculations, shows that
can determine
distances, and it is conjectured that this is best possible up to constants. A celebrated result of Guth and Katz, discussed in this previous blog post, shows that
will determine at least
distances. Note that the lower bound
here is far larger, and in fact comparable to the total number
of distances available, thus expressing the belief that the “local” condition that every four points determine at least five distances forces the global collection distances to be almost completely distinct. In fact, in one of the papers posing the problem, Erdös made the even stronger conjecture that the set
must contain a subset
of cardinality
for which all the
distances generated by
are distinct.
A paper of Dumitrescu came close to resolving this problem. Firstly, the number of ways in which four points could fail to determine five distinct distances was classified in that paper, with the four-point configurations necessarily being one of the following eight patterns:
-
: An equilateral triangle plus an arbitrary vertex.
-
: A parallelogram.
-
: An isosceles trapezoid (four points on a line,
, where
, form a degenerate isosceles trapezoid).
-
: A star with three edges of the same length.
-
: A path with three edges of the same length.
-
: A kite.
-
: An isosceles triangle plus an edge incident to a base endpoint, and whose length equals the length of the base.
-
: An isosceles triangle plus an edge incident to the apex, and whose length equals the length of the base.
Given that the grid determine only
distances, one could seek a counterexample to this by finding a set of
points in the grid
that avoided all of the eight patterns
.
Dumitrescu then counted how often each of the patterns occured inside the grid
. The answer is:
-
does not occur at all. (This is related to the irrationality of
.)
-
occurs
times.
-
occurs
times.
-
occurs
times.
-
occurs
times.
-
occurs
times.
-
occurs
times.
-
occurs
times.
Using this and a standard probabilistic argument, Dumitrescu then established the following “near miss” to a negative answer to the above problem:
Theorem 2 (First near miss) Ifis sufficiently large, then there exists a subset of
of cardinality
which avoids all of the patterms
.
In particular, this generates a set of points with
distances that avoids seven out of the eight required forbidden patterns; it is only the parallelograms
that are not avoided, and are the only remaining obstacle to a negative answer to the problem.
Proof: Let be a small constant, and let
be a random subset of
, formed by placing each element of
with an independent probability of
. A standard application of Hoeffding’s inequality (or even the second moment method) shows that this set
will have cardinality
with high probability if
is large enough. On the other hand, each of the
patterns
has a probability
of lying inside
, so by linearity of expectation, the total number of such patterns inside
is
on the average. In particular, by Markov’s inequality, we can find a set
of cardinality
with only
such patterns. Deleting all of these patterns from
, we obtain a set
of cardinality
, which is
if
is a sufficiently small constant. This establishes the claim.
Unfortunately, this random set contains far too many parallelograms (
such parallelograms, in fact) for this deletion argument to work. On the other hand, in earlier work of Thiele and of Dumitrescu, a separate construction of a set of
points in
that avoids all of the parallelograms
was given:
Theorem 3 (Second near miss) Forlarge, there exists a subset
of
of cardinality
which contains no parallelograms
. Furthermore, this set is in general position: no three points in
are collinear, and no four are concyclic. As a consequence, this set
in fact avoids the three patterns
(the pattern in
is concyclic, and the pattern
does not occur at all in the grid).
Proof: One uses an explicit algebraic construction, going back to an old paper of Erdös and Turán involving constructions of Sidon sets. Namely, one considers the set is a prime between
and
(the existence of which is guaranteed by Bertrand’s postulate). Standard Gauss sum estimates can be used to show that
has cardinality
. If
contained four points that were in a parallelogram or on a circle, or three points in a line, then one could lift up from
to the finite field plane
and conclude that the finite field parabola
also contained four points in a parallelogram or a circle, or three points on a line. But straightforward algebraic calculations can be performed to show that none of these scenarios can occur. For instance, if
were four points on a parallelogram that were contained in a parabola, this would imply that an alternating sum of the form
Given that we have one “near-miss” in the literature that avoids , and another “near-miss” that avoids
, it is natural to try to combine these two constructions to obtain a set that avoids all eight patterns
. This inspired the following problem of Dumitrescu (see Problem 2 of this paper):
Problem 4 Does the setin (1) contain a subset of cardinality
that avoids all eight of the patterns
?
Unfortunately, this problem looked difficult, as the number-theoretic task of counting the patterns in
looked quite daunting.
This ends the survey of the prior literature on this problem. Can you guess the missing ingredient needed to resolve the problem? I will place the answer below the fold.
The Erdös problem site was created last year, and announced earlier this year on this blog. Every so often, I have taken a look at a random problem from the site for fun. A few times, I was able to make progress on one of the problems, leading to a couple papers; but the more common outcome is that I play around with the problem for a while, see why the problem is difficult, and then eventually give up and do something else. But, as is common in this field, I don’t make public the observations that I made, and the next person who looks at the same problem would likely have to go through the same process of trial and error to work out what the main obstructions that are present are.
So, as an experiment, I thought I would record here my preliminary observations on one such problem – Erdös problem #385 – to discuss why it looks difficult to solve with our current understanding of the primes. Here is the problem:
Problem 1 (Erdös Problem #385) Letwhere
is the least prime divisor of
. Is it true that
for all sufficiently large
? Does
as
?
This problem is mentioned on page 73 of this 1979 paper of Erdös (where he attributes the problem to an unpublished work of Eggelton, Erdös, and Selfridge that, to my knowledge, has never actually appeared), as well as briefly in page 92 of this 1980 paper of Erdös and Graham.
At first glance, this looks like a somewhat arbitrary problem (as many of Erdös’s problems initially do), as the function is not obviously related to any other well-known function or problem. However, it turns out that this problem is closely related to the parity barrier in sieve theory (as discussed in this previous post), with the possibility of Siegel zeroes presenting a particular obstruction. I suspect that Erdös was well aware of this connection; certainly he mentions the relation with questions on gaps between primes (or almost primes), which is in turn connected to the parity problem and Siegel zeroes (as is discussed recently in my paper with Banks and Ford, and in more depth in these papers of Ford and of Granville).
Let us now explore the problem further. Let us call a natural number bad if
, so the first part of the problem is asking whether there exist bad numbers that are sufficiently large. We unpack the definitions:
is bad if and only if
for any composite
, so placing
in intervals of the form
we are asking to show that
It is now natural to try to understand this problem for a specific choice of interval as a function of
. If
is large in the sense that
, then the claimed covering property is automatic, since every composite number less than or equal to
has a prime factor less than or equal to
. On the other hand, for
very small, in particular
, it is also possible to find
with this property. Indeed, if one takes
to lie in the residue class
, then we see that the residue classes cover all of
except for
, and from Linnik’s theorem we can ensure that
is prime. Thus, to rule out bad numbers, we need to understand the covering problem at intermediate scales
.
A key case is when for some
. Here, the residue classes
for
sieve out everything in
except for primes and semiprimes, and specifically the semiprimes that are product of two primes between
and
. If one can show for some
that the largest gap between semiprimes in say
with prime factors in
is
, then this would affirmatively answer the first part of this problem (and also the second). This is certainly very plausible – it would follow from a semiprime version of the Cramér conjecture (and this would also make the more precise prediction
) – but remains well out of reach for now. Even assuming the Riemann hypothesis, the best upper bound on prime gaps in
is
, and the best upper bound on semiprime gaps is not significantly better than this – in particular, one cannot reach
for any
. (There is a remote possibility that an extremely delicate analysis near
, together with additional strong conjectures on the zeta function, such as a sufficiently quantitative version of the GUE hypothesis, may barely be able to resolve this problem, but I am skeptical of this, absent some further major breakthrough in analytic number theory.)
Given that multiplicative number theory does not seem powerful enough (even on RH) to resolve these problems, the other main approach would be to use sieve theory. In this theory, we do not really know how to exploit the specific location of the interval or the specific congruence classes used, so one can study the more general problem of trying to cover an interval
of length
by one residue class mod
for each
, and only leaving a small number of survivors which could potentially be classified as “primes”. The discussion of the small
case already reveals a problem with this level of generality: one can sieve out the interval
by the residue classes
for
, and leave only one survivor,
. Indeed, thanks to known bounds on Jacobsthal’s function, one can be more efficient than this; for instance, using equation (1.2) from this paper of Ford, Green, Konyagin, Maynard, and myself, it is possible to completely sieve out any interval of sufficiently large length
using only those primes
up to
. On the other hand, from the work of Iwaniec, we know that sieving up to
is insufficient to completely sieve out such an interval; related to this, if one only sieves up to
for some
, the linear sieve (see e.g., Theorem 2 of this previous blog post) shows that one must have at least
survivors, where
can be given explicitly in the regime
by the formula
These lower bounds are not believed to be best possible. For instance, the Maier–Pomerance conjecture on Jacobsthal’s function would indicate that one needs to sieve out primes up to in order to completely sieve out an interval of length
, and it is also believed that sieving up to
should leave
survivors, although even these strong conjectures are not enough to positively resolve this problem, since we are permitted to sieve all the way up to
(and we are allowed to leave every prime number as a survivor, which in view of the Brun–Titchmarsh theorem could permit as many as
survivors).
Unfortunately, as discussed in this previous blog post, the parity problem blocks such improvements from taking place from most standard analytic number theory methods, in particular sieve theory. A particularly dangerous enemy arises from Siegel zeroes. This is discussed in detail in the papers of of Ford and of Granville mentioned previously, but an informal discussion is as follows. If there is a Siegel zero associated to the quadratic character of some conductor , this roughly speaking means that almost all primes
(in certain ranges) will be quadratic non-residues mod
. In particular, if one restricts attention to numbers
in a residue class
that is a quadratic residue, we then expect most numbers in this class to have an even number of prime factors, rather than an odd number.
This alters the effect of sieving in such residue classes. Consider for instance the classical sieve of Eratosthenes. If one sieves out for each prime
, the sieve of Eratosthenes tells us that the surviving elements of
are simply the primes between
and
, of which there are about
many. However, if one restricts attention to
for a quadratic residue class
(and taking
to be somewhat large compared to
), then by the preceding discussion, this eliminates most primes, and so now sieving out
should leave almost no survivors. Shifting this example by
and then dividing by
, one can end up with an example of an interval
of length
that can be sieved by residue classes
for each
in such a manner as to leave almost no survivors (in particular,
many). In the presence of a Siegel zero, it seems quite difficult to prevent this scenario from “infecting” the above problem, creating a bad scenario in which for all
, the residue classes
for
already eliminate almost all elements of
, leaving it mathematically possible for the remaining survivors to either be prime, or eliminated by the remaining residue classes
for
.
Because of this, I suspect that it will not be possible to resolve this Erdös problem without a major breakthrough on the parity problem that (at a bare minimum) is enough to exclude the possibility of Siegel zeroes existing. (But it is not clear at all that Siegel zeroes are be the only “enemy” here, so absent a major advance in “inverse sieve theory”, one cannot simply assume GRH to run away from this problem).
— 0.1. Addendum: heuristics for Siegel zero scenarios —
This post also provides a good opportunity to refine some heuristics I had previously proposed regarding Siegel zeroes and their impact on various problems in analytic number theory. In this previous blog post, I wrote
“The parity problem can also be sometimes be overcome when there is an exceptional Siegel zero … [this] suggests that to break the parity barrier, we may assume without loss of generality that there are no Siegel zeroes.”
On the other hand, it was pointed out in a more recent article of Granville that (as with the current situation), Siegel zeroes can sometimes serve to enforce the parity barrier, rather than overcome it, and responds to my previous statement with the comment “this claim needs to be treated with caution, since its truth depends on the context”.
I actually agree with Granville here, and I propose here a synthesis of the two situations. In the absence of a Siegel zero, standard heuristic models in analytic number theory (such as the ones discussed in this post) typically suggest that a given quantity of interest in number theory (e.g., the number of primes in a certain set) obey an asymptotic law of the form
However, the presence of a Siegel zero tends to “magnetize” the error term by pulling most of the fluctuations in a particular direction. In many situations, what this means is that one can obtain a refined asymptotic of the form
The implications of this refined asymptotic then depend rather crucially on how the Siegel correction term is aligned with the main term, and also whether it is of comparable order or lower order. In many situations (particularly those concerning “average case” problems, in which one wants to understand the behavior for typical choices of parameters), the Siegel correction term ends up being lower order, and so one ends up with the situation described in my initial blog post, where we are able to get the predicted asymptotic in the Siegel zero case. However, as pointed out by Granville, there are other situations (particularly those involving “worst case” problems, in which some key parameter can be chosen adversarially) in which the Siegel correction term can align to completely cancel (or to highly reinforce) the main term. In such cases, the Siegel zero becomes a very concrete manifestation of the parity barrier, rather than a means to avoid it. (There is a tiny chance that there may be some sort of “repulsion” phenomenon in which having no semiprimes in
for one value of
somehow generates semiprimes in
for another value of
, which would allow one to solve the problem without having to directly address the Siegel issue, but I don’t see how two such intervals could “communicate” in order to achieve such a repulsion effect.)
The following problem was posed by Erdös and Graham (and is listed as problem #437 on the Erdös problems website):
Problem 1 Letbe integers. How many of the partial products
,
,
,
can be squares? Is it true that, for any
, there can be more than
squares?
If one lets denote the maximal number of squares amongst such partial products, it was observed in the paper of Erdös and Graham that the bound
is “trivial” (no proof was provided, but one can for instance argue using the fact that the number of integer solutions to hyperelliptic equations of the form
for fixed
is quite sparse, and in fact finite for
thanks to Siegel’s theorem), and the problem then asks if
.
It turns out that this problem was essentially solved (though not explicitly) by a recently published paper of Bui, Pratt, and Zaharescu, who studied a closely related quantity introduced by Erdös, Graham, and Selfridge (see also Problem B30 of Guy’s book), defined for any natural number
as the least natural number
such that some subset of
, when multiplied together with
, produced a square. Among the several results proven about
in that paper was the following:
Theorem 2 (Bui–Pratt–Zaharescu, Theorem 1.2) Forsufficiently large, there exist
integers
such that
.
The arguments were in fact quite elementary, with the main tool being the theory of smooth numbers (the theory of hyperelliptic equations is used elsewhere in the paper, but not for this particular result).
If one uses this result as a “black box”, then an easy greedy algorithm argument gives the lower bound
Theorem 3 (Bounds for) As
, we have the lower bound
and the upper bound
In particular, for any
, one has
for sufficiently large
.
The purpose of this blog post is to record this modification of the argument, which is short enough to present immediately. For a large , let
denote the quantity
To prove the lower bound on , which is a variant of Theorem 2. The key observation is that given any
-smooth numbers
, some non-trivial subcollection of them will multiply to a square. This is essentially Lemma 4.2 of Bui–Pratt–Zaharescu, but for the convenience of the reader we give a full proof here. Consider the multiplicative homomorphism
defined by
From (1), (2) we can find sequences of
-smooth numbers
in
, with each sequence being to the right of the previous sequence. By the above observation, each sequence contains some non-trivial subcollection that multiplies to a square. Concatenating all these subsequences together, we obtain a single sequence
with at least
partial products multiplying to a square, giving the desired lower bound on
.
Next, we prove the upper bound on . Suppose that a sequence
has
partial products
that are squares for some
. Then we have
a square for all
(with the convention
). The key observation (essentially Lemma 3.4 of Bui–Pratt–Zaharescu) is that, for each
, one of the following must hold:
- (i) At least one of the
is
-smooth.
- (ii) At least one of the
is divisible by
for some prime
.
- (iii)
.
From (1) we see that the number of for which (i) occurs is at most
. From the union bound we see that the number of
for which (ii) occurs is at most
The upper bound arguments seem more crude to the author than the lower bound arguments, so I conjecture that the lower bound is in fact the truth: .
I’ve just uploaded to the arXiv my paper “Dense sets of natural numbers with unusually large least common multiples“. This short paper answers (in the negative) a somewhat obscure question of Erdős and Graham:
Problem 1 Is it true that ifis a set of natural numbers for which
goes to infinity as
, then the quantity
also goes to infinity as
?
At first glance, this problem may seem rather arbitrary, but it can be motivated as follows. The hypothesis that (1) goes to infinity is a largeness condition on ; in view of Mertens’ theorem, it can be viewed as an assertion that
is denser than the set of primes. On the other hand, the conclusion that (2) grows is an assertion that
becomes significantly larger than
on the average for large
; that is to say, that many pairs of numbers in
share a common factor. Intuitively, the problem is then asking whether sets that are significantly denser than the primes must start having lots of common factors on average.
For sake of comparison, it is easy to see that if (1) goes to infinity, then at least one pair of distinct elements in
must have a non-trivial common factor. For if this were not the case, then the elements of
are pairwise coprime, so each prime
has at most one multiple in
, and so can contribute at most
to the sum in (1), and hence by Mertens’ theorem, and the fact that every natural number greater than one is divisible by at least one prime
, the quantity (1) stays bounded, a contradiction.
It turns out, though, that the answer to the above problem is negative; one can find sets that are denser than the primes, but for which (2) stays bounded, so that the least common multiples in the set are unusually large. It was a bit surprising to me that this question had not been resolved long ago (in fact, I was not able to find any prior literature on the problem beyond the original reference of Erdős and Graham); in contrast, another problem of Erdős and Graham concerning sets with unusually small least common multiples was extensively studied (and essentially solved) about twenty years ago, while the study of sets with unusually large greatest common divisor for many pairs in the set has recently become somewhat popular, due to their role in the proof of the Duffin-Schaeffer conjecture by Koukoulopoulos and Maynard.
To search for counterexamples, it is natural to look for numbers with relatively few prime factors, in order to reduce their common factors and increase their least common multiple. A particularly simple example, whose verification is on the level of an exercise in a graduate analytic number theory course, is the set of semiprimes (products of two primes), for which one can readily verify that (1) grows like but (2) stays bounded. With a bit more effort, I was able to optimize the construction and uncover the true threshold for boundedness of (2), which was a little unexpected:
Theorem 2
The proofs are not particularly long or deep, but I thought I would record here some of the process towards finding them. My first step was to try to simplify the condition that (2) stays bounded. In order to use probabilistic intuition, I first expressed this condition in probabilistic terms as
It is then natural to try a random construction, in which one sieves out the natural numbers by permitting each natural number to survive with a probability resembling
, in order to get the predicted behavior for
. Performing some standard calculations, this construction could ensure (2) bounded with a density a little bit less than the one stated in the main theorem; after optimizing the parameters, I could only get something like
It then remained to improve the lower bound construction to eliminate the losses in the exponents. By deconstructing the proof of the upper bound, it became natural to consider something like the set of natural numbers
that had at most
prime factors. This construction actually worked for some scales
– namely those
for which
was a natural number – but there was some strange “discontinuities” in the analysis that prevented me from establishing the boundedness of (2) for arbitrary scales
. The basic problem was that increasing the number of permitted prime factors from one natural number threshold
to another
ended up increasing the density of the set by an unbounded factor (of the order of
, in practice), which heavily disrupted the task of trying to keep the ratio (2) bounded. Usually the resolution to these sorts of discontinuities is to use some sort of random “average” of two or more deterministic constructions – for instance, by taking some random union of some numbers with
prime factors and some numbers with
prime factors – but the numerology turned out to be somewhat unfavorable, allowing for some improvement in the lower bounds over my previous construction, but not enough to close the gap entirely. It was only after substantial trial and error that I was able to find a working deterministic construction, where at a given scale one collected either numbers with at most
prime factors, or numbers with
prime factors but with the largest prime factor in a specific range, in which I could finally get the numerator and denominator in (2) to be in balance for every
. But once the construction was written down, the verification of the required properties ended up being quite routine.
I’ve just uploaded to the arXiv my paper “On product representations of squares“. This short paper answers (in the negative) a (somewhat obscure) question of Erdös. Namely, for any , let
be the size of the largest subset
of
with the property that no
distinct elements of
multiply to a square. In a paper by Erdös, Sárközy, and Sós, the following asymptotics were shown for fixed
:
-
.
-
.
-
.
-
for
.
-
for
.
-
for
.
In the end, the argument turned out to be relatively simple; no advanced results from additive combinatorics, graph theory, or analytic number theory were required. I found it convenient to proceed via the probabilistic method (although the more combinatorial technique of double counting would also suffice here). The main idea is to generate a tuple of distinct random natural numbers in
which multiply to a square, and which are reasonably uniformly distributed throughout
, in that each individual number
is attained by one of the random variables
with a probability of
. If one can find such a distribution, then if the density of
is sufficienly close to
, it will happen with positive probability that each of the
will lie in
, giving the claim.
When , this strategy cannot work, as it contradicts the arguments of Erdös, Särközy, and Sós. The reason can be explained as follows. The most natural way to generate a triple
of random natural numbers in
which multiply to a square is to set
However, the situation changes for larger . For instance, for
, we can try the same strategy with the ansatz


Recent Comments