In analytic number theory, it is a well-known phenomenon that for many arithmetic functions of interest in number theory, it is significantly easier to estimate logarithmic sums such as
than it is to estimate summatory functions such as
(Here we are normalising to be roughly constant in size, e.g. as .) For instance, when is the von Mangoldt function , the logarithmic sums can be adequately estimated by Mertens’ theorem, which can be easily proven by elementary means (see Notes 1); but a satisfactory estimate on the summatory function requires the prime number theorem, which is substantially harder to prove (see Notes 2). (From a complex-analytic or Fourier-analytic viewpoint, the problem is that the logarithmic sums can usually be controlled just from knowledge of the Dirichlet series for near ; but the summatory functions require control of the Dirichlet series for on or near a large portion of the line . See Notes 2 for further discussion.)
Viewed conversely, whenever one has a difficult estimate on a summatory function such as , one can look to see if there is a “cheaper” version of that estimate that only controls the logarithmic sums , which is easier to prove than the original, more “expensive” estimate. In this post, we shall do this for two theorems, a classical theorem of Halasz on mean values of multiplicative functions on long intervals, and a much more recent result of Matomaki and RadziwiÅ‚Å‚ on mean values of multiplicative functions in short intervals. The two are related; the former theorem is an ingredient in the latter (though in the special case of the Matomaki-RadziwiÅ‚Å‚ theorem considered here, we will not need Halasz’s theorem directly, instead using a key tool in the proof of that theorem).
We begin with Halasz’s theorem. Here is a version of this theorem, due to Montgomery and to Tenenbaum:
Theorem 1 (Halasz-Montgomery-Tenenbaum) Let be a multiplicative function with for all . Let and , and set
Then one has
Informally, this theorem asserts that is small compared with , unless “pretends” to be like the character on primes for some small . (This is the starting point of the “pretentious” approach of Granville and Soundararajan to analytic number theory, as developed for instance here.) We now give a “cheap” version of this theorem which is significantly weaker (both because it settles for controlling logarithmic sums rather than summatory functions, it requires to be completely multiplicative instead of multiplicative, it requires a strong bound on the analogue of the quantity , and because it only gives qualitative decay rather than quantitative estimates), but easier to prove:
Note that now that we are content with estimating exponential sums, we no longer need to preclude the possibility that pretends to be like ; see Exercise 11 of Notes 1 for a related observation.
To prove this theorem, we first need a special case of the Turan-Kubilius inequality.
Informally, this lemma is asserting that
for most large numbers . Another way of writing this heuristically is in terms of Dirichlet convolutions:
This type of estimate was previously discussed as a tool to establish a criterion of Katai and Bourgain-Sarnak-Ziegler for Möbius orthogonality estimates in this previous blog post. See also Section 5 of Notes 1 for some similar computations.
Proof: By Cauchy-Schwarz it suffices to show that
Expanding out the square, it suffices to show that
We just show the case, as the cases are similar (and easier). We rearrange the left-hand side as
We can estimate the inner sum as . But a routine application of Mertens’ theorem (handling the diagonal case when separately) shows that
and the claim follows.
Remark 4 As an alternative to the Turan-Kubilius inequality, one can use the Ramaré identity
(see e.g. Section 17.3 of Friedlander-Iwaniec). This identity turns out to give superior quantitative results than the Turan-Kubilius inequality in applications; see the paper of Matomaki and RadziwiÅ‚Å‚ for an instance of this.
We rearrange the left-hand side as
We now replace the constraint by . The error incurred in doing so is
which by Mertens’ theorem is . Thus we have
From Mertens’ theorem, the expression in brackets can be rewritten as
and so the real part of this expression is
By (1), Mertens’ theorem and the hypothesis on we have
for any . This implies that we can find going to infinity such that
and thus the expression in brackets has real part . The claim follows.
The Turan-Kubilius argument is certainly not the most efficient way to estimate sums such as . In the exercise below we give a significantly more accurate estimate that works when is non-negative.
Exercise 5 (Granville-Koukoulopoulos-Matomaki)
- (i) If is a completely multiplicative function with for all primes , show that
as . (Hint: for the upper bound, expand out the Euler product. For the lower bound, show that , where is the completely multiplicative function with for all primes .)
- (ii) If is multiplicative and takes values in , show that
for all .
Now we turn to a very recent result of Matomaki and Radziwiłł on mean values of multiplicative functions in short intervals. For sake of illustration we specialise their results to the simpler case of the Liouville function , although their arguments actually work (with some additional effort) for arbitrary multiplicative functions of magnitude at most that are real-valued (or more generally, stay far from complex characters ). Furthermore, we give a qualitative form of their estimates rather than a quantitative one:
Theorem 6 (Matomaki-RadziwiÅ‚Å‚, special case) Let be a parameter going to infinity, and let be a quantity going to infinity as . Then for all but of the integers , one has
A simple sieving argument (see Exercise 18 of Supplement 4) shows that one can replace by the Möbius function and obtain the same conclusion. See this recent note of Matomaki and Radziwiłł for a simple proof of their (quantitative) main theorem in this special case.
Of course, (4) improves upon the trivial bound of . Prior to this paper, such estimates were only known (using arguments similar to those in Section 3 of Notes 6) for unconditionally, or for for some sufficiently large if one assumed the Riemann hypothesis. This theorem also represents some progress towards Chowla’s conjecture (discussed in Supplement 4) that
as for any fixed distinct ; indeed, it implies that this conjecture holds if one performs a small amount of averaging in the .
Below the fold, we give a “cheap” version of the Matomaki-Radziwiłł argument. More precisely, we establish
for any fixed .
Note that (5) improves upon the trivial bound of . Again, one can replace with if desired. Due to the cheapness of Theorem 7, the proof will require few ingredients; the deepest input is the improved zero-free region for the Riemann zeta function due to Vinogradov and Korobov. Other than that, the main tools are the Turan-Kubilius result established above, and some Fourier (or complex) analysis.
— 1. Proof of theorem —
We now prove Theorem 7. We first observe that it will suffice to show that
for any smooth supported on (say) and respectively, as the claim follows by taking and to be approximations to and respectively and using the triangle inequality to control the error.
We need some quantities that go to infinity reasonably fast; more specifically we take
and so from Lemma 3 and the triangle inequality we have
for any fixed , which implies that
since the inner sum is . The claim then follows from the triangle inequality.
Since and , our task is now to show that
I will (perhaps idiosyncratically) adopt a Fourier-analytic point of view here, rather than a more traditional complex-analytic point of view (for instance, we will use Fourier transforms as a substitute for Dirichlet series). To bring the Fourier perspective to the forefront, we make the change of variables and , and note that , to rearrange the previous claim as
Introducing the normalised discrete measure
it thus suffices to show that
where now denotes ordinary (Fourier) convolution rather than Dirichlet convolution.
From Mertens’ theorem we see that has total mass ; also, from the triangle inequality (and the hypothesis ) we see that is supported on and obeys the pointwise bound of . Thus we see that the trivial bound on is by Young’s inequality. To improve upon this, we use Fourier analysis. By Plancherel’s theorem, we have
where are the Fourier transforms
From Plancherel’s theorem we have
Since the derivative of is bounded by , a similar application of Plancherel also gives
so the contribution of those with or is acceptable. Also, from the definition of we have
and so from the prime number theorem we have when ; since , we see that the contribution of the region is also acceptable. It thus suffices to show that
whenever and . But by definition of , we may expand as
so by smoothed dyadic decomposition it suffices to show that
whenever . We replace the summation over primes with a von Mangoldt function weight to rewrite this as
Performing a Fourier expansion of the smooth function , it thus suffices to show the Dirichlet series bound
as and (we use the crude bound to deal with the contribution). But this follows from the Vinogradov-Korobov bounds (who in fact get a bound of as ); see Exercise 43 of Notes 2 combined with Exercise 4(i) of Notes 5.
Remark 8 If one were working with a more general completely multiplicative function than the Liouville function , then one would have to use a duality argument to control the large values of (which could occur at a couple more locations than ), and use some version of Halasz’s theorem to also obtain some non-trivial bounds on at those large values (this would require some hypothesis that does not pretend to be like any of the characters with ). These new ingredients are in a similar spirit to the “log-free density theorem” from Theorem 6 of Notes 7. See the Matomaki-Radziwiłł paper for details (in the non-cheap case).