Lecture 2: stereographic projection and first branch cuts
1. Stereographic projection
We’ve discussed the idea that it’s often helpful to think of the complex numbers as forming a plane, which we can formalize via identifying with . We will use this perspective freely, e.g. identify functions with the corresponding function by .
Today, we’ll go one step further, and “complete” the complex plane to get the “extended complex plane” by adding a point at infinity, i.e. . How should we think of this space?
We claim we should think of it as a sphere, “folding up” the plane and joining all the “edges” (infinitely far from the origin) to a single point . We’ll make sense of this by writing down an explicit bijection between and the unit sphere .
The idea is this. Let be the “north pole” of the sphere, and any other fixed point. We can draw a line passing through these two points; since the only point on the sphere with -coordinate equal to is , for this line is not parallel with the --plane (i.e. the plane ) and so intersects it at some point . We claim that the map is a bijection from to ; so we can think of as corresponding to , giving a bijection . (Note that this is really about the plane , and doesn’t a priori involve the complex structure!)
Before proving this, let’s think about some examples to get a feel for this mapping. If we took to be the south pole, , then the line through and passes through the --plane at . If is itself on this plane, then (or rather, if , then ). As approaches , the corresponding point tends to infinity, at least in absolute value, so it makes some sense to say should correspond to the point at infinity.
A slightly more subtle example is the point on the sphere. The line connecting this to the north pole always has -coordinate , so the corresponding point in the complex plane is actually real, i.e. it has imaginary part . Actually computing where this lands though is harder; we’ll come back to this when we’ve found a formula for in terms of .
How should we go about proving that this gives the desired bijection? The key step in fact is finding a formula for . Points on the line connecting and are given by ; we’re looking for the point on this line with last coordinate zero, i.e. , so . Therefore is given by the first two coordinates: . In the complex language,
We can now construct an inverse, showing that this is in fact a bijection. Given , we want to find satisfying such that and . Recalling the notation , we multiply the defining equation for the sphere by with , , and , we get
and solving for gives
Using the equations above, we solve to get
So we can invert the transformation.
Returning to our example , we find
In addition to giving us a new way to think of the (extended) plane, this correspondence also satisfies some nice properties. For example, longitudinal lines on the sphere map to straight lines on the plane, while latitudinal lines map to circles in the plane. More generally, every circle on the sphere maps to either a circle or a line in the plane, and vice versa lines and circles on the plane correspond to circles on the sphere; when we add the point at infinity, we can think of straight lines in the plane as circles passing through the point at infinity. The extended complex numbers, thought of as a sphere, are often called the Riemann sphere.
2. Squares and square roots
When dealing with real functions, there’s a standard way to visualize them, namely by graphing them: we put the domain on one axis and the codomain on the other, and plot the graph of the function on the resulting two-dimensional space. When we’re working with complex functions, both the domain and codomain are given by the complex plane, so the resulting space would be four-dimensional, which is much harder to visualize; so we need another approach. There are many ways of doing this, such as using color or graphing the real and imaginary parts separately. We want to introduce another: graphing how the values of the function change as the inputs change.
Let’s work with the function . This is easiest to think about in polar coordinates: if , then
so squares the modulus and doubles the argument.
Graphically, squaring the modulus is straightforward enough in terms of scaling, so let’s think about points on the unit circle. Here as moves around the unit circle, i.e. as goes from to , moves around the unit circle twice as fast: the argument goes from (equivalently, zero) to (again) in this same period. So if we just wanted to go around the unit circle, from to , we should take from to .
In particular, if we wanted to find an inverse for , we would need make a restriction something like this. That is: writing with , we can find a square root of given by
We note though that something weird happens near the ray : if for some very small positive , then is very close to , and in this model its square root is very close to . However, if we instead looked at , which is also very close to but on the other side of the real axis, we would find that its square root is which is very close to . So this is quite far from , even as . In other words, this square root is not continuous!
Now, it only fails to be continuous near this ray ; elsewhere everything is fine. You might point out that this failure of continuity is only due to our particular choice of , we could have chosen a different parametrization; but that would just move the discontinuity somewhere else.
To avoid having to think of this as a discontinuity, we make a “branch cut” in the complex plane along the negative real axis. If we think of this as a boundary, so we no longer think of points like and as close to each other, then we can now describe our square root function as a continuous function on this slit plane.
Observe that our square root has image in complex numbers with argument between and . Translating back to Cartesian coordinates, this is equivalent to having positive (or at least nonnegative111There is some subtlety when the real part is exactly zero: then we include the positive imaginary axis but not the negative one, i.e. for .) real part. Restricted to nonnegative real numbers, this gives the positive square root. But we could also take the negative square root, which in the complex setting has image in complex numbers with argument less than or equal to or greater than . To make this nicer, we could translate by and say these have argument between and .
These options for the square root function, which is a priori multivalued, are called its branches; let’s call the “positive” branch we wrote down above , given by , and the other branch , which is given by . Each of these naturally lives on the slit complex plane we described above. If we wanted to be able to cross the negative real axis, we would have to combine these two branches somehow.
Let’s return to the example above, where . Here was close to . As , so from the positive imaginary side, this approaches ; but when becomes negative, so crosses the negative real axis, then jumps to near , as we saw before. However, is then near , i.e. near the value of on the other side! So if we wanted to define a single inverse to , we would take our two copies of the slit plane and glue them together: the top edge of the slit on the first plane is glued to the bottom edge on the second plane, and vice versa.
This space is a little hard to imagine; by thinking about how you deform this space, you can shape it into a sphere with two punctures. We can see this more concretely as follows: the function gives a bijection between this space and (we have to exclude to make sure the polar representation is well-defined, and indeed at zero the square root is only single-valued), and by stereographic projection we can think of as .
This is the Riemann surface for the square root function. We’ll see more examples of Riemann surfaces next week.
3. The complex exponential
Finally, let’s return to our definition of the exponential function in the complex setting. If , then we set
where is as usual for a real number . We saw when studying polar representations that this satisfies the additivity property : on the real component, this is a standard property, and on the imaginary component this means that rotating by and then by is the same as rotating by . We can likewise confirm properties such as
We see that the exponential function is easiest to define in terms of Cartesian coordinates. On the other hand, its output is easiest to understand in terms of polar coordinates: while computing its real and imaginary parts involves the use of trigonometric functions, we can quickly see that while . So when we repeat our idea of comparing how changes in the domain translate to changes in the codomain, we use changes in Cartesian coordinates, i.e. in the real and imaginary part, on the domain, and changes in polar coordinates on the codomain. Changes in the real part scale the modulus by an exponential factor, so horizontal lines in the complex plane map to rays in the codomain; and changes in the imaginary part shift the argument, so vertical lines map to circles.
Unlike the square root function, the exponential function is not multivalued, so we don’t need to worry about its Riemann surface. However, it is not injective, unlike in the real setting: for example, . More generally, is periodic with period : that is, for all . This also means that for every integer , so the preimage of contains infinitely many points; so its inverse function will be (very) multivalued! We’ll worry about this more next time.