Showing posts with label qed. Show all posts
Showing posts with label qed. Show all posts

2020-08-10

Eighth, Ninth, Tenth, and Eleventh Papers

My eighth, ninth, tenth, and eleventh papers have been published! These require subscriptions to read, so here are alternate links to older preprints for the eighth, ninth, tenth, and eleventh papers, respectively (which have most of the same content, with some minor changes to explanations, citations, and figures relative to the published versions). As with my previous papers, in the interest of explaining these ideas in a way that is easy to understand, I am using the ten hundred most used words in English (except for the two lines that came before this one), as put together from the XKCD Simple Writer. I will use numbers sometimes without completely writing them out, use words for certain names of things without explaining further, and explain less used words when they come up. Keep reading to see what comes next. While these papers aren't as closely related to each other as the previous three, there are enough relations that I'm putting them together in a single post. These papers need a lot more math (note: "math" isn't one of the ten hundred words) than the papers before, and because they need a lot of thinking to get, I actually won't say as much about them.

On another note, this is a milestone for me because these are the last papers from my PhD in which I was a leading author. I still have one more review paper left to be published, but as that has been submitted to the journal and as I'm not the leading author, I don't really need to worry about that at this point. (Of course, once that is published, I will write a blog post summarizing it, though as it is a review paper, that summary will probably be quite short.) Thus, I am truly done with the work from my PhD, and can fully shift my mindset away from physics toward thinking about problems in transportation policy, as I will do in my postdoctoral research at UC Davis.

2020-01-13

Fifth, Sixth, and Seventh Papers

My fifth, sixth, and seventh papers have been published! These require subscriptions to read, so here are alternate links to older preprints for the fifth, sixth, and seventh papers, respectively (which have most of the same content, with some minor changes to explanations, citations, and figures relative to the published versions). As with my previous papers, in the interest of explaining these ideas in a way that is easy to understand, I am using the ten hundred most used words in English (except for the two lines that came before this one), as put together from the XKCD Simple Writer. I will use numbers sometimes without completely writing them out, use words for certain names of things without explaining further, and explain less used words when they come up. Keep reading to see what comes next. I'm putting these three papers together in a single post because they form a trilogy of sorts, all having to do with finding the biggest number for how much heat, through light, can go from one body to another when they are really close together, or can go from one body into outer space. These papers need a lot more math (note: "math" isn't one of the ten hundred words) than the papers before, and because they need a lot of thinking to get, I actually won't say as much about them.

The fifth paper is called "T Operator Bounds on Angle-Integrated Absorption and Thermal Radiation for Arbitrary Objects", and is in volume 123, issue 5 of Physical Review Letters. This is the one that has to do with how much heat, through light, can go from one body to outer space. People knew before that the number for how much heat really big bodies can put through light into outer space grows like the surface area of the body, but for really small bodies it grows like the space of the whole body (volume), and they were not sure how these two things join in between. This paper lets people figure out what the most heat is that can go from a body through light into outer space no matter what the largest shape the body can sit in, and shows how to join the things that people knew before for middle-size bodies of different shapes. (Another press release from my department can be found here.)

The sixth paper is called "Fundamental limits to radiative heat transfer: Theory", and is in volume 101, issue 3 of Physical Review B, while the seventh paper is called "Fundamental Limits to Radiative Heat Transfer: The Limited Role of Nanostructuring in the Near-Field", and is in volume 124, issue 1 of Physical Review Letters. Those two papers go together, so I'll write about them together. The sixth paper is about the math behind figuring out the biggest number for heat, through light, to go between two bodies. The seventh paper shows that heat, through light, going between two big flat bodies that are close together can be pretty close to the biggest number possible, so making the shapes of the bodies less simple than just flat surfaces is of no use.

2019-11-05

Fourth Paper: "Impact of nuclear vibrations on van der Waals and Casimir interactions at zero and finite temperature"

My fourth paper has been published! It is in volume 5, issue 11 of Science Advances, which is an open-access journal, so anyone can read it. As with my previous papers, in the interest of explaining these ideas in a way that is easy to understand, I am using the ten hundred most used words in English (except for the two lines that came before this one), as put together from the XKCD Simple Writer. I will use numbers sometimes without completely writing them out, use words for certain names of things without explaining further, and explain less used words when they come up. Keep reading to see what comes next.

In papers that came before this one, I looked at how to do a better job of figuring out the van der Waals (vdW) forces, which are the forces that let geckos (small animals with hard skin over which your finger can slip easily) stick to anything no matter what it is made of, between molecules, which are the little things that make up most of the stuff we see and are in turn made of smaller things called atoms; I also looked at how to do a better job of figuring out how heat (through light) goes between different molecules, especially when they are near larger bodies, and that needed me to do a better job of considering how molecules can make changes on each other through light, and that means that I need to better consider how the full atoms within molecules move toward and away from each other in a way that repeats itself. In this paper, I used the new way from the paper that came before of considering how molecules can make changes on each other through light to show what happens with vdW forces between molecules and larger bodies, especially when they aren't as cold as possible but are as hot as a room you might go into each day. It turns out that for small or long thin molecules made of carbon atoms, our work can do a pretty good job of showing how being as hot as a room can make vdW forces look very different. For some kinds of large thin sheets of atoms, like boron nitride in which every other atom is boron or nitrogen, our work still does a pretty good job because the electrons, which are the parts of the atoms that are the smallest, lightest, and move around the most, are still pretty close to the centers of the atoms. On the other hand, for other kinds of large thin sheets of atoms, like graphene in which every atom is carbon, our work has some more problems, because electrons in graphene can move around a lot more than our work might make you think, and the ways in which those electrons change how the rest of the atoms move around when they are as hot as a room (instead of as cold as possible) makes vdW forces harder to figure out than our work can say. This means we still need to do more work to better figure out how vdW forces look for those kinds of sheets.

2018-08-01

Third Paper: "Phonon-Polariton Mediated Thermal Radiation and Heat Transfer among Molecules and Macroscopic Bodies: Nonlocal Electromagnetic Response at Mesoscopic Scales"

My third paper has been published! It is in volume 121, issue 4 of Physical Review Letters, and an older preprint of it is available too for those who don't have access to academic journals (it has all of the same figures and ideas, though it is missing a few sentences of further explanation as well as a couple of new citations that were inserted for the final publication). As with my previous papers, in the interest of explaining these ideas in a way that is easy to understand, I am using the ten hundred most used words in English (except for the two lines that came before this one), as put together from the XKCD Simple Writer. I will use numbers sometimes without completely writing them out, use words for certain names of things without explaining further, and explain less used words when they come up. Keep reading to see what comes next.

In the paper that came before this one, I looked at how to do a better job of figuring out the van der Waals (vdW) forces, which are the forces that let geckos (small animals with hard skin over which your finger can slip easily) stick to anything no matter what it is made of, between molecules, which are the little things that make up most of the stuff we see and are in turn made of smaller things called atoms. I tried to figure out how these forces look at distances where the fact that molecules are made of atoms is important, but if those molecules are near much larger bodies, the fact that the larger bodies are made of atoms and molecules should be less important; it turns out that at those distances, how fast light goes matters a lot, and using ways to figure out these forces exactly instead of using easier ways to figure out those forces makes a big change in what those forces are. That paper was able to show how to bring together all of these different ideas from considering large and small bodies in a single way where none of those ideas can be ignored. That would be important when considering new kinds of molecules like graphene, which is made of a lot of carbon atoms in a thin sheet, or really long molecules like DNA or those found in foods, when those molecules are near larger bodies that we make.

This paper looks at the same sorts of molecules, but not at vdW forces anymore. Instead, in this paper, I look at how heat (through light) goes between different molecules, especially when they are near larger bodies. For that, I need to do a better job of considering how molecules can make changes on each other through light, and that means that I need to better consider how the full atoms within molecules move toward and away from each other in a way that repeats itself. By doing that, I can now show how heat through light goes between different molecules, whether they are close together or far away from each other. When people considered heat going through light between larger bodies, they found that the heat would keep growing as the bodies came closer together, and that growing wouldn't stop; from knowing how things work every day, we know that once bodies come close enough, they touch each other, and the growing stops at some point. In this paper, I've shown that when molecules come close together, the heat grows for a while, but if they come close enough, that growing does stop, so I've been able to show what distance we can say two molecules touch each other, so that heat going between molecules happens through them touching instead of through light. This is really important for things like graphene, which are used as part of things used for making power from the sun by getting its light and heat, and also for making new things that can become part of computers made of really small things like molecules that work because of heat going between different parts.

2017-07-03

Second Paper: "Unifying Microscopic and Continuum Treatments of van der Waals and Casimir Interactions"

My second paper has been published! It is in volume 118, issue 26 of Physical Review Letters, and an older preprint of it is available too for those who don't have access to academic journals (it has all of the same figures and ideas, though it is missing a few sentences of further explanation as well as a couple of new citations that were inserted for the final publication). As with my first paper, in the interest of explaining these ideas in a way that is easy to understand, I am using the ten hundred most used words in English (except for the two lines that came before this one), as put together from the XKCD Simple Writer. I will use numbers sometimes without completely writing them out, use words for certain names of things without explaining further, and explain less used words when they come up. Keep reading to see what comes next.

2016-07-18

Classical Damping of Gases and Oscillators

I was on vacation last week, and during some quiet time, I randomly happened to be thinking about explanations for damping in physical systems. I remember learning in ELE 456 — Quantum Optics, from last spring, that the phenomenological linear damping of a classical oscillator could be derived by coupling a quantum oscillator to a thermal bath of quantum oscillators; each linear oscillator is microscopically undamped, but by treating the bath through statistical thermodynamics, the coupling of the oscillator in question to a bath essentially produces a linear damping coefficient dependent on the spectrum of the bath (and the coupling too). Microscopically, the quantization of energy levels in a linear oscillator makes it easy to interpret how discrete excitations can move from one oscillator to another coupled oscillator, but I was wondering if quantum mechanics is really necessary to explain damping. Follow the jump to see an extremely rough sketch of ideas that may (or may not) justify the use of classical mechanics by itself. (Added after finishing: this turns out to be a rambling and possibly ultimately pointless post with a much clearer and more self-consistent explanation linked at the end, so for the time being, humor me.)

2016-04-15

Generals Impending

I briefly thought of doing a review or another longer post this month, but I realized that studying for my generals will require too much of my time and concentration to allow for that. Instead, I'll keep this post as an update about my upcoming generals. My generals will have two parts: the first is a standard research seminar where I get to talk about some of the stuff that I've done over the last year and more, while the second is an oral exam where the three committee members get to ask me more fundamental questions. The oral exam is tricky to prepare for because those questions could in principle be about anything; that said, from what I understand, the committee members tend to ask about things related to my research topics, my coursework, or other basic things that they expect someone in my field to understand, so I basically have to study a broad range of material and hope for the best. The research seminar is a little better, because I have a better sense of my own research than at least two out of the three committee members (the excluded member being my advisor); I just have to make sure that I know what I'm talking about (and on that same note that I don't start making stuff up), so this will require me to go a little broad but more deep into the fundamentals underlying my research.

I said in a post from two months ago that I'm working on projects involving nanoscale wetting as well as more accurate modeling of the optical response of electrons in nanoscale metal structures. Right now, I'm not so sure about the future of the second project, because there still seems to be a lot of controversy about how to properly account for boundary effects in finite metal systems, and our group is not really in the business of making those [more fundamental] decisions; instead, we'd like to use a more well-established model of optical response and hope to show some new and interesting results using novel techniques in computational electromagnetics. Given that, my research seminar is almost exclusively going to focus on the first project, which is a much better-posed and better-developed problem and has consequently led to very interesting results. (I can't really give more details until we put out a publication.)

Anyway, at this point I'm waiting to be over with generals, while studying hard for them as the days count down. I promise that I'll have a more typical post next month, after I finish both parts of generals.

2016-02-24

Research and Generals

I was intending to post a review this month, but I got too busy and didn't have the time to do that. Instead, I'll keep this as a short update. Right now, I'm quite occupied with doing research on two projects: one has to do with wetting at the nanoscale, and the other has to do with modeling the optical response of electrons in metals and semiconductors. They may seem divorced from each other at first, but they are both electromagnetic phenomena, and if I consider Casimir forces or heat transfer from metals or semiconductors, they both involve fluctuational phenomena as well. Concurrently with research, I am preparing for my general examinations (often called qualifying examinations elsewhere) which are coming up in April or May. Anyway, hopefully I'll have a little time next month to put a review or another sort of post like that out.

2014-02-25

Green's Functions and Correlations

I had the idea of writing this post a couple of weeks ago, but I didn't feel like I had enough stuff to write here at that time. Now I do, so here goes. (Also, here's hoping that inputting LaTeX into this post works once more.)

When I took 18.03 — Differential Equations in 2010 fall, one of the topics covered was linear time-invariant systems. The general system of interest was $Lu(t) = f(t)$ where $L$ is a linear time-invariant operator. The technique of course is to find a weight function $w(t)$ where $Lw(t) = \delta(t)$, and once that is done, the solution is $u(t) = \int_{-\infty}^{\infty} f(t') w(t - t') dt'$ which is a convolution of the input $f$ with the weight $w$. The professor mentioned that it is essentially akin to inverting the operator $L$, but while I could see the general utility in this method, I never quite understood why it might be considered inversion on any deeper level.

Last semester, I took 8.07 — Electromagnetism II, and there we discussed Green's functions a little more in the context of electromagnetism & electrodynamics. In a static situation, the Green's function comes up in solving the Poisson equation $\nabla^2 \phi = -\rho$. In this case, $\nabla^2 G(\vec{x}, \vec{x}') = -\delta(\vec{x} - \vec{x}')$ is solved by the familiar potential of a unit point charge $G(\vec{x}, \vec{x}') = \frac{1}{4\pi |\vec{x} - \vec{x}'|}$. I started to see a little more clearly why this worked, because if a general charge distribution was some superposition of point charges, then a general potential distribution should be the same superposition of point charge potentials. However, it still wasn't entirely clear to me how this was "inversion" per se. Follow the jump to see what changed.

2013-08-27

Particles in the Continuous Quantum Field

The last thing I discussed in the last post was about the energy eigenstates of the continuous field. The ground state $|0\rangle$ classically corresponds to there being no displacement in the chain at any spatial index $x$ and quantum mechanically corresponds to each oscillator for each normal mode index $k$ being in its ground state, while the first excited state $|k\rangle = a^{\dagger} (k)|0\rangle$ for a given $k$ classically corresponds to a traveling plane wave normal mode of wavevector $k$ and quantum mechanically corresponds to only the oscillator at the given normal mode index $k$ being in its first excited state (and all others being in their ground states). The excited state $|k\rangle$ has energy $E = \hbar v|k|$ above the ground state and overall momentum $p = \hbar k$ above the ground state. This post will discuss what the second and higher excited states are. Follow the jump to see more.

2013-08-26

Operators and States of the Continuous Quantum Field

In my last post about intuiting and visualizing quantum field theory, I discussed the diagonalization of the Hamiltonian and overall momentum and how they become operators. In this post I'm going to discuss more the meanings of the operators and associated quantum states of this field. Follow the jump to see more.

2013-08-24

Diagonalizing and Quantizing the Continuous Field Hamiltonian

In my previous post I discussed the intuition behind the classical acoustic field in one dimension. Now I'm going to talk about diagonalizing the Hamiltonian and making the step into quantum field theory. Follow the jump to see what it's like.

2013-08-23

Classical Discrete and Continuum Fields

I've been reading various documents about quantum field theory over the last several weeks, specifically about the canonical quantization of quantum fields. In doing so, I've come to realize that quantum mechanics has a lot of crazy math and even crazier physical interpretations, and I just took that for granted, but now those things are coming back to haunt me in quantum field theory. It is very hard for me to wrap my head around, and I feel like I could use a lot more help in visualizing and intuiting what certain concepts in canonical quantization mean. This will be the first of a few posts which are outlets for me to gather my thoughts and put them out there for you all to see and correct; this one will be about classical fields.

I feel like the easiest quantum field system to study is the phonon. It is a spin-0 bosonic system, so it can be described by a scalar field. Furthermore, said field can be restricted to one dimension, which simplifies the math even further. This means that taking the continuum limit becomes a bit easier than in three dimensions. Follow the jump to see how it goes.

2013-04-15

Harmonic Oscillator from Fields not Potentials

I finally started writing my paper for 8.06 yesterday. Before that, though, I had asked a couple questions about the topic to my UROP supervisor, whose primary area of expertise is actually in QED and Casimir problems. I was asking him why the book The Quantum Vacuum by Peter Milonni uses the magnetic potential $\vec{A}$ instead of the electromagnetic fields $\vec{E}$ and $\vec{B}$ to expand in Fourier modes and derive the harmonic oscillator Hamiltonian. He said that is just a matter of choice, and in fact the same derivation can be done using the fields rather than the potential; moreover, he encouraged me to try this out for myself, and I did just that. Lo and behold, the right answer popped out by modifying the derivation in that book to use the fields and the restrictions of the Maxwell equations instead of the potential and the Coulomb gauge choice; follow the jump to see what it looks like. I'm going to basically write out the derivation in the book and show how at each point I modify it.

2013-04-09

Charge Conservation and Legendre Transformations

As a follow-up (sort of, but not exactly) to my previous post on the matter, I would like to post a few updates and new questions, using Einstein summation throughout for convenience. The first has to do with why $p^{\mu} = \int T^{(0, \mu)} d^3 x$ is a Lorentz-contravariant vector. Apparently Noether's theorem says that if some Noether current $J^{\mu}$ generates a symmetry and satisfies \[ \partial_{\mu} J^{\mu} = 0 \] then the quantity \[ q = \int J^{(0)} d^3 x \] called the Noether charge is Lorentz-invariant and conserved. The first is not easy to show, but apparently some E&M textbooks do it for the example of electric charge. The second is fairly easy to show: using the condition that $\int \partial_{j} J^{j} d^3 x = \int_{\partial \mathbb{R}^3} J^{j} d\mathcal{S}_{j} = 0$ (from the divergence theorem applied to all of Euclidean space) in conjunction with $\partial_{\mu} J^{\mu} = 0$, the result $\dot{q} = 0$ follows.

As an example, let us consider the generator of rotations and Lorentz boosts for a general energy distribution: that is the 3-index angular momentum tensor \[ M^{\mu \nu \sigma} = x^{\mu} T^{\nu \sigma} - x^{\nu} T^{\mu \sigma} .\] Given that $\partial_{\nu} T^{\mu \nu} = 0$ then $\partial_{\sigma} M^{\mu \nu \sigma} = (\partial_{\sigma} x^{\mu})T^{\nu \sigma} + x^{\mu} \partial_{\sigma} T^{\nu \sigma} - (\partial_{\sigma} x^{\nu})T^{\mu \sigma} - x^{\nu} \partial_{\sigma} T^{\mu \sigma}$ $= \delta_{\sigma}^{\; \mu} T^{\nu \sigma} - \delta_{\sigma}^{\; \nu} T^{\mu \sigma} = T^{\nu \mu} - T^{\mu \nu} = 0$. Therefore the 3-index angular momentum is a proper Noether current. Its corresponding conserved charge is the 2-index angular momentum integrated over spatial directions: \[ L^{\mu \nu} = \int M^{(\mu \nu, 0)} d^3 x \] (except for perhaps a sign because $M^{\mu \nu \sigma}$ is antisymmetric in its indices) and it should be easy now to show that $\dot{L}^{\mu \nu} = 0$, which is cool. The only remaining question I have is whether it is more correct to say $L^{\mu \nu} = x^{\mu} p^{\nu} - x^{\nu} p^{\mu}$ where $p^{\mu} = \int T^{(\mu, 0)} d^3 x$ as before or if the better definition is the one integrating $M^{(\mu \nu, 0)}$ over space.

Now I have an even bigger question looming ahead of me though. The Noether current generating spacetime translational symmetry is exactly the stress-energy tensor derivable as the Legendre transformation of the Lagrangian. The term involving the conjugate momenta is easy, but the term involving the Lagrangian is confusing. For a scalar field $\phi$ (and for a vector field this is easily replaced with $A^{\sigma}$), what I have seen is \[ T^{\mu \nu} = \frac{\partial \mathcal{L}}{\partial (\partial_{\mu} \phi)} \partial^{\nu} \phi - \mathcal{L} B^{\mu \nu} .\] The problem is that the tensor $B^{\mu \nu}$ seems to depend on either the field $\phi$ used or on the notation consistently used. Sometime $B^{\mu \nu} = \delta^{\mu \nu}$, while other times $B^{\mu \nu} = \eta^{\mu \nu}$. I'm not really sure which it is supposed to be, as sometimes for scalar fields $B = \delta$ is used, while for the electromagnetic field $B = \eta$ is used, and sometimes the notation isn't even that consistent. The issue is that either one would properly specify a Lorentz-contravariant 2-index tensor, but only one of them actually defines the translational symmetry Noether current properly. Which one is it? The issue appears to be akin to the problem of two grammatically correct sentences where one carries meaning and makes sense while the other makes no sense at all (e.g. "colorless green ideas sleep furiously").

2013-03-27

Hamiltonian Density and the Stress-Energy Tensor

As an update to a previous post about my adventures in QED-land for 8.06, I emailed my recitation leader about whether my intuition about the meaning of the Fourier components of the electromagnetic potential solving the wave equation (and being quantized to the ladder operators) was correct. He said it basically is correct, although there are a few things that, while I kept in mind at that time, I still need to keep in mind throughout. The first is that the canonical quantization procedure uses the potential $\vec{A}$ as the coordinate-like quantity and finds the conjugate momentum to this field to be proportional to the electric field $\vec{E}$, with the magnetic field nowhere to be found directly in the Hamiltonian. The second is that there is a different harmonic oscillator for each mode, and the number eigenstates do not represent the energy of a given photon but instead represent the number of photons present with an energy corresponding to that mode. Hence, while coherent states do indeed represent points in the phase space of $(\vec{A}, \vec{E})$, the main point is that the photon number can fluctuate, and while classical behavior is recovered for large numbers $n$ of photons as the fluctuations of the number are $\sqrt{n}$ by Poisson statistics, the interesting physics happens for low $n$ eigenstates or superpositions thereof in which $a$ and $a^{\dagger}$ play the same role as in the usual quantum harmonic oscillator. Furthermore, the third issue is that only a particular mode $\vec{k}$ and position $\vec{x}$ can be considered, because the electromagnetic potential has a value for each of those quantities, so unless those are held constant, the picture of phase space $(\vec{A}, \vec{E})$ becomes infinite-dimensional. Related to this, the fourth and fifth issues are, respectively, that $\vec{A}$ is used as the field and $\vec{E}$ as its conjugate momentum rather than using $\vec{E}$ and $\vec{B}$ because the latter two fields are coupled to each other by the Maxwell equations so they form an overcomplete set of degrees of freedom (or something like that), whereas using $\vec{A}$ as the field and finding its conjugate momentum in conjunction with a particular gauge choice (usually the Coulomb gauge $\nabla \cdot \vec{A} = 0$) yields the correct number of degrees of freedom. These explanations seem convincing enough to me, so I will leave those there for the moment.

Another major issue that I brought up with him for which he didn't give me a complete answer was the issue that the conjugate momentum to $\vec{A}$ was being found through \[ \Pi_j = \frac{\partial \mathcal{L}}{\partial (\partial_t A_j)} \] given the Lagrangian density $\mathcal{L} = \frac{1}{8\pi} \left(\vec{E}^2 - \vec{B}^2 \right)$ and the field relations $\vec{E} = -\frac{1}{c}\partial_t \vec{A}$ & $\vec{B} = \nabla \times \vec{A}$. This didn't seem manifestly Lorentz-covariant to me, because in the class 8.033 — Relativity, I had learned that the conjugate momentum to the electromagnetic potential $A^{\mu}$ in the above Lagrangian density would be the 2-index tensor \[ \Pi^{\mu \nu} = \frac{\partial \mathcal{L}}{\partial (\partial_{\mu} A_{\nu})} .\] This would make a difference in finding the Hamiltonian density \[ \mathcal{H} = \sum_{\mu} \Pi^{\mu} \partial_t A_{\mu} - \mathcal{L} = \frac{1}{8\pi} \left(\vec{E}^2 + \vec{B}^2 \right). \] I thought that the Hamiltonian density would need to be a Lorentz-invariant scalar just like the Lagrangian density. As it turns out, this is not the case, because the Hamiltonian density represents the energy which explicitly picks out the temporal direction as special, so time derivatives are OK in finding the momentum conjugate to the potential; because the Lagrangian and Hamiltonian densities looks so similar, it looks like both could be Lorentz-invariant scalar functions, but deceptively, only the former is so. At this point, I figured that because the Hamiltonian and (not field conjugate, but physical) momentum looked so similar, they could arise from the same covariant vector. However, there is no "natural" 1-index vector with which to multiply the Lagrangian density to get some sort of covariant vector generalization of the Hamiltonian density, though there is a 2-index tensor, and that is the metric. I figured here that the Hamiltonian and momentum for the electromagnetic field could be related to the stress-energy tensor, which gives the energy and momentum densities and fluxes. After a while of searching online for answers, I was quite pleased to find my intuition to be essentially spot-on: indeed the conjugate momentum should be a tensor as given above, the Legendre transformation can then be done in a covariant manner, and it does in fact turn out that the result is just the stress-energy tensor \[ T^{\mu \nu} = \sum_{\mu, \xi} \Pi^{\mu \xi} \partial^{\nu} A_{\xi} - \mathcal{L}\eta^{\mu \nu} \] (UPDATE: the index positions have been corrected) for the electromagnetic field. Indeed, the time-time component is exactly the energy/Hamiltonian density $\mathcal{H} = T_{(0, 0)}$, and the Hamiltonian $H = \sum_{\vec{k}} \hbar\omega(\vec{k}) \cdot (\alpha^{\star} (\vec{k}) \alpha(\vec{k}) + \alpha(\vec{k}) \alpha^{\star} (\vec{k})) = \int T_{(0, 0)} d^3 x$. As it turns out, the momentum $\vec{p} = \sum_{\vec{k}} \hbar\vec{k} \cdot (\alpha^{\star} (\vec{k}) \alpha(\vec{k}) + \alpha(\vec{k}) \alpha^{\star} (\vec{k}))$ doesn't look similar just by coincidence: $p_j = \int T_{(0, j)} d^3 x$. The only remaining point of confusion is that it seems like the Hamiltonian and momentum should together form a Lorentz-covariant vector $p_{\mu} = (H, p_j)$, yet if the stress-energy tensor respects Lorentz-covariance, then integrating over the volume element $d^3 x$ won't respect transformations in a Lorentz-covariant manner. I guess because the individual components of the stress-energy tensor transform under a Lorentz boost and the volume element does as well, then maybe the vector $p_{\mu}$ as given above will respect Lorentz-covariance. (UPDATE: another issue I was having but forgot to write before clicking "Publish" was the fact that only the $T_{(0, \nu)}$ components are being considered. I wonder if there is some natural 1-index Lorentz-convariant vector $b_{\nu}$ to contract with $T_{\mu \nu}$ so that the result is a 1-index vector which in a given frame has a temporal component given by the Hamiltonian density and spatial components given by the momentum density.) Overall, I think it is interesting that this particular hang-up was over a point in classical field theory and special relativity and had nothing to do with the quantization of the fields; in any case, I think I have gotten over the major hang-ups about this and can proceed reading through what I need to read for the 8.06 paper.

2013-03-21

Time and Temperature are Complex

In a post from a few days ago, I briefly mentioned the notion of imaginary time with regard to angular momentum. I'd like to go into that a little further in this post.

In 3 spatial dimensions, the flat (Euclidean) metric is $\eta_{ij} = \delta_{ij}$, which is quite convenient, as lengths are given by $(\Delta s)^2 = (\Delta x)^2 + (\Delta y)^2 + (\Delta z)^2$ which is just the usual Pythagorean theorem. When a temporal dimension is added, as in special relativity, the coordinates are now $x^{\mu} = (ct, x_{j})$, and the Euclidean metric becomes the Minkowski metric $\eta_{\mu \nu} = \mathrm{diag}(-1, 1, 1, 1)$ so that $\eta_{tt} = -1$, $\eta_{(t, j)} = 0$, and $\eta_{ij} = \delta_{ij}$. This means that spacetime intervals become $(\Delta s)^2 = -(c\Delta t)^2 + (\Delta x)^2 + (\Delta y)^2 + (\Delta z)^2$, which is the normal Pythagorean theorem only if $\Delta t = 0$. In general, time coordinate differences contribute negatively to the spacetime interval. In addition, Lorentz transformations are given by a hyperbolic rotation by a [hyperbolic] angle $\alpha$ equal to the rapidity given by $\frac{v}{c} = \tanh(\alpha)$. This doesn't look quite the same as normal Euclidean geometry. However, a transformation to imaginary time, called a Wick rotation, can be done by setting $\tau = it$, so $x^{\mu} = (ic\tau, x_{j})$, $\eta_{\mu \nu} = \delta_{\mu \nu}$, $(\Delta s)^2 = (c\Delta t)^2 + (\Delta x)^2 + (\Delta y)^2 + (\Delta z)^2$ as in the usual Pythagorean theorem, and the Lorentz transformation is given by a real rotation by an angle $\theta = i\alpha$ (though I may have gotten some of these signs wrong so forgive me) where $\alpha$ is now imaginary. Now, the connection to the component $L_{(0, j)}$ of the angular momentum tensor should be more clear.

I first encountered this in the class 8.033 — Relativity, where I was able to explore this curiosity on a problem set. That question and the accompanying discussion seemed to say that while this is a cool thing to try doing once, it isn't really useful, especially because it does not hold true in general relativity with more general metrics $g_{\mu \nu} \neq \eta_{\mu \nu}$ except in very special cases. However, as it turns out, imaginary time does play a role in quantum mechanics, even without the help of relativity.

Schrödinger time evolution occurs through the unitary transformation $u = e^{-\frac{itH}{\hbar}}$ satisfying $uu^{\dagger} = u^{\dagger} u = 1$. This means that the probability that an initial state $|\psi\rangle$ ends after time $t$ in the same state is given by the amplitude (whose square is the probability [density]) $\mathfrak{p}(t) = \langle\psi|e^{-\frac{itH}{\hbar}}|\psi\rangle$. Meanwhile, assuming the states $|\psi\rangle$ form a complete and orthonormal basis (though I don't know if this assumption is truly necessary), the partition function $Z = \mathrm{trace}\left(e^{-\frac{H}{k_B T}}\right)$, which can be expanded in the basis $|\psi\rangle$ as $Z = \sum_{\psi} \langle\psi|e^{-\frac{H}{k_B T}}|\psi\rangle$. This, however, is just as well rewritten as $Z = \sum_{\psi} \mathfrak{p}\left(t = -\frac{i\hbar}{k_B T}\right)$. Hence, quantum and statistical mechanical information can be gotten from the same amplitudes using the substitution $t = -\frac{i\hbar}{k_B T}$, which essentially calls temperature a reciprocal imaginary time. This is not really meant to show anything more deep or profound about the connection between time and temperature; it is really more of a trick stemming from the fact that the same Hamiltonian can be used to solve problem in quantum mechanics or equilibrium statistical mechanics.

As an aside, it turns out that temperature, even when measured in an absolute scale, can be negative. There are plenty of papers of this online, but suffice it to say that this comes from a more general statistical definition of temperature. Rather than defining it (as it commonly is) as the average kinetic energy of particles, it is better to define it as a measure of the probability distribution that a particle will have a given energy. Usually, particles tend to be in lower energy states more than in higher energy states, and as a consequence, the temperature is positive. However, it is possible (and has been done repeatedly) under certain circumstances to cleverly force the system in a way that causes particles to be in higher energy states with higher probability than in lower energy states, and this is exactly the negative temperature. More formally, $\frac{1}{T} = \frac{\partial S}{\partial E}$ where $E$ is the energy and $S$ is the entropy of the system, which is a measure of how many different states the system can possibly have for a given energy. For positive temperature, if two objects of different temperatures are brought into contact, energy will flow from the hotter one to the colder, cooling the former and heating the latter until equal temperatures are achieved. For negative temperature, though, if an object with negative temperature is brought in contact with an object that has positive temperature, each object tends to increase its own entropy. Like most normal objects, the latter does this by absorbing energy, but by the definition of temperature, the former does this by releasing energy, meaning the former will spontaneously heat the latter. Hence, negative temperature is hotter than positive temperature; this is a quirk of the definition of reciprocal temperature, so really what is happening is that absolute zero on the positive side is still the coldest possible temperature, absolute zero on the negative side is now the hottest temperature, and $\pm \infty$ is in the middle.

This was really just me writing down stuff that I had been thinking about a couple of months ago. I hope this helps someone, and I also await the day when TV newscasters say "complex time brought to you by..." instead of "time and temperature brought to you by...".

2013-03-20

Nonzero Electromagnetic Fields in a Cavity

The class 8.06 — Quantum Physics III requires a final paper, written essentially like a review article of a certain area of physics that uses quantum mechanics and that is written for the level of 8.06 (and not much higher). At the same time, I have also been looking into other possible UROP projects because while I am quite happy with my photonic crystals UROP and would be pleased to continue with it, that project is the only one I have done at MIT thus far, and I would like to try at least one more thing before I graduate. My advisor suggested that I not do something already done to death like the Feynman path integrals in the 8.06 paper but instead to do something that could act as a springboard in my UROP search. One of the UROP projects I have been investigating has to do with Casimir forces, but I pretty much don't know anything about that, QED, or [more generally] QFT. Given that other students have successfully written 8.06 papers about Casimir forces, I figured this would be the perfect way to teach myself what I might need to know to be able to start on a UROP project in that area. Most helpful thus far has been my recitation leader, who is a graduate student working in the same group that I have been looking into for UROP projects; he has been able to show me some of the basic tools in Casimir physics and point me in the right direction for more information. Finally, note that there will probably be more posts about this in the near future, as I'll be using this to jot down my thoughts and make them more coherent (no pun intended) for future reference.

Anyway, I've been able to read some more papers on the subject, including Casimir's original paper on it as well as Lifshitz's paper going a little further with it. One of the things that confused me in those papers (and in my recitation leader's explanation, which was basically the same thing) was the following. The explanation ends with the notion that quantum electrodynamic fluctuations in a space with a given dielectric constant, say in a vacuum surrounded by two metal plates, will cause those metal plates to attract or repel in a manner dependent on their separation. This depends on the separation being comparable to the wavelength of the electromagnetic field (or something like that), because at much larger distances, the power of normal blackbody radiation (which ironically still requires quantum mechanics to be explained) does not depend on the separation of the two objects, nor does it really depend on their geometries, but only on their temperatures. The explanation of the Casimir effect starts with the notion of an electromagnetic field confined between two infinite perfectly conducting parallel plates, so the fields form standing waves like the wavefunctions of a quantum particle in an infinite square well. This is all fine and dandy...except that this presumes that there is an electromagnetic field. This confused me: why should one assume the existence of an electromagnetic field, and why couldn't it be possible to assume that there really is no field between the plates?

Then I remembered what the deal is with quantization of the electromagnetic field and photon states from 8.05 — Quantum Physics II. The derivation from that class still seems quite fascinating to me, so I'm going to repost it here. You don't need to know QED or QFT, but you do need to be familiar with Dirac notation and at least a little comfortable with the quantization of the simple harmonic oscillator.

Let us first get the classical picture straight. Consider an electromagnetic field inside a cavity of volume $\mathcal{V}$. Let us only consider the lowest-energy mode, which is when $k_x = k_y = 0$ so only $k_z > 0$, stemming from the appropriate application of boundary conditions. The energy density of the system can be given as \[H = \frac{1}{8\pi} \left(\vec{E}^2 + \vec{B}^2 \right)\] and the fields that solve the dynamic Maxwell equations \[\nabla \times \vec{E} = -\frac{1}{c} \frac{\partial \vec{B}}{\partial t}\] \[\nabla \times \vec{B} = \frac{1}{c} \frac{\partial \vec{E}}{\partial t}\] as well as the source-free Maxwell equations \[\nabla \cdot \vec{E} = \nabla \cdot \vec{B} = 0\] can be written as \[\vec{E} = \sqrt{\frac{8\pi}{\mathcal{V}}} \omega Q(t) \sin(kz) \vec{e}_x\] \[\vec{B} = \sqrt{\frac{8\pi}{\mathcal{V}}} P(t) \cos(kz) \vec{e}_y\] where $\vec{k} = k_z \vec{e}_z = k\vec{e}_z$ and $\omega = c|\vec{k}|$. The prefactor comes from normalization, the spatial dependence and direction come from boundary conditions, and the time dependence is somewhat arbitrary. I think this is because the spatial conditions are unaffected by time dependence if they are separable, and the Maxwell equations are linear so if a periodic function like a sinusoid or complex exponential in time satisfies Maxwell time evolution, so does any arbitrary superposition (Fourier series) thereof. That said, I'm not entirely sure about that point. Also note that $P$ and $Q$ are not entirely arbitrary, because they are restricted by the Maxwell equations. Plugging the fields into those equations yields conditions on $P$ and $Q$ given by \[\dot{Q} = P\] \[\dot{P} = -\omega^2 Q\] which looks suspiciously like simple harmonic motion. Indeed, plugging these electromagnetic field components into the Hamiltonian [density] yields \[H = \frac{1}{2} \left(P^2 + \omega^2 Q^2 \right)\] which is the equation for a simple harmonic oscillator with $m = 1$; this is because the electromagnetic field has no mass, so there is no characteristic mass term to stick into the equation. Note that these quantities have a canonical Poisson bracket $\{Q, P\} = 1$, so $Q$ can be identified as a position and $P$ can be identified as a momentum, though they are actually neither of those things but are simply mathematical conveniences to simplify expressions involving the fields; this will become useful shortly.

Quantizing this yields turns the canonical Poisson bracket relation into the canonical commutation relation $[Q, P] = i\hbar$. This also implies that $[E_a, B_b] \neq 0$, which is huge: this means that states of the photon cannot have definite values for both the electric and magnetic fields simultaneously, just as a quantum mechanical particle state cannot have both a definite position and momentum. Now the fields themselves are operators that depend on space and time as parameters, while the states are now vectors in a Hilbert space defined for a given mode $\vec{k}$, which has been chosen in this case as $\vec{k} = k\vec{e}_z$ for some allowed value of $k$. The raising and lowering operators $a$ and $a^{\dagger}$ can be defined in the usual way but with the substitutions $m \rightarrow 1$, $x \rightarrow Q$, and $p \rightarrow P$. The Hamiltonian then becomes $H = \hbar\omega \cdot \left(a^{\dagger} a + \frac{1}{2} \right)$, where again $\omega = c|\vec{k}|$ for the given mode $\vec{k}$. This means that eigenstates of the Hamiltonian are the usual $|n\rangle$, where $n$ specifies the number of photons which have mode $\vec{k}$ and therefore frequency $\omega$; this is in contrast to the single particle harmonic oscillator eigenstate $|n\rangle$ which specifies that there is only one particle and it has energy $E_n = \hbar \omega \cdot \left(n + \frac{1}{2} \right)$. This makes sense on two counts: for one, photons are bosons, so multiple photons should be able to occupy the same mode, and for another, each photon carries energy $\hbar\omega$, so adding a photon to a mode should increase the energy of the system by a unit of the energy of that mode, and indeed it does. Also note that these number eigenstates are not eigenstates of either the electric or the magnetic fields, just as normal particle harmonic oscillator eigenstates are not eigenstates of either position or momentum. (As an aside, the reason why lasers are called coherent is because they are composed of light in coherent states of a given mode satisfying $a|\alpha\rangle = \alpha \cdot |\alpha\rangle$ where $\alpha \in \mathbb{C}$. These, as opposed to energy/number eigenstates, are physically realizable.)

So what does this have to do with quantum fluctuations in a cavity? Well, if you notice, just as with the usual quantum harmonic oscillator, this Hamiltonian has a ground state energy above the minimum of the potential given by $\frac{1}{2} \hbar\omega$ for a given mode; this corresponds to having no photons in that mode. Hence, even an electrodynamic vacuum has a nonzero ground state energy. Equally important is the fact that while the mean fields $\langle 0|\vec{E}|0\rangle = \langle 0|\vec{B}|0\rangle = \vec{0}$, the field fluctuations $\langle 0|\vec{E}^2|0\rangle \neq 0$ and $\langle 0|\vec{B}^2|0 \rangle \neq 0$; thus, the electromagnetic fields fluctuate with some nonzero variance even in the absence of photons. This relieves the confusion I was having earlier about why any analysis of the Casimir effect assumes the presence of an electromagnetic field in a cavity by way of nonzero fluctuations even when no photons are present. Just to tie up the loose ends, because the Casimir effect is introduced as having the electromagnetic field in a cavity, the allowed modes are standing waves with wavevectors given by $\vec{k} = k_x \vec{e}_x + k_y \vec{e}_y + \frac{\pi n_z}{l} \vec{e}_z$ where $n_z \in \mathbb{Z}$, assuming that the cavity bounds the fields along $\vec{e}_z$ but the other directions are left unspecified. This means that each different value of $\vec{k}$ specifies a different harmonic oscillator, and each of those different harmonic oscillators is in the ground state in the absence of photons. You'll be hearing more about this in the near future, but for now, thinking through this helped me clear up my basic misunderstandings, and I hope anyone else who was having the same misunderstandings feels more comfortable with this now.

2010-01-17

Science Bowl a Learning Experience? Not Anymore

Our school's team (of which I am a member) lost the regional competition in the semifinals yesterday.
Yeah, yeah. I can accept that.
What saddens me though is that after having done this for 3 years, I will never be able to do it again; I can never redeem myself.
[sob]
But the thing I'm hearing from my parents and others is that it's all a "learning experience".
It is. But that's not all it is. To say that it is all a "learning experience" is to totally miss the point.
If I wanted a learning experience, I could have done problems from a book, signed up for a class, or attended a super-special seminar.
I have a bit of a competitive streak (though definitely not as much as some people I know). I wanted to win, especially after last year's similar defeat in the regional semifinals.
Hence, I did Science Bowl. Quantum ElectroDynamics (QED).
Furthermore, Science Bowl requires one to learn a bunch of trivia that isn't really useful (until much higher-level applications) outside of Science Bowl itself. While I won't say that all that I have learned has suddenly become for nought, it saddens me that I don't get another chance at this.
I lost. If it was a learning experience, I could use the lessons of my failures in the competition the next time around.
Except, for me, there is no next time around. This means that Science Bowl can no longer be a learning experience for me.
So, the people who call it just a "learning experience" (to console me or whatever) are being willfully blind to the other half of the competition - the competition.
That said, I would like to see our future teams do well. I plan to talk to (and maybe even coach a little, given our actual coach's frequent absences) the 2 team members (both in 10th grade) who will be on the team next year on what to do then.
Hopefully, that will work out.