The contradiction in our dark matter

The contradiction in our dark matter

It is straightforward to falsify theories that make clear predictions, should they happen to be wrong. It is impossible to falsify ideas that involve invisible components.

Specific theories of dark matter (e.g., WIMPs) can be seen to be increasingly unlikely, but can they be falsified outright? If we decide that WIMPs have been practically falsified – to all intents and purposes – that doesn’t mean dark matter is wrong, just that we’ve spent the past four decades building dozens of giant experiments costing billions of dollars and frustrating thousands of careers barking up the wrong tree. The unseen forest that is the ‘dark sector’ is vast; there are limitless opportunities to bark up other wrong trees.

How could we know if the entire dark sector is a non-entity? On many occasions, I’ve had colleagues say to me that they would “only consider MOND as a last resort.” OK, so when do we know we’ve reached that point? After an eternity of searching for unseen, undetected non-entities seems a little late.

This is the bind we’re in. Most of us (scientists working in the field) are unwilling to consider something as radical-seeming as MOND until dark matter has been falsified. But dark matter cannot be falsified. So we patch up our hypotheses accordingly, rinse and repeat with every new crisis#. This has been going on for nearly my entire career, to the point that I now see junior scientists who seem to think this is how science is supposed to work. Why wouldn’t they? They’ve known nothing else.

I empathize. I started from the same place: there has to be dark matter, it has to be non-baryonic, it is almost certainly a new particle and that new particle is almost certainly a WIMP. It was the hardest thing to realize and accept that we could be wrong, wrong, wrong, and wrong on all counts. I worked incredibly hard to avoid that conclusion. I take solace only in the fact that at every epoch in human history we’ve had a cosmology that we were absolutely sure$ was right but that later turned out not to be. Maybe we’re a blessed generation who finally got it right. Or maybe we’re just the latest in a long line of brilliant savants who fooled themselves into thinking we understand more than we actually do.

I remain unwilling to say that dark matter has been falsified because I don’t think it is falsifiable. That, in itself, is usually considered to be a bad thing for a scientific theory. Perhaps at some future point it will be portrayed that way for dark matter, but in the present I’ve heard plenty of scientists pretend like it is somehow a good thing. Certainly it makes a fertile playground for theorists, so I guess good/bad is a matter of perspective. I suppose an infinite forest of trees in a conveniently unobservable dark sector is an irresistible opportunity for a dog! with an insatiable bladder.

I do, however, think that dark matter has been practically falsified. If I’m wrong about that, it is possible to demonstrate. To be clear, I don’t just mean WIMPs. I mean dark matter as an explanation for the mass discrepancies observed in extragalactic systems. No kind of dark matter* can suffice. I came to this conclusion before I really knew anything about MOND, which is why I was receptive to it. Others haven’t had that experience, so aren’t.

I don’t expect other scientists to accept that an entire paradigm is wrong because I say so. I do, however, expect them to acknowledge that the dark matter hypothesis should be falsifiable. It is incumbent on each scientist to establish for themselves criteria by which that conclusion could be reached, should it be appropriate.

For me, the problem is the contradiction present in the dynamical data for galaxies. We simultaneously require galaxy disks to be maximal and also for them not to be maximal. The only way to avoid this contradiction is to engage in fine-tuning: one must build a model in which everything works out just so. Since dark matter is not otherwise falsifiable, a requirement for fine-tuning& is pretty much the worst thing we can say about it.

The contradiction as framed by others

The contradiction that concerns me has certainly been noticed by others. Usually, they choose to come down on one side or the other. The problem is that they’re both right.

On the one hand, it is clear observationally that the luminous mass matters to the dynamics of galaxies. For example, Swaters et al. note that

the luminous mass dominates the gravitational potential in the central regions, even in low surface brightness dwarf galaxies

which is to say, disk galaxies are maximal. We’ll explore what that means below, but you can see the effect by eye:

Rotation curves color coded by galaxy surface brightness. There is an almost perfect rainbow variation that is plain to see. The gravitational potential traced by the rotation curve correlates well with the distribution of stellar mass. (Adapted from McGaugh 2020.)

On the other hand, there are no residuals from the Tully-Fisher relation.

Residuals from the Tully-Fisher relation as a function of size at a given mass. Blue points are star-dominated galaxies; cyan points are gas-dominated. Compact galaxies are to the left, diffuse ones to the right. The red dashed line [δlog(V)/δlog(R) = -1/2] is what should happen with maximal disks: more compact galaxies should rotate faster at a given mass. This is what Newton predicts, but galaxies didn’t get the memo. (Adapted from McGaugh 2005.)

The inference from this observation is that galaxy disks cannot be maximal. As Courteau & Rix put it,

The case of δlog(V)/δlog(R) = -0.5 expected% for a maximal disk is ruled out

which is to say, disk galaxies are not maximal – even the high surface brightness (HSB) disks that dominated their sample.

They say this because maximal disks – or even those that contribute noticeably to the mass budget at small radii – predict deviations from the Tully-Fisher relation. They did this by intentionally focusing on the point where the stars should contribute the most. I’ve shown similar things many times, so here let’s let LCDM advocate Frank van den Bosch demonstrate it:

Model galaxies (left) of the same mass but with different spin parameters (van den Bosch 2001). In this model, spin determines size and surface brightness: lower spin galaxies are more compact and have higher surface brightness. This causes them to evince higher velocities that should cause residuals from Tully-Fisher, but these are not observed (right, McGaugh 2021).

So galaxies cannot be maximal. Only they have to be to have the observed correlation between the luminous mass distribution and the kinematics, as seen in the first figure above. Well, which is it? Are all galaxies maximal? Or none? Or can they both^ be right?

Maximal disks

Maximal disk is a technical term that is well known to people who work on the subject but not outside that relatively narrow&& field. So what do we mean by this term?

The rotation curve of NGC 6946. The solid blue line shows the expected rotation curve for the visible mass including both stars (dashed line: stellar disk; dot-dashed line: bulge) and atomic gas (dotted line). The stars suffice to describe the inner rotation curve, illustrating the case of maximum disk. The red points are the rotation curve measured optically (Daigle et al. 2006; Epinat et al. 2008); the orange points that measured with radio interferometric observations of the gas (Boomsma et al. 2008; I’ve suppressed the overlap for clarity). The thin line is the prediction of MOND; the black line is the implied dark matter distribution.

In essence, a maximal disk is one in which the stars provide practically all the mass at small radii. The depiction of NGC 6946 above illustrates maximum disk. Sure, the rotation curve flattens out at large radii and we need to invoke dark matter. But the observed stars explain the amplitude and shape of the inner rotation curve quite well. In this case there is a compact bulge at the center of the galaxy that causes a sharp rise in the rotation curve right from R = 0. (This is an example of Renzo’s Rule.) The disk (plus bulge) is maximal in the sense that we cannot attribute any more mass to them without exceeding the observed rotation curve.

The amplitude of the portion of the rotation curve due to the stars depends on their mass-to-light ratio. While this cannot exceed maximum disk, it could be lower. So conceivably, the good match to the shape of the observed rotation curve is a chimera, and really this galaxy is dark matter dominated. That is a possibility many seem to have embraced, but while it might work for the disk, it does not work for the bulge. As we suppress the contribution of the stars as Courteau & Rix argue we must, then the rotation curve looks more and more like that of the dark matter halo alone. That goes up and flattens out (and ultimately must turn over again somewhere beyond the edge of the data) but it has no features.

The rotation curve of NGC 6946 as above but with the stellar mass reduced by a factor of two. This disk is submaximal, contributing less than the implied dark matter.

One thing the rotation curve due to the dark matter halo cannot do is go up then down then up again. Yet that is exactly what it needs to do to explain the inner peak inside 1 kpc if the bulge component is not maximal. I suppose we could have a smaller dark matter halo inside the main dark matter halo that does this, but its mass distribution would have to be practically identical to that of the bulge. That’s insane. Why would we invoke a second dark matter halo when the stars are right there?

In this case, the stars have the right mass for what we expect from stellar populations. There was a long debate historically about whether the optical band mass-to-light ratios for maximum disk were consistent with those expected from stellar population synthesis models. For a long time they looked close but a bit high. This difference has pretty much gone away now that we have access to near-infrared data: the two are consistent, and having a stellar mass-to-light ratio much below the maximum disk value for high surface brightness disks becomes problematic from a population perspective.

Indeed, the 3.6 micron M*/L = 0.37 M/L in the case depicted for NGC 6946 with a maximum disk. That’s reasonable but on the low side for what is plausible for the stars in a mature spiral galaxy like this. Halving that strains credulity, so there is no room for a second inner halo, or even for the cusp predicted** for the primary cold dark matter halo. Stellar mass really does seem to dominate in the inner parts, just as Swaters et al. said.

LSB galaxies

The issue that confounded me was whether the low surface brightness (LSB) galaxies I was working on were maximal or not. My inital expectation was that LSB galaxies would be stretched out versions of HSB galaxies. I expected them to shift off of the Tully-Fisher relation and follow the line δlog(V)/δlog(R) = -0.5. They did not do that. If I didn’t have them be maximal, I found that I could explain pretty much any slope other than the one observed (δlog(V)/δlog(R) = 0). That required fine-tuning to perfectly balance the lesser contribution of stars in LSB galaxies which we had to back fill with dark matter just so. I spent ages running around in circles trying to make that work. Every time I thought I had succeeded, I realized I had assumed something that made it so: tautologies abound.

If we want to explain the shapes of rotation curves as seen up top, we need the stars to contribute to the gravitational potential. For that to work for LSB galaxies, we have to turn maximum disk up to eleven:

The low surface brightness, gas rich dwarf DDO 154. The lines have the same meaning as above, with the case of maximum disk in the left panel and that natural for stellar populations at right. The difference in disk mass is a factor of ten.

A crazy-high stellar mass-to-light ratio is what happens if we just ignore what we know about stars and just focus on the kinematics. But we do know a lot about stars. Population models indicate stellar masses that are very submaximal. Even boosting the mass-to-light ratio doesn’t get us very far. LSB galaxies aren’t really maximal in the same sense as HSB galaxies, and there is even less room for the expected cuspy halos that are already problematic when the stellar contribution is small.

Fine tuning is unavoidable

Even if we ignore what we know about stars, we still have a fine-tuning problem. The lack of a shift in the Tully-Fisher relation with either surface brightness or radial size implies that disks are all the same mass surface density. So we observe a wide range of surface brightness, but the surface mass density is always the same. That makes no sense, and is just another example of squeezing the toothpaste tube: we can make a model look OK from one perspective as long as we don’t look from another.

Worse, we still need to explain the role of the luminous mass in LSB galaxies. These are dark matter dominated at almost all radii, and yet the distribution of the observed stars and gas is predictive of the kinematics. This is a contradiction to Newtonian dynamics. The only theory that does this right – and predicted it a priori – is MOND. But that’s too horrible to contemplate, so we shield our eyes and ignore%% one or the other set of inconvenient facts. As a result, the field has become moribund, and will remain so until we free ourselves of our invisible demons.


#We’ve experienced so many crises that we seem no longer able to recognize new ones. JWST observations of high redshift galaxies follows a well-worn trajectory: an observation that contradicts the standard model is made, much huffing and puffing ensues, the theorists get to work constructing implausible models, these are accepted as patching up the hypothesis (whether satisfactory or not), and the field moves on as if nothing happened.

$To give one historical example, prior to Hubble’s discoveries in the 1920s, it was thought that the Milky Way was the entire universe. Certainly there were no other galaxies comparable to the Milky Way:

“No competent thinker, with the whole of the available evidence before him, can now, it is safe to say, maintain any single nebula to be a star system of coordinate rank with the Milky Way. A practical certainty has been attained that the entire contents, stellar and nebular, of the sphere belong to one mighty aggregation.” [i.e., the Milky Way]

-Agnes Mary Clerke in The System of the Stars (1890)

!It used to be that one would not claim a detection of dark matter until all astrophysical alternatives had been exhausted. Now it seems to be the fad to claim a detection first on the off-chance it works out later. I already peed on that tree! It’s mine!

*Excepting some sort of hybrid “dark matter” that is invented to do what ordinary dark matter cannot. By ordinary I mean CDM, WDM, SIDM, and every other variation on particle physics that simply invents new mass with no consideration of how the observed galaxy dynamics comes about. That would include primordial black holes and various macroscopic DM ideas (e.g., MACHOS, strange nuggets). Coming up with half-baked ideas for new particle dark matter is big business these days, but any idea not informed by observed astrophysics (which are most of them) is doomed to fail.

Examples of hybrid dark matter that are informed by observed astrophysics include dipolar dark matter and superfluid dark matter. Regardless of whether these specific cases are viable, the point is that the observed dynamics are a fundamental aspect if nature and require a commiserate explanation. Simply throwing in some extra mass with some fine-tuned feedback models can never provide a satisfactory explanation. Note that coming up with extra mass is mostly done by particle phenomenologists while feedback models are built by numerical astrophysicists. There is very little overlap between these communities; they pretty much just take it on faith that since dark matter has to exist, the part they don’t know about will magically work out.

&The classic example of fine-tuning in the sense that I mean is the Ptolemaic model of epicycles and deferents. If one adds enough of these and tunes them just so, anything can be fit. Note that epicycles are not explicitly falsifiable for this reason; we rejected them because they got ridiculously complicated and there turns out to be a more parsimonious explanation. The same thing holds now for dark matter and MOND.

%This slope is expected because Newton teaches us that V2 = GM/R. δlog(V)/δlog(R) = -0.5 follows from taking the logarithm of this at fixed mass. Galaxies are observed to span a large range of radius at a given mass, but not a corresponding range in circular velocity.

^Yes, they can both be right, but not with dark matter. Only MOND naturally explains both observations simultaneously.

Also, for the hyper vigilant, Courteu and Rix (1999) use a slightly different definition of velocity than I do in this Tully-Fisher residual plot. I went through all that in McGaugh & de Blok (1998) and in McGaugh (2005) and it makes no difference to the discussion here.

&&A vote we held at a conference on disk dynamics in Rome in 2000. The question of whether disks were maximal was posed; most people voted no based on the statistical lack of residuals from Tully-Fisher. After the vote, one of the dissenters noted that those who voted in favor of maximal disks were the people who actually worked in the subject. Those of us with other concerns were persuaded by the statistical evidence because we didn’t engage with the details of real, individual galaxies in the same way.

**The NFW halo famously gets the inner shape of the rotation curve wrong (the cusp-core problem), but it is also wrong at intermediate radii and at large radii. Other than that it’s great.

%%A common excuse I here for this behavior is that galaxies are “small” and nonlinear – complicated entities that we can never hope to understand, so whatever they do can be ignored as irrelevant. As a scientific argument, that’s pathetic. Galaxies should be complicated in LCDM, but in observational reality they’re kinematics are sufficiently simple that they obey a single effective force law. That’s one thing they should not do, just as a complicated set of epicycles and deferents shouldn’t always add up to the inverse square law.

The blinding influence of the Bullet Cluster

The blinding influence of the Bullet Cluster

A few posts back, the issue of the Bullet Cluster came up in the comments. I was still working my way up to addressing that in a full post, so I limited myself to making a sociological observation based on my experience:

People who invoke the bullet cluster in response to queries about MOND are usually doing so to deflect from the need to engage with it.

That is a general observation that was not aimed at anyone in particular, but it was made in response to a comment citing Don Lincoln saying this, so Dr. Lincoln hopped into the comments to reply:

Or…and bear with me here…some people actually find the Bullet Cluster (and the DF2/DF4 situation) to be persuasive.
FWIW, I’m a physicist, working in the field. And my views have changed over the years. In the early 1990s, I was pretty sure that dark matter was MACHOs. When MACHO, OGLE and all them disproved that conjecture, I then strongly favored MOND (broadly defined…not necessarily Milgrom’s conjecture, but rather the much vaguer paradigm that inertia or gravity needed some improved understanding). However, in the modern day, Bullet and Dragon Fly have again changed my leaning.
Yes, yes, the Bullet Cluster is moving awfully fast. Yes, yes, it’s a problem for LCDM. But this isn’t the same level of problem as it poses for modified physics. For LCDM, it is simply an unusual entry on the tail of a known velocity distribution, while the problems it poses for MOND-ish issues is more central.
Now, this is your page and you are allowed to have an echo chamber and sycophants…no problem. But to characterize those who disagree with you as somehow not being thoughtful is just…sloppy. And dismissive. And patronizing.
The broad community might be wrong. But, you know what? So could you.
FWIW, it will be difficult to convince me of >>ANY<< solution without DM being produced in particle accelerators. Indirect measurements are background-prone. And direct measurements, while super helpful, will tell us where to look in accelerators. (Think the DAMA debacle.)
My fear is that dark matter is real, but it only interacts gravitationally or far weaker than the weak force. If that’s the case, our grandkids will be having this argument.

Don Lincoln, posting as Science Guy, 4 June 2026

In my experience, Dr. Lincoln is one of the more reasonable people connected to this debate. He says some things that are fair, and also some things that are revealing of the current sociology, so it is worth exploring point by point.

Or…and bear with me here…some people actually find the Bullet Cluster (and the DF2/DF4 situation) to be persuasive.

The sentence starts with sociology: the “bear with me” trope is an assertion of reasonableness before it is demonstrated, which it may or may not be – often it is employed by people who think they’re being reasonable when really not so much. But in this case, yes: lots of people find the Bullet Cluster to be persuasive. My complaint is not that. It is that they cite the Bullet Cluster as an excuse to not think further about MOND and stifle debate. That has been my lived experience.

The Bullet Cluster achieved the status of a totem object long ago:

I myself find clusters persuasive. As I’ve written repeatedly, they pose a real problem for MOND. The Bullet Cluster is just one example, and an extremely weird case at that. The universe is big, so there’s always a unicorn somewhere: astronomers long ago learned not rely too much on the weirdest object in a category. So I’m more impressed when people cite clusters in general as a problem for MOND, both because that’s true, and because it evinces awareness of the subject beyond the totem that has become the go-to code word for dismissing a predictively successful paradigm without understanding it.

Dr. Lincoln also cites “the DF2/DF4 situation.” I’m not sure why that comes up if the Bullet Cluster by itself is entirely persuasive. This is a topic very much in my expertise, and I do not find it all that persuasive. I spent some time reading up on this situation – the data keep changing – in hopes of saying something quantitative here, and that left me feeling it was even less persuasive than I had given it credit for.

The DF2/DF4 (and now DF9) galaxies have become prominent examples of galaxies that appear to lack dark matter. Such objects are a problem for MOND, because you can imagine stripping away dark matter (though it is hard to do) but you can’t switch off the force law. However, the prominence of these particular objects has more to do with advertising (a ridiculous amount of attention has been devoted to these few weird objects) than with how convincing they are. That’s not to say they aren’t important, just that their importance is exaggerated.

These DF galaxies are a good example of cognitive dissonance in action. People give more weight to evidence that supports what they already believe, and less to that which supports things that don’t. In this case, the other evidence is all the other galaxies in the universe. This is a classic case of missing the forest for a few outlying trees.

I used to spend a lot of time fact-checking claims# to falsify MOND. So when DF2 was first announced as a problem for MOND with great fanfare, I went to check. It was not. Indeed, had I known of this object’s in advance, I could have used MOND to correctly predict its velocity dispersion as I did with Crater 2 and the 30+ dwarf satellites of Andromeda. This rather quenched my enthusiasm for fact-checking every claim, so I vowed* not to spend more time doing so.

This is a damned-if-you-do, damned-if-you-don’t situation. If incorrect claims are left uncontested, the community seems to assume they are correct. If one spends the time to write a paper, you are diverted from your own research. Once published, much of the community remains unaware of the rebuttal, or chooses to believe their preferred narrative (another example of cognitive dissonance).

So before writing this post, I found myself breaking my vow. These are interesting objects, and there is now a new one (DF9). All three of these dwarfs have unusually large diameters and low velocity dispersions, about 8 km/s. That’s pretty much what we expect for the stars we see. No dark matter, no MOND. Even though it is really weird to find galaxies without dark matter in a universe made of dark matter that requires dark matter to make galaxies, it is even worse for MOND, if true.

There are devils in the details. The data for DF2 have changed repeatedly, both its distance and velocity dispersion. Which version to believe? Working my way through the literature, I found the statement “We find an instrumental resolution σinst = 0.375 Å (13.0 km s−1)” and decided to stop right there. A rule of thumb in this business is that you shouldn’t try to measure a velocity dispersion smaller than your instrumental resolution because, well, you can’t resolve it. 8 km/s is smaller than 13 km/s. Now, in principle, you can tease more information out of the data, but that’s hard to do. In the best case, the two dispersions add in quadrature, so to infer 8 km/s, what you’ve really observed is √(82+132) = 15 km/s. That’s not much different from the instrumental resolution, and I’ve seen plenty of claims where that obscured a correct MOND prediction (e.g., Cetus). For DF2, we expect an intrinsic velocity dispersion of ~13 km/s, which is sensitive to the distance that keeps changing and to the EFE of the rough neighborhood in which these dwarfs find themselves. That corresponds to observing √(2*132) = 18 km/s. So we have to be able to distinguish between 15 and 18 km/s. That can be done, but it is a lot less clean than the difference between 8 and 13 km/s sounds.

There are other corrections for broadening and binaries, so the above is the sanitized version. Binaries are hard to correct for in these unresolved objects. In nearby ultrafaint dwarfs where we measure velocity dispersions one star (or unresolved binary) at a time, the correction can be dramatic. Boötes III provides a rececnt example, having: “a velocity dispersion of σv=1.69+1.03−0.85 km s−1, about six times smaller than the previously reported 10.7±3.5 km s−1.” That takes it from having lots of dark matter, completely inconsistent with MOND, to bang on what MOND predicts. So you can perhaps appreciate my relutance to put too much credence in every claim made about measurements of these ultrafaint/ultradiffuse galaxies. This is the hardest place to work, and my experience has been that as the data improve, so too does agreement with MOND.

Is DF2 even a significant problem? Accepting the updated numbers as stated, corrected (and perhaps overcorrected) to reflect all the above effects, Keim et al. report two independent measurements for DF2 that give 6.3+3.1-3.7 and 9.2+3.8-4.5 km/s. In our paper we found that MOND predicts 13.4+4.8-3.7 km/s. Those are all consistent within the stated uncertainties: the one-sigma error bar of the lower measurement overlaps with that of the prediction,& and the higher measurement is in as good agreement with MOND as could be expected given the uncertainty on both measurement and prediction. Dark matter advocates would be lauding such agreement to high heaven if they could% make this prediction.

That’s a long segue for a parenthetical comment by Dr. Lincoln. It takes a lot of words to address five, and it isn’t for me to judge if the interpretation he seems to take for granted is better or worse than the alternative I describe. Even so, there’s a saying about this that might apply.

OK, let’s return to Dr. Lincoln’s comments. In case you’d forgotten – which I almost did, having spent so much time trying to track down DF2 data – this is a post about the sociology illustrated by Dr. Lincoln’s comments.

my views have changed over the years. In the early 1990s, I was pretty sure that dark matter was MACHOs. When MACHO, OGLE and all them disproved that conjecture, I then strongly favored MOND (broadly defined…)

My views have changed over the years too. In the early 1990s, I was completely sure that the dark matter was WIMPs. Had to be. I remember joking with other astronomers about how futile the search for MACHOs would be. As Rob Kennicutt put it in 1995 at IAU 171: “What are the MACHO people doing? Have they never heard of Big Bang Nucleosynthesis?” I recall shrugging and thinking that these projects wouldn’t detect dark matter, but they would provide a great variable star database. And so it came to pass.

We have here two different recollections of the same time period. Both are valid, but which is a fair representation of the community? Could be both. Gradually I’ve come to recognize that there were [at least] two distinct communities working on this issue, one in physics and one in astronomy. They don’t need to be echo chambers to develop mutually exclusive attitudes; the networks of communication can be broad and yet largely distinct. There is also a temporal aspect. I’m told by astronomers a few years older than myself that baryonic dark matter (particularly brown dwarfs) was favored in the late 1970s; indeed, that it seemed pretty obvious at that time. This changed quickly and was mostly (though never totally) supplanted by non-baryonic dark matter by the mid-1980s. Even that depended on the sub-community – those of us more concerned with cosmology rejected baryonic dark matter as nonviable while it remained reasonable to those of us more concerned with the dynamics of individual galaxies. I had a foot in both camps, and it metaphorically tore me apart.

In the mid-1990s, I was wrestling with our new data for low surface brightness galaxies. It did not make sense in terms of dark matter. Any kind of dark matter. I began to fear that the entire dark matter paradigm was no longer viable, a conclusion I fought tooth and nail to avoid. I worked much harder trying to save dark matter than I ever did subsequently working on MOND. More importantly, I think I was only receptive to MOND^ because I was already deeply concerned for the viability of dark matter. There lies the great schism: to me, dark matter was already practically falsified. No one else had that horrible, visceral experience – it was like losing a dear friend – so most of the community glibly ignored the surprising successes of MOND.

Maybe I was wrong to doubt dark matter? This is why I challenge other scientists to state their own criteria for its falsification. I don’t expect them to accept something so important just because I say so. But I do say so, and for good reasons – reasons most of them seem to be unaware of, so we’re starting from very different places. But if dark matter is a scientifically valid physical hypothesis, then it should be falsifiable. How can we tell if it is wrong?

I’m old enough to remember when cuspy dark matter halos were an absolute prediction of cold dark matter. That prediction failed, yet we find ways for CDM to persist. If we gave up on CDM as easily as we give up on MOND, we would have stopped talking about it thirty years ago.

Right. Where were we?

Yes, yes, the Bullet Cluster is moving awfully fast. Yes, yes, it’s a problem for LCDM. But this isn’t the same level of problem as it poses for modified physics. For LCDM, it is simply an unusual entry on the tail of a known velocity distribution, while the problems it poses for MOND-ish issues is more central.

I agree with this and I don’t. The issue is more central for MOND because a theory that seeks to supplant dark matter appears to need dark matter. So why bother?

It’s a good point, and as I’ve said over and over and over again, I find the situation in clusters profoundly dissatisfactory for MOND. Perhaps it even falsifies it. But a loss for MOND isn’t an automatic win for non-baryonic dark matter (which also requires new physics), so we bother because of everything else MOND does right that dark matter does not.

A common misconception here is that the unseen mass in MOND is necessarily non-baryonic. That’s a logical fallacy that stems from the sloppiness of the term “dark matter” which many of us equate with non-baryonic dark matter. Non-baryonic dark matter requires new physics. MOND requires new physics. So it sounds like we need double-new physics here when in fact we “just” need some undetected baryons$ – something we need in both paradigms:

The agreement or mismatch between baryonic mass and observed velocity in LCDM (top) and MOND (bottom). As discussed before, the accounting of baryonic mass LCDM is better in galaxy clusters while MOND is better in all other gravitationally bound extragalactic systems.

What about the collision velocity? How bad a problem this is for LCDM depends on how unusual an entry it is in the tail of a known velocity distribution. There are many assessments of this; they range from “pretty unlikely” to “oh hell no.” For example, Lee & Komatsu found “the probability of finding 3000 km s-1 in (2-3)R 200 is between 3.3 × 10-11 and 3.6 × 10-9.” That’s worse than an unusual entry, that’s “oh hell no.” Various other workers hedge this way and that so it comes across as less improbable, but the most optimistic assessment I saw recently (and cannot now relocate) is that there is only a 10% chance of the Bullet Cluster existing in the volume of the universe that contains it. That’s over the whole sky, not just the 5% of the sky that has been surveyed in a way that could find it. So pretty unlikely, but not necessarily fatal. Ascencio et al. attempt to quantify this, finding that the “Bullet Cluster is in 2.78σ tension” with LCDM – that’s bad, but not fatal – but that including El Gordo results in a “combined tension” “estimated as 6.43σ”. That exceeds the usual 5σ threshold for fatal.

There is the obvious temptation to believe the assessment we prefer. There are analyses in which the Bullet Cluster collision speed doesn’t seem all that bad, though it never looks right. I’ve worked on this too, and we found it was pretty much impossible, and it looks to me like the more favorable analyses may be squeezing the toothpaste tube to make the collision lees improbable by shifting that to improbability in other parameters. One I recall requiring a ridiculously perfect, head-on bullseye collision.

In contrast, the high collision speed is entirely natural in MOND. That’s what happens as a consequence of the enhanced long-range force law. It just falls out, no muss, no fuss. So it isn’t just improbable in LCDM; it happens because of MOND, just like every other surprising result for the past forty years. This is why I say galaxy clusters ruin everything. We can’t say they favor one paradigm or the other without under-weighting some aspect of the data.

I guess that was enough science, because now we get to sociology:

to characterize those who disagree with you as somehow not being thoughtful is just…sloppy.

I may be wrong, but I am pretty much the antithesis of sloppy. Sloppy is describing me of characterizing those who disagree with me as not being thoughtful. I do not doubt the intellectual energy of people working on dark matter. I do doubt how deeply most of them have allowed themselves to think about MOND. I know this problem well, as I suffered it myself early on. MOND is too horrible to contemplate. Surely we can’t be that wrong! So we find some reason not to do so. The most common reason such people cite is the Bullet Cluster. It is my lived experience that lots of scientists (though certainly not all) fall into that category. If anybody finds that insulting, then they should do better to not be that person. I find it strange if not surprising that a scientist is offended at being challenged to think more about a topic about which he is, to use his word, dismissive.

FWIW, it will be difficult to convince me of >>ANY<< solution without DM being produced in particle accelerators. Indirect measurements are background-prone. And direct measurements, while super helpful, will tell us where to look in accelerators. (Think the DAMA debacle.)
My fear is that dark matter is real, but it only interacts gravitationally or far weaker than the weak force. If that’s the case, our grandkids will be having this argument.

OK, so this makes sense, but just let me note that it assumes that the solution to the mass discrepancy problem is a novel form of dark matter that can be created in an accelerator. That would be great if it could be done. All plausible dark matter particle candidates for which such a detection is plausible are pretty much excluded at this point, and building a bigger accelerator promises zero surety of success.

More generally, I don’t accept that the solution is a novel form of particle dark matter. That’s very much not in evidence! But I do share his fear of a non-interacting particle. I call that the Angel Particle because, indeed, we can argue about how many angels can dance in the core of a neutron star forever and ever. I have a greater fear, that by attending only to the failings of MOND while ignoring its successes – and a lot of scientists are guilty of that – we help to usher in a new age of dark epicycles.


I will not be available to respond to comments for a while, so I have deactivated them for this post.


#Initially it came as a surprise how often claims to falsify MOND did not hold water. That has since become my experience with the vast majority of such claims. There is a powerful temptation to see in the data what one wants to see.

*For the most part, I’ve kept to this vow. It is hard not to check what MOND predicts when one hears claims that “it can’t do this!” when my experience, over and over again, is that usually it can. Not always, but usually.

&I decline the opportunity to refine the prediction by chasing the continually changing distance estimate. Regardless of how close the current numbers are to the underlying truth, this is certainly not a five sigma exclusion of MOND.

%The prediction cannot be made with dark matter. I’ve tried. Indeed, I’ve tried many things – there are multiple paths by which one might attempt to do this, and they do not yield a consistent answer. Worse, the one thing that is clear is that for low mass galaxies like these, there is a lot of scatter in DM halo mass at a given luminosity. The natural prediction from this is that galaxies of the same (low) luminosity have very different velocity dispersions. Yet the opposite is observed; luminosity is strongly predictive of kinematics. So we can fault MOND for predicting 3 km/s for Antlia 2 when 6 km/s is observed because it makes a prediction. We don’t fault LCDM because it does not make a comparably precise prediction to test.

^Unlike Dr. Lincoln, I mean MOND specifically. I considered other modified gravity theories, but they fail. It is very obvious when the force law is wrong. Only MOND worked at the time and continues to work today. That it even comes close is telling us something profound, a lesson the community steadfastly refuses to learn.

$I can hear from across the ocean Dr. Kroupa shouting that they’ve solved this problem.

Missing baryons: LCDM and MOND compared

Missing baryons: LCDM and MOND compared

In the last few posts we’ve discussed the local missing baryon problem in extragalactic objects spanning over ten orders of magnitude in mass from tiny dwarfs to rich clusters of galaxies. This discussion has so far been entirely in the context of LCDM. So – how does LCDM compare with MOND?

As a refresher, these are the data we’re trying to understand:

The Extended Baryonic Tully-Fisher Relation (BTFR) for extragalactic objects. Rotating galaxies are shown as circles; objects dominated by pressure support as squares. Adapted from Fig. 3 of McGaugh et al. (2026)

The flat rotation speed Vf is an indicator of the dynamical mass – that of the dark matter halo and all the baryons it contains in LCDM, and that of all the (presumptively baryonic) mass in MOND. In LCDM, it would be satisfactory for the baryon fraction of each object, mb = Mb/M200, to be equal to the cosmic baryon fraction (fb = 0.157 according to Planck). For MOND, what you see is supposed to be what you get, so the baryon fraction should be one.

As we saw previously, mb = fb for rich clusters of galaxies. There is no local missing baryon problem for galaxy clusters: a satisfactory result. However, as we look at smaller systems, observations depart from this ideal. They do so systematically, with our accounting of baryons falling progressively shorter of our expectation as we examine progressively lower mass objects. This deficit is illustrated by the gray region here:

The baryonic mass fraction as a function of baryonic mass. The horizontal line is the cosmic baryon fraction fb = 0.157; the shaded region depicts the quantity of baryons that are missing. Adapted from Fig. 4 of McGaugh et al. (2026)

Everything is fine for clusters at the high mass end (Mb > 1014 M), and many people reasonably interpret that as corroboration of LCDM. For lower mass groups and bright galaxies, there is a deficit of a factor of two or three: an issue, but nothing too concerning by the standards of extragalactic astronomy, so this is widely ignored outside the community that works on it. The implicit assumption is that it’ll work out. But the magnitude of the problem continues to grow for smaller objects, becoming already an order of magnitude for intermediate mass galaxies. Not tiny dwarfs, just middle of the road spirals. The smallest mass dwarfs are worse off yet, missing over 90% of the baryons, approaching 98% or 99%. That is not satisfactory.

Making a straight-up comparison with MOND is a little tricky because the concept of a baryon fraction is a non-sequitor. There is no dark matter halo to compare against. Instead, we return to the concept of the velocity factor. In LCDM, we relate the observed flat rotation speed to that of the total dynamical mass through Vf = fvV200. Indeed, we can ask what velocity factor we need to explain away the missing baryon problem: maybe there are no missing baryons, just a systematic divergence of the observed Vf from the halo V200. This can’t work, but it is useful to think about and provides a direct comparison with MOND.

In MOND, Mb = AVf4 where A is the normalization& of the BTFR. We can thus define an equivalent to the velocity factor, the residual velocity, taken here to be the ratio of the observed velocity to that expected for the observed mass, ΔM = Vf,obs/Vf,pred. If the mass is a good predictor of the flat velocity, then ΔM = 1. This leads to

Figure 8 from McGaugh et al. (2026): The velocity factor in ΛCDM (top panel) and the residual velocity in MOND (bottom panel) as a function of baryonic mass. The gray region illustrates where each theory gets it wrong. The limits of this log-log plot are identical so that the areas of the shaded regions are directly comparable.

This is a straight-up comparison between the theories. Both theories suffer a missing baryon problem, but at different scales. The magnitude of each problem is indicated by the area of the shaded regions. (There is a dearth of data in our study* from 1013 < Mb < 1014 M, so we’ll just ignore that here.)

LCDM is spot on for clusters over the range 1014 < Mb < 1015 M: fv = 1 suffices to explain the data. Outside of that range, fv must increase systematically to make up for what we previously attributed to missing baryons. In effect, we’re making the dark matter halos smaller so that the baryon fraction works out. As noted before, this can’t work, as rotation curve fits restrict the viable range of the velocity factor to 1 < fv < 1.4, but we need it to grow to fv = 5. That’s silly: at that point, the dark matter halo is contributing so little to the observed dynamics that we wouldn’t infer its existence at all.

MOND is spot on over the range 5 x 105 < Mb < 5 x 1012 M: the data are consistent with ΔM = 1. It falls short for rich clusters, where the observed mass of baryons in the intracluster medium (ICM) and the stars in galaxies predicts only ~80% of the observed velocity. This is the residual mass discrepancy in MOND.

For perspective, it helps to plot the linear baryon fraction. The astronomical scales of astronomical data oblige us to use logarithmic scales in many circumstances, but this may lead one to under-appreciate the scale of the issue. So here is the baryon fraction again, in both LCDM and MOND, this time with a linear scale:

The baryon fraction in LCDM (top) and MOND (bottom) as a function of mass. The scatter is an artifact of the propagation of errors when dividing one large, uncertain number (baryonic mass) by another large, uncertain number raised to a power (Vf3 in the top panel, Vf4 in the bottom). The data and their intrinsic scatter are the same but the scatter looks worse in the bottom panel because of the extra power of Vf. (I ran out of patience translating every single datum; some of the least accurate data fall off the edge of this plot.)

Individual galaxies and groups of galaxies are missing a lot of baryons in LCDM. This is not a subtle problem. It is not explained by simulations, nor am I aware of a satisfactory% explanation. Worse, the apparent reason that we infer all these missing baryons is because the BTFR looks like the Mb ~ Vf4 of MOND rather than the M200 ~ V2003 of LCDM. With dark matter, we can accommodate pretty much any power law, or none at all – a lot of scatter would be more natural. So why did it have to be MOND? Even in ignorance of MOND the data pose a fine-tuning problem for LCDM. But it isn’t just a fine-tuning problem; it is a fine-tuning that arises because of MOND. To be successful, a LCDM model must be tuned to look like MOND. If it doesn’t, it’s wrong. If it does, why should we prefer a fine-tuned model to the theory that predicted the correct behavior in the first place?

MOND is not perfect here: it suffers a missing baryon problem in rich clusters. Since Mb ~ Vf4, predicting only ~80% of the observed velocity translates to missing ~60% of the mass. That’s a lot! But it could be worse: if, like Zwicky, we had done this experiment before the advent of X-ray observatories, we would be unaware of the mass of gas in the ICM, and infer that MOND was missing practically all (~96%!) the mass. That would seem utterly ridiculous, and we would conclude that MOND is wrong when much of the problem would have been that we were missing an important reservoir of baryons. Perhaps we still are. I do not like this possibility – there is still a lot of ground to make up, and I am not aware of a satisfactory solution. I guess I’m just a skeptic that way.

If we think the residual mass discrepancy problem MOND suffers in rich clusters is serious and perhaps fatal, should we not also conclude the same from the local missing baryon problem in LCDM?

But the bullet cluster double-secret falsifies MOND!

Let’s examine that assertion in the context of what we learned above.

The Bullet Cluster, which is made up of two galaxy clusters that collided a few billion years ago. The pink is the ICM observed by the Chandra X-ray Observatory. JWST provides the image of the many galaxies and also provides the data to map the mass through gravitational lensing (blue). Note that most of the mass indicated by lensing is centered on the galaxies, not the ICM. Image: NASA, ESA, CSA, STScI, CXC; Science: James Jee (Yonsei University/UC Davis), Sangjun Cha (Yonsei University), Kyle Finner (IPAC at Caltech)

The bullet cluster is composed of two clusters that collided and passed through one another. The collision segregated the gas of the ICM (pink above) from the galaxies. This happens because gas is diffuse and collisional. The gas of the two clusters can’t help smacking into each other, slowing down and forming the shock front visible in the shape of the gas of the smaller cluster on the right. Galaxies, on the other hand, have lots of empty space between them. They are collisionless and pass right by each other. In doing so, they are slowed less than the gas, getting ahead of it, leading to the separation that we observe.

OK, cool. The argument one usually hears against MOND based on this is that the baryonic mass in gas outweighs that in galaxies, so the lensing signal should be centered on the gas: the blue should align with the pink, not with the galaxies. Instead, we see the opposite, so the mass has to be dark matter.

This would be a good argument if the gas were all of the baryonic mass. This is a common assumption that makes sense in LCDM, where the baryon fraction checks out, so most people seem to stop thinking at that point. But each theory needs to be considered in its own context, and it cannot be the case in pure# MOND that we see all the baryons## in the picture above. That’s what we learned above. It may be unsatisfactory, but we knew this already before the bullet cluster was discovered (e.g., Sanders & McGaugh 2002). So the only new thing we learn from this aspect of the bullet cluster is that if there is an additional reservoir of baryonic mass, it is collisionless. It didn’t collide like the gas, it passed through like the galaxies. There are lots of candidate baryonic objects that fit that requirement: brown dwarfs, neutron stars, black holes, very small rocks^. There is no requirement that the unseen mass be non-baryonic; we do not need the new physics of a new dark matter particle from beyond the Standard Model of particle physics on top of the new physics of MOND.

Now, as I think I’ve made clear, I am very uncomfortable with the apparent requirement that there is lots of undetected baryonic mass in clusters. If I were the MOND partisan that lots of people seem to assume I am, then I guess I’d portray this as a bold prediction. The dark baryons have to be there, and we should be turning all possible resources to detecting them, rather like we have for WIMPs. But I’m not that person. I am also not a person who sees this missing baryon problem for MOND as automatically worse than the missing baryon problem for LCDM. There is a much bigger deficit to be made up in LCDM, in many more systems### of very different types over a larger dynamic range in mass. The missing baryon problem in LCDM looks worse to me than that in MOND. Yet the community attitude seems to be largely unaware of it. Those who are seem mostly to presume that it’ll work out. Maybe, but this should not be accepted by assumption, it needs to be demonstrated. It has yet to be.

If you think the missing baryon problem in clusters is a terrible problem for MOND, then you should be similarly worried that LCDM evinces the same kind of problem – one that is objectively larger in amplitude. It seems that, having accepted that there is dark matter, people don’t much care what it is. I do. The dark matter paradigm has obliged us to abandon parsimony. Not only does LCDM need two novel substances, dark matter and dark energy, it requires two kinds of dark matter: baryonic dark matter and non-baryonic dark matter.

There is a communal failure of objectivity about this. The thought process is both transparent and simple: MOND doesn’t explain clusters; it requires dark matter. Therefore dark matter#### exists and it is silly to think about MOND. That would make sense if it weren’t a logical fallacy. Instead, it provides a permission structure to remain ignorant of what MOND gets right. I get that; there’s a lot to know. But I would also suggest that ignorance does not provide a strong basis for drawing scientific conclusions, especially for a subject so rife with confirmation bias and cognitive dissonance.


&The normalization is related to Newton’s constant and Milgrom’s constant through A = ζ/(a0G) where ζ is a factor of order unity that depends on the geometry of the system. It is one for spheres, and always approaches the limit ζ → 1 at sufficiently large radii, but observations are usually obtained at radii where the flattened geometry of disk galaxies is relevant, so in practice ζ ≈ 0.8. This can be derived from the geometry (all purely conventional; nothing to do with MOND) or one can obtain it empirically by comparing A = 50 M km-4 s4 from fitting the BTFR to data for galaxies with known a0; for a0 = 1.2 x 10-10 m s-2, (a0G)-1 = 63 M km-4 s4, so ζ =A(a0G) = 50/63 = 0.8.

*There remains room for improvement for poor clusters (here I call 1013 < Mb < 1014 M objects “poor clusters” because astronomical terminology can always be made worse). A particular issue is the quantity of intracluster gas, which dominates rich clusters (and is readily detected in X-rays), but seems to be absent in the smallest groups. There has to be a transition in between, but is it smooth so that all poor clusters have the same amount, or is there a huge variation in ICM mass among poor clusters? I have seen anecdotal indications that poor clusters that are detected in X-rays extend the trend of rich clusters while those that aren’t don’t, as if the residual mass discrepancy MOND evinces in clusters is somehow related to the presence of X-ray gas.

%There are lots of unsatisfactory explanations. Some sound more plausible than others, but all fail to engage with the underlying prompt: why do the data look like MOND if we live in a universe made of dark matter?

#It is possible that the problem MOND faces in clusters might not be one of missing mass, but rather it could be an indication of a deeper theory that is not exactly like pure MOND.

##If there is additional mass in clusters, it doesn’t necessarily have to be baryonic. It could, in part, be neutrinos or sterile neutrinos or other more exotic beasts of the unknown meagerie of our enormous universe. However, there is no requirement that the unseen mass be anything other than mundane, ordinary matter.

^Though an amusing thought, very small rocks do not make a viable candidate dark matter object any more than witches float because they weigh the same as a duck.

###I have heard otherwise brilliant scientists dismiss the successes of MOND as a fluke. MOND has made too many successful predictions for that to be a reasonable assertion; it is a good example of what Putnam meant by “no miracles.” Yet the same scientists will cite the consistency of the baryon fraction in clusters to the cosmic baryon fraction as something that cannot be a fluke, ergo LCDM must be right. So which fluke is worse? I do not have patience to list all of MOND’s successful predictions here, though there are many reviews that do so and there will be a long paper soon that does more. What I will note here, having just done the exercise, is that the cluster baryon fraction is more likely to be a fluke. In order to estimate a baryonic mass for each cluster, we extrapolate the so-called beta profile that describes the distribution of X-ray gas. That’s a reasonable thing to do, and when we do it, we get an answer that is satisfactory in LCDM. However, it is not a small extrapolation. We are inferring a lot of baryonic mass at large radii from the fit of the beta profile at smaller radii. That’s the obvious thing to do, and I think it is probably correct, but it is also something that could go badly wrong. We experimented with other plausible gas mass profiles, and the answer can vary a lot, often leading to considerably fewer baryons than the cosmic fraction. That would be bad for LCDM, and also make the problem MOND suffers (too few baryons) worse, so it doesn’t help anything. But if there is a fluke here, it is more likely to be the coincidence of the cluster baryon fraction with the cosmic baryon fraction than is the consistency of the observed BTFR with the prediction of MOND for most of the rest of the universe.

####This is where sloppy terminology leads to a logical fallacy: people equate “dark matter” with non-baryonic cold dark matter. The latter is a subset of the former; the unseen mass in MOND need not be the same as the non-baryonic stuff that we commonly assume the dark matter is.

The local missing baryon problem

The local missing baryon problem

Last time, we started talking about the data in the recent paper The Baryonic Mass-Halo Mass Relation of Extragalactic Systems. Here, we’ll put on our dark matter hat, and use the data to make an accounting of the mass – both the dark matter and the baryons in all their various forms. From this conventional perspective we will obtain a method for relating what we see to what we don’t. In the context of LCDM cosmology, this provides an alternative approach to abundance matching. It also provides a test: are the two consistent?

The conventional picture we have in mind is a baryonic galaxy residing in a dark matter halo bathed in a background of intergalactic matter.


Fig. 1 of McGaugh et al. (2026): Conceptual elements of a galaxy: the stars (yellow/blue) and atomic gas (green) of NGC 6946 (Spitzer 3.6µ and 21 cm data: F. Walter et al. 2008) are shown embedded in an extended dark matter halo (black). The dark matter density decreases continuously with radius so the halo has no hard edge, but for convenience we adopt the common convention that the radius r200 marks the boundary of the dark matter halo and the dividing line between the circumgalactic medium (CGM) and the intergalactic medium (IGM; orange). The stars and atomic gas illustrated here appear within r < 20 kpc while r200 ≈ 220 kpc (not shown to scale).

I’ve talked here about the stars and gas a lot because that’s what we see. These are the essential components that define a galaxy and comprise the mass that correlates with rotation velocity to make the baryonic Tully-Fisher relation (BTFR). I’ve talked a bit about the stuff between the galaxies, the intergalactic medium (IGM), but I don’t think I’ve previously had cause to talk much about the circumgalactic medium (CGM). As the name implies, this is gas in the vicinity of a galaxy, but not in the galaxy itself – at least not the part we can readily see. In the notional picture above, the distinction between the CGM and the IGM is the boundary of the dark matter halo that nominally demarcates gravitationally bound from unbound material.

Notional is doing a lot of work here. There’s a lot of gas in the IGM, and some of it is certainly in the vicinity of galaxies, so in that regard counts as circum-galactic. But there’s no hard and fast distinction between these components just as there’s no hard edge to a dark matter halo. Our brains don’t like that, so we impose notional boundaries and proceed as if these are meaningful.

Proceeding thus, we expect our dark matter halo* to contain its fair share of the cosmic baryon fraction, fb = Mb/M200 = 0.157 according to the Planck flavor of LCDM cosmology. We can test this by adding up all the baryons and comparing that to the total mass enclosed by r200. This is straightforward for the stars and gas we see, but not for the stuff we don’t see – both dark matter and the gas in the CGM.

There are some measurements of the CGM, but these tend to be statistical in nature (if we stack data for a bunch of galaxies, we sorta see something), not the precise, individual, galaxy-by-galaxy measurements that we have for the stars and atomic gas. The stars and atomic gas are the mass in the extended Tully-Fisher relations we discussed previously, and are the bulk of the normal material in the galaxies we see. The bulk of the CGM lies at much larger radii, beyond the stars and atomic gas, but within the notional edge of the dark matter halo, as depicted above. Since we don’t measure it directly in individual galaxies, we’re gonna leave the mass of the CGM as an open question rather than something to be included in the sum of known baryonic mass.

The situation is even murkier for the dark matter, which we don’t see at all, so we don’t have a good way to measure the “total” mass of dark matter halos. This isn’t even a well-defined quantity in principle since halos are not expected to have a hard edge. Conventionally, we adopt the mass within a radius that contains a density two hundred times the cosmic critical density, r200, as the notional edge. There are obscure historical reasons for this choice that I do not have the patience to describe. One could make other choices, arguably better choices, but r200 is the most common choice used in the literature so we’ll stick with it here. The halo mass is the mass enclosed by this radius, M200. If one goes through the math, it turns out that the circular speed of a test particle, V200, orbiting at r200 scales with the Hubble parameter [h = H0/(100 km/s/Mpc)] such that V200 = h r200 when V200 is in km/s and r200 is in kpc. The dynamical mass (rV2/G) can then be written

M200=(3.3×105Mkm3s3)V2003.M_{200} = (3.3 \times 10^5\;\mathrm{M}_{\odot}\,\mathrm{km}^{-3}\,\mathrm{s}^3)\,V_{200}^3.

That is a lot of huffing and puffing to get a way to relate the halo mass to something we can (kinda sorta) measure. The flat rotation velocity Vf has always been taken as the signature of the dark matter halo. One therefore expects V200 ~ Vf. Indeed, these quantities cannot differ by much if dark matter is what explains flat rotation curves. However, the notional radius of the dark matter halo where V200 occurs is much larger, by roughly an order of magnitude or more, than the radius where Vf is measured. So they need not be identical, depending on the halo model. So to relate what we measure to what we’d like to know we define a little ol’ fudge factor, fv, such that:

Vf=fvV200V_f = f_v V_{200}

If a rotation curve stays flat indefinitely (as our empirical experience suggests), fv = 1. If instead dark matter halos behave as they should in LCDM, then the rotation speed should gradually decline as we approach the halo’s edge so that fv > 1. How much greater?

One way to estimate the fudge factor fv is to fit dark matter halo models to data. This process does not directly measure V200, but it does provide an estimate of that quantity based on the data available a smaller radii. One can do this for as many halo models as one has the patience to consider. For example, here are the results for two common halo models, the traditional pseudo-isothermal halo first adopted to explain flat rotation curves and the CDM-expected NFW halo:

Figure 2 from McGaugh et al. (2026): The observed flat velocity Vf as it relates to the fitted V200 for pseudo-isothermal (left panel) and NFW (right panel) halos (Li et al. 2020). Filled points have formal uncertainties <20% in V200; open points are less accurate fits. The solid line shows Vf = V200. The gray line in the right panel shows Equation (2a) of Katz et al. (2019), which corresponds roughly to fv ≈ 1.4.

The result for pseudo-isothermal halos is consistent with fv = 1, as expected – this model was adopted to make flat rotation curves. There is nevertheless some scatter. This typically happens because the observed rotation is not observed to be flat over a large enough range of radii to enforce flatness further out (as often happens in dwarf galaxies) or because the stars account for so much of the mass over the observed range that the inferred dark matter component is still rising (as often happens in bright, high surface brightness galaxies). This sort of haziness is inevitable when one only measures the inner few percent of the notional virial radius.

The result for NFW halos is approximately fv = 1.4, albeit with a lot more scatter. This happens for the same reasons as above, with the additional problem that the dark matter profile in real galaxies rarely looks like NFW. Of all the many halo models considered by Li et al. (2020), NFW consistently performs the worst. One is forcing a fit of a function that would rather not. One signature of this misfit is the occurrence of very large V200 for dwarf galaxies with small Vf. Taken literally, this would mean that some of the smallest dwarf galaxies reside in dark matter halos that outweigh those of giants like the Milky Way. This seems absurd, and it is. For example, by this approach, the dwarf galaxy NGC 3109 residing just outside the Local Group outweighs the Local Group and both its giants, Andromeda and the Milky Way, put together. But it is pretty clear from the local velocity field that the entire Local Group is not orbiting this little dwarf.

The estimation of huge V200 for galaxies with small Vf happens because of the cusp-core problem. The density cusp predicted by NFW expects a curved shape for the inner rotation curve while the data show a more gradual, quasi-linear rise. Any decent fitting program will realize that it can make a curve look like a straight line if it stretches it out enough, so it does exactly this by making the halo very large. That sorta fits the data, but it makes no physical sense. Between this systematic effect and the large scatter induced by the other effects discussed above, one is better off inferring V200 from Vf with a fixed fudge factor. So we’ll do that, leaving the exact value of fv as an open question, but noting that for most objects it almost certainly resides in the narrow range

1fv1.4.1 \le f_v \le 1.4.

That’s a lot of words to say the observed flat rotation speed gives us our best kinematic estimator or the dark matter halo mass. In this context, bear in mind the small scatter in the extended Tully-Fisher relations. This contrasts with the large scatter seen in the fits above. This strongly implies that Vf is more closely tied to the underlying mass^ than are the model-specific halo fits to the entire rotation curve. That might seem counterintuitive given that Vf is only a portion of the rotation curve (albeit a well-defined portion). However, it makes more sense when one considers that rotation curve fits must consider the contribution of stars as well as dark matter. Since the stellar mass-to-light ratio is never perfectly known, there is a degeneracy between the two that contributes to the scatter seen above. That variation is not real, it’s just an artifact of the fitting procedure. But when we get to large radii, beyond the confounding effects of the stellar population, the signature of the dominant mass becomes apparent in the flat rotation speed.

We saw above that we expect the halo mass M200 to correlate with V200. We observe that baryonic mass Mb correlates with the flat rotation velocity Vf. The natural assumption is that the stuff we see is proportional to the total (mostly dark) mass while the observed flat velocity is a property of the halo. Hence Mb ~ M200 and Vf ~ V200. This simple argument has been the basis for many papers claiming to explain the Tully-Fisher relation over the course of many years. This would be entirely satisfactory if it weren’t so completely wrong.

Here we need to introduce another fudge factor, mb, that relates the mass we see to the halo that spawned each galaxy:

Mb=mbM200M_b = m_b\,M_{200}

The obvious assumption is that mb is a constant for all galaxies, in which case Tully-Fisher follows because Mb ~ M200 ~ V2003 and V200 ~ Vf. The wee problem is that this predicts a Tully-Fisher relation with slope 3: Mb ~ Vf3 when we observe one with slope 4: Mb ~ Vf4. In order to reconcile these two, our new fudge factor cannot be a constant. Worse, we need to fine tune it to transform the predicted power law into the observed one: mb ~ Vf. That… doesn’t make any sense.

We can refrain from thinking and plunge ahead to simply plot the baryon fraction. While we’re at it, let’s also plot the stellar mass fraction m* = M*/M200 because that is more commonly discussed in the literature. (Often stellar masses are available for galaxies without the corresponding gas mass measurements.) These fractions have to be increasing functions of circular velocity, or equivalently, mass (mb ~ Vf ~ Mb1/4):

Figure 4 from McGaugh et al. (2026): The stellar mass fraction as a function of stellar mass (top) and the baryonic mass fraction as a function of baryonic mass (bottom). Data and symbols as in Figure 3 with the additional distinction that large squares in the top panel represent the sum of the stellar mass of all galaxies in a group or cluster while small squares are the stellar mass of the brightest galaxy only. The horizontal line is the cosmic baryon fraction fb = 0.157 (Planck Collaboration et al. 2020). The colored lines in the top panels show the stellar mass–halo mass relations from abundance matching given by B. P. Moster et al. (2013; dashed–dotted green line), P. S. Behroozi et al. (2013; dashed–triple dotted pink line), and A. V. Kravtsov et al. (2018; red dashed line). The black line in the lower panel is mb = fb tanh(Mb/M0)1/4 where fb is the cosmic baryon fraction (0.157) and M0 = 5 x 1013 M.

To be specific, I’ve computed the halo mass assuming fv = 1. Different assumptions just slide the data up and down; the trend persists. This is discussed more in the paper if you’re interested in such details.

This gives a nifty way to relate what we can see to what we can’t. There’s a simple formula:

mb=fbtanh(MbM0)1/4m_b = f_b \tanh\left(\frac{M_b}{M_0}\right)^{1/4}

where fb = 0.157 is the cosmic baryon fraction and and M0 = 5 x 1013 M is the scale where the function bends, transitioning from the Mb ~ Vf4 of the BTFR that holds over most of the mass range to the mb = fb of rich galaxy clusters. The precise value of the turnover mass is not well constrained, as it happens in the one place that is not well sampled by the available data. Indeed, there is nothing special about the functional form; it is simply a choice that transitions nicely from one regime to the other. There’s no physics in it&. Still, this is a useful way to estimate the halo mass of pretty much any extragalactic object just by summing up its observed baryonic mass.

Indeed, this kinematic mass-matching relation is better than the widely used abundance matching relations in that it has less scatter. Abundance matching generally relies on stellar mass; that results in more scatter for the same reasons discussed for Tully-Fisher. This is particularly apparent at the low mass end of the top panel above, where galaxies of the same circular velocity (halo mass) have very different stellar masses. This goes away when baryonic mass is used instead.

There is reasonable agreement between abundance matching and kinematics at intermediate masses. The lines representing various abundance matching relations parallel the kinematic data. The offsets that are apparent can be cured by an appropriate choice of fv. Always a free parameter to the rescue there is.

At the high mass end, things go amiss again. Partly this is because abundance matching relations reference the stellar mass of the “central” galaxy. The picture is that each halo contains one central galaxy with many satellite galaxies in subhalos, so what matters is the stellar mass of the central. This is overly simplistic: galaxy clusters are messy, the brightest galaxy isn’t necessarily at the center, and most have substructure with multiple groups rather than a single hierarchy. Besides that, the stellar mass tells you little about the halo mass without further environmental context: a galaxy with M* ~ 4 x 1011 M could reside in halo masses spanning a couple of orders of magnitude.

Setting aside the issue of centrals, there is a serious tension for individual high mass galaxies. The stellar mass fraction suggested by kinematics keeps going up where that of abundance matching turns over. This is due to the linearity of the Tully-Fisher relation compared to the knee in the Schechter function shape of the stellar mass function. The two don’t match up, as discussed previously. This same tension has long been with us; in the ’90s we were concerned with the difference between “the luminosity function normalization” and “the Tully-Fisher normalization.” This tension never went away. Still, the tension between abundance matching and kinematics doesn’t seem tragic, and might be remedied with some appropriate finagling of both the baryon fraction and the velocity fudge factor.

But where are all the baryons? They’re all accounted for in clusters, which reach the cosmic baryon fraction. But in no other system is the checksum complete. There is a missing baryon problem locally in each and every dark matter halo below the cluster scale. To confound matters further, there is a fine-tuning problem: the amount of missing baryons scales precisely with the amount of observed baryons.

The logarithmic plot above may understate the magnitude of the problem. To clarify this, we can plot the ratio of missing-to-observed baryons on a linear scale, at least in part:

Figure 7 from McGaugh et al. (2026)The ratio of missing-to-observed baryonic mass as a function of baryonic mass. Data and symbols are the same as above. The ratio is linear in the bottom half of the diagram, then switches to logarithmic in the top half. Spiral galaxies are shown twice: once with fv = 1.0 (solid blue circles) and again with fv = 1.4 (small open circles). The Milky Way is the yellow point at the top of the gray band, which shows the range from zero CGM to that required to explain all of the locally missing baryons when fv = 1. Stars represent the CGM measurements of Milky Way–mass galaxies by Miller & Bregman (2015), Bregman et al. (2022), and Zhang et al. (2026) from bottom to top. These suffice to explain the missing baryons provided that fv ≈ 1.4. This explanation becomes progressively less plausible for lower mass galaxies.

The scatter blows up when we plot linear ratios; this is an artifact of error propagation. Nevertheless, it is helpful to see that the local missing baryon problem is not subtle. It is already a factor of ~2 for groups and ~3 for bright galaxies. It’s not as if we’ve misplaced a few percent of the baryons. Most of the baryons that should be associated with galaxy dark matter halos are not in evidence.

This problem has been known for a while, but doesn’t seem to be acknowledged to be a problem. Not all baryons need condense down into the central galaxy; some might be left behind, still mixed in with the dark matter halo. The widespread assumption seems to be that the missing baryons are probably in the CGM.

Accounting for the missing baryons with gas in the CGM almost works in bright galaxies like the Milky Way where we need “only” a factor of a few. Recent estimates suggest that the CGM is comparable in mass to the stars, or even somewhat more. These are very uncertain, as this mass is dispersed in diffuse gas over an enormous volume, and the total mass estimates often involve large extrapolations: the CGM is detected most readily nearby the central galaxy, but most of its implied mass is way far out near r200. Accepting these estimates at face value leads to the star symbols in the plot above. This makes the checksum complete provided the halo is not too massive, as happens if fv ≈ 1.4. This is what we expect for NFW halos, so it might work out if those were viable. However, there is a bigger issue.

The local missing baryon problem gets progressively worse for lower mass galaxies. For 1010 M galaxies – not all that much smaller than the Milky Way (Mb = 7 x 1010 M), the problem isn’t a factor of two or three: there are ~6 baryons missing for every one that is observed. For 109 M galaxies, the deficit is an order of magnitude. For even lower mass galaxies, the difference is so large we have to abandon the linear plot lest the interesting parts for bright galaxies get scrunched into invisibility. By the time we get to small dwarf galaxies of 106 M, the ratio of missing-to-observed baryons approaches 100:1. It is not plausible to imagine that the CGM of dwarf galaxies explains this deficit. (And yes, we’ve looked.)

A common explanation for this variation is that low mass dark matter halos have shallower potential wells, so have a harder time holding onto their baryons. Supernova can drive material out of galaxies; these go off with the same energy regardless of the galaxy they’re in so they may be more effective at blowing baryons out of lower mass systems. There is sufficient energy (IF properly% distributed) to completely unbind the baryons, so they might wind up in the IGM, defeating any hope of completing the checksum. This is the sort of argument that sounds clever but fails to address the real problem. The difficulty isn’t just ridding ourselves of these meddlesome baryons, it is getting rid of exactly the right amount each and every time.

As awkward as it is to realize that most of the baryons that should be in low mass halos are not in evidence, it is not difficult to imagine ways in which this might happen, like the aforementioned supernova-driven galactic winds. The more dire aspect of the problem is the fine-tuning. Galaxies of the same observed baryonic mass are always missing the same amount of baryons, whether that’s a factor of 2 or 10 or 100. If the visible parts of a dwarf galaxy are only 1% of the available baryons, you’d expect a lot of scatter. Sometimes a halo of that mass might have 2% or even 3% of its baryons condense to the parts we see. That would show up in the scatter in a way it does not: galaxies of the same circular velocity (halo mass) have the same baryonic mass every time. They don’t vary by factors of two (or more). So while we can build models that makes the baryon fraction just so, the fact that we can write a simple equation for it with practically zero scatter is profoundly uncomfortable.

An extra bit of weirdness is that in LCDM, galaxies are built hierarchically by merging small objects into large ones. This poses a teleological problem. Consider a small halo at high redshift. If it remains alone, then it it will contain a dwarf galaxy at low redshift that has a low baryon fraction. But if it mergers into a larger system, then by the current time that larger system has to have a larger baryon fraction. In effect, a low mass halo has to know where it will end up some billions of years in the future. Will it remain alone and unmerged? Better blow out all those baryons! Will it merge into a larger system? Better hang on to the right amount of baryons. Does that system merge into a still larger object? Hope it held onto even more baryons, in exactly the right amount at every step along dozens of mergers.

I can imagine all this happening in a stochastic fashion with the net result being that more massive systems wind up with a higher baryon fraction, at least on average. I cannot give credence to this process resulting in the small observed scatter. As people are always telling me, “galaxies are complicated.” Indeed, they should be – in LCDM. But in reality they’re not! They obey simple scaling laws, laws that do not follow naturally from LCDM.

The local missing baryon problem encapsulates one of the fine-tuning problems that has never been satisfactorily explained. This alone would be considered fatal for most theories. For LCDM, it is just another problem to be addressed through the eternal tweaking of models and simulations.


*Strictly speaking, M200 refers to all mass within r200, baryons as well as dark matter. I’m going to call it halo mass anyway, because that’s what we mean, the baryons are a small fraction of the total, and because that’s what everybody does in the literature. If we make some other choice for the definition of the mass of the halo, MΔ, then the inferred baryon fraction of an objects scales by M200/MΔ. The cosmic baryon fraction does not care what choice we make, so the implicit assumption is that one asymptotes to the cosmic fraction if one gets far enough out, irrespective of what rΔ we adopt. While this is a sensible assumption – individual objects must merge into the larger cosmos at some point – there is no guarantee that the universe cooperates. For example, the baryon fraction in galaxies declines with increasing radius, but that in galaxy clusters increases with radius. I’ve seen hints that it doesn’t really settle down to the cosmic (or any particular) value. These are only hints – considerable extrapolation is involved – so we’ll ignore this inconvenience and assume that the baryon fractions of individual objects do in fact converge to the cosmic value far enough out.

^It makes the most sense if the underlying total mass is the observed baryonic mass.

&I made a very similar fit in McGaugh et al. (2010) but didn’t publish it because there was no physics in it. Since then the field has been awash in abundance matching relations that were similarly fit sans physics. There has been much ink spilled justifying it post-facto with feedback, but I have refrained from this exercise in intellectual onanism.

%It is common to assume in simulations that a large fraction (50 – 100%) of the energy from supernovae is returned to the surrounding gas. This process is not resolved in cosmological simulations, all the energy return happens as part of the “subgrid” physics, so the feedback efficiency is set, in practice, to make things work out as well as possible.

Observationally, most of the SN energy finds its way out along the path of least resistance where the density of the surrounding gas is smallest (“chimneys”). This process couples to the surrounding gas with only a few percent efficiency.

Yep, it’s a religion

Yep, it’s a religion

I have been concerned for years that dark matter was morphing from legitimate science into a cold, dark religion. I have been reluctant to put it that way, because there are lots of scientists who work on dark matter that have not fallen entirely down that rabbit hole and who continue to make valuable contributions working in that context. But a recent experience reminded me that my concerns were not misplaced, and there are plenty of scientists who have fallen irredeemably down this rabbit hole. No matter what answer the future holds to be correct, many current scientists will have gone to their graves in denial of it.

Where is the boundary between science and religion? It is hard to assess where the borderline is. But it is easy to see when people are far over the line – so far over that it doesn’t really matter where exactly the line is. One can attend any conference on the subject to find people who unabashedly assert that dark matter exists without question. Not just that acceleration discrepancies have been amply demonstrated empirically, but that the only possible interpretation is dark matter. If asked whether this invisible mass is in the room with us now, they will enthusiastically# answer yes! Since dark matter has not been detected in the laboratory, this assertion is an expression of faith – the hallmark of religion – not of an established scientific fact. What we have established is that there are discrepancies between what we see and what we get when we assume Newtonian gravity (or GR, if needed). What we don’t know is whether the cause of these discrepancies is some form of invisible mass (dark matter) or if the equations we employ are inadequate (modified gravity [or more generally, dynamics]).

Indeed, these days many people will assert that dark matter has already been detected, usually citing astronomical evidence that used to be considered too feeble to merit a Nobel prize. Funny how repeating a mantra long enough morphs an aspiration into accepted reality. Modern physics is not providing a strong falsification of the supposition that science is a social construct.

A prominent example of an observation of the sky that is frequently cited as absolutely requiring cold dark matter is the acoustic power spectrum of the cosmic microwave background. Quoting clayton from a few years ago:

the primary reason to believe in the phenomenon of cold dark matter is the very high precision with which we measure the CMB power spectrum, especially modes beyond the second acoustic peak. There is a stone-cold, qualitative, crystal clear prediction of CDM about the relative sizes of the second and third peaks that modified gravity profoundly and irredeemably gets wrong: it thinks the third peak should be relatively larger* than the second… whereas CDM thinks they should be about the same

I would accept that this were conclusive proof of dark matter if this were the unique prediction of dark matter: that there was no other way to do it, so all other approaches were indeed irredeemable. (Quite the strong language, eh?) The problem is that CDM is not the one unique was to fit these data. Skordis & Zlosnik showed that it is possible to write a modified gravity theory that also fits the CMB data:

CMB power spectrum observed by Planck fit by AeST (Skordis & Zlosnik 2021).

This does not prove the AeST theory of Skordis & Zlosnik is correct, but it does demonstrate that it is possible to write a modified gravity theory that does indeed do what it is frequently asserted to be impossible for a modified gravity theory to do. I’ve heard of a couple of other theories that can also do this (the relativistic Khronon theory of Blanchet and nonlocal MOND as discussed by Deffayet & Woodard), so clearly this success is not uniquely limited to cold dark matter, or even a particular modified gravity theory. The work of Skordis & Zlosnik (2021) was known and in the literature before clayton made the assertion above in late 2022, so either he wasn’t paying attention (likely) or is convinced that it is impossible so doesn’t even consider the possibility (also likely). The former just says we’re all too busy, but the latter is a mark of religious thinking: my god is the only god, thou shalt have no other hypotheses before& me.

Many people are very impressed with the quality of the LCDM fit to the CMB. That is indeed very good, but there are enough free parameters that we were going to get a fit to any physically plausible power spectrum. If not, we’ve never been shy about making up new parameters. (Evolving dark energy, anyone? How about a running power spectrum? There’s a whole bag of possibilities!) What I’ve been more impressed with is the consistency of the fit to the CMB data with the many independent constraints on conventional cosmology. Or at least it was, until it wasn’t.

The Hubble tension has gotten steadily worse (in terms of statistical significance), and it really does not look like local measurements are to blame, nor is it the only tension. People seem to miss that it is the CMB-fitted value of the Hubble constant that has evolved over time to spoil the concordance that got us to believe in LCDM in the first place. But if the CMB is the cornerstone of your religion, all other data must inevitably be at fault and can be ignored: there is an entire community of cosmologists who choose to believe the best-fit Planck cosmology to the exclusion of all other data. It’s like the bad old days of the Hubble tension all over again, with the physics community choosing to believe the lower value of H0 because it makes more sense for the aspects of cosmology that they care about while those in the astronomical community who actually measure H0 find a persistently higher value.

A real tension in LCDM implies the need for new physics of the unknown variety. One doesn’t want to go there if it can be helped. I didn’t consider MOND until I was already concerned for the viability of dark matter. There are real problems for the paradigm that its more intense advocates simply deny, brush aside without real thought, or choose to remain ignorant of. When they are confronted with a problem, they are pretty creative about making stuff up on the spot. Anything to avoid having to confront the unspeakable – another hallmark of religion.

For example, cold dark matter is scale free. That’s foundational to the hypothesis. So the existence of an acceleration scale in the kinematic data is anathema to CDM. When I first pointed this contradiction out, there were a variety of assertions to the effect of “does too!” One example is provided by Kaplinghat & Turner, who claim to show “how Milgrom’s law comes about in the cold dark matter theory of structure formation.” That would, indeed, be ideal, and is a requirement for any theory to be successful.

Wee problem: they demonstrat no such thing. CDM is scale free, yet K&T claim that it explains Milgrom’s Law, which is predicated on the existence of an acceleration scale. Well, which is it? Is CDM scale free? Or does it explains the acceleration scale? We can’t have it both ways: their very premise is self-contradictory. It is absurd on its face.

The acceleration scale is defined by baryons, for which K&T have no model. To connect baryons with dark matter, they make a hand-waving argument about galaxies reaching a0 at the edge of their disks. This is not even a concept of a model and does not begin to suffice as an explanation for many reasons, a prominent one being that low surface brightness galaxies have accelerations less than a0 everywhere:

Centripetal acceleration curves color coded by galaxy surface brightness. Low surface brightness galaxies (blue colors) have low (sub-a0) accelerations everywhere: there is no edge at which they reach a0. (Adapted from McGaugh 2020.)

Milgrom pointed out this and many other shortcomings of their scenario, so I feel no need to elaborate further. Milgrom eviscerated their paper so thoroughly that the proper course of action would have been to retract it. Instead, they simply never acknowledge the criticism, and persist to this day in pushing it as some sort of valid scientific explanation. It is not; it does not withstand even mild critical scrutiny. But it doesn’t need to: it reassures the faithful that all is well. They hear what they want to hear without questioning its veracity. That’s another hallmark of religion.

I have refrained from saying these things in the past because I’m too nice. For example, a few years ago I started then abandoned the draft text below, which I simply cut & paste:


One of the things that attracted me to a career in science is the notion of objectivity. I grew up for a time in the bible belt, where people earnestly believed things that were obviously untrue, even to the eyes of a small child. On the occasions that I had the temerity to point out the obvious, the contradictions posed by facts never had an impact on their belief system. Rather, it inevitably earned me a warning that I was going to hell. No few of these people seemed to think it was their religious duty to send me there prematurely, or at least to make life on Earth a living hell.

Scientists eschew such behavior, but are also human, so often engage in it anyway. I’ve encountered it a lot. I get it; I went through the same denial, grief, and anger over the prospect of losing my good friend cold dark matter. The stages of grief never brought something back from the dead, but it has engendered a lot of blame-the-messenger.

Here’s an example, from a review by Mike Turner:

Excerpted from Turner (2021).

There is a lot of misinformation packed into this short paragraph.

The first clue is right there at the beginning, in red: the heading “False starts.” This is false framing, a classic tool of propagandists. It starts from the outset by asserting that the topic to be discussed is wrong at a level of knowledge so common it requires no justification. This is not the way one starts an objective discussion, much less a scientific one.

Turner then misconstrues what Milgrom did. He didn’t notice the scale a0 in the data, for which there was scant evidence at the time. Rather, Milgrom made the obvious statement that the inference of dark matter relied on the assumption that dynamics, as encapsulated by the laws of inertia and gravity, is the same on the very different scales of galaxies as in the solar system where they were established, so we ought to consider if dynamics might change in some way. He quickly excluded a size dependence as a possibility. How he settled on acceleration is beyond the scope of this post, and not for me to say. Neither is it for Turner to say.

After a brief and incomplete description of what MOND is, Turner allows that “this one-parameter model fits all the rotation-curve data”. Even in making this admission, he chooses to call it a model rather than a theory. A model is something specific you build in the context of a theory, like a halo model in CDM. MOND is more than that.

Turner quickly moves on without contemplating any meaning that rotation curves might hold. Let’s pause to consider that.

First, I would not say that MOND fits all the rotation curve data. It fits most galaxies, but there are a minority of weird cases that are not well fit. The weird cases inevitably don’t make sense in terms of dark matter either, so on the whole I interpret this to be the usual price of dealing with astronomical data – some of it is just goofy. Setting such cases aside, I can and have fit the same data with all sorts of dark matter halo models. MOND requires fewer parameters, which is important, but the difference isn’t in the fitting. The difference is in predictive ability. I can use MOND to predict the dynamics of galaxies a priori, and have done so many times. I cannot use any flavor of dark matter theory to do the same, and it’s not for lack of trying.

The predictive power of MOND must be telling us something, even if it is something about the nature of dark matter or the process of galaxy formation. There are many papers written on this, some deep and profound, others absurd and banal. Turner cites none of them, nor displays any awareness that such work exists. I would venture to guess that is because acknowledging such work would imply that there is something to debate here, something he would apparently rather not admit.


That’s where I left off. It’s exhausting deciphering other people’s false assertions. Moreover, I just don’t like criticizing other people, no matter how richly they deserve it. (Turner has never refrained from criticizing me in ad hominem terms: on one occasion$ he showed my picture to an audience and called me “the enemy.”) A large segment of the particle physics and cosmology community appears to think this way, and has succumbed to a scientific version of bible thumping in which you can assert any absurd thing so long as it falls within the framework of the holy LCDM. They really need to find something better to do.

I had hoped we were past this, but I heard a talk last week that was exactly in this mode. To paraphrase, the talk went

We’re sure dark matter exists. We have been sure about it for decades. In that time, we have been repeatedly proven wrong about what it is. Rather than re-think our paradigm in the face of these repeated failures, we double down yet again on the existence of this invisible, undetected mass, asserting aggressively% that it must be true while eliding or misrepresenting the evidence that it is not. This enables us to make up a whole lot of exciting new possibilities for what the dark matter might be and conceive of ever more grandiose experiments to continue not to detect it. You must believe in dark matter!

This was not a science talk so much as an indoctrination session. It was as if I had stumbled into a revivalist tent where some hothead was preaching to the choir. This is the kind of talk that misled an entire generation into wasting their careers at the bottom of a mine shaft searching for WIMPs. At least WIMPs were a well-motivated hypothesis; this kind of talk could lead a new generation down an even greater variety of garden paths.

I am well aware that I might fall prey to this attitude myself. That’s why I set criteria by which I would change my mind: detect dark matter already, or at least provide a satisfactory explanation as to how MOND comes about. Neither of those criteria have been met. There are claims to do the latter, but so far these are just variations on models I tried and found to fail long ago. If I thought these could work, I would have said so. At the same time, I don’t see any dark matter advocates taking up the challenge to specify what would change their minds. When I ask them what could falsify dark matter, I get dumbfounded looks – the deer-in-the-headlight face one gets when the immediate response why would you even ask that? is checked by a distant memory that scientific theories are supposed to be falsifiable.

Personally, I found it humbling to encounter MOND in my own data. I too thought we understood the universe with dark matter. But who ordered this? Certainly not me: my own conventional, dark-matter based predictions were falsified. No one else working in the context of dark matter had got it right at the time either. Only Milgrom ordered this.

And what is this? There is a direct connection between what we see and what we get. Even in ignorance of MOND, the radial acceleration relation encodes a one-to-one relation between the distribution of baryons and the effective force. This is so direct that one can right down a single equation connecting the two:

gobs=F(gN/a0)gN.g_{obs} = F(g_N/a_0)\,g_N.

The observed acceleration is a simple function of that predicted by Newton for the stars and gas that we see. There is no mention of unseen mass; everything is specified by what we can see is there.

I’ve sometimes heard astronomers complain about the reductionist ethos of physics, trying to cram all the complexity of the entire universe into a theory of everything. But here it is appropriate: there is a single, apparently universal force-law at work in galaxies. That’s telling us something profound. And yet if questioned about this, the physicists are the ones who will complain that galaxies are complicated, so they should be exempted from having to explain them. Galaxies should be complicated – in LCDM. But they’re observed not to be, in the sense that a single equation suffices to describe their kinematics. The problem isn’t that galaxies are inexplicably complicated, it’s that they should be but aren’t.

I am deeply disappointed that many scientists apparently lack the physical intuition to immediately recognize the import of the simple relation between what we see and what we get. It is the same sort of thing Newton noticed in the solar system: everything happens as if the gravitational force is proportional to the product of the masses and the inverse square of their separation. He didn’t understand why at the time, and was criticized for indulging in magical thinking: how can there be action at a distance? But that’s what the data were saying, and the same applies now. We might not yet understand the why, but that the data look as if MOND is what’s happening in this universe.


#The framing has morphed over the years. A recent advent is that some people have started proactively asserting that invisible mass is in the room with us now in order to avoid having to answer it as a question that makes them sound like loonies.

*He means the third peak should be smaller than the second, not larger, if by “it” he means modified gravity with the baryon density expected from big bang nucleosynthesis, which was the hypothesis that correctly predicted the first-to-second peak ratio but does indeed get the second-to-third peak ratio wrong. Funny how the CMB community was able to completely ignore the successful prediction for several years, but were then suddenly all over the latter failure. The third peak falsifies the ansatz on which that particular prediction was built, not the entire concept of modified gravity. This would be like asserting that all possible forms of dark matter are excluded because we haven’t yet detected WIMPs. It is a classic failure of objectivity, which is another hallmark of faith-based argumentation: we know His name is [insert favorite deity], not [insert any other deity].

&Or after me. Dark matter was my first hypothesis, and I’m here to tell you that True Believers do not suffer second hypotheses or those who stray from the fold. I guess that’s why so many scientists who are MOND-curious keep it on the down low. Wise, perhaps (that’s why tenure needs to be a thing), but hardly the ideal of the open and free exchange of scientific ideas.

$I wasn’t there, but one audience member (not someone I knew) thought it was so over the top that he told me about it, sharing a link with a video. (I did not retain that link, and doubt the hosting conference website is still active.)

%Argument weak here. RAISE VOICE!

The well of inspiration

The well of inspiration

AP News recently ran a story on the interface of religious inspiration and the search for dark matter. I thought about commenting on it, then thought better of it, but now here I am again. I know some but not all of the people who are quoted, all good scientists. The article ends with a nice quote from Jennifer Wiseman:

“Studying the deep universe may make us feel insignificant,” Wiseman said. “But it also gives us a sense of unity that we’re all on the same planet. … The hope is we get a sense of joy, humility and love from these contemplations.”

I’ve known Dr. Wiseman since undergrad days, before either of us were Ph.Ds. We don’t agree on much that is specific to religion, but somehow manage to be friends anyway, and I completely agree with the sentiments she expresses here.

There is a tendency to portray religion as being at odds with science, and that can certainly happen. But they share more in common than this trope implies. Both arise from a deep wellspring in the human spirit, the same wellspring: the desire to know. How does the universe work? How did it come to be so? Why? and on and on. Where they differ is in approach: when encountering an unknown, especially things that are fundamentally unknowable (e.g., does God exist?), religion asserts an answer and asks that we accept it on faith. Faith is anathema to the scientific process; one is, in principle, to search for the truth through observation and experimentation.

Religion and science come into conflict when scientific knowledge encroaches on religion’s known unknowns. There was a time when the structure of the heavens must have seemed unknowable; the realm of the sacred, safe from access by mundane human knowledge. So it is easy to see how Galileo came into conflict with the Church despite being a faithful member:

Church: Scripture teaches us that the Earth is the unmoving center of mortal corruption; the heavens above us are perfect and unchanging.

Galileo: I can see spots on the sun and mountains on the moon.

Church: We are the center about which the heavens rotate every day, clearly the center of all rotation.

Galileo: I see satellites circling Jupiter.

Church: God is great; He could do one thing one day and an entirely different thing another day*.

Galileo: It appears that the Earth revolves around the Sun, never the other way around.

Church: Do you see these instruments of torture?

This is a tongue-in-cheek portrayal of a serious historical incident, but my point is that conflict arises when scientific knowledge encroaches onto turf that had formerly been the exclusive province of religion. Addressing our spiritual need to know can be inspirational, but it can also be a let down. Giving specific answers to formerly unknowable questions can be a bit of a mood killer. Is that all there is?

Another downside of scientific knowledge is that it can never be complete. There is always something left unknown, and we are not supposed to fill that void with something we take on faith. Yet we are profoundly uncomfortable with not knowing. A little extrapolation at and beyond the fringes of current knowledge is natural, and can sometimes provides a useful way forward in driving new discoveries. But it can get carried away, e.g., string theory. So extraordinary caution is also warranted: the temptation to fill in the blank is where empiricism transitions to theology.

Understanding that we know a lot – as we do at this juncture in the history of science – and yet still don’t know something important, like what most of the universe is made of, is hard to accept. It becomes even harder when admitting something we thought we understood is wrong, or is less complete than we thought. We know most of the mass in the universe is made of non-baryonic cold dark matter&. Don’t we?

The dynamical evidence for acceleration discrepancies is abundant. That these must be caused by non-baryonic cold dark matter is less clear. This is where cosmology intrudes. Cosmology has always been the nexus where science and religion meet, with philosophical imperatives often obscuring essential observational facts. Before Kepler, orbits had to be circular. After Inflation, the density parameter had to be one. (We meant Ωm = 1, not the modern weak-sauce version Ωm + ΩΛ = 1.) The mass density is larger than the baryon density from BBNm > Ωb), so there has to be non-baryonic dark matter. That such a substance is also required to fit the acoustic power spectrum of the CMB amplifies our faith in the existence of such stuff.

We think we’ve solved cosmology, and that solution requires non-baryonic cold dark matter. So to admit that maybe we were wrong about dark matter just because it persistently remains undetected and provides unsatisfactory explanations of many astronomical observations and is consistently outperformed in predictive capacity by an alternative theory is to admit that we don’t understand as much about the universe as we thought. That’s really, really hard. Our unwillingness to admit that maybe we have been wrong about such an important issue is where human nature kicks in to blur the line between scientific knowledge and religious faith. It’s much easier to ignore those nagging doubts% and have faith that we were right all along.

Cosmology works so well that dark matter has to exist. I’ve head this sentiment expressed over and over by many different scientists, yet this is an assertion of faith. A more conservative statement is that cosmology as we currently conceive it works if, and only if, an appropriate form of non-baryonic dark matter exists with the required cosmic density. If not, then we need a new model – something that presumably^ stems from a more general underlying theory.

Since both religion and science arise from the same desire to know, it is easy to have faith that we know more than we actually do.

This is the sticking point we’ve hit in the dark matter debate. We’ve been calling the acceleration discrepancy the dark matter problem for so long that this linguistic mistake has morphed into an absolute certainty that invisible mass exists. It has become a matter of faith.


*Of course an omnipotent deity could do that. Apparently He chooses not to do so, sparking the schism between theists, for whom God actively intervenes in miraculous ways in real world affairs, and deists, who view God more as the great watchmaker, setting creation in motion but not interfering with its operation. Deism is conducive to science, as the search to identify the rules by which the universe works is to gain insight – however remote – into the mind of God. I suspect this attitude informed Einstein’s complaint against quantum mechanics: “God does not play dice with the universe.”

&I certainly thought so before I didn’t.

%The social pressure to conform to the preferred cosmology is enormous. I know I’m only making it worse for myself. But an explanation that omits MOND is a lie of omission. Here science and religion certainly overlap in the sentiments expressed by the seventeenth century cleric Paul Gerhardt:

“When a man lies, he murders some part of the world.”

Assertions that somehow “feedback” explains MOND are simply a numerical form of magical thinking; an excuse to not have to explain the inexplicable. By eliding MOND, we murder a part of the world.

^Here I am extrapolating at the fringes of knowledge.

Paradigm Shifts in Modern Astrophysics

Paradigm Shifts in Modern Astrophysics

I see that I’ve been posting once a month so far in 2026. I’ve lots to say but no time to say it. Some of it good, some of it bad, maybe sometime I’ll get around to it. No guarantees. On the good side, I’ve been working on a big project or two; may have something to say about those soon. I’ve also been meaning to write about the Planet 9 anomaly for months stretching into years now. Fascinating stuff related to MOND but not something I’ve worked on myself. On the bad side, I’ve been obliged to waste yet more time on my university administration’s insistence on merging our department into physics based on a snap decision made by a disinterested leader who employed all the forethought typically reserved for bombing a random country in the Middle East.

So I have had no time for novel posts lately, and today is no different. However, I thought readers of this blog would appreciate the post Paradigm Shifts in Modern Astrophysics: Applying Thomas Kuhn’s The Structure of Scientific Revolutions to Dark Matter at Heritage Diner that was pointed out to me by Moti Milgrom. Since I wouldn’t have seen it had he not mentioned it, perhaps that’s the case for you as well. I’m not gonna re-post it verbatim – you can read it there yourself – but I am going to offer a running commentary with a few observations, both personal and historical. So bring it up in a separate browser window and let’s read along…

This post riffs off of Kuhn’s The Structure of Scientific Revolutions as it pertains to dark matter and MOND. If you’re not familiar with it, Kuhn’s work on the philosophy of science is foundational to the way in which a lot of physical scientists approach their field (whether they realize it or not). Philosophers of science have done a lot more since then, but I’m not going to attempt to go there. I will look back to Popper* to note that I’ve heard Kuhn depicted as being some sort of antithesis to Popper. I don’t see it that way. To be pithy, Popper tells us how science should be done while Kuhn tells us how it is done. Who could have imagined that a human endeavor would be messy in practice and not always live up to its ideal?

I’m not sure how to do this; I guess I’ll excerpt relevant quotes and riff off those. The basic thesis is that dark matter is on the brink of a Kuhnian paradigm shift.

We are living through exactly that moment in modern astrophysics.

I certainly hope so! This moment in the history of science is taking a long damn time. A century ago, we went from “classical physics explains everything” to “quantum mechanics, WTF?’ in the space of about a decade. I’ve been working on matters related to MOND for over thirty years now, dark matter longer than that, and of course Milgrom started more than a decade before I did.

The essay discusses the “cartography of collapse,” which includes crisis and revolution:

The third stage is crisis — triggered when anomalies accumulate beyond the paradigm’s absorptive capacity. And the fourth is revolution, in which a new framework displaces the old not through incremental persuasion but through a gestalt shift, what Kuhn famously described as seeing the same duck-rabbit drawing and suddenly recognizing a rabbit where you had always seen a duck.

This resonated with me because I had exactly this experience. I started my career as much a believer in dark matter as anyone. I was barely aware that MOND existed (this seems to remain a common condition). But it reared its ugly head in my own data for low surface brightness galaxies. Try as I might – and I tried mighty hard, for a long time – I could not reconcile how the shapes of rotation curves depended on surface brightness as they should according to Newton while simultaneously lying exactly on the Tully-Fisher relation without any hint of dependence on surface brightness+. I could explain one or another, but not both simultaneously – at least, not without engaging in some form of tautology that made it so. I came up with a lot of those, and that has been a full-time occupation for many theorists ever since.

For me, this gradually became a genuine crisis. I pounded my head against the wall for months. Then, as I was wrestling with this problem, I happened to attend a talk by Milgrom. I almost didn’t go. I remember thinking “modified gravity? Who wants to hear about that?” But I did, and in a few short lines on the board, Milgrom derived from MOND exactly the result I found so confusing in terms of dark matter. This chance meeting in Middle Earth (Cambridge, UK) changed how I saw the universe. The change wasn’t immediate – it had to ferment a while – but ultimately I found myself asking myself over and over how this stupid theory could have its predictions come true when there was so much evidence for dark matter. Finally I realized that the evidence for dark matter assumes that gravity is normal; really it was just evidence of a discrepancy, and it could be that the assumption was at fault. That realization was sudden: where I’d always seen a duck, suddenly I could also see a rabbit.

Most scientists have not had this experience. What constitutes a crisis serious enough to contemplate a paradigm change is a highly personal matter of judgement. It happened in my data, so I took it seriously, but others didn’t care. So I made predictions for their data. Some of those came true, but they rejected the evidence of their own data. It just could not be so! At what point does a mere problem amount to a true anomaly?

Part of the sociological issue is that the dark matter paradigm has been in a constant state of crisis since its inception. The reasons vary over time. Sometimes valid solutions have been found to the crisis du jour, other times we’ve chosen to just live with it. It is much easier to live with a bad solution than to rethink one’s entire world view.

The problem with being in a constant state of crisis makes is that it seems like nothing can ever be a genuine crisis. Every foundational change is just another new normal. We complain, say it can’t be so, argue, offer bad ideas, reject them, get used to them, then eventually accept that one of them maybe isn’t so bad, so that must be what is going on. After a few years It is Known and people convince themselves that we expected just that all along.

It takes a lot of evidentiary weight for a paradigm to change, and it takes a lot of time for that to accumulate. But, as Kuhn recognized, mere facts are not enough. Humans and their attitudes matter. As Feyerabend noted,

The normal component [i.e. the accepted paradigm and its adherents] is large and well entrenched. Hence, a change of the normal component is very noticeable. So is the resistance of the normal component to change. This resistance becomes especially strong and noticeable in periods where a change seems to be imminent.

P. Feyerabend in Criticism and Growth of Knowledge

The post correctly points out that dark matter itself was an anomaly going back to Zwicky in 1933. This is often depicted^ as the first detection of dark matter, but it was also noted by Oort in 1932. Zwicky was aware of Oort’s work and cited him, but they’re very different results. Oort was worried about a factor of ~2 discrepancy in stellar dynamics in our local chunk of the Milky Way; Zwicky discovered a discrepancy of a factor of ~1000 in the Coma cluster of galaxies. These both imply the need for unseen mass, but the results are not at all the same. In retrospect, Oort’s discrepancy is a subtle detection of a flat rotation curve while Zwicky’s discrepancy was (at least) two distinct discrepancies: what we now consider the usual cosmic dark matter, but also missing baryons: most of the normal matter in clusters is in the hot, diffuse intracluster medium, not in the stars in the galaxies that Zwicky could see and account for. The modern discrepancy is only a factor of ~6, which is rather less than 1,000. (The distance scale also played a role in exaggerating Zwicky’s result.)

This all seemed crazy in the 1930s, even in the immediate aftermath of the quantum revolution. Consequently, Zwicky’s work was mostly ignored$. The subject of dark matter didn’t really take off until the 1970s. Considerable credit goes (rightly) to Vera Rubin, though many others made essential contributions – just on the subject of rotation curves, Albert Bosma, Mort Roberts, and Seth Shostak all made important contributions, the relative importance of which depends on who you ask.

An important aspect of scientific revolutions is persistence. Vera was persistent. She was fond of relating the story of showing her first (1970) flat rotation curve of Andromeda to Alan Sandage, only to have him dismiss it as “the effect of looking at a bright galaxy.” What the heck did that mean? Nothing, of course – it is the sort of stupid thing that smart people say when confronted with the inconceivable. So Vera persisted, and by the end of the decade had shown that flat rotation curves were the rule, not some strange exception. They became accepted as a de facto law of nature, and the dark matter interpretation was solidly in place by 1980.

The scientific community absorbed this anomaly not by questioning Newtonian gravity or Einstein’s general relativity, but by proposing an invisible scaffolding — a halo of non-luminous, non-interacting matter surrounding every galaxy. Dark matter became not a crisis but a patch.

Indeed, this seemed the most appropriate (scientifically conservative) course of action at the time, as summarized in this exchange (also from the early 1980s):

To emphasize the essence of what is said here:

Tohline: I might be so bold as to suggest that the validity of Newton’s law should now be seriously questioned.

Rubin: The point you raise is worth keeping in mind although I believe most of us would rather alter Newtonian gravitational theory only as a last resort.

This was a very reasonable attitude, at the time. But I’ve heard the phrase “only as a last resort” many times now over the course of many years from many different scientists. At what point have we reached the last resort? In the case of dark matter, once we’ve convinced ourselves that invisible mass has to exist, how can we possibly disabuse ourselves of that notion, should it happen to be wrong?

In Kuhnian terms the last resort is reached when the weight of anomalies in the standard paradigm become too great to sustain. But that point is never reached for many die-hard adherents. Whatever the right answer about dark matter turns out to be, I’m sure many brilliant people will go to their graves in denial. Hence the more cynical phrase

Science progresses one funeral at a time.%

But does it? What if the adherents of an ingrained but incorrect paradigm breed faster than they go away? I’ve seen True Believers train graduate students who’ve gone on to train students of their own. Each generation seems to accept without serious examination the inadequate explanations for the anomalies made by their antecedents, so the weight of the anomalies doesn’t accumulate; instead, each one gets swept separately under the proverbial rug and forgotten. Forgetting is important: when new anomalies come to light, hands are waved and new explanations are promulgated; no one chekcs if the new explanations contradict the previous generation of explanations. What passed before is a solved problem, and we need never speak of it again.

This is not a recipe for a scientific revolution, but for a thousand years of dark epicycles.

Returning to the post,

By the late 1980s and early 1990s, dark matter had been formally incorporated into the reigning cosmological framework. Lambda-CDM — where Lambda refers to the cosmological constant (a proxy for dark energy) and CDM stands for Cold Dark Matter — became the standard model of cosmology.

The essence of this statement is correct but some of the details are not. Dark matter was widely accepted by 1980. That’s still a little before my time, but my impression is that the magnitude of the discrepancy was at first a factor of two, so it could simply have been normal baryons that were hard to see. However, the discrepancy rapidly snowballed to an order of magnitude, so we needed something non-baryonic. This was happening simultaneously with talk of supersymmetry and grand unified theories in particle physics that could readily provide new particles to be candidates for the dark matter, leading to the shotgun marriage of particle physics and cosmology, two communities that had had little to do with each other before then, and which still make an odd couple. Cosmology as traditionally practiced by astronomers needed dark matter but didn’t much care what it was; particle physics was all about the possibility of new particles but didn’t care about the details of the astronomical evidence.

To rephrase the above quote, I think it is fair to say that “by the late 1980s and early 1990s, cold dark matter had been formally incorporated into the reigning cosmological framework.” But that framework was not yet LCDM, it was Ωm = 1 SCDM. The Lambda only came to prominence by the end of the 1990s, as I’ve related elsewhere. This process is depicted by many scientists as a revolution in itself, and in many regards it was. The cosmological constant had been very far out of favor; rehabilitating it was a grueling experience and no trivial matter. But it wasn’t really a scientific revolution in the sense that Kuhn meant: our picture didn’t fundamentally change, we just learned to accept a parameter& that was already there but that we didn’t like.

The post goes on to note the absence of dark matter detections:

This silence is itself an anomaly… as the silence deepens, the null result itself becomes harder to dismiss.

This is correct, and yet… Physicists have built many experiments that have achieved extraordinary sensitivities. If cold dark matter was composed of WIMPs as originally hypothesized, we would have detected them long ago. Initially, the reaction was to modify WIMPs. Did we say the cross-section would be 10-39 cm2? We meant 10-44 cm2. When that was excluded, we slid the cross section still lower, but people also started giving themselves permission to think the unthinkable. By unthinkable I mean a particle that can’t be detected, not modified gravity. That’s more unthinkable. So the anomaly isn’t dismissed, but it is treated with less gravity than it should be, and certainly with less import than a positive detection would have been granted. Did we say WIMPs? We didn’t mean just WIMPs. It could be anything. (They damn well meant WIMPs and only WIMPs#. Anyone who tells you otherwise is gaslighting*% you, and probably themselves.)

The post goes on to talk about MOND. It gives me too much credit for the gravitational lensing work. This was done by Tobias Mistele, and our work is based on that of Brouwer et al. But it is correct to note that these data are a problem for the dark matter paradigm. Rotation curves remain flat beyond where dark matter halos should end. If correct, this is a genuine anomaly. Perhaps in some distant future it will be recognized** as such in retrospect; at present it seems mostly to be ignored.

It goes on to talk about the JWST observations. Yeah, that part is correct. The community seems to be in the usual process of gaslighting itself into denial of the anomaly. For the first two years after JWST started returning images of the deep universe, people were aghast. How can this be so? It was all anyone could talk about. But then the unexpected became the new normal. Hands were waved, star formation was accepted to be absurdly efficient, and people accepted the impossible. I no longer hear the talk of how problematic the JWST observations are; this chatter simply stopped.

Anomalies don’t weigh a paradigm down if we don’t accept that they’re anomalies. But I’ve lived through the revolution, it’s hard to see a positive outcome while it is still ongoing. For it is certainly true that

What waits on the other side of the dark matter revolution — if that is what is coming — we cannot yet know.

The future is the unknown territory. We don’t know, and can never know, if dark matter doesn’t exist – it is impossible to prove the negative. But we do know MOND works much better than it should in a universe made of dark matter. That demands a scientific explanation that is still wanting. But MOND by itself is not a complete answer, so we are like the parable of the blind men and the elephant, each sensing a different part of reality but as yet unable to see the whole.

Still, there is reason for optimism. The article closes by noting that

Kuhn’s deepest insight was not that science changes. It is that the change, when it comes, is never merely technical. It is a reorganization of the world itself — the universe seen suddenly whole in a configuration it has always had, but that we had simply lacked the paradigm to perceive.

Not knowing how things ultimately work out is good, actually. One way or the other, there is still fundamental science to be done. We have not reached the stage of looking for our discoveries in the sixth place of decimals.


*Trivia I just learned looking at Popper’s wikipedia page: he was spending his last days in London around the same time I was a postdoc in Cambridge just starting to struggle with the scientific and philosophical implications of the dark matter-MOND miasma.

Unrelated trivia: I was at a workshop in Jerusalem early in the century but missed the opportunity to meet Jacob Bekenstein because I was too shy to bother the great man.

+If you do not find this confusing, you are not thinking clearly.

^A nice, brief summary of this early history is related by Einasto. This is the first place I’ve seen the citation to Opik (1915) written out. I’ve only heard mentioned verbally before, so I’ll have to try looking that up later.

The full story is way more complicated than this sounds, and still gets debated off and on. The amplitude of the Oort discrepancy is much smaller today. Locally, the 3D density of mass seems to be accounted for by known stars, gas, and stellar remnants (which were still a new thing in the 1930s). So this Oort limit shows no discrepancy. There remains a modest discrepancy in the 2D dynamical surface density. It appears to me to boil down to the vertical restoring force having a (sometimes ignored) term that depends on the gradient of the rotation curve. Were that falling in a normal Newtonian way, there would be no discrepancy. But it isn’t; this deviation from Newton in the radial direction leads to the Oort discrepancy in the vertical direction. Instead of being as negative as Newton predicts, dV/dR is close to zero, hence my description of this as an indirect detection of a[n almost] flat rotation curve. (dV/dR = -1.7 km/s/kpc, so not exactly zero, but a lot closer to zero than Newton without dark matter would have it be.) The vertical discrepancy is nevertheless much reduced, now being well below a factor of two.

$To his apparently great embitterment. He had some choice things to say about astronomers of his time. I am inclined to suspect that those who praise Zwicky the loudest today would have been among those he had reason to complain about had they been contemporaries.

%This is attributed to Planck, but he had a lot more nuanced things to say about it in his Nobel Prize lecture.

&Einstein disavowed the cosmological constant as his “greatest blunder,” so one argument against it was (for a long time) that it should never have been a part of the theory of General Relativity in the first place. I wonder how things might have gone had that been the case – that he had never introduced Lambda. Perhaps then the data that led to us accepting Lambda would have required a genuine revolution, but it isn’t obvious that we would have accepted it (we might still be debating it), nor is it apparent that LCDM is what comes out of such a revolution. But we don’t get to do that experiment: the Great Man had suggested Lambda, so it was OK to bring it back: we weren’t wrecking his theory by introducing a crazy new entity, we were just admitting an unlikely (antigravity-like) component thereof.

#Or axions! Or warm or self-interacting dark matter. Or macros nee strange nuggets! Or or or… Sure, there have been lots of ideas for what the dark matter could be. But when we say that “by the late 1980s and early 1990s, cold dark matter had been formally incorporated into the reigning cosmological framework” what the vast majority of scientists working on the topic (including myself) meant was that CDM == WIMPs. We were aggressively derisive of other ideas, and these are only dredged up again now because of the experimental non-detection of WIMPs. WIMPs are still a better dark matter candidate than the others for the same reasons that we were derisive of the others back in the day. We haven’t been looking as hard for the others, so comparable experimental limits do not yet exist. To quote myself,

The concept of dark matter is not falsifiable. If we exclude one candidate, we are free to make up another one. After WIMPs, the next obvious candidate is axions. Should those be falsified, we invent something else. (Particle physicists love to do this. The literature is littered with half-baked dark matter candidates invented for dubious reasons, often to explain phenomena with obvious astrophysical causes. The ludicrous uproar over the ATIC and PAMELA cosmic ray experiments is a good example.)

McGaugh (2008)

*%An easy way to deflate such gaslighting is to ask why so many experiments have been built to search for WIMPs but not all these other allegedly great dark matter candidates. After a pause and dismayed stare, you’ll probably get an answer about “looking under the lamp post” because that’s where it is possible to make detections. That’s sorta true, but it isn’t the real reason. The real reason is that we all drank the Kool-Aid of the WIMP miracle, so genuinely believed that the dark matter had to be WIMPS, not merely that they were a convenient experimental target. (I did not chug the kool-aid as hard as the people who based entire careers on building WIMP detection experiments, but I did buy into the idea to the exclusion of other possibilities for dark matter – as did most everyone else.)

**In retrospect, Galileo’s observations of the angular size and phases of Venus were utterly fatal to the geocentric paradigm. That’s easy to say now; at the time it was just another piece of evidence.

Very thin galaxies

Very thin galaxies

The stability of spiral galaxies was a foundational motivation to invoke dark matter: a thin disk of self-gravitating stars is unstable unless embedded in a dark matter halo. Modified dynamics can also stabilize galactic disks. A related test is provided by how thin such galaxies can be.

Thin galaxies exist

Spiral galaxies seen edge-on are thin. They have a typical thickness – their short-to-long axis ratio – of q ≈ 0.2. Sometimes they’re thicker, sometimes they’re thinner, but this is often what we assume when building mass models of the stellar disk of galaxies that are not seen exactly* edge-on. One can employ more elaborate estimators, but the results are not particularly sensitive to the exact thickness so long as it isn’t the limit of either razor thin (q = 0) or a spherical cow (q = 1).

Sometimes galaxies are very thin. Behold the “superthin” galaxy UGC 7321:

UGC 7321 as seen in optical colors by the Sloan Digital Sky Survey.

It also looks very thin in the infrared, which is the better tracer of stellar mass:

Fig. 1 from Matthews et al (1999): H-band (1.6 micron) image of UGC 7321. Matthews (2000) finds a near-IR axis ratio of 14:1. That’s super thin (q = 0.07)!

UGC 7321 is very thin, would be low surface brightness if seen face-on (Matthews estimates a central B-band surface brightness of 23.4 mag arcsec-2), has no bulge component thickening the central region, and contains roughly as much mass in gas as stars. All of these properties dispose a disk to be fragile (to perturbations like mergers and subhalo crossings) and unstable, yet there it is. There are enough similar examples to build a flat galaxy catalog, so somehow the universe has figured out a way for galaxy disks to remain thin and dynamically cold# for the better part of a Hubble time.

We see spiral galaxies at various inclinations to our line of sight. Some will appear face on, others edge-on, and everything in between. If we observe enough of them, we can work out what the intrinsic distribution is based on the projected version we see.

First, some definitions. A 3D object has three principle axes of lengths a, b, and c. By convention, a is the longest and c the shortest. An oblate model imagines a galaxy like a frisbee: it is perfectly round seen face-on (a = b); seen edge-on q = c/a. More generally, an object can be triaxial, with a ≠ b ≠ c. In this case, a galaxy would not appear perfectly round even when seen perfectly face-on^ because it is intrinsically oval (with similar axis lengths a ≈ b but not exactly equal). I expect this is fairly common among dwarf Irregular galaxies.

The observed and intrinsic distribution of disk thicknesses

Benevides et al. (2025) find that the distribution of observed axis ratios q is pretty flat. This is a consequence of most galaxies being seen at some intermediate viewing angle. One can posit an intrinsic distribution, model what one would see at a bunch of random viewing angles, and iterate to extract the true distribution in nature, which they do:

Figure 6 from Benevides et al. (2025): Comparison between the observed (projected) q distribution and the inferred intrinsic 3D axis ratios for a subsample of dwarfs in the GAMA survey with M=109109.5M. The observed shapes are shown with the solid black line and are used to derive an intrinsic c/a (long-dashed) and b/a (dotted) distribution when projected. Solid color lines in each panel corresponds to the q values obtained from the 3D model after random projections. Note that a wide distribution of q values is generated by a much narrower intrinsic c/a distribution. For example, the blue shaded region in the left panel shows that an observed 5% of galaxies with q<0.2 requires 41% of galaxies to have an intrinsic c/a<0.2 for an oblate model. Similarly, for a triaxal model (right panel, red curve) 43% of galaxies are required to be thinner than c/a=0.2. The additional freedom of ba in the triaxial model helps to obtain a better fit to the projected q distribution, but the changes mostly affect large q values and changes little the c/a frequency derived from highly elongated objects.

That we see some thin galaxies implies that they they have to be common, as most of them are not seen edge-on. For dwarf$ galaxies of a specific mass range, which happens to include UGC 7321, Benevides et al. (2025) infer a lot% of thin galaxies, at least 40% with q < 0.2. They also infer a little bit of triaxiality, a ≈ b.

The existence and numbers of thin dwarfs seems to come as a surprise to many astronomers. This is perhaps driven in part by theoretical expectations for dwarf galaxies to be thick: a low surface brightness disk has little self-gravity to hold stars in a narrow plane. This expectation is so strong that Benevides et al. (2025) feel compelled to provide some observed examples, as if to say look, really:

Figure 8 – images of real galaxies from Benevides et al. (2025): Examples of 10 highly elongated dwarf galaxies with q0.2 and M=107108.5M. They resemble thin edge-on disks and can be found even among the faintest dwarfs in our sample. Legends in each panel quote the stellar mass, the shape parameter q, as well as the GAMA identifier. Objects are sorted by increasing M, left to right.

As an empiricist who has spent a career looking at low mass and low surface brightness galaxies, this does not come as a surprise to me. These galaxies look normal. That’s what the universe of late type dwarf$ galaxies looks like.

Edge-on galaxies in LCDM simulations

Thin galaxies do not occur naturally in the hierarchical mergers of LCDM (e.g., Haslbauer et al. 2022), where one would expect a steady bombardment by merging masses to mess things up. The picture above is not what galaxy-like objects in LCDM simulations look like. Scraping through a few simulations to find the flattest galaxies, Benevides et al. (2025) find only a handful of examples:

Figure 11 – images of simulated galaxies from Benevides et al. (2025): Edge-on projection of examples of the flattest galaxies in the TNG50 simulation, in different bins of stellar mass.

Note that only the four images on the left here occupy the same stellar mass range as the images of reality above. These are as close as it gets. Not terrible, but also not representative&. The fraction of galaxies this thin is a tiny fraction of the simulated population whereas they are quite common in reality. Here the two are compared: three different surveys (solid lines) vs. three different simulations (dashed lines).

Figure 9 from Benevides et al. (2025): Fraction of galaxies that are derived to be intrinsically thinner than c/a0.2 as a function of stellar mass. Thick solid lines correspond to our observational samples while dashed lines are used to display the results of cosmological simulations. Different colors highlight the specific survey or simulation name, as quoted in the legend. In all observational surveys, the frequency of thin galaxies peaks for dwarfs with M109M, almost doubling the frequency observed on the scale of MW-mass galaxies. Thin galaxies do not disappear at lower masses: we infer a significant fraction of dwarf galaxies with M<109M to have c/a<0.2. This is in stark contrast with the negligible production of thin dwarf galaxies in all numerical simulations analyzed here.

Note that the thinnest galaxies in nature are dwarfs of mass comparable to UGC 7321. Thin disks aren’t just for bright spirals like the Milky Way with log(M*) > 10.5. They are also common*$ for dwarfs with log(M*) = 9 and even log(M*) = 8, which are often gas dominated. In contrast, the simulations produce almost no galaxies that are thin at these lower masses.

The simulations simply do not look like reality. Again. And again, etc., etc., ad nauseam. It’s almost as if the old adage applies: garbage in, garbage out. Maybe it’s not the resolution or the implementation of the simulations that’s the problem. One could get all that right, but it wouldn’t matter if the starting assumption of a universe dominated by cold dark matter was the input garbage.

Galaxy thickness in Newton and MOND

Thick disks are not merely a product of simulations, they are endemic to Newtonian dynamics. As stars orbit around and around a galaxy’s center, they also oscillate up and down, bobbing in and out of the plane. How far up they get depends on how fast they’re going (the dynamical temperature of the stellar population) and how strong the restoring force to the plane of the disk is.

In the traditional picture of a thin spiral galaxy embedded in a quasi-spherical dark matter halo, the restoring force is provided by the stars in the disk. The dark matter halo is there to boost the radial force to make the rotation curve flat, and to stabilize the disk, for which it needs to be approximately spherical. The dark matter halo does not contribute much to the vertical restoring force because it adds little mass near the disk plane. In order to do that, the halo would have to be very squashed (small q) like the disk, in which case we revive the stability problem the halo was put there to solve.

This is why we expect low surface brightness disks to be thick. Their stars are spread thin, the surface mass density is low, so the restoring force to the disk should be small. Disks as thin as UGC 7321 shouldn’t be possible unless they are extremely cold*# dynamically – a situation that is unlikely to persist in a cosmogony built by hierarchical merging. The simulations discussed above corroborate this expectation.

In MOND, there is no dark matter halo, but the modified force should boost the vertical restoring force as well as the radial force. One thus expects thinner disks in MOND than in Newton.

I pointed this out in McGaugh & de Blok (1998) along with pretty much everything else in the universe that people tell me I should consider without bothering to check if I’ve already considered. Here is the plot I published at the time:

Figure 9 of McGaugh & de Blok (1998): Thickness q = z0/h expected for disks of various central surface densities σ0. Shown along the top axis is the equivalent B-band central surface brightness μ0 for ϒ* = 2. Parameters chosen for illustration are noted in the figure (a typical scale length h and two choices of central vertical velocity dispersion ςz). Other plausible values give similar results. The solid lines are the Newtonian expectation and the dashed lines that of MOND. The Newtonian and MOND cases are similar at high surface densities but differ enormously at low surface densities. Newtonian disks become very thick at low surface brightness. In contrast, MOND disks can remain reasonably thin to low surface density.

There are many approximations that have to be made in constructing the figure above. I assumed disks were plane-parallel slabs of constant velocity dispersion, which they are not. But this suffices to illustrate the basic point, that disks should remain thinner&% in MOND than in Newton as surface density decreases: as one sinks further into the MOND regime, there is relatively more restoring force keep disks thin. To duplicate this effect in Newton, one must invent two kinds of dark matter: a dissipational kind of dark matter that forms a dark matter disk in addition to the usual dissipationless cold dark matter that makes a quasi-spherical dark matter halo.

The idea of the plot above was to illustrate the trend of expected thickness for galaxies of different central surface brightness. One can also build a model to illustrate the expected thickness as a function of radius for a pair of galaxies, one high surface brightness (so it starts in the Newtonian regime at small radii) and one of low surface brightness (in the MOND regime everywhere). I have chosen numbers** resembling the Milky Way for the high surface brightness galaxy model, and scaled the velocity dispersion of the low surface brightness model so it has very nearly the same thickness in the Newtonian regime. In MOND, both disks remain thin as a function of radius (they flare a lot in Newton) and the lower surface brightness disk model is thinner thanks to the relatively stronger restoring force that follows from being deeper in the MOND regime.

The thickness of two model disks, one high surface brightness (solid lines) and the other low surface brightness (dashed lines), as a function of radius. The two are similar in Newton (black), but differ in MOND (blue). The restoring force to the disk is stronger in MOND, so there is less flaring with increasing radius. The low surface brightness galaxy is further in the MOND regime, leading naturally to a thinner disk.

These are not realistic disk models, but they again suffice to illustrate the point: thin disks occur naturally in MOND. Low surface brightness disks should be thick in LCDM (and in Newtonian dynamics in general), but can be as thin as UGC 7321 in MOND. I didn’t aim to make q ≈ 0.1 in the model low surface brightness disk; it just came out that way for numbers chosen to be reasonable representations of the genre.

What the distribution of thicknesses is depends on the accretion and heating history of each individual disk. I don’t claim to understand that. But the mere existence of dwarf galaxies with thin disks is a natural outcome in MOND that we once again struggle to comprehend in terms of dark matter.


*Seeing a galaxy highly inclined minimizes the inclination correction to the kinematic observations [Vrot = Vobs/sin(i)] but to build a mass model we also need to know the face-on surface density profile of the stars, the correction for which depends on 1/cos(i). So as a practical matter, the competition between sin(i) and cos(i) makes it difficult to analyze galaxies at either extreme.

#Dynamically cold means the random motions (quantified by the velocity dispersion of stars σ) are small compared to ordered rotation (V) in the disk, something like V/σ ≈ 10. As a disk heats (higher σ) it thickens, as some of that random motion goes in the vertical direction perpendicular to the disk. Mergers heat disks because they bring kinetic energy in from random directions. Even after an object is absorbed, the splash it made is preserved in the vertical distribution of the stars which, once displaced, never settle back into a thin disk. (Gas can settle through dissipation, but point masses like stars cannot.)

^Oval distortions are a major source of systematic error in galaxy inclination estimates, especially for dwarf Irregulars. It is an asymmetric error: a galaxy with a mild oval distortion can be inferred to have an inclination (i > 0) even when seen face-on (i = 0), but it can never have an inclination more face-on (i < 0) than exactly face-on. This is one of the common drivers of claims that low mass galaxies fall off the Tully-Fisher relation. (Other common problems include a failure to account for gas mass, bad distance estimates, or not measuring Vflat.)

$In a field with abominable terminology, what is meant by a “dwarf” galaxy is one of the worst offenders. One of my first conference contributions thirty years ago griped about the [mis]use of this term, and matters have not improved. For this particular figure, Benevides et al. (2025) define it to mean galaxies with stellar masses in the range 9 < log(M*) < 9.5, which seems big to me, but at least it is below the mass of a typical L* spiral, which has log(M*) ~ 10.5. For comparison, see Fig. 6 of the review of Bullock & Boylan-Kolchin (2017), who define “bright dwarfs” to have 7 < log(M*) < 9, and go lower from there, but not higher into the regime that we’re calling dwarf right now. So what a dwarf galaxy is depends on context.

%Note that the intrinsic distribution peaks below q = 0.2, so arguably one should perhaps adopt as typical the mode of the distribution (q ≈ 0.17).

&Another way in which even the thin simulated objects are not representative of reality is that they are dynamically hot, as indicated by the κrot parameter printed with the image. This is the fraction of kinetic energy in rotation. One of the more favorable cases with κrot = 0.67 corresponds to V/σ = 2.5. That happens in reality, but higher values are common. Of course, thin disks and dynamical coldness go hand in hand. Since the simulations involve a lot of mergers, the fraction of kinetic energy in rotation is naturally small. So I’m not saying the simulations are wrong in what they predict given the input physics that they assume, but I am saying that this prediction does not match reality.

*$The fraction of thin galaxies observed by DESI is slightly higher than found in the other surveys. Having looked at all these data, I am inclined to suspect the culprit is image quality: that of DESI is better. Regardless of the culprit for this small discrepancy between surveys, thin disks are much more common in reality than in the current generation of simulations.

*#There seems to be a limit to how cold disks get, with a minimum velocity dispersion around ~7 km/s observed in face-on dwarfs when the appropriate number, according to Newton, would be more like 2 km/s, tops. I remember this number from observations in the ’80s and ’90s, along with lots of discussion then to the effect of how can it be so? but it is the new year and I’m feeling too lazy to hunt down all the citations so you get a meme instead.


&%In an absolute sense, all other things being equal, which they’re not, disks do become thicker to lower surface brightness in both Newton and MOND. There is less restoring force for less surface mass density. It is the relative decline in restoring force and consequent thickening of the disk that is much more precipitous in Newton.

**For the numerically curious, these models are exponential disks with surface density profiles Σ(R) = Σ0 e-R/Rd. Both models have a scale length Rd = 3 kpc. The HSB has Σ0 = 866 M pc-2; this is a good match to the Eilers et al. (2019) Milky Way disk; see McGaugh (2019). The LSB has Σ0 = 100 M pc-2, which corresponds roughly to what I consider the boundary of low surface brightness, a central B-band surface brightness of ~23 mag. arcsec-2. For the velocity dispersion profile I also assume an exponential with scale length 2Rd (that’s what supposed to happen). The central velocity dispersion of the HSB is 100 km/s (an educated guess that gets us in the right ballpark) and that of the LSB is 33 km/s – the mass is down by a factor of ~9 so the velocity dispersion should be lower by a factor of 9\sqrt{9}. (I let it be inexact so the solid and dashed Newtonian lines wouldn’t exactly overlap.)

These models are crude, being single-population (there can be multiple stellar populations each with their own velocity dispersion and vertical scale height) and lacking both a bulge and gas. The velocity dispersion profile sometimes falls with a scale length twice the disk scale length as expected, sometimes not. In the Milky Way, Rd ≈ 2.5 or 3 kpc, but the velocity dispersion falls off with a scale length that is not 5 or 6 kpc but rather 21 or 25 kpc. I have also seen the velocity dispersion profile flatten out rather than continue to fall with radius. That might itself be a hint of MOND, but there are lots of different aspects of the problem to consider.

Has dark matter been detected in the Milky Way?

Has dark matter been detected in the Milky Way?

If a title is posed as a question, the answer is usually

No.

There has been a little bit of noise that dark matter might have been detected near the center of the Milky Way. The chatter seems to have died down quickly, for, as usual, this claim is greatly exaggerated. Indeed, the claim isn’t even made in the actual paper so much as in the scuttlebutt# related to it. The scientific claim that is made is that

The halo excess spectrum can be fitted by annihilation with a particle mass mχ 0.5–0.8 TeV and cross section συ (5–8)×1025cm3s1 for the bb¯ channel.

Totani (2025)

What the heck does that mean?

First, the “excess spectrum” refers to a portion of the gamma ray emission detected by the Fermi telescope that exceeds that from known astrophysical sources. This signal might be from a WIMP with a mass in the range of 500 – 800 GeV. That’s a bit heavier than originally anticipated (~100 GeV), but not ridiculous. The cross-section is the probability for an interaction with bottom quarks and anti-quarks. (The Higgs boson can decay into b quarks.)

Astrophysical sources at the Galactic center

There is a long-running issue with the interpretation of excess signals as dark matter. Most of the detected emission is from known astrophysical sources, hence the term “excess.” There being an excess implies that we understand all the sources. There are a lot of astrophysical sources at the Galactic center:

The center of the Milky Way as seen by the South African MeerKAT radio telescope with a close up from JWST. Image credit: NASA, ESA, CSA, STScI, SARAO, S. Crowe (UVA), J. Bally (CU), R. Fedriani (IAA-CSIC), I. Heywood (Oxford).

As you can see, the center of the Galaxy is a busy place. It is literally the busiest place in the Galaxy. Attributing any “excess” to non-baryonic dark matter is contingent on understanding all of the astrophysical sources so that they can be correctly subtracted off. Looking at the complexity of the image above, that’s a big if, which we’ll come back to later. But first, how does dark matter even come unto a discussion of emission from the Galactic center?

Indirect WIMP detection

Dark matter does not emit light – not directly, anyway. But WIMP dark matter is hypothesized to interact with Standard Model particles through the weak nuclear force, which is what provides a window to detect it in the laboratory. So how does that work? Here is the notional Feynman diagram:

Conceivable Interactions between WIMPs (X) and standard model particles (q). The diagram can be read left to right to represent WIMPs scattering off of atomic nuclei, top to bottom to represent WIMPs annihilating into standard model particles, or bottom to top to represent the production of dark matter particles in high energy collisions.

The devious brilliance of this Feynman diagram is that we don’t need to know how the interaction works. There are many possibilities, but that’s a detail – that central circle is where the magic happens; what exactly that magic is can remain TBD. All that matters is that it can happen (with some probability quantified by the interaction cross-section), so all the pathways illustrated above should be possible.

Direct detection experiments look for scattering of WIMPs off of nuclei in underground detectors. They have not seen anything. In principle, WIMPs could be created in sufficiently high-energy collisions of Standard Model particles. The LHC has more than adequate energy to produce dark matter particles in this way, but no such signal has been seen$. The potential signal we’re discussing here is an example of indirect detection. There are a number of possibilities for this, but the most obvious^ one follows from WIMPs being their own anti-particles, so they occasionally meet in space and annihilate into Standard Model particles.

The most obvious product of WIMP annihilations is a pair of gamma rays, hence the potential for the Fermi gamma ray telescope to detect their decay products. Here is a simulated image of the gamma ray sky resulting from dark matter annihilations:

Simulated image from the via Lactea II simultion (Fig. 1 of Kuhlen et al. 2008).

The dark regions are the brightest, where the dark matter density is highest. That includes the center of the Milky Way (white circle) and also sub-halos that might contain dwarf satellite galaxies.

Since we don’t really know how the magic interaction happens, but have plenty of theoretical variations, many other things are also possible, some of which might be cosmic rays:

Fig. 3 of Topchiev et al. (2017) illustrating possible decay channels for WIMP annihilations. Gamma rays are one inevitable product, but other particles might also be produced. These would be born with energies much higher than their rest masses (~100 GeV, while electrons and positrons have masses of 0.5 MeV) so would be moving near the speed of light. In effect, dark matter could be a source of cosmic rays.

The upshot of all this is that the detection of an “excess” of unexpected but normal particles might be a sign of dark matter.

Sociology: different perspectives from different communities

A lot hinges on the confidence with which we can disentangle expected from unexpected. Once we’ve accounted for the sources we already knew about, there are always new sources to be discovered. That’s astronomy. So initially, the communal attitude was that we shouldn’t claim a signal was due to dark matter until all astrophysical signals had been thoroughly excluded. That never happened: we just kept discovering new astrophysical sources. But at some point, the communal attitude transformed into one of eager credulity. It was no longer embarrassing to make a wrong claim; instead, marginal and dubious claims were made eagerly in the hopes of claiming a Nobel prize. If it didn’t work out, oh well, just try again. And again and again and again. There is apparently no shame in claiming to see the invisible when you’re completely convinced it is there to be seen.

This switch in sociology happened in the mid to late ’00s as people calling themselves astroparticle& physicists became numerous. These people were remarkably uninterested in astrophysics or astrophysical sources in their own right but very interested in dark matter. They were quick to claim that any and every quirk in data was a sign of dark matter. I can’t help but wonder if this behavior is inherited from the long drought in interesting particle collider results, which gradually evolved into a propensity for high energy particle phenomenologists to leap on every two-sigma blip as a sign of new physics, dumping hundreds of preprints on arXiv after each signal of marginal significance was announced. It is always a sprint to exercise the mental model-building muscles and make up some shit in the brief weeks before the signal inevitably goes away again.

Let’s review a few examples of previous indirect dark matter detection claims.

Cosmic rays from Kaluza-Klein dark matter – or not

This topic has a long and sordid history. In the late ’00s, there were numerous claims of an excess in cosmic raysATIC saw too many electrons for the astrophysical background, and and PAMELA saw an apparent rise in the positron fraction, perhaps indicating a source with a peak energy around 620 GeV. (If the signal is from dark matter, the rest mass of the WIMP is imprinted in the energy spectrum of its decay products.) The combination of excess electrons and extra positrons seemed fishy enough* to some to point to new physics: dark matter. There were of course more sober analyses, for example:

Fig. 3 from Aharonian et al. (2009): The energy spectrum E3 dN/dE of cosmic-ray electrons measured by H.E.S.S. and balloon experiments. Also shown are calculations for a Kaluza-Klein signature in the H.E.S.S. data with a mass of 620 GeV and a flux as determined from the ATIC data (dashed-dotted line), the background model fitted to low-energy ATIC and high-energy H.E.S.S. data (dashed line) and the sum of the two contributions (solid line). The shaded regions represent the approximate systematic error as in Fig. 2.

A few things to note about this plot: first, the data are noisy – science is hard. The ATIC and H.E.S.S. data are not really consistent – one shows an excess, the other does not. The excess is over a background model that is overly simplistic – the high energy astrophysicists I knew were shouting that the apparent signal could easily be caused by a nearby pulsar##. The advocates for a detection in the astroparticle community simply ignored this point, or if pressed, asserted that it seemed unlikely.

One problem that arose with the dark matter interpretation was that there wasn’t enough of it. Space is big and the dark matter density is low, so it is hard to get WIMPs together to annihilate. Indeed, the expected signal scales as the square of the WIMP density, so is very sensitive to just how much dark matter is lurking about. The average density in the solar neighborhood needed to explain astronomical data is around 0.3 to 0.4 GeV cm-3; this falls short of producing the observed signal (if real) by a factor of ~500.

An ordinary scientist might have taken this setback as a sign that he$$ was barking up the wrong tree. Not to be discouraged, the extraordinary astroparticle physicists started talking about the “boost factor.” If there is a region of enhanced dark matter density, then the gamma ray/cosmic ray signal would be boosted, potentially by a lot given the density-squared dependence. This is not quite as crazy as it sounds, as cold dark matter halos are predicted to be lumpy: there should be lots of sub-halos within each halo (and many sub-sub halos within those, right the way down). So, what are the odds that we happen to live near enough to a subhalo that could result in the required boost factor?

The odds are small but nonzero. I saw someone at a conference in 2009 make a completely theoretical attempt to derive those odds. He took a merger tree from some simulation and calculated the chance that we’d be near one of these lumps. Then he expanded that to include a spectrum of plausible merger trees for Milky Way-mass dark matter halos. The noisier merger histories gave higher probabilities, as halos with more recent mergers tend to be lumpier, having had a fresh injection of subhalos that haven’t had time to erode away through dynamical friction into the larger central halo.

This was all very sensible sounding, in theory – and only in theory. We don’t live in any random galaxy. We live in the Milky Way and we know quite a bit about it. One of those things is that it has had a rather quiet merger history by the standards of simulated merger trees. To be sure, there have been some mergers, like the Gaia-Enceladus Sausage. But these are few and far between compared to the expectations of the simulations our theorist was considering. Moreover, we’d know if it weren’t, because mergers tend to heat the stellar disk and puff up its thickness. The spiral disk of the Milky Way is pretty cold dynamically, which places limits on how much mass has merged and when. Indeed, there is a whole subfield dedicated to the study of the thick disk, which seems to have been puffed up in an ancient event ~8 Gyr ago. Since then it has been pretty quiet, though more subtle things can and do happen.

The speaker did not mention any of that. He had a completely theoretical depiction of the probabilities unsullied by observational evidence, and was succeeding in persuading those who wanted to believe that the small probability he came up with was nevertheless reasonable. It was a mixed audience: along with the astroparticle physicists were astronomers like myself, including one of the world’s experts on the thick disk, Rosy Wyse. However, she was too polite to call this out, so after watching the discussion devolve towards accepting the unlikely as probable, I raise my hand to comment: “We know the Milky Way’s merger history isn’t as busy as the models that give a high probability.” This was met with utter incredulity. How could astronomy teach us anything about dark matter? It’s not like the evidence is 100% astronomical in nature, or… wait, it is. But no, no waiting or self-reflection was involved. It rapidly became clear that the majority of people calling themselves astroparticle physicists were ignorant of some relevant astrophysics that any astronomy grad student would be expected to know. It just wasn’t in their training or knowledge base. Consequently, it was strange and shocking&& for them to learn about it this way. So the discussion trended towards denial, at which point Rosy spoke up to say yes, we know this. Duh. (I paraphrase.)

The interpretation of the excess cosmic ray signal as dark matter persisted a few years, but gradually cooler heads prevailed and the pulsar interpretation became widely accepted to be more plausible – as it always had been. Indeed, claiming cosmic rays were from dark matter became almost disreputable, as it richly deserved to be. So much so that when the AMS cosmic ray experiment joined the party late, it had essentially zero impact. I didn’t hear anyone advocating for it, even in whispers at workshops. It seemed more like its Nobel laureate PI just wanted a second Nobel prize, please and thank you, and even the astroparticle community felt embarrassed for him.

This didn’t preclude the same story from playing out repeatedly.

Gamma rays from WIMPs – or not

In the lead-up to a conference on dark matter hosted at Harvard in 2014, there were claims that the Fermi telescope – the same one that is again in the news – had seen a gamma ray line around 126 GeV that was attributed to dark matter. This claim had many red flags. The mass was close to the Higgs particle mass, which was kinda weird. The signal was primarily seen on the limb of the Earth, which is exactly where you’d expect garbage noise to creep in. Most telling, the Fermi team itself was not making this claim. It came from others who were analyzing their data. I am no fan of science by big teams – they tend to become bureaucratic behemoths that create red tape for their participants and often suppress internal dissent** – but one thing they do not do is leave Nobel prizes unanalyzed in their data. The Fermi team’s silence in this matter was deafening.

In short, this first claim of gamma rays from dark matter looked to be very much on the same trajectory as that from cosmic rays. So I was somewhat surprised when I saw the draft program for the Harvard conference, as it had an entire afternoon session devoted to this topic. I wrote the organizers to politely ask if they really thought this would still be a thing by the time the conference happened. One of them was an enthusiastic proponent, so yes.

Narrator: it was not.

By the time the conference happened, the related claims had all collapsed, and all the scientists invited to speak about it talked instead about something completely different, as if it had never been a thing at all.

X-rays from sterile neutrinos – or not

Later, there was the 3.5 keV line. If one squinted really hard at X-ray data, it looked like there might sorta kinda be an unidentified line. This didn’t look particularly convincing, and there are instances when new lines have been discovered in astronomical data rather than laboratory data (e.g., helium was first recognized in the spectrum of the sun, hence the name; also nebulium, which was later recognized to be ionized oxygen), so again, one needed to consider the astrophysical possibilities.

Of course, it was much more exciting to claim it was dark matter. Never mind that it was a silly energy scale, being far too low mass to be cold dark matter (people seem to have forgotten*# the Lee-Weinberg limit, which requires mX > 2 GeV); a few keV is rather less than a few GeV. No matter, we can always come up with an appropriate particle – in this case, sterile neutrinos*$.

If you’ve read this far, you can see how this was going to pan out.

Gamma rays from WIMPs again, maybe maybe

So now we have a renewed claim that the Fermi excess is dark matter. Given the history related above, the reader may appreciate that my first reaction was Really? Are we doing this again?

“Many people have speculated that if we knew exactly why the bowl of petunias had thought that we would know a lot more about the nature of the Universe than we do now.”

― Douglas Adams, The Hitchhiker’s Guide to the Galaxy

This is different from the claim a decade ago. The claimed mass is different, and the signal is real, being part of the mess of emission from the Galactic center. The trick, as so often the case, is disentangling the dark matter signal from the plausible astrophysical sources.

Indeed, the signal is not new, only this particular fit with WIMP dark matter is. There had, of course, been discussion of all this before, but it faded out when it became clear that the Fermi signal was well explained by a population of millisecond pulsars. Astrophysics was again the more obvious interpretation*%. Or perhaps not: I suppose if you’re part of a community convinced that dark matter exists who is spending an enormous amount of time and resources looking for a signal from dark matter and whose basic knowledge of astrophysics extends little beyond “astronomical data show dark matter exists but are messy so there’s always room to play” then maybe invoking an invisible agent from an unknown dark sector seems just as plausible as an obvious astrophysical source. Hmmm… that would have sounded crazy to me even back when, like them, I was sure that dark matter had to exist and be made of WIMPs, but here we are.

Looking around in the literature, I see there is still a somewhat active series of papers on this subject. They split between no way and maybe.

For example, Manconi et al. (2025) show that the excess signal has the same distribution on the sky as the light from old stars in the Galaxy. The distribution of stars is asymmetrical thanks to the Galactic bar, which we see at an angle somewhere around ~30 degrees, so one end is nearer to us than the other, creating a classic “X/peanut” shape seen in other edge-on barred spiral galaxies. So not only is the spectrum of the signal consistent with millisecond pulsars, it has the same distribution on the sky as the stars from which millisecond pulsars are born. So no way is this dark matter: it is clearly an astrophysical signal.

Not to be dissuaded by such a completely devastating combination of observations, Muru et al. (2025) argue that sure, the signal looks like the stars, but the dark matter could have exactly the same distribution as the stars. They cite the Hestia simulations of the Local Group as an example where this happens. Looking at those, they’re not as unrealistic as many simulations, but they appear to suffer the common affliction of too much dark mass near the center. That leaves the dark matter more room to be non-spherical so maybe be lumpy in the same was as the stars, and also provide a higher annihilation signal from the high density of dark matter. So they say maybe, calling the pulsar and dark matter interpretations “equally compelling.”

Returning to Totani’s sort-of claimed detection, he also says

This cross section is larger than the upper limits from dwarf galaxies and the canonical thermal relic value, but considering various uncertainties, especially the density profile of the MW halo, the dark matter interpretation of the 20 GeV “Fermi halo” remains feasible.

Totani (2025)

OK, so there’s a lot to break down in this one sentence.

The canonical thermal relic value is kinda central to the whole WIMP paradigm, so needing a value higher than that is a red flag reminiscent of the need for a boost factor for the cosmic ray signal. There aren’t really enough WIMPs there to do the job unless we juice their effectiveness at making gamma rays. The juice factor is an order of magnitude here: Steigman et al. (2012) give 2.2 x 10-26 cm3s-1 for what the thermal cross-section should be vs. the (5-8) x 10-25 cm3s-1 suggested by Totani (2025).

It is also worth noting that one point of Steigman’s paper is that as a well-posed hypothesis, the WIMP cross section can be calculated; it isn’t a free parameter to play with, so needing the cross-section to be larger than the upper limits from dwarf galaxies is another red flag. If this is indeed a dark matter signal from the Galactic center, then the subhalos in which dwarf satellites reside should also be visible, as in the simulated image from via Lactea above. They are not, despite having fewer messy astrophysical signals to compete with.

So “remains feasible” is doing a lot of work here. That’s the scientific way of saying “almost certainly wrong, but maybe? Because I’d really like for it to work out that way.”

The dark matter distribution in the Milky Way

One of the critical things here is the density of dark matter near the Galactic center, as the signal scales as the square of the density. Totani (2025) simply adopts the via Lactea simulation to represent the dark matter halo of the Galaxy in his calculations. This is a reasonable choice from a purely theoretical perspective, but it is not a conservative choice for the problem at hand.

What do we know empirically? The via Lactea simulation was dark matter only. There is no stellar disk, just a dark matter halo appropriate to the Milky Way. So let’s add that halo to a baryonic mass model of the Galaxy:

The rotation curve of the via Lactea dark matter halo (red curve) combined with the Milky Way baryon distribution (light blue line). The total rotation (dark blue line) overshoots the data.

The important part for the Galactic center signal is the region at small radius – the first kpc or two. Like most simulations, via Lactea has a cuspy central region of high dark matter density that is inconsistent with data. This overshoots the equivalent circular velocity curve from observed stellar motions. I could fix the fit above by reducing the stellar mass, but that’s not really an option in the Milky Way – we need a maximal stellar disk to explain the microlensing rate towards the center of the Galaxy. The “various uncertainties, especially the density profile of the MW halo” statement elides this inconvenient fact. Astronomical uncertainties are ever-present, but do not favor a dark matter signal here.

We can subtract the baryonic mass model from the rotation curve data to infer what the dark matter distribution needs to be. This is done in the plot below, where it is compared to the via Lactea halo:

The empirical dark matter halo density profile of the Milky Way (blue line) compared to the via Lactea simulation (red line).

The empirical dark matter density profile of the Milky Way does not continue to rise inwards as steeply as the simulation predicts. It shows the same proclivity for a shallower core as pretty much every other galaxy in the sky. This reduced density of dark matter in the central couple of kpc means the signal from WIMP annihilation should be much lower than calculated from the simulated distribution. Remember – the WIMP annihilation signal scales as the square of the dark matter density, so the turn-down seen at small radii in the log-log plot above is brutal. There isn’t enough dark matter there to do what it is claimed to be doing.

Cry wolf

There have now been so many claims to detect dark matter that have come and gone that it is getting to be like the fable of the boy who cried wolf. A long series of unpersuasive claims does not inspire confidence that the next will be correct. Indeed, it has the opposite effect: it is going to be really hard to take future claims seriously.

It’s almost as if this invisible dark matter stuff doesn’t exist.


Note added: Jeff Grube points out in the comments that Wang & Duan (2025) have a recent paper showing that the dark matter signal discussed here also predicts an antiproton signal that is already excluded by AMS data. While I find this unsurprising, it is an excellent check. Indeed, it would have caused me to think again had the antiproton signal been there: independent corroboration from a separate experiment is how science is supposed to work.


#It has become a pattern for advocates of dark matter to write a speculative paper for the journals that is fairly restrained in its claims, then hype it as an actual detection to the press. It’s like “Even I think this is probably wrong, but let’s make the claim on the off chance it pans out.”

$Ironically, a detection from a particle collider would be a non-detection. The signature of dark matter produced in a collision would be an imbalance between the mass-energy that goes into the collision and that measured in detected particles coming out of it. The mass-energy converted into WIMPs would escape the detector undetected. This is analogous to how neutrinos were first identified, though Fermi was reluctant to make up an invisible, potentially undetectable particle – a conservative value system that modern particle physicists have abandoned. The 13,000 GeV collision energy of the LHC is more than adequate to make ~100 GeV WIMPs, so the failure of this detection mode is telling.

^A less obvious possibility is spontaneous decay. This would happen if WIMPs are unstable and decay with a finite half-life. The shorter the half-life, the more decays, and the stronger the resulting signal. This implies some fine-tuning in the half-life – if it is much longer than a Hubble time, then it happens so seldom it is irrelevant; if it is shorter than a Hubble time, then dark matter halos evaporate and stable galaxies don’t exist.

&Astroparticle physics, also known as particle astrophysics, is a relatively new field. It is also an oxymoron, being a branch of particle physics with only aspirational delusions of relevance to astrophysics. I say that to be rude to people who are rude to astronomers, but it is also true. Astrophysics is the physics of objects in the sky, and as such, requires all of physics. Physics is a broad field, so some aspects are more relevant than others. When I teach a survey course, it touches on gravity, electromagnetism, atomic and molecular quantum mechanics, nuclear physics, and with the discovery of exoplanets, increasingly on geophysics. Particle physics doesn’t come up. It’s just not relevant, except where it overlaps with nuclear physics. (As poorly as particle physicists think of astronomers, they seem to think even less of nuclear physicists, whom they consider to be failed particle physicists (if only they were smart enough!) and nuclear physicists hate them in return.) This new field of astroparticle physics seems to be all about dark matter as driven by early universe cosmology, with contempt for everything that happens in the 13 billion years following the production of the relic radiation seen as the microwave background. Anything later is dismissed as mere “gastrophysics” that is too complicated to understand so cannot possibly inform fundamental physics. I guess that’s true if one chooses to remain ignorant of it.

*Fishy results can also indicate something fishy with the data. I had a conversation with an instrument builder at the time who pointed out that PAMELA had chosen to fly without a particular discriminator in order to save weight; he suggested that its absence could explain the apparent upturn in positrons.

##There is a relatively nearby pulsar that fits the bill. It has a name: Geminga. This illustrates the human tendency to see what we’re looking for. The astroparticle community was looking for dark matter, so that’s what many of them saw in the excess cosmic ray signal. High energy astrophysicists work on neutron stars, so the obvious interpretation to them was a pulsar. One I recall being particularly scornful of the dark matter interpretation when there was an obvious astrophysical source. I also remember the astroparticle people being quick to dismiss the pulsar interpretation because it seemed unlikely to them for one to be so close but really they hadn’t thought about it before: that pulsars could do this was news to them, and many preferred to believe the dark matter interpretation.

$$All the people barking were men.

&&This experience opened my eyes to the existence of an entire community of scientists who were working on dark matter in somewhat gratuitous ignorance of the astronomical evidence for dark matter. To them, the existence of the stuff had already been demonstrated; the interesting thing now was to find the responsible particle. But they were clearly missing many important ingredients – another example is disk stability, a foundational reason to invoke dark matter that seems to routinely come as a surprise to particle physicists. This disconnect is part of what motivated me to develop an entire semester course on dark matter, which I’ve taught every other year since 2013 and will teach again this coming semester. The first time I taught it, I worried that there wasn’t enough material for a whole semester. Now a semester isn’t enough time.

**I had a college friend (sadly now deceased) who was part of the team that discovered the Higgs. That was big business, to the extent that there were two experiments – one to claim the detection, and another on the same beam to do the confirmation. The first experiment exceeded the arbitrary 5σ threshold to claim a 5.2σ detection, but the second only reached 4.9σ. So, in all appropriateness, he asked in a meeting if they could/should really announce a detection. A Nobel prize was on the line, so the answer was straightforward: Do you want a detection or not? (His words.)

*#Rather than forget, some choose to fiddle ways around the Lee-Weinberg limit. This has led to the sub-genre of “light dark matter” which means lightweight, not luminous. I’d say this was the worst name ever, but the same people talk about dark photons with a straight face, so irony continues to bleed out.

*$Ironically, a sterile neutrino has also been invoked to address problems in MOND.

*%I was amused once to see one of the more rabid advocates of dark matter signals of this type give an entire talk hyping the various possibilities only to mention pulsars at the end with a sigh, admitting that the Fermi signal looked exactly like that.

The odd primordial halo of the Milky Way

The odd primordial halo of the Milky Way

The mass distribution of dark matter halos that we infer from observations tells us where the dark matter needs to be now. This differs form the mass distribution it had to start, as it gets altered by the process of galaxy formation. It is the primordial distribution that dark matter-only simulations predict most robustly. We* reverse-engineer the collapse of the baryons that make up the visible Galaxy to infer the primordial distribution, which turns out to be… odd.

The Gaia rotation curve and the mass of the Milky Way

As we discussed a couple of years ago, Gaia DR3 data indicate a declining rotation curve for the Milky Way. This decline becomes more steep, nearly Keplerian, in the outskirts of the Milky Way (17 < R < 30 kpc). This is may or may not be consistent with data further out, which gets hard to interpret as the LMC (at 50 kpc) perturbs orbits and the observed motions may not correspond to orbits in dynamical equilibrium. So how much do the data inform us about the gravitational potential?

Milky Way rotation curve (various data) including Gaia DR3 (multiple analyses). Also shown is the RAR model (blue line) that was fit to the terminal velocities from 3 < R < 8.2 kpc (gray points) and predates other data illustrated here.

I am skeptical of the Keplerian portion of this result (as discussed at length at the time) because other galaxies don’t do that. However, I am a big fan of listening to the data, and the people actually doing the work. Taken at face value, the Gaia data show a Keplerian decline with a total mass around 2 x 1011 M. If correct, this falsifies MOND.

How does dark matter fare? There is an implicit assumption made by many in the community that any failing of MOND is an automatic win for dark matter. However, it has been my experience that observations that are problematic for MOND are also problematic for dark matter. So let’s check.

Short answer: this is really weird in terms of dark matter. How weird? For starters, most recent non-Gaia dynamical analyses suggest a total mass closer to 1012 M, a factor of five higher than the Gaia value. I’m old enough to remember when the accepted mass was 2 x 1012 M, an order of magnitude higher. Yet even this larger mass is smaller than suggested by abundance matching recipes, which give more like 4 x 1012 M. So somewhere in the range 2 – 40 x 1011 M.

The Milky Mass has been adjusted so often, have we finally hit it?

The guy was all over the road. I had to swerve a number of times before I hit him.

Boston Driver’s Handbook (1982 edition)&

If it sounds like we’re all over the map, that’s because we are. It is very hard to constrain the total mass of a dark matter halo. We can’t see it, nor tell where it ends. We infer, indirectly, that the edge is way out beyond the tracers we can see. Heck, even speaking of an “edge” is ill-defined. Theoretically, we expect it to taper off with the density of dark matter falling as ρ ~ r-3, so there is no definitive edge. Somewhat arbitrarily,** we adopt the radius that encloses a density 200 times the average density of the universe as the “virial” radius. This is all completely notional, and it gets worse, as the process of forming a galaxy changes the initial mass distribution. What we observe today is the changed form, not the primordial initial condition for which the notional mass is defined.

Adiabatic compression during galaxy formation

To form a visible galaxy, baryons must dissipate and sink to the center of their parent dark matter halo. This process changes the mass distribution and alters the halo from its primordial state. In effect, the gravity of the sinking baryons drags some dark matter along# with them.

The change to the dark matter halo is often called adiabatic compression. The actual process need not be adiabatic, but that’s how we approximate it. We’ve tested this approximation with detailed numerical simulations, and it works pretty well, at least if you do it right (there are boring debates about technique). What happens makes sense intuitively: the response of the primordial halo to the infall of baryons is to become more dense at the center. While this makes sense physically, it is problematic for LCDM as it takes an NFW halo that is already too dense at the center to be consistent with data and makes it more dense. This has been known forever, so opposing this is one thing feedback is invoked to do, which it may or may not do, depending on how it really works. Even if feedback can really turn a compressed cusp into a core, it is widely to expected to be important only in low mass galaxies where the gravitational potential well isn’t too deep. It isn’t supposed to be all that important in galaxies as massive as the Milky Way, though I’m sure that can change as needed.

There are a variety of challenges to implementing an accurate compression computation, so we usually don’t bother: the standard practice is to assume a halo model and fit it to the data. That will, at best, given a description of the current dark matter halo, not what it started as, which is our closest point of comparison with theory. To give an example of the effect, here is a Milky Way model I built a decade ago:

Figure 13 from McGaugh (2016)Milky Way rotation curve from the data of Luna et al. (2006, red points) and McClure-Griffiths & Dickey (2007, gray points) together with a bulgeless baryonic mass model (black line). The total rotation is approximately fit (blue line) with an adiabatically compressed NFW halo (solid green line) using the procedure implemented by Sellwood & McGaugh (2005). The primordial halo before compression is shown as the dashed line. The parameters of the primordial halo are a concentration c = 7 and a mass M200 = 6 x 1011 M. Fitting NFW to the present halo instead gives c = 14, M200 = 4 x 1011 M, so the difference is appreciable and depend on the quality and radial extent of the available data.

The change from the green dashed line to the solid green line is the difference compression makes. That’s what happens if a baryon distribution like that of the Milky Way settles in an NFW halo. The inferred mass M200 is lower and the concentration c higher than it originally was – and it is the original version that we should compare to the expectations of LCDM.

When I built this model, I considered several choices for the bulge/bar fraction: something reasonable, something probably too large, and something definitely too small (zero). The model above is the last case of zero bulge/bar. I show it because it is the only case for which the compression procedure worked. If there is a larger central concentration of baryons – i.e., a bulge and/or a bar – then the compression is greater. Too great, in fact: I could not obtain a fit (see also Binney & Piffl and this related discussion).

The calculation of the compression requires knowledge of the primordial halo parameters, which is what one is trying to obtain. So one has to guess an initial state, run the code, check how close it came, then iterate the initial guess. This is computationally expensive, so I was just eyeballing the fit above. Pengfei has done a lot of work to implement a method that iteratively computes the compression and rigorously fits it to data. So we decided to apply it to the newer Gaia DR3 data.

Fitting the Gaia rotation curve with adiabatically compressed halos

We need two inputs here: one, the rotation curve to fit, and two, the baryonic distribution of the Milky Way. The latter is hard to specify given our location within the Milky Way, so there are many different estimates. We tried a dozen.

Another challenge of doing this is deciding which data rotation curve data to fit. We chose to focus on the rotation curve of Jiao et al. (2023) because they made estimates of the systematic as well as random errors. The statistics of Gaia are so good it is practically impossible to fit any equilibrium model to them. There are aspects of the data for which we have to consider non-equilibrium effects (spiral arms, the bar, “snails” from external perturbations) so the usual assumptions are at best an approximation, plus there can always be systematic errors. So the approach is to believe the data, but with the uncertainty estimate of Jiao et al. (2023) that includes systematics.

For a halo model, we started with the boilerplate LCDM NFW halo$. This doesn’t fit the data. Indeed, all attempts to fit NFW halos fail in similar ways for all of the different baryonic mass models we tried. The quasi-Keplerian part of the Gaia rotation curve simply cannot be fit: the NFW halo inevitably requires more mass further out.

Here are a few examples of the NFW fits:


Fig. A.3 from Li et al. (2025). Fits of Galactic circular velocities using the NFW model implementing adiabatic halo contraction using 3 baryonic models. [Another 9 appear in the paper.] Data points with errors are the rotation velocities from Jiao et al. (2023), while open triangles show the data from Eilers et al. (2019), which are not fitted. [The radius ranges from 5 to 30 kpc.] Blue, purple, green and black solid lines correspond to the contributions by the stellar disk, central bar, gas (and dust if any), and compressed dark matter halo, respectively. The total contributions are shown using red solid lines. Black dashed lines are the inferred primordial halos.

LCDM as represented by NFW suffers the same failure mode as seen in MOND (plot at top): both theories overshoot the Gaia rotation curve at R > 17 kpc. This is an example of how data that are problematic for MOND are also problematic for dark matter.

We do have more freedom in the case of dark matter. So we tried a different halo model, Einasto. (For this and many other halo models, see Pengfei’s epic compendium of dark matter halo fits.) Where NFW has two parameters, a concentration c and mass M200, Einasto has a third parameter that modulates the shape of the density profile%. For a very specific choice of this third parameter (α = 0.17), it looks basically the same as NFW. But if we let α be free, then we can obtain a fit. Of all the baryonic models, the RAR model+compressed Einasto fits best:


Fig. 1 from Li et al. (2025). Example of a circular velocity fit using the McGaugh19$$ model for baryonic mass distributions. The purple, blue, and green lines represent the contributions of the bar, disk, and gas components, respectively. The solid and dashed black lines show the current and primordial dark matter halos, respectively. The solid red line indicates the total velocity profile. The black points show the latest Gaia measurements (Jiao et al. 2023), and the gray upward triangles and squares show the terminal velocities from (McClure-Griffiths & Dickey 2007, 2016), and Portail et al. (2017), respectively. The data marked with open symbols were not fit because they do not consider the systematic uncertainties.

So it is possible to obtain a fit considering adiabatic compression. But at what price? The parameters of the best-fit primordial Einasto halo shown above are c = 5.1, M200 = 1.2 x 1011 M, and α = 2.75. That’s pretty far from the α = 0.17 expected in LCDM. The mass is lower than low. The concentration is also low. There are expectation values for all these quantities in LCDM, and all of them miss the mark.


Fig. 2 from Li et al. (2025). Halo masses and concentrations of the primordial Galactic halos derived from the Gaia circular velocity fits using 12 baryonic models. The red and blue stars with errors represent the halos with and without adiabatic contraction, respectively. The predicted halo mass-concentration relation within 1 σ from simulations (Dutton & Macciò 2014) is shown as the declining band. The vertical band shows the expected range of the MW halo mass according to the abundance-
matching relation (Moster et al. 2013). The upper and lower limits are set by the highest stellar mass model plus 1 σ and the lowest stellar mass model minus 1 σ, respectively.

The expectation for mass and concentration is shown as the bands above. If the primordial halo were anything like what it should be in LCDM, the halo parameters represented by the red stars should be where the bands intersect. They’re nowhere close. The same goes for the shape parameter. The halo should have a density profile like the blue band in the plot below; instead it is more like the red band.


Fig. 3 from Li et al. (2025). Structure of the inferred primordial and current Galactic halos, along with predictions for the cold and warm dark matter. The density profiles are scaled so that there is no need to assume or consider the masses or concentrations for these halos. The gray band indicates the range of the current halos derived from the Gaia velocity fits using the 12 baryonic models, and the red band shows their corresponding primordial halos within 1σ. The blue band presents the simulated halos with cold dark matter only (Dutton & Macciò 2014). The purple band shows the warm dark matter halos (normalized to match the primordial Galactic halo) with a core size spanning from 4.56 kpc (WDM5 in Macciò et al. 2012) to 7.0 kpc, corresponding to a particle mass of 0.05 keV and lower.

So the primordial halo of the Milky Way is pretty odd. From the perspective of LCDM, the mass is too low and the concentration is too low. The inner profile is too flat (a core rather than a cusp) and the outer profile is too steep. This outer steepness is a large part of why the mass comes out so low; there just isn’t a lot of halo out there. The characteristic density ρs is at least in the right ballpark, so aside from the inner slope, the outer slope, the mass, and the concentration, LCDM is doing great.

What if we ignore the naughty bits?

It is really hard for any halo model to fit the steep decline of the Gaia rotation curve at R > 17 kpc. Doing so is what makes the halo mass so small. I’m skeptical about this part of the data, so do things improve if we don’t sweat that part?

Ignoring the data at R > 17 kpc allows the mass to be larger, consistent with other dynamical determinations if not quite with abundance matching. However, the inner parts of the rotation curve still prefer a low density core. That is, something like the warm dark matter halo depicted as the purple band above rather than NFW with its dense central cusp. Or self-interacting dark matter. Or cold dark matter with just-so feedback. Or really anything that obfuscates the need to confront the dangerous question: why does MOND perform better?


*This post is based on the recently published paper by my former student Pengfei Li, who is now faculty at Nanjing University. They have a press release about it.

&A few months after reading this in the Boston Driver’s Handbook, this exact thing happened to me.

**This goes back to BBKS in 1986 when the bedrock assumption was that the universe had Ωm = 1, for which the virial radius was 188 times the critical density. 200 was close enough, and stuck, even though for LCDM the virial radius is more like an overdensity close to 100, which is even further out.

#This is one of many processes that occur in simulations, which are great for examining the statistics of simulated galaxy-like objects but completely useless for modeling individual galaxies in the real universe. There may be similar objects, but one can never say “this galaxy is represented by that simulated thing.” To model a real galaxy requires a customized approach.

$NFW halos consistently perform worse in fitting data than any other halo model, of which there are many. It has been falsified as a viable representation of reality so many times that I can’t recall them all, and yet they remain the go-to model. I think that’s partly thanks to their simplicity – it is mathematically straightforward to implement – and to the fact that is what simulations predict: LCDM halos should look like NFW. People, including scientists, often struggle to differentiate simulation from reality, so we keep flogging the dead horse.

%The density profile of the NFW halo model asymptotes to power laws at both small and large radii: ρ → r-1 as r → 0 and ρ → r-3 as r → ∞. The third parameter of Einasto allows a much wider ranges of shapes.

Einasto profiles. Einasto is observationally indistinguishable from NFW for α = 0.17, but allows many other shapes.

$$The McGaugh19 model user here is the one with a reasonable bulge/bar. This dense component can be fit in this case because we start with a halo model with a core rather than a cusp (closer to α = 1 than to the α = 0.17 of NFW/LCDM).