Sister blog of Physicists of the Caribbean. Shorter, more focused posts specialising in astronomy and data visualisation.

Wednesday, 19 February 2025

Nobody Ram Pressure Strips A Dwarf !

Very attentive readers may remember a paper from 2022 claiming, with considerable and extensive justification, to have detected a new class of galaxian object : the ram pressure dwarf. These are similar to the much more well-known tidal dwarf galaxies, which form when gravitational encounters remove so much gas from galaxies that the stripped material condenses into a brand new object. Ram pressure dwarfs would be essentially similar, but result from ram pressure stripping instead of tidal encounters. A small but increasing number of objects in Virgo seem to fit the bill for this quite nicely, as they don't match the scaling relations for normal galaxies very well at all.

This makes today's paper, from 2024, a little late to the party. Here the authors are also claiming to have discovered a new class of object, which they call a, err... ram pressure dwarf. From simulations.

I can't very well report this one without putting my sarcastic hat on. So you discovered the same type of object but two years later and only in a simulation eh ? I see. And you didn't cite the earlier papers either ? Oh.

And I also have to point out an extremely blatant "note to self" that clearly got left in accidentally. On the very first page :

Among the ∼60 ram-pressure-stripped galaxies belonging to this sample, ionized gas powered by star formation has been detected (R: you can get ionized gas that is not a result of star formation as well, so maybe you could say how they have provided detailed information about the properties of the ionized gas, its dynamics, and star formation in the tails instead) in the tentacles.

No, that's not even the preprint. That's the full, final, published journal article !

Okay, that one made me giggle, and I sympathise. Actually I once couldn't be bothered to finish looking up the details of a reference so I put down "whatever the page numbers are" as a placeholder... but the typesetter fortunately picked up on this ! 

What does somewhat concern me at a (slightly) more serious level, though, is that this got through the publication process. Did the referee not notice this ? I seem to get picked up on routinely for the most minor points which frankly amount to no more than petty bitching, so it does feel a bit unfair when others aren't apparently having to endure the same level of scrutiny.

Right, sarcastic hat off. In a way, that this paper is a) late and b) only using simulations is advantageous. It seems that objects detected initially in observational data have been verified by theoretical studies fully independently of the original discoveries. That gives stronger confirmation that ram pressure dwarfs are indeed really a thing.

Mind you, I think everyone has long suspected in the back of their minds that ram pressure dwarfs could form. After all, why not ? If you remove enough gas, it stands to reason that sometimes part of it could become gravitationally self-bound. But it's only recently that we've had actual evidence that they exist, so having theoretical confirmation that they can form is important. That puts the interpretation of the observational data on much stronger footing.

Anyway, what the authors do here is to search one of the large, all-singing, all-dancing simulations for candidates where this would be likely. They begin by looking for so-called jellyfish galaxies, in which ram pressure is particularly strong so that the stripped gas forms distinct "tentacle" structures. They whittle down their sample to ensure they have no recent interactions with other galaxies, so that the gas loss should be purely due to ram pressure and not tidal encounters. Of the three galaxies in their sample which meet this criteria, they look for stellar and gaseous overdensities within their sample and find one good ram pressure dwarf candidate, which they present here.

By no means does this mean that such objects are rare. Their criteria for sample selection is deliberately strict so they can be extremely confident of what they've found. Quite likely there are many other candidates lurking in the data which they didn't find only because they had recent encounters with other galaxies, which would mean they weren't "purely" resulting from ram pressure. I use the quotes because determining which factor was mainly responsible for the gas loss can be extremely tricky. And simulation resolution limits mean there could be plenty of smaller candidates in there. The bottom line is that they've got only one candidate because they demand the quality of that candidate be truly outstanding, not because they're so rare as to be totally insignificant.

And that candidate does appear to be really excellent and irrefutable. It's a clear condensation of stars and gas at the end of the tentacle that survives for about a gigayear, with no sign of any tidal encounters being responsible for the gas stripping. It's got a total stellar mass of about ten million solar masses, about ten times as much gas, and no dark matter – the gas and stars are bound together by their own gravity alone. The only weird thing about it is the metallicity, which is extraordinarily large, but this appears to be an artifact of the simulations and doesn't indicate any fundamental problem.

In terms of the observational candidates, this one is similar in size but at least a hundred times more massive. Objects that small would, unfortunately, be simply unresolvable in the simulations because it doesn't have nearly enough particles. But this is consistent with this object being just the tip of a much more numerous iceberg of similar but smaller features. Dedicated higher resolution simulations might be able to make better comparisons with the observations, until someone finds a massive ram pressure dwarf in observational data.

I don't especially like this paper. It contains the phrase "it is important to note" no less than four times, it says "as mentioned previously" in relation to things never before mentioned, it describes the wrong panels in the figures, and it has many one-sentence "paragraphs" that make it feel like a BBC News article if the writer was unusually technically competent. But all of these quibbles are absolutely irrelevant to the science presented, which so far as I can tell is perfectly sound. As to the broader question of whether ram pressure dwarfs form a significant component of the galaxy population, and indeed how they manage to survive without dark matter in the hostile environment of a cluster... that will have to await further studies.

How To Starve Your Hedgehog

Today, two papers on hedgehogs quenched galaxies. It'll make more sense later on, but only slightly.

"Quenched" is just a bit of jargon meaning that galaxies have stopped forming stars, if not completely, then at least well below their usual level. There are a whole bunch of ways this can happen, but they all mostly relate to environment. Basically you need some mechanism to get the gas out of galaxies where it then disperses. In clusters this is especially easy because of ram pressure stripping, where the hot gas of the cluster itself can push out gas within galaxies. In smaller groups the main method would be tidal interactions, though this isn't as effective.

What about in isolation ? There things get tricky. Even the general field is not a totally empty environment : there are other galaxies present (just not very many) and external gas (just of very low density). But you also have to start to consider what might have happened to galaxies there over the whole of time, because conditions were radically different in the distant past.

To cut a long story short, what we find is that giant galaxies seem to have formed the bulk of their stars way back in a more exciting era when things were just getting started. Dwarf galaxies in the field, on the other hand, are still forming stars, and in fact their star formation rate has been more or less permanently constant.

This phenomena is called downsizing, and for a long time had everyone sorely puzzled : naively, giant galaxies ought to assemble more slowly, so were presumed to have taken longer to assemble their stellar population, whereas dwarfs should form more quickly. Simplifying, this was due to host of problems in the details of the physics of the models, and as far as I know it's generally all sorted out now. Small amounts of gas can, in fact, quite happily maintain a lower density for longer, hence dwarfs form stars more slowly but much more persistently.

Dwarfs are, of course, much more susceptible to environmental gas-loss removal processes than giants, and indeed dwarfs in clusters are mostly devoid of gas (except for recent arrivals). And so conversely, any dwarfs which have lost their gas in the field are unexpected, because there's nothing very much going on out there : all galaxies of about the same mass should have about the same level of star formation. There's no reason that some of them should have lost their gas and others held on to it - it should be an all-or-nothing affair.

That's why isolated quenched galaxies are interesting, then. On to the new results !


The first paper concentrates on a single example which they christen "Hedgehog", because "hedgehogs are small and solitary animals" and also presumably because "dw1322m2053" is boring, and cutesy acronyms are old hat. Wise people, I approve.

This particular hedgehog galaxy is quite nearby (2.4 Mpc) and extremely isolated, at least 1.7 Mpc from any nearby massive galaxies. That puts it at least four times further away than expected from the region of influence of any groups, based on their masses. It's a classic quenched galaxy, "red and dead", smooth and structureless, with no detectable star formation at all.

It's also very, very small. They estimate the stellar mass at around 100,000 solar masses, whereas for more typical dwarf galaxies you could add at least two or three zeros on to that. Now that does mean they can't say definitively if its lack of star formation is a really significant outlier, simply because for an object this small, you wouldn't expect much anyway. But in every respect it appears consistent with being a tiny quenched galaxy, so the chance that it has any significant level of star formation is remote.

How could this happen ? There are a few possibilities. While it's much further away from the massive groups than where you'd normally expect to see any effect from them, simulations have shown that it's just possible to get quenched galaxies this far out. But this is extraordinarily unlikely, given that they found this object serendipitously. They also expect these so-called "backsplash" galaxies (objects which have passed through a cluster and out the other side*) to be considerably larger than this one, because they would have formed stars for a prolonged time, right up until the point they fell into the cluster.

* I presume and hope this is a men's urinal reference.

Another option is simply that the star formation in small galaxies might be self-limiting, with stellar winds and supernovae able to eject the gas. This, they say, is only expected to be temporary (since most of the gas should fall back in after a while), so again the chances of finding something like this are pretty slim. But I'd have liked more details about this, since I would expect that for galaxies this small - and it really is tremendously small - the effects of feedback could be stronger than for more typical, more massive galaxies. Maybe stellar winds and explosions could permanently eject much more of the gas, although on the other hand galaxies this small would have fewer massive stars capable of this.

Similarly another possibility, which I don't think they mention, is quenching due to ram pressure in the field. Again, for normal dwarf galaxies, this is hardly a promising option. For ram pressure to work effectively, you need gas of reasonably high density and galaxies moving at significant speeds, neither of which happens in the field. But, studies have shown that galaxies in the field do experience (very) modest amounts of gas loss which correlates with the distance from the large-scale filaments. Ordinarily this is not really anything substantial, but for galaxies this small, it might be. Since a galaxy this small just won't have much gas to begin with, and removing it will be easy because it's such a lightweight, what would normally count as negligible gas loss might be fatal for a tiddler like this.

The most interesting option is reionisation. When the very first stars were formed, theory says, there were hardly any elements around except hydrogen and helium and a smattering of others. Heavier elements allow the gas to cool and therefore condense more efficiently, so today's stars are comparative minnows. But with none of this cooling possible, the earliest stars were monsters, perhaps thousands of times more massive than the Sun. They were so powerful that they reionised the gas throughout the Universe, heating it so that cooling was strongly suppressed, at least for a while. In more massive galaxies gravity eventually overcame this, but in the smallest galaxies it could be halted forever.

Hedgehog, the authors say, is right on the limit where quenching by reionisation is expected to be effective. If so then it's a probe of conditions in the very early universe, one which is extremely important as it's been used a lot to explain why we don't detect nearly as many dwarf galaxies as theory otherwise predicts*. The appealing thing about this explanation is the small size and mass of the object, which isn't predicted by other mechanisms.

* They do mention that the quenched fraction of galaxies in simulations rises considerably at lower masses, but how much of this is due to reionisation is unclear.

This galaxy isn't quite a singular example, but objects like this one are extremely rare. Of course ideally we'd need a larger sample, which is where the second paper comes in.


This one is a much more deliberate attempt to study quenched galaxies, though not necessarily isolated. What they're interested in is our old friends, Ultra Diffuse Galaxies, those surprisingly large but faint fluffy objects that often lack dark matter. In this paper the authors used optical spectroscopy to target a sample of 44 UDGs, not to measure their dynamics (the spectroscopic measurements are too imprecise for that) but to get their chemical composition. With this they can identify galaxies in a post-starburst phase, essentially just after star formation has stopped. That kind of sample should be ideal for identifying where and when galaxies get quenched.

I'm going to gloss over a lot of careful work they do to ensure their sample is useful and their measurements accurate. The sample size is necessarily small because UDGs are faint, and their own data finds that some of the distance estimates were wrong so a few candidates weren't actually UDGs after all. Their final result of 6 post-starburst UDGs doesn't sound much, and indeed it isn't, but these kinds of studies are still in their very early days and you have to start somewhere.

Even with the small size, they find two interesting results. First, the fraction of quenched UDGs is around 20%, much higher than the general field population. The stellar masses are a lot higher than the Hedgehog but still small compared to most dwarfs though, so this one needs to be treated with a bit of caution but it's definitely interesting. Second, while most quenched UDGs do appear to result from environmental effects, a few are indeed isolated. Which is a bit weird and unexpected. UDGs in clusters might form by gas loss of more "typical" galaxies, but this clearly can't work in the field, so why only a select few should lose gas isn't clear at all.


What all this points to isn't all that surprising, though in a somewhat perverse sense : it underscores that we don't fully understand the physics of star formation. The authors of the second study favour stellar feedback as being responsible for a temporary suppression of star formation. If this is common and repeated, with galaxies experiencing many periods of star formation interspersed with lulls, that could also make Hedgehog a bit less weird - if, say, it's forming/not forming stars for roughly the same total amount of time, then it wouldn't be so strange to detect it during a quenched phase. And of course the lower dark matter content of UDGs surely also has some role to play in this, although what that might be is anyone's guess.

As usual, more research is needed. At this point we just need more data, both observational and simulations. That we're still finding strange objects that're hard to explain isn't something to get pessimistic about though. We've learned a lot, but we're still figuring out just much further we have to go before we really understand these objects.

Monday, 17 February 2025

Sports Stars Can Save Humanity

I know, I know, I get far less than my proverbial five-a-day so far as reading papers goes. Let me try and make some small amends.

Today, a brief overview of a couple of visualisation papers I read while I was finishing off my own on FRELLED, plus a third which is somewhat tangentially related.


The first is a really comprehensive review of the state of astronomical visualisation tools in 2021. Okay, they say it isn't comprehensive, which is strictly speaking true, but that would be an outright impossible task. In terms of things at a product-level state, with useable interfaces, few bugs and plenty of documentation, this is probably as close as anyone can realistically get.

Why is a review needed ? Mainly because with the "digital tsunami" of data flooding our way, we need to know which tools already exist before we go about reinventing the wheel. As they say, there are data-rich but technique-poor astronomers and data-poor but technique-rich visualisation experts, so giving these groups a common frame of reference is a big help. And as they say, "science not communicated is science not done". The same is true for science ignored as well, of which I'm extremely guilty... you can see from the appallingly-low frequency of posts here how little time I manage to find for reading papers. 

So yeah, having everything all together in one place makes things very much easier. They suggest a dedicated keyword in papers "astrovis" to make everything easier to find. As far as I know this hasn't been adopted anywhere, but it's a good idea all the same.

Most of the paper is given to summarising the capabilities of assorted pieces of software, some of which I still need to check out properly (and yes, they include mine, so big brownie points to them for that !). But they've also thought very carefully about how to organise all this into a coherent whole. For them there are five basic categories for their selected tools : data wrangling (turning data into something suitable for general visualisation), exploration, feature identification, object reconstruction, and outreach. They also cover the lower-level capabilities (e.g. graph plotting, uncertainty visualisation, 2D/3D, interactivity) without getting bogged-down in unproductively pigeon-holing everything. 

Perhaps the best bit of pigeon-unholing is something they quote from another paper : the concept of explornation, an ugly but useful word meaning the combination of exploration and explanation. This, I think, has value. It's possible to do both independently, to go out looking at stuff and never getting any understanding of it at all, or conversely to try and interpret raw numerical data without ever actually looking at it. But how much more powerful is the combination ! Seeing can indeed be believing. The need for good visualisation tools is not only about making pretty pictures (although that is a perfectly worthwhile end in itself) but also in helping us understand and interpret data in different ways, every bit as much as developing new techniques for raw quantification. 

I also like the way they arrange things here because we too often tend to ignore tools developed for different purposes other than our own field of interest. And they're extraordinarily non-judgemental, both about individual tools and different techniques. From personal experience it's often difficult to remain so aloof, to avoid saying, "and we should all do it this way because it's just better". Occasionally this is true, but usually what's good for one person or research topic just isn't useful at all for others.

On the "person" front I also have to mention that people really do have radically different preferences for what they want out of their software. Some, inexplicably, genuinely want everything to do be done via text and code and nothing else, with only the end result being shown graphically. Far more, I suspect, don't like this. We want to do everything interactively, only using code when we need to do something unusual that has to be carefully customised. And for a long time astronomy tools have been dominated too much by the interface-free variety. The more that's done to invert the situation, the better, so far as I'm concerned.


The second paper presents a very unusual overlap between the world of astronomy and... professional athletes. I must admit this one languished in my reading list for quite a while because I didn't really understand what it was about from a quick glance at the abstract or text, mostly because of my own preconceptions : I was expecting it to be about evaluating the relative performance of different people at source-finding. Actually this is (almost) only tangential to the main thrust of the paper, though it's my own fault for misreading what they wrote.

Anyway, professional sports people train themselves and others by reviewing their behaviour using dedicated software tools. One of the relatively simple features that one of these (imaginatively named "SPORTSCODE") has is the ability to annotate videos. This means that those in training can go back over past events and see relevant features, e.g. an expert can point out exactly what and where something of interest happened – and thereby, one hopes, improve their own performance.

What the authors investigate is whether astronomers can use this same technique, even using the same code, to accomplish the same thing. If an expert marks on the position of a faint source in a data cube, can a non-expert go back and gain insight into how they made that identification ? Or indeed if they mark something they think is spurious, will that help train new observers ? The need for this, they say, is that ever-larger data volumes threaten to make training more difficult, so having some clear strategy for how to proceed would be nice. They also note that medical data, where the stakes are much, much higher, relies on visual extraction, while astronomical algorithms have traditionally been... not great. "Running different source finders on the same data set rarely generates the same set of candidates... at present, humans have pattern recognition and feature identification skills that exceed those of any automated approach."

Indeed. This is a sentiment I fully endorse, and I would advocate using as much visual extraction as possible. Nevertheless, my own tests have found that more modern software can approach visual performance in some limited cases, but a full write-up on that is awaiting the referee's verdict.

While this paper asks all the right questions, it presents only limited answers. I agree that it's an interesting question as to whether source finding is a largely inherent or learned (teachable) skill, but most of the paper is about the modifications they made to SPORTSCODE and its setup to make this useful. The actual result is a bit obvious : yes indeed, annotating features is useful for training, and subjectively this feels like a helpful thing to do. I mean... well yeah, but why would you expect it to be otherwise ? 

I was hoping for some actual quantification of how users perform before and after training – to my knowledge nobody has ever done this for astronomy. We muddle through training users as best we can, but we don't quantify which technique works best. That I would have found a lot more interesting. As it is, it's an interesting proof of concept, and it asks all the right questions, but the potential follow-up is obvious and likely much more interesting and productive. I also have to point out that FRELLED comes with all the tools they use for their training methods, without having to hack any professional athletes (or their code) to get them to impart their pedagogical secrets.


The final paper ties back into the question of whether humans can really outperform algorithms. I suppose I should note that these algorithms are indeed truly algorithms in the traditional, linear, procedural sense, and nothing at all to do with LLMs and the like (which are simply no good at source finding). What they try to do here is use the popular SoFiA extractor in combination with a convolutional neural network. SoFiA is a traditional algorithm, which for bright sources can give extremely reliable and complete catalogues, but it doesn't do so well for fainter sources. So to go deeper, the usual approach is to use a human to vet its initial catalogues to reject all the likely-spurious identifications.

The authors don't try to replace SoFiA with a neural network. Instead they use the network to replace this human vetting stage. Don't ask me how neural networks work but apparently they do. I have to say that while I think this is a clever and worthwhile idea, the paper itself leaves me with several key questions. Their definition of signal to noise appears contradictory, making it hard to know exactly how well they've done : it isn't clear to me if they're really used the integrated S/N (as they claim) or the peak S/N (as per their definition). The two numbers mean very different things. It doesn't help that the text is replete with superlatives, which did annoy me quite a bit.

The end result is clear enough though, at least at a qualitative level : this method definitely helps, but not as much as visual inspection. It's interesting to me that they say this can fundamentally only approach but not surpass humans. I would expect that a neural network could be trained on data containing (artificial) sources so faint a human wouldn't spot them, but knowing they were there, the program could be told when it found them and thereby learn their key features. If this isn't the case, then it's possible we've already hit a fundamental limit, that when humans start to dig into the noise, they're doing about as well as it's ever possible to do by any method. When you get to the faintest features we can find, there simply aren't any clear traits that distinguish signal from noise. Actually improving, in any significant way, on human vision, might be a matter of a radically different approach... but it might even be an  altogether hopeless challenge.

And that's nice, isn't it ? Cometh the robot uprising, we shall make ourselves useful by doing astronomical source-finding under the gentle tutelage of elite footballers. 

Or not, because that algorithms can be thousands of times faster can more than offset their lower reliability levels, but that's another story.

Phew ! Three papers down, several hundred more to go.

Saturday, 11 January 2025

Turns out it really was a death ray after all

Well, maybe.

Today, not a paper but an engineering report. Eh ? This is obviously not my speciality at all, in any way shape or form. In fact reading this only revealed to me even further the tremendous depths of my own ignorance regarding materials science and engineering practises. The former is something I never cared for at undergraduate level and the latter is something about which I know literally nothing. Naturally, I wouldn't normally even glance at a report like this, except that it's about a topic that's personally important to me : why Arecibo collapsed.

There's an okay-but-short press release version here. It's interesting to see the extent of the deconstruction at the site, which was already well advanced in 2021; I couldn't find a more recent photo. Otherwise the Gizmodo version is the 30-second read and not much else. For this post I read most of the full 113 page report, which really is "jaw dropping", at least in parts, as Gizmodo described it. Unsurprisingly there are fairly hefty tracts where my eyes glazed over, but there's still plenty in here that's accessible and understandable to non-engineers like me.

In a nutshell, Arecibo collapsed due to a combination of factors, two of which are predictable enough but the third is something nobody expected. The first two are inadequate maintenance and the impact of hurricane Maria. But it's important not to oversimplify, as these are intimately bound with the third : the effects of the radar transmitter. This is not quite a case where one can simply say, "if they'd just done their jobs properly then it'd still be standing today", thought the report does contain some damming stuff.

Going through this linearly would end up being a shorter version of the report, which wouldn't really help anyone. If you want that level of detail you should go through it yourself; it's thorough to the point of going back to hand-written notes from the earliest days of the telescope. I have to say, though, that it's also highly repetitive in parts and in my view somewhat self-contradictory in places – but as it says, this is a preprint and still subject to editorial revision. Anyway, rather than doing a blow-by-blow breakdown, let me extract some broader lessons here.


Safety is not the same as redundancy

Probably the most general lesson is, I presume, obvious to anyone with an engineering background. But to an outsider me the distinction between safety and redundancy was interesting just because it makes a lot of intuitive sense but I'd never heard of it before. Safety, apparently, refers to the breaking point of any particular element. For example a cable with a safety factor of two could support twice as much as its current load before it would snap. Redundancy, on the other hand, is about how many elements could fail before the whole structure would come crashing down. Arecibo's three towers, they say, don't provide redundancy because a single failure would inevitably mean a total collapse (compare with the six of FAST).

Of course it's very unlikely that even a single tower would ever fail because their safety factor was massive, so redundancy there was unnecessary (at least regarding any failure of the concrete towers themselves). The same can't be said for the metal cables, where the safety factor generally seems to have been about a factor two or a bit less, in accordance with standard design practises – still plenty, but with a need for redundancy just in case. The report stops short of saying that there were any actual design flaws in the telescope, but does not that it obviously would have been better if there had been more towers. Safety factors, they say, were not the issue, although I think I detect some inconsistency here. Where they do issue an outright criticism, for example, is that while the original cable system had redundancy, this was no longer true with the 1997 upgrade that added the 900-tonne Gregorian dome and altered the cable system. Which is a little bit in contradiction to their claims that the telescope didn't suffer from design faults. It's a bit muddled.


Poor maintenance contributed to but did not cause the collapse

At least this is the overall gist I got. There's plenty of criticism levelled here, but it's hard to disentangle how serious the maintenance problems really were. As I read it, a more diligent maintenance program probably could have prevented the collapse, but this is partly with the benefit of hindsight – the failures which occurred were unprecedented (see below) but should have been spotted all the same. Of particular concern is that there wasn't enough knowledge transfer during the telescope's two changes of management (I'll speak from first-hand experience in declaring that management changes should be avoided for a host of other reasons; I went through one such change at Arecibo and God knows what the staff must have felt like when a second took place only a few years later). In addition, and probably worst of all, is that the post-Hurricane Maria repair efforts were both much too slow, taking months to even get started and would have lasted for years, and targeting a cable which never failed. Major repairs needed to happen far sooner but there was also a need to identify the failures more accurately.

The failures were of the cable sockets rather than the cables themselves. In these "spelter sockets", there's normally some degree of cable pullout after construction is complete and the structure assumes its full load : these sockets are widely used so this is known to be absolutely normal and no cause for concern. But the report is somewhat ambiguous as to whether the extra pullout which happened could have been noticed. Sometimes it sounds quite damning in describing the extra movement as "clear" but elsewhere describes it as "not accessible by visual inspection". The amount of movement we're talking about, until the point of the collapse itself, was small, of the order 1cm or so. It certainly isn't something you'd spot from a casual glance, but you could measure it by hand easily enough with a ruler. Not noticing this, if I understand things correctly, meant that the cables were estimated to still have their original high safety factors whereas in fact they were much lower. They say this "should have raised the highest alarm level, requiring urgent action". Perhaps most damming, they also say that it is "highly unlikely" that this excessive pullout went unnoticed. They also note that there was a lack of good documentation of maintenance records and procedures.

The contribution of the recent hurricanes, especially Maria, was extremely significant in precipitating the collapse. In fact, "absent Maria, the Committee believes the telescope would still be standing today". Pullout from the sockets shortly after installation is entirely normal as the structure takes the weight, but after that, any further movement isn't normal at all. This did in fact happen, and should have been spotted – but even this, as we'll see, apparently wasn't enough to bring down the telescope by itself.

One final point is that Arecibo wasn't well liked by backend management. I often had the impression of a behind-the-scenes mood of delenda est Arecibo  or at the very least, that that was what some staff members sincerely believed was happening even if it wasn't true. The report notes that a 2006 NSF report recommended closing Arecibo by 2011 if other funding sources couldn't be found, which I found truly bizarre. This was less than ten years after a major upgrade and exactly at the point the biggest surveys were just beginning. As to why anyone would think that closing it at that particular moment was a good idea, I'm truly at a loss. Nothing about it, even with some familiarity of the large-than-life politics behind the place, has ever made a lick of sense to me. 

This is not, I hasten to add, any suggestion of deliberately shoddy maintenance; inasmuch as that was inadequate, there is no need to attribute that to anything besides an incompetently low budget. One strikingly simple recommendation in the report is that funding sources for site operations (e.g. science and development) and maintenance be entirely separate, so there is no chance of any conflict of interest or competition for resources which are essential to both.


The failure was unprecedented

The final and most interesting point of the report, the big headline message, is that Arecibo may have failed because of its radar transmitter. The report is emphatic, and repeats almost ad nauseum, that the kind of socket failures seen here have never before occurred in a century of operations of identical sockets used in bridges and other structures around the world. The damage from the hurricanes was significant, but not enough by itself to explain the failure. There is a crucial missing factor here.

The explanation suggested in the report is electroplasticity. In laboratory conditions, material creep (stretching) can be induced by electrical currents, apparently directly because of the energy released by the flow of electrons. As they note, in the lab this has been found under much higher currents operating for much shorter times, but could presumably work at lower currents if sustained for much longer periods. If correct, this would be Arecibo's final first, another effect of its unique nature. Such currents, they hypothesise, would have been induced by the powerful 1MW radar transmitter used for zapping asteroids and other Solar System objects. This would explain why the cables failed while still having apparently high safety factors, and possibly account for why the failures occurred in some of the youngest cables with no evidence of manufacturing defects (and weren't even the ones with the highest load). It would also, of course, explain why no such other socket failures have ever been seen. Hardly anything else has this combination of radar transmitter and spelter sockets, let alone in tropical conditions in an earthquake zone.

The report goes quite deep into the technical details of electroplasticity. Interestingly, it notes that even less powerful sources can induce currents in human skin that can be directly sensed a few hundred feet from the transmitter. The problem is that understanding the effects of these currently requires highly detailed simulations accounting for the complicated structure of Arecibo's cables, the exact path the current would follow, and using data on low, long-term currents that at present doesn't exist. The most obvious deficiency seems to me that they don't estimate just how long the radar was ever transmitting for. Sure, it was up there for decades, but it wasn't used routinely : regularly, to be sure, but not daily. This is something where a crude estimate should be relatively easy by searching the observing records; even the schedule of what was planned (which didn't always match what was actually done, usually because the err, radar broke) would give a rough indication.




If the report is correct, then there's little need for concern about other structures. The report strongly disagrees that Arecibo points to a need to revise safety standards for spelter sockets more generally; unless your bridge is in the path of 1MW S-band radar transmitter, you can carry on with your morning commute as usual. Well, that's good. Clearly, regardless of electroplasticity, something happened here that was truly exceptional, and not worth worrying about whether it will ever be repeated. Not unless you're an engineer, at any rate.

Whether electroplasticity really was the cause I'm not qualified to judge. Talking to someone older and wiser, the opinion was "they had to come up with something". I don't disagree with that – there just isn't enough data here to say anything for certain. It could be electroplasticity, or it could be something the committee just didn't think of. More analysis of the surviving hardware, along with more studies and simulations, is badly needed.

The broader lesson I would take from all this is that you can run things on a shoestring for a while, but you can't keep trying to do less with more indefinitely. Yes, I'm coloured by my political biases, but austerity to survive a short-term hit is very different to austerity as a way of life : one is manageable, the other isn't. Such a policy does far more harm than good. Yes, you save a little money immediately, but you ultimately lose an awful lot more a little further down the line. So if you're going to fund things, fund things as properly as you can. Incorporate redundancy thinking into managerial practises as well as engineering standards. Have teams large enough to survive the loss of several members. Hire separate observing support staff rather than expecting scientists to do everything. 

Finally, don't expect people to work for meagre compensation (and I'm here not thinking just financially but in other benefits, for example high pay is useless with long hours and/or low holiday time) just because they enjoy their job. Not even the most wildly enthusiastic, energy-driven fanatic can operate at 110% for long. Just because someone is uncomfortable doesn't mean they're working extra-hard. Part of America's puritan hangover appears to be in thinking that work = bad => people who are suffering are good workers. In the end this just leads to everyone hating their job are wanting to overthrow the system but having no clue what to replace it with. Far better to reverse the thinking and presume that those who are happy and comfortable are the best workers. 

This has taken a rather political turn, but it's not unmotivated by my experiences at Arecibo. One notorious manager was definitely of the ilk who believe that more work good, less work bad. Thankfully this is not a mentality I've encountered much in Europe. And, in my view, understanding this isn't just good for us as people, but actually as scientists in getting the work done we want to do. By all means, take a liberal approach : let those who want to work obsessively, who actively thrive because of it, do so, but don't presume the same conditions produce the same results from different people. They don't. As with good software interface design, in the end, solving these issues is just as important for the science we want to do as the scientific problems themselves. Soft issues produce hard results.

Monday, 5 August 2024

Giants in the deep

Here's a fun little paper about hunting the gassiest galaxies in the Universe.

I have to admit that FAST is delivering some very impressive results. Not only is it finding thousands upon thousands of galaxies – not so long ago, the 30,000 HI detections of ALFALFA was leagues ahead of everything else, this has already been surpassed – but in terms of data quality too it looks like it's delivering. This paper exploits that to the extreme with the FAST Ultra Deep Survey, FUDS.

Statistically, big, all-sky surveys are undeniably the most useful. With a data set like that, anyone can search the catalogue and see if anything was detected at any point of interest, and at the very least they can get an upper limit of the gas content of whatever they're interested in. Homogeneity has value too. But of course, with any new telescope you can always go deeper, as long as you're prepared to put in the observing time. That can let you find ever-fainter nearby sources, or potentially sources at greater distances. Or indeed both.

It's the distance option being explored in this first FUDS paper. Like previous ultra-deep surveys from other telescopes, FUDS tales a pencil-beam approach : incredibly sensitive but only over very small areas. Specifically it's about 12 times more sensitive than AGES but in an area almost 50 times smaller (or, if you prefer, 44 times more sensitive than ALFALFA but in an area 1,620 times smaller). This paper looks at their first of six 0.72 square degree fields, concentrating on the HI detections at redshifts at around 0.4, or a lookback time of about 4 Gyr. Presumably they have redshift coverage right down to z=0, but they don't say anything about that here.

They certainly knock off a few superlatives though. As well as being arguably the most distant direct detection of HI (excluding lensing) they also have, by a whisker, the most massive HI detection ever recorded – just shy of a hundred billion solar masses. For comparison, anything above ten billion is considered a real whopper.

All this comes at a cost. It took 95 hours of observations in this one tiny field and they only have six detections at this redshift. On the other hand, there's really just no other way to get this data at all (with the VLA it would take a few hundred hours per galaxy). Theoretically one could model how much HI would be expected in galaxies based on their optical properties and do much shorter, targeted observations which would be much more efficient. But this redshift is already high enough that optically the galaxies look pretty pathetic, not because they're especially dim but simply because they're so darn far away. So there just isn't all that much optical data to go on.

As you might expect, these six detections tend to be of extraordinarily gas-rich galaxies, with correspondingly high star formation rates. While they're consistent with scaling relations from local galaxies, their number density is higher than the local distribution of gas-rich galaxies would predict. That's probably they're most interesting finding, that we might be seeing the effects of gas evolution (albeit at a broad statistical level) over time. And it makes sense. We expect more distant galaxies to be more gas-rich, but exactly how much has hitherto been rather mysterious : other observations suggest that galaxies have been continuously accreting gas to replenish at least some of what they've consumed. For the first time we have some actual honest-to-goodness data* about how this works.

* Excluding previous results from stacking. These have found galaxies at even higher redshifts, but since they only give you the result in aggregate and not for individual galaxies, they're of limited use.

That said, it's probably worth being a bit cautious as to how well they can identify the optical counterparts of the HI detections. At this distance their beam size is huge, a ginormous 1.3 Mpc across ! That's about the same size as the Local Group and not much smaller than the Virgo Cluster. And they do say that in some cases there may well be multiple galaxies contributing to the detection. 

A particular problem is here is the phenomenon of surface brightness dimming. The surface brightness of a galaxies scales as (1+z)4. For low redshift surveys like AGES, z is at most 0.06, so galaxies appear only about 25% dimmer than they really are. But at z=0.4 this reaches a much more worrying factor of four. And the most HI-rich galaxy known (apart from those in this sample), Malin 1, is itself a notoriously low surface brightness object, so very possibly there's more galaxies contributing to the detections than they've identified here. It would be interesting to know if Malin 1 would be optically detectable at this distance...

On the other hand, one of their sources has the classic double-horn profile typical of ordinary individual galaxies. This is possible but not likely to arise by chance alignment of multiple objects : it would require quite a precise coincidence both in space and velocity. So at least some of their detections are very probably really of individual galaxies, though I think it's going to take a bit more work to figure out exactly which ones.

It's all quite preliminary so far, then. Even so, it's impressive stuff, and promises more to come in the hopefully near future.

Thursday, 1 August 2024

Going through a phase

Still dealing with the fallout from EAS 2024, this paper is one I looked up because someone referenced in in a talk. It caught my attention because the speaker mentioned how some galaxies have molecular gas but not atomic, and vice-versa. But we'll get to that.

I've no idea who the author is but the paper strongly reminds me of my first paper. There I was describing an HI survey of part of the Virgo cluster. This being my first work, I described everything in careful, meticulous detail, being sure to consider all exceptions to the general trends and include absolutely everything of any possible interest to anyone, insofar as I could manage it. Today's paper is an HI survey of part of the Fornax cluster and it is similarly careful and painstaking. If this paper isn't a student's first or early work, if it's actually a senior professor... I'm going to be rather embarrassed.

Anyway, Fornax is another nearby galaxy cluster, a smidgen further than Virgo at 20 compared to 17 Mpc. It's nowhere near as massive, probably a factor ten or so difference, but considerably more compact and dense. It also has less substructure (though not none) : the parlance being "more dynamically evolved" meaning that it's had more time to settle itself out, though it's not quite finished assembling itself yet. Its velocity dispersion of ~400 km/s is quite a bit smaller than than the >700 km/s of Virgo, but like Virgo, it too has hot X-ray gas in its intracluster space.

This makes it a natural target for comparison. It should be similar enough that the same basic processes are at work in both clusters : both should have galaxies experiencing significant tidal forces from each other, and galaxies in both clusters should be losing gas through ram pressure stripping. But the strengths of these effects should be quite different, so we should be able to see what difference this makes for galaxy evolution.

The short answer is : not all that much. Gas loss is gas loss, and just as I found in Virgo, the correlations are (by and large) the same regardless of how the gas is removed. I compared the colours of galaxies as a function of their gas fraction; here they use the more accurate parameter of true star formation rate, but the finding is the same. 

The major overall difference appears to be a survival bias. In Virgo there are lots of HI-detected and non-detected galaxies all intermingled. While there is a significant difference in their preferred locations, in Fornax this is much stronger : there are hardly any HI-detected galaxies inside the cluster proper at all. Most of the detections appear to be on the outskirts or even beyond the approximate radius of the cluster entirely. Exact numbers are a bit hard to come by though : the author's give a very thorough review of the state of affairs but don't summarise the final numbers. Which doesn't really matter, because the difference is clear.

What detections they do have, though, are quite similar to those in Virgo. They cover a range of deficiencies, meaning they've lost different amounts of gas. And that correlates with a similar change in star formation rate as seen in Virgo and elsewhere. They also tend to have HI extensions and asymmetrical spectra, showing signs that they're actively in the process of losing gas. Just like streams in Virgo, the total masses in the tails aren't very large, so they still follow the general scaling relations. 

So far, so standard. All well and good, nothing wrong with standard at all. They also quantify that the galaxies with the lowest masses tend to be the most deficient, which is not something I saw in Virgo, and is a bit counter-intuitive : if a galaxy is small, it should more easily lose so much gas as to become completely undetectable, so high deficiencies can only be detected in the most massive galaxies. But in Fornax, where the HI-detections may be more recent arrivals and ram pressure is weaker, this makes sense. They also quantify that the detections are likely infalling into the cluster for the first time in a now-standard phase diagram* which demonstrates this extremely neatly.

* Why they call them this I don't know. They plot velocity relative to the systemic velocity of the cluster as a function of clustercentric distance, and have nothing at all to do with "phases" in the chemical sense of the word.

The one thing I'd have liked them to try and this point would be stacking the undetected galaxies to increase the sensitivity. In Virgo this emphatically didn't work : it seems that there, galaxies which have lost so much gas as to be below the detection threshold have really lost all their gas entirely. But since Fornax dwarfs are still detectable even at higher deficiencies, then the situation might be different here. Maybe some of them are indeed just below the threshold for detectability, in which case stacking might well find that some still have gas.

Time to move on to the feature presentation : galaxies with different gas phases. Atomic neutral hydrogen, HI, is thought to be the main reservoir of fuel for star formation in a galaxy. The fuel tank analogy is a good one : the petrol in the tank isn't powering the engine itself. For that, you need to allow the gas to cool to form molecular gas, and it's this which is probably the main component for actual star formation.

There are plenty of subtleties to this. First, there's some evidence that HI is also involved directly in star formation : scaling relations which include both components have a smaller scatter and better correlation than ones which only use each phase separately. Second, galaxies also have a hot, low density component extending out to much greater distances. If the molecular gas is the fuel actually in the engine and the HI is what's in the tank, then this corona is what's in the petrol station, or, possibly, the oil still in the ground. And thirdly, cooling rates can be strongly non-linear : left to itself, HI gas will pretty much mind its own business and take absolutely yonks to cool into a molecular state.

Nevertheless this basic model works well enough. And what they find here is that while most galaxies have nice correlations between the two phases – more atomic gas, more molecular gas – some don't. Some have lots of molecular gas but no detectable HI. Some have lots of HI but no detectable molecular gas. What's going on ? Why are there neat relations most of the time but not always ?

Naively, I would think that CO without HI is the harder to explain. The prevailing wisdom is that gas starts off very hot indeed and slowly cools into warm HI (10,000 K or so) before eventually cooling to H2 (perhaps 1,000 K but this can vary considerably, and it can also be much colder). Missing this warm phase would be weird.

And if we were dealing with a pure gas system then the situation would indeed be quite bewildering. But these are galaxies, and galaxies in clusters no less. What's probably going on is something quite mundane : these systems are, suggest the authors, ones which have been in the cluster for a bit longer. There's been time for the ram pressure to strip the HI, which tends to be more extended and less tightly bound, leaving behind the H2 – which hasn't yet had the time to fully transform into stars. So all the usual gas physics is still in play, it's just there's been this extra complication of the environment to deal with.

What of the opposite case – galaxies with HI but no H2 ? How can you consume the H2 without similarly affecting the HI ? There things might be more interesting. They suggest several options, none of which are mutually exclusive. It could be that the HI is only recently acquired, perhaps from a tidal encounter or a merger. The former sounds more promising to me : gas will be preferentially ripped off from the outskirts of a galaxy where it's less tightly bound, and here there's little or no molecular gas. Such atomic gas captured by another galaxy may simply not have had time to cool into the molecular phase, whereas in a merger I would expect there to be some molecular gas throughout the process. 

Tidal encounters could have a couple of other roles, one direct, one indirect. The direct influence is that they might be so disruptive that they keep the gas density low, meaning its cooling rate and hence molecular content remains low (the physics of this would be complicated to explore quantitatively but it works well as a hand-waving explanation). The indirect effect is that gas at a galaxy's edge should be of lower metallicity : that is, purer and less polluted by the products of star formation. The thing is, we don't detect H2 directly but use CO as a tracer molecule. Which means that if the gas has arrived from the outskirts of a galaxy, it may be CO-dark. There could be some molecular gas present, it's just that we can't see it. Of course, to understand which if any of these mechanisms are responsible is a classic (and well justified) case of "more research is needed".

Tuesday, 23 July 2024

EAS 2024 : The Other Highlights

What's this ? A second post on the highlights of the EAS conference ? Yes ! This year I've been unusually diligent in actually watching the online talks I didn't get to see in person. Thankfully these are available for three months after the conference, long enough to actually manage to watch them but also, crucially, short enough to provide an incentive to bother. And I remembered a couple of interesting things from the plenaries that I didn't mention last time but which may be of interest to a wider audience.


Aliens ? There are hardly any talks which dare mention the A-word at astronomical conferences, but one of the plenaries on interstellar asteroids dared to go there. The famous interstellar visitor with the unpronounceable name of ʻOumuamua (which is nearly as bad as that Icelandic volcano that shut down European airspace a few years ago) got a lot of attention because Avi Loeb insists it must be an alien probe. He's wrong, and his claims to have found bits of it under the ocean have been utterly discredited. Still, our first-recorded visitor on a hyperbolic trajectory did do some interesting things. After accounting for the known gravitational forces, its rotation varies in a way that's inconsistent with gravity at the 10-sigma level. The speaker said that the only other asteroids and comets known to do this have experienced obvious collisions or have obvious signs of outgassing, neither of which happened here. He took the "alien" idea quite seriously.

Ho hum. No comment.


Time-travelling explosions. The prize lecture was by best-PhD student Lorenzo Gavassino, who figured out that our equations for hydrodynamics break down at relativistic velocities. Normally I would find this stuff incomprehensible but he really was a very good speaker indeed. And the main results is that they break down in a spectacular way. You might be familiar with simultaneity breaking, where events look different to observers at different speeds. Well, says Lorenzo, this happens to fluids moving at relativistic speeds in dramatic fashion : one observer should see a small amount of heat propagating at faster than the speed of light while another would see some energy travelling backwards through time. The result should be massive (actually, infinite) instabilities and the spontaneous formation of singularities. Accretion discs in simulations ought by rights to explode, and God knows what should happen to neutron stars.

The reason that this doesn't happen appears to be a numerical artifact which effectively smooths over this (admittedly small) amount of leakage. But what we need to do to make the equations rigorously correct, and how that would affect our understanding of these systems, isn't yet known.


Ultra Diffuse Galaxies may be tidal dwarfs in disguise. Another really interesting PhD talk was on how UDGs might form in clusters. When galaxies interact in low density environments, they can tear off enough gas and stars to form so-called tidal dwarfs. The key features of these mini-galaxies is that they don't have any dark matter (which is too diffuse to be captured in an interaction like this) and short-lived, usually re-merging with one of their parents in, let's say, 1 Gyr or so. But what if the interaction happens near the edge of a cluster ? Well, then the group can disperse and its members separated as they fall in, so the TDG won't merge with anything. Ram pressure will initially increase its star formation, increasing the stellar content in its centre and making it more compact, before eventually quenching it due to simple lack of gas. So there should be a detectable trend in these galaxies, from more compact to more diffuse going outwards from the cluster centre, all lacking in dark matter.

Of course this doesn't explain UDGs in isolated environments, but there's every reason to think that UDGs might be formed by multiple different mechanisms. A bigger concern was that the simulations didn't seem to include the other galaxies in the cluster, so the potentially very destructive effects of the tidal encounters weren't included. But survivorship bias was very much acknowledged : all galaxies, she said, get more compact closer to the centre, but not all survive at all. It's a really intriguing idea and definitely one to watch.


Even more about UDGs ! These were a really hot topic this year and whoever decided to schedule the session to be in one of the smaller rooms was very foolish, because it was overflowing. A few hardy souls stood at the back, but most gave up due to the poor air conditioning. Anyway, a couple of extra points. You might remember that I wasn't impressed by early claims than NGC 1052-DF4, one of the archetypes of galaxies without dark matter, had tidal tails. Well, I was wrong about that. New, deeper data clearly shows that it does have extended features beyond its main stellar disc. Whether that really indicates tidal disruption... well, I'll read the paper on that. And its neighbour DF2 remains stubbornly tail-less.

The other point is a new method for measuring distances to UDGs by looking at the stellar velocity dispersion of their globular clusters. This was the work of a PhD student who found that there's a relationship between this dispersion and the absolute brightness of the parent galaxy. Getting dispersion of the clusters is still challenging, requiring something like 20 hours on the VLT... but this is a far cry from the 100 HST orbits needed for the dispersion of the main stellar component of the galaxy itself. Apparently this work on the dispersion within individual clusters, so even one would be enough. They tested this on DF2 and DF4 and found a distance of....16 Mpc, right bang in the middle of the 13 and 20 Mpc claims that have been plagued with so much controversey.

Ho hum. No comment.


Fountains of youth and death. Some galaxies which today are red and dead appear to have halted their star formation very early on, but why ? One answer presented here quite decisively was due to AGN – i.e. material expelled from the enormous energies of a supermassive black hole in the centre of the galaxy. Rather unexpectedly it seems that most of this gas is neutral with only a small fraction being ionised, and detections of these neutral outflows are now common. In fact this may even be the main mechanism for quenching at so-called "cosmic noon" (redshifts of 1-2) when star formation peaked. Well, we'll see.

The other big talking point about fountains of ejected material was how galaxies replenish their gas. Here I learned two things I wish someone had told me years ago because they're very basic and I should probably have known them anyway. First, by comparing star formation rates with the mass of gas, one can estimate the gas depletion time, which is just a crude measure of how long the gas should last. And at low redshift this is suspiciously low, about a billion years. Does this mean we're in the final stages of star formation ? This is still about 10% or so of the lifetime of the Universe so it's never seemed all that suspicious to me.

The problem is that this depletion time has remained low at all redshifts. It's not that galaxies are suspiciously close to the end, it's that they should have already stopped forming stars and run out of gas long ago. Star formation can be estimated in different ways with no real constraint on distance, though gas content is a bit harder – we can't do neutral hydrogen in the distant Universe, but we can absolutely do molecular and ionised gas. Despite the many caveats of detail there's a very strong consensus that galaxies simply must be refuelling from somewhere.

One of those models has been the so-called galactic fountain. Galaxies expel gas due to stellar winds and supernovae, some of which escapes but most of which falls back to the disc. Now this is obvious as to how it explains why star formation keeps going in individual, local parts of the disc where the depletion time is too short, but how this explains the galaxy overall has never been clear to me. What might be going on is that the cold clouds of ejected gas (which look like writhing tendrils in the simulations) act as condensation sites as they move through the hot corona and fall back. Here gas in the hot, low density corona of the galaxy can cool, with the simulations saying that this mass of gas can be very significant. So the galaxy tops up its fuel tank from its own wider reservoir. It will of course eventually run out completely, but not anytime soon.

This is a compelling idea but there are two major difficulties, one theoretical and one observational. The theoretical problem is that the details of simulations really matter, especially resolution. If this is too low, clouds might appear to last much longer than they do in reality. One speaker presented simulations showing that this mechanism worked very well indeed while another showed that actually the clouds should tend to evaporate before they ever make it back to the disc, so this wouldn't be a viable mechanism at all. On the other hand, neither used a realistic corona : if it's actually not the smooth and homogenous structure they assume it to be, this could totally change the results.

The observational difficulty is that these cold gas clouds are just not seen anywhere. This is harder to explain but may depend on the very detailed atomic physics : maybe the clouds are actually warmer and more ionised than the predictions, or maybe colder and molecular. Certainly we know there can be molecular gas which is very hard to detect because it doesn't contain any of the tracer molecules we usually use; H2 is hard to detect directly so we usually use something like CO. 


And with that, I really end my summaries of EAS 2024, and return to regular science.

Thursday, 11 July 2024

ChatGPT Is Not A Source Extractor

When ChatGPT-4o came along I was pretty keen to try out its shiny new features, especially since some of the shine has rubbed off the chatbots of late. Oh, ignore the hype trains completely : those who are saying it's going to cause the apocalypse or usher in the Utopian end of history are equally deluded. I'm talking about actual use cases for LLMs. This situation remains pretty much as it has been since they were first unleashed. That is...
  • Decent enough if you want free-form discussions (especially if you need new ideas and don't care too much about factual accuracy)
  • Genuinely actually very useful indeed for coding (brilliant at doing boiler-plate work, a serious time-saver !)
  • Largely crap if you need facts, and even worse if you need those facts to be reliably accurate
Pretending that those first two are unimportant is in my view quite silly, legitimate concerns about energy expenditure notwithstanding. But that third one... nothing much seems to have shifted on that at all. They're still plagued with frequent hallucinations, since they're not grounded in anything so that they have no internal distinction between a verifiable, observable truth and the CPU-equivalent of a random brain-fart firing of the neurons.

Unfortunately GPT-4o just seems to extend this into a multi-modal world, giving results basically consistent with my earlier tests of chatbots. But I was intrigued by its apparent accuracy when supplying image files. It seemed to be, albeit from limited testing, noticeably more accurate when asked questions about image files than, say, PDFs. So I had a passing thought : could I use ChatGPT-4o to find sources in my data ?

Spoiler : no. It doesn't work.

It's not possible to share the chat itself because it contains images, but basically what I did was this. I uploaded an image of a typical data set I would customarily trawl look looking for galaxies. The very short version is that the HI detections of galaxies typically look like elongated blobs, sometimes appearing saturated and sometimes as mere enhancements in the noise. You can find a much more thorough explanation on my website, but that's the absolute basics. For example, in the image below, there are seven very obvious detections and one which is a bit fainter. 

I began by giving ChatGPT a detailed description of the image and the task at hand. This is the kind of thing that takes a few minutes to explain to a new observer; the actual training of data inspection can take a few days, but the explanations need be only very short indeed. And finding the bright sources is trivial : almost anyone can do that almost immediately. The bright galaxies are inherently obvious in the data when presented like this. Even if you have no idea what the axes labels refer to, it's clear that some parts of the image are very different to the others.

I asked ChatGPT to mark the location of the sources or otherwise describe their position. It didn't mark them but instead gave descriptions. Its world coordinates weren't precise enough to verify what it had identified, however, being limited to only values directly readable in the image and not doing any interpolation. I also gave it a broad alpha-numeric grid (A-J along the x-axis and 1-6 along the y-axis), but this was too coarse to properly confirm what it thought it had found. 

Its results were ambiguous at best. Even with this coarse grid it was clear some of its results were simply wrong. So I did what I'd do with new observers. I marked the sources with red outlines and numbers, uploaded the new image and described what I'd done, so it would have some kind of reference image. I also described the sources in more detail, e.g. which ones were bright and which were faint, and whether they extended into adjacent cells.

Next I gave it a new image with a finer grid (A-O and 1-14). This time, two sources (out of the ten or so visible) were reported correctly while the rest were wrong.  By mistake, I missed out the "D" cell in the coordinate labels, but ChatGPT reported a source at D4 ! Its revised claims were still wrong though, with once again getting only two correct.

This wasn't going well. I decided to dial it back and try something simpler. Maybe ChatGPT was able to "see" the features but not was accurately reading the coordinates, or perhaps hallucinating its answers and so mangling its results. So now I uploaded an image devoid of any coordinates and asked it for a simple count of the number of bright blobs. It got the answer right ! Okay, better... I asked it if it could mark the locations directly on the image, but it said it couldn't edit images. Instead it suggested giving the coordinates of the sources as a percentage of the axis length from the top left. Fair enough, but when comparing its reported coordinates it had again two near-misses and got all the rest simply wrong.

Finally I decided to check if at least the reported number count wasn't just a fluke. I uploaded three images in one file (thus circumventing OpenAI's painfully-limited restrictions on the free plan), each labelled with a number, and asked for the number of sources in each. It got one right and the rest wrong. It also gave descriptions of where it thought the sources were (i.e. upper left, middle, that sort of thing) and these were all wrong. Then, rather surprisingly and quite unprompted, it decided that it actually could edit images to mark the positions after all. The result came back :


Well... it's less than stellar. 

The upshot is that nothing much has changed about chatbot use cases at all. Good for discussions,  useless for facts. Whether it is "seeing" the images in some sense I don't know : possibly at some level it does recognise the sources but hallucinates both when trying to mark them and describe their positions, or possibly it's just making stuff up and nothing else. The latter seems rather unlikely though. Too often in other tests it was capable of giving results from figures in PDFs and image files which could not have been obtained from reading any of the text, that required actually "looking" at the images. 

Regardless of what it's actually doing, in terms of using ChatGPT as a source extractor, it's a non-starter. It doesn't matter why it gets things wrong, for practical application it only matters that it does. Maybe there's something capable under there, maybe there isn't. For now it's just an energy-intensive way of getting the wrong answers. Well, I could have done that anyway !

Monday, 8 July 2024

EAS 2024 : The Highlights

For the last major conference I went to, I combined the science and travel reports together. This was easy because Cardiff to me is not an exotic destination so it hardly needs much description. But this year's EAS conference was in Padova, and my explorations of the city itself and neighbouring Venice easily required their own dedicated post over on Physicists of the Caribbean. That, however, is a sideshow. Here I should say something about the main reason I was there, i.e. the SCIENCE !

This will only be brief. I have a number of things I want to look up in detail, but for now, here's what I leaned from the conference itself.


Euclid Is Mega Awesome. Like, seriously awesome. I have this chronic bad habit of not paying attention to new telescopes when they're being proposed or even during construction : who know when they'll launch or whether they'll reach their design spec ? Worst of all, and perhaps a better reason, is that their marketing is often terrible. It's all public outreach about their main target mission. For example as far as I knew Euclid was some specialist thing for probing dark energy presumably by measuring something very specific and niche, which just goes to show how little attention I ever give new instruments.

Actually it's more like Hubble on steroids. It's got comparable resolution (though not quite Hubble level) but with a vastly larger field of view and exquisite optics that gives a uniformly high quality image across the whole field. This makes it a fantastic, game-changing instrument for the low surface brightness universe. Want to find faint stellar streams, tiny dwarf galaxies and other exotic phenomena ? Euclid to the rescue ! If they'd sold it as more of a survey instrument and not this dark energy thing... well, quite possibly they did and I just wasn't listening. 

Whoops ! Luckily for me, its main survey will be almost all-sky and openly available, so I won't have to do anything except look at the data when it's available.


Cosmology Isn't Dead Yet. There are lots of press releases about the discovery of disc-like and massive galaxies in the early Universe, only a few hundred million years after the Big Bang, which is supposedly not long enough for them to have formed. The picture according to the experts is a bit more subtle than this popular description though. Some of these objects might be a problem - they really might, and not in a "but probably not" way : there's every chance there's something going on here that we don't understand. Whether that's the cosmology itself, i.e. the structure and nature of the universe, or just the detailed physics of the gas and star formation... that's where things get more suspect.

For example, it's not, it turns out, that simulations don't predict such objects. They do. It's just that it appears that they predict fewer than the numbers observed. In the Illustris simulations apparently about 10% of the relevant comparison sample are discy in the early simulated Universe, which isn't a lot but isn't insignificant either. The problem is that it's unclear if this is in conflict with the observations because the numbers are still small, and we aren't sure about the observational biases. For mass the situation is much worse : when deriving the mass using synthetic observational procedures (that is, transforming the simulations into the kind of data observers would process, and then using their methods to guestimate the stellar mass), results vary by up to three orders of magnitude away from the true value. So claims that there are too many massive galaxies in the early Universe can be safely ignored.

Except, there's a major caveat. Mass is a derived parameter with many uncertainties, but straightforward luminosity (i.e. brightness) is a direct observable. And here there does appear to be a conflict (sorry, tension) with observations.  There are potential solutions here as well though. For example a more "bursty" mode of star formation, rather than the more continuous process generally assumed, as well as accounting for nebular continuum emission and a top-heavy IMF (that is, forming more big stars in proportion to small ones than in the nearby universe, which could happen because their chemistry would be completely different), might be enough to solve all this. As some wise old sage commented, complicated problems often have complex, multi-parameter solutions rather than one big single "change this and all will be well" moment. Which unfortunately means that figuring this out is going to take time.

EDIT : I almost forgot. Back in 2017 there was a paper claiming that galaxies in the early Universe show declining rotation curves, indicating they were far less dark matter dominated than today. I was skeptical of this, as reported here with some follow-up here. And from conference results it seems that these results are heavily dependent on observation time : too short and indeed the result is declining curves, but observed for longer than they flatten considerable. Exactly why this should be I don't know, but it appears the original authors rowed back on their initial claims somewhat in a 2020 paper. It seems that the original results were something of an oversimplification, if not just simply wrong.


There Are No Dark Galaxies. New HI surveys are reaching lower column (or surface) densities, that is, how much gas they detect per unit area, than I was aware. In fact they're comparable to Arecibo but with the added benefit of much higher resolution. The penalty is that this takes enormous amounts of observing time but this is largely compensated for by large fields of view.

The results so far I think are still mainly "watch this space" except for individual objects. But one interesting finding is that there's a distinct lack of optically dark galaxies which have gas but no stars. This is something I've been working on for many years with Arecibo data, but MHONGOOSE's much improved resolution combined with its sensitivity means we would already have expected to see something if they exist in numbers of any significance. Thousands of detections of normal galaxies already (over 6,000 in fact - compare this to the 31,000 from ALFALFA which took many years to achieve !) but no hint of anything dark that isn't explicable by another mechanism... though of course, this has some caveats, but a significant population there appears none.


Gas Accretion Definitely Happens. But we still don't know when or how ! Obviously galaxies acquire gas at some point or they'd never form any stars, but while there's much indirect evidence for this, direct observations remain hugely unconvincing. Mergers have been ruled out as a significant source of the gas, there just aren't enough objects to do this. Cold accretion, where cold atomic HI falls into galaxies along streams, remains a distinct possibility but unobserved, even with the newest and most sensitive instruments.

Hot accretion (from the hot gas large-scale cosmic web) also remains a possibility but while large-scale bridges of X-ray gas have been detected, everyone was much more circumspect about claiming these as detections of the web itself than they have been in the past. And quite properly too, because any individual detection can always be challenged as an interaction. To claim a detection of the web, we'd really need to see it ubiquitously, with multiple strands connecting multiple clusters. Interestingly, HI intensity mapping has thus far got a statistical detection, but not to the point of being able to do actual imaging yet.


Radio Halos Do Funny Things. Not only does the X-ray gas trace giant, megaparsec-scale structures linking entire galaxy clusters, but so does the much lower-energy radio emission. There are different kinds of radio halos, with Mpc-scale giants to ~100 kpc scale "minihalos". And these aren't the same component, with the density profiles of the two showing a distinct change of slope. Then there are radio relics, the result apparently of shocks in the intracluster medium producing giant arcs of radio emission.

In one particularly noteworthy case, one of these giant radio lobes appears to be interacting with a galaxy. This shows a tail which is distinctly different from most. Where galaxies lose gas by ram pressure, the tails tend to be well-collimated and decrease in brightness at greater distances. This one instead gets both wider and brighter, indicating a different physical processes is at work. Like classical ram pressure, however, it seems to have caused an initial increase in the star formation rate of the galaxy. This is interesting to me because I've always assumed these features were of too low a density level to have an impact on anything, being an interesting way to trace the dynamics of an environment but being themselves no more than tracers, not interactors.


Ultra Diffuse Galaxies Are Still A Thing. There seems now to be a quite firm consensus : UDGs are two populations, one of "puffy" dwarves of low total mass that have become somehow extended, and the other of "failed", much more massive galaxies not predicted by any simulations. Several people made that last point and nobody raised any arguments, though exactly how massive they are (and how numerous this population is) wasn't clearly stated. Even so, to my mind if you want a serious challenge to cosmology, forget the attention-hogging results from JWST and look at these much nearer objects !

Likewise, the view seems to be that UDGs are indeed (following results from a few years ago) not actually all that large - but they are flatter : their light profiles are basically constant and then suddenly truncated at their edges, whereas normal galaxies have more complicated and varied profiles. There still seems to be some disagreement, however, as to whether UDGs therefore represent extreme examples of regular dwarf galaxies or are genuine outliers which are qualitatively different from the main galactic population. 

The dynamics of the more massive objects to me suggests the latter, but there's still not a clear answer to this. Their globular cluster populations are also highly diverse, with some having none at all and others having far more than expected. One very interesting set of observations by Pierre-Alain Duc shows stellar tails from globular clusters in the inner regions of the UDGs, likely merging with the main body of the galaxy - that's how good our observations have become ! Considering their whole set of properties, it remains understandably difficult to decide what the hell UDGs actually are.

To me the most interesting individual object presented in the whole conference was a UDG by Pavel Mancera PiƱa, he of the "UDGs have no dark matter" fame. Regular readers will know I was at first enthusiastic about this result, then became a bit more skeptical, but finally I've settled (?) back to my original stance : the results seem secure enough that they can't be attributed to observational errors or improper corrections. Now Pavel has found an object which is especially weird. Like the others, it's isolated. Its HI map shows very neatly circular contours, and its kinematics are consistent with no dark matter at all. The only way it can have a dark halo is if the concentration is very low - much lower than standard cold dark matter predicts, but explicable with self-interacting dark matter... 

And because it's isolated, explaining it with modified gravity is hard. If gravity rather than matter governs dynamics, then all isolated objects of the same mass and radius should show the same kinematic properties. That they don't might well argue that such notions are simply wrong, even as objects like this one indicate that the standard model of dark matter itself has flaws.


Well, of course much remains to be done in all those categories, especially the last. My own talk I'm pleased to say went down very well, I got lots of questions about the AGES dark clouds and had some nice discussions afterwards. My poster (should be accessible for a few months, I think) may have sunk without trace, but no matter. And the conference slogan "Where Astronomers Meet" may be the most blandest thing that ever did bland, but the sessions themselves were full of interesting stuff.

I'm old enough to remember conferences of a different era of more, shall we say, "robust conversations". I've not seen any actual arguments (except in some much smaller events) in many years now. Disagreements still arise but they're altogether gentler. Sometimes I miss the spectacle of hearing a good row from a safe distance, but perhaps, as long as we still do interesting science and still have fruitful discussions, this new way is better.

Thursday, 27 June 2024

The galaxies that quenched backwards

Today's paper is about contrarian galaxies that don't play by the rules. Most galaxies, left to their own devices, build up a big bulge of stars in the middle as the gas density is initially highest there. But this high star formation rate burns through its fuel very quickly and soon pitters out, "quenching" the star formation in the centre. Meanwhile the disc happily ambles along forming stars at a stately, steady pace, unhurried but longer-lasting.

There are some, though, which appear to do just the opposite. In clusters this is easy. Ram pressure preferentially removes the outer gas first, leaving the most stubborn gas remaining in the centre to carry on forming stars while the disc slowly reddens and dies. What's weirder is that there appear to be some galaxies like this in environments where ram pressure can't be playing any role. Ram pressure requires a hot intergalactic medium, which is pretty much only found at any significant levels in massive clusters. Everywhere else it should be far too weak to cause this kind of damage*.

* It probably doesn't have zero role to play outside clusters though. Very small satellite dwarfs can experience significant ram pressure from the hot gas of their parent galaxies, and there's some evidence that larger galaxies can experience at least a little gas-loss due to ram pressure in large-scale filaments. But probably not anywhere enough to account for galaxies like this. What they more likely experience is only starvation, where their outermost, thinnest gas is removed but nothing from within their denser discs. This means they can't replenish their gas as it gets eaten up by star formation. 

This paper is about a particular sort of these kind of galaxies, which they give the ugly name of "BreakBRDs" : Break Bulge Red Discs. The "break" refers to a particular spectral line indicating that there was star formation very recently in the bulges. "Blue Bulges Red Discs" might have been easier, but technically the bulges aren't actually blue, so this wouldn't be right. Even so, BBRDS would be a better acronym. Or heck, I'll just call 'em backwards galaxies.

I have to say I both like and dislike this paper. On the one hand it's very careful, thorough, and doesn't draw any overblown conclusions from the limited data. On the other, some of the discussion is long-winded, non-committal, and rather tedious considering that in the end the conclusion is so indecisive. I think there's some really great discussion on each individual scenario proposed to explain the backwards galaxies, but the collective whole becomes at times very confusing. It feels a bit like a paper written by committee. The main problem is nothing to do with the science at all, but the structure : there isn't a good unifying framework to tie all the different scenarios together. 

Here I shall try and simplify and disentangle things a bit. 

They begin with a nice overview of the observational difficulties of establishing what's going on. Some results support this classical picture of outside-in quenching where ram pressure (or other environmental effects) strip the outer gas first. But others find that this happens more in low density environments, exactly the opposite of what's expected ! Still others show something entirely different, where quenching happens irrespective of position in the galaxy, i.e. the whole disc quenches everywhere, all at once. Even simulations aren't much help, with galaxies similar to their BBRDs being found but with no clear mechanism responsible for what happened.

The particularly interesting thing about BBRDs, i.e. backwards quenchers, is that they exist over a wide range of stellar masses and apparently in all environments. Unfortunately they don't say anything else about the environments(s) of their particular sample, which is my only scientific quibble with the paper. Anyway what they do here is look at a sample of about a hundred or so which have HI gas measurements, using a combination of existing and their own new observations.

Their main result is actually quite simple, but you need to be aware of the colour-magnitude diagram first. Very simply, there are :

  • Galaxies which are blue, have lots of gas, and are forming stars as per usual (the so-called star formation main sequence, or more simply the blue cloud).
  • Galaxies which are red, don't have any gas, and aren't forming any stars : the red sequence.
  • Galaxies which are intermediate between red and blue, have a bit of gas, and are forming stars more slowly : the green valley (or transition region).
Of course, regular readers here will know that it isn't as simple as that, but this'll do for what's needed here.

They show that BBRDs and GV (green valley) galaxies have about the same gas fractions regardless of their stellar mass, both considerably lower than those in the blue cloud. But given their star formation rates, BBRDs would burn through their current gas content very much faster than GV galaxies. That is, for the same amount of gas, BBRDs are consuming their gas far more efficiently than GV galaxies. They're forming stars at rates typical of blue cloud galaxies, even though they've got way less gas. GV galaxies, by contrast, are forming stars even slowly than those on the main sequence.

How do they do this ? Or, conversely, what suppresses the star formation in GV galaxies ?

This is not at all easy to answer. For the GV galaxies, their gas depletion times seem to be low because of their extremely low star formation rates, which more than compensates for their lower gas contents. For the BBRDs, their rapid gas depletion times are easier to explain, as their star formation rate remains normal despite their low gas contents.

But what's behind it all ? There are three main options.


1) The final stages of quenching

These BBRD galaxies could be experiencing the end stages of gas removal due to some external mechanism, as is usually the case in clusters. Perturbations to the gas could drive inflows into the centres of the galaxies, resulting in an increase in gas density and thus an increase in star formation rate. Unfortunately they don't have the detailed maps of the gas needed to say if this is the case or not, but here it would have been useful for them to say something more about environment.


2) Accretion

It could be that these galaxies were in fact already fully quenched and are experiencing something of a renaissance. Accretion is a somewhat controversial topic because it's hard to ever determine that it's really happening with any certainty, but clearly galaxies have got to get their gas from somewhere. One possibility is so-called cold accretion from the cosmic web, where the hot, diffuse extragalactic background cools and condenses, falling into the potential wells of galaxies along streams.

Reading between the lines I think this is their favoured scenario. There are several reasons to think the gas isn't doing what it's supposed to be doing. For a few galaxies they have well-resolved kinematics of the gas and stars, and here it tends to be either misaligned with and/or highly distorted compared to the stars. They've only got such data for the innermost regions, but the spectral profiles of the gas (measured over much larger scales) also suggest it's significantly asymmetric. And they're also systematically offset form the Tully-Fisher relation : that is, the gas kinematics has broader line widths that typical galaxies of the same mass. 

All this is what you'd expect if the gas had only recently arrived rather than evolving along with the galaxy throughout its history. What they can't say is whether this indicates a temporary revival of the galaxy or a full resuscitation. They might, potentially, be rejuvenating themselves back onto the star formation main sequence, or they may have just accumulated a little bit of gas and will soon burn out once again. My guess is that latter is more likely but this would mean galaxies would have extraordinarily complicated histories.


3) Mergers

The most dramatic way a galaxy can accumulate more gas is by gobbling up other galaxies whole. This is their least favoured scenario as there are no signs of the remnants of the encounter you'd expect, e.g. long stellar tails. They can't rule it out though, as it's possible the star formation is persisting long after the tails have dispersed, but resolved gas measurements could help verify this idea.


I've simplified the discussion here considerably but this is roughly what it boils down to; they themselves, in my opinion, are too focused on the details of the different processes rather than painting a clear picture of the main differences. Fortunately, they've got VLA time to resolve the gas in a pilot sample, so this looks like a solvable problem. As they say, "it is important to remember the results of McKay et al. (in preparation)"... well, asking to remember results that don't exist yet is a new one on me, but I'll try.

Making Shit Up

Today's paper is one that fits into a very rare category where I'm prepared to say : this should not have been accepted by the refe...