Please or to access all these features

Feminism: Sex and gender discussions

They scraped Mumsnet again.

314 replies

ArabellaSaurus · 08/11/2025 17:26

archive.ph/e0u3Z

https://bulletin.appliedtransstudies.org/article/4/1-3/7/

Another data scrape. I'd say it's also defamatory against Mumsnet.

I've archived.

Article is a load of tedious wank, as you'd expect.

'In this study, however, we excavate what it means to write like a GC by analyzing how GC forum users rely on reactionary language and deploy storytelling practices in ways that calcify their anti-trans ideologies as personal and natural while rendering transgender people as anti-feminist, dangerous, and monstrous. To identify how GC groups perform political mythmaking and construct extremist identities, we undertook a computationally assisted discursive analysis of two popular GC forums: Ovarit and Mumsnet’s “Feminism: Sex & Gender” board (abbreviated to “FSG”). Through comparative platform discourse analysis, we analyzed over 80k posts and comments scraped from Ovarit and over 60k posts and comments scraped from Mumsnet (Burgess and Matamoros-Fernández 2016; Lewis and Marwick 2017)'

The only mildly amusing thing about it is the name of the Journal.

BATS.

“I Took a Deep Breath and Came Out as GC”: Gender Critical Storytelling, Radicalization, and Discursive Practice on Ovarit and Mumsnet

Following the closure of the anti-trans subreddit r/GenderCritical, gender critical (GC) internet users have migrated to more obscure, invite-only spaces. A side-effect of this GC dispersal is that activity in online anti-trans spaces has become increa...

https://bulletin.appliedtransstudies.org/article/4/1-3/7/

OP posts:
New posts on this thread. Refresh page
Thread gallery
11
HildegardP · 17/11/2025 22:35

CForCake · 17/11/2025 14:22

Apologies if it's already been asked, but how good are these IT tools at understanding who said what? E.g. can they distinguish:

a feminist reporting being insulted by trans activist (They called us X)

vs

a feminist insulting trans activists (I think that trans activists are X)

Absolutely dreadful.
I got suspended from X for "violent speech" - what was my appalling word crime? I used the phrase "shot yourself in the foot there" when the person to whom I was replying had trotted out a line of which they were mighty proud but that in fact undermined their case. The automated moderation is incapable of weighing context or doing anything beyond recognising the verb "to shoot".

AstonScrapingsNameChange · 17/11/2025 14:35

I'm not sure.

But if the people conducting the research have already decided that it's a given that sex realist women are hateful bigoted transphobes, there's not much chance of the research producing anything unbiased, regardless off the AIs capabilities

CForCake · 17/11/2025 14:22

Apologies if it's already been asked, but how good are these IT tools at understanding who said what? E.g. can they distinguish:

a feminist reporting being insulted by trans activist (They called us X)

vs

a feminist insulting trans activists (I think that trans activists are X)

DrBlackbird · 13/11/2025 17:19

@NorthernBogbean you include several arguments and seem to roll them all into it's important to be able to scrutinise web content independently/there shouldn’t be obstacles that gatekeep public content from public scrutiny.

There are few issues that I have with your argument. Who gets to decide what is ‘public interest’ and who is the public here and why is it in their interest to read somebody else’s interpretation of what’s been said by people online? Including cherrypicking isolated comments, possibly out of context. You are sweeping a huge amount of ‘online data’ into one basket and deeming it public interest. Losing context. Bundling up security threats posted online, with flu predictions, with mothers posting concerns about abusive partners, with real estate sales with anything and everything. There has to be a more nuanced approach in determining what is / is not public interest but ultimately that data still belongs to the users posting.

Here, for example, an American researcher has scraped data from a British owned platform that is contrary to their T&Cs. It costs money to run the platform, hire the mods, keep it running etc. The researchers have been free to take isolated comments and ignore whole themes that run across the platform threads. Likewise closer to home Aston scraped MN data without requesting consent for its sandbox and at least one of its researchers has ‘analysed’ (I use that word loosely) based on presumptions of posters motivation and explicitly ignoring the concerns raised. That is ethical how? I’m struggling to see how somebody’s definition of public interest overrides the interests of both us the users generating content and the platform owners legal rights.

Plus, this argument conflates two issues. Universities, media companies, journalists and even governments are increasingly subject to the power of global internet technology barons to withdraw their technologies. For these reasons, I think the principle of public scrutiny of the public internet is a prime concern.

Yes, big tech firms have power, they can withdraw services such as when Musk temporarily withdrew starlink from the Ukrainian govt. The ability to do so is worrying. But data scraping ordinary users content is not going to help illustrate that power. Conjecture is doing too much heavy lifting here. To analyse tech firms you’d need access to decision making on algorithmic development, which will never happen. Finally there are papers written on the ethics of using internet sources and they do not align with your view that anyone ought to be free to use online data for research and journal publication.

HildegardP · 12/11/2025 21:36

I'd call their peer review process mere intellectual frottage were it not so glaringly apparent that everyone involved in that dreck is 5 beers short of a sixpack.

Veilsofmorning · 12/11/2025 21:17

Agreed - difficult to understand how it passed any meaningful scrutiny or ethics review in any serious academic institution

BunfightBetty · 12/11/2025 18:50

FlirtsWithRhinos · 12/11/2025 18:24

I get what you are saying, but "we can't enforce it perfectly so there's no point in making it illegal" is not an argument that carries much weight round here.

Perhaps if women had not lost confidence in the ability or appetite of universities to apply ethical standards to trans-biased researchers and activists your theoretical position on how things should work would carry more weight.

As things are, we are not going to rely on models of academic engagement that fundamentally assume good intent.

I do not believe these researchers have any intention other than to back up their own biased edifices, and they should be treated not as genuine academics but as exactly the type of bad actors the MN T&Cs aim to target. I hope these nasty little men get their misogynist arses sued to hell and back.

Agree, I was absolutely gobsmacked that the ‘research’ question this scraping was used for could pass muster in any academic institution. The bias and question-begging was wild. A question that biased would never have been ok’d by any of the academic institutions I’ve studied with.

FlirtsWithRhinos · 12/11/2025 18:24

NorthernBogbean · 12/11/2025 17:54

I wasn't going to keep responding on this thread because I know the p.o.v. I'm advocating for isn't popular here, and I've never wanted to derail or disrupt threads.

I just want to say this and will duly shut up, IMV private or other ownership interests on the public internet shouldn't be able to gatekeep public content from public scrutiny. Think of all the sites on the public internet that should be accessible to scrutiny.

Creating T&Cs will maybe protect a public site from ripping for commercial use if the site has an appetite for lawsuits but they won't IMV protect from public interest journalism or academic research where approved protocols laid down for research are properly in place and complied with. I am familiar with those but I no longer work in HEIs.

Research protocols provide guidelines, risk assessments and hurdles for ethical approval. If researchers in universities don't conduct studies / handle data properly and as agreed, they will / should be internally sanctioned. If the university hasn't taken enough care to ensure compliance, it can be prosecuted. It depends on the case as to what the consequences might be. Because of the principles of public interest and academic freedom in democracies, the parameters of study aren't simply decided by institutions / publications / people who might be studied.

The new protocol element is the way the public internet, specifically social media, combines citizens and publication. So protocols for data gathering fall somewhere between studying publications / media and studying groups of people. If you set up a private group hosted on the internet, which no-one can see the content of, unless signed up with whatever data the organisers require, you have an expectation of privacy, just as you would if you met in person. But if the content of a site is publicly visible to consume, contributors are not private in the same way. Contributors are the monetised content, and their position isn't like private people communicating just with one another. The dilemma is how researchers can follow protocols of defaulting to anonymity for a general public, in this case site users. This is the debate.

My view is that if data is handled well, further anonymising of users is possible. The biggest risk for breaching anonymity will always be the content that users themselves publicly post. Sites can offer various aids to anonymity and erasure of content (but the content may still be permanent) and it's their responsibility not to allow users' private data they hold to be accessible. The arguments about the status of user-contributors can go on, but a site can't be both a public resource, a monetised public entertainment or a public influence, and also private.

It's not possible to acquire individual consent from user-content contributors, an unidentifiable quantity of people contributing over time, as it is from people involved in old-fashioned field studies. You could argue this means you can't study sites, but I disagree and I don't think I'm out of line with ethics panels on that. IMV site or platform-owners consent or lack of consent should not decide whether it can be scrutinised or studied.

The technologies which enable the public internet and social media aren't in the hands of its users, or even their governments, so it's important to be able to scrutinise web content independently. Public web content isn't accessible to study as easily as printed or other physical record just by the nature of the medium. It can take a lot longer and far more bodies to data gather from an ever-moving, multifaceted source of information. The biggest obstacle to getting the data therefore is time / funding. This is why researchers use ways of speeding up data collection like software. Using software to access private / hidden data is hacking and is illegal. Using software tools to retrieve and catalogue public content isn't but what you do with it is variously restricted. Academic researchers have access to commercial made-for-purpose software which requires user licenses. This removes some independent control of data from the researchers and is potentially problematic for that reason. Use of self-built open source software tools to do the same job shouldn't impede data gathering IMV, it's not ethically different from doing it by hand - the tools are not the problem so much as what is done with the data - but that doesn't mean university managements won't self-censor as they are increasingly nervous of being subject to litigation, being increasingly commercial themselves and fearful of funding cuts.

Most barriers to research come down to money / resources as well as confidence in being supported by governments and the law - the allocation of funding and validation of institutions, or threat of its withdrawal, can have a freezing effect on research. Universities, media companies, journalists and even governments are increasingly subject to the power of global internet technology barons to withdraw their technologies. For these reasons, I think the principle of public scrutiny of the public internet is a prime concern.

Aside from bad or illegal use of data, the issue with sites being studied isn't the production of papers like the one that generated this thread but the lack of motivation for people researching in unis right now to produce better, challenging studies from a wide variety of angles. Maybe working academics on MN can do that, hopefully unhampered by 'T&Cs'.

I get what you are saying, but "we can't enforce it perfectly so there's no point in making it illegal" is not an argument that carries much weight round here.

Perhaps if women had not lost confidence in the ability or appetite of universities to apply ethical standards to trans-biased researchers and activists your theoretical position on how things should work would carry more weight.

As things are, we are not going to rely on models of academic engagement that fundamentally assume good intent.

I do not believe these researchers have any intention other than to back up their own biased edifices, and they should be treated not as genuine academics but as exactly the type of bad actors the MN T&Cs aim to target. I hope these nasty little men get their misogynist arses sued to hell and back.

NorthernBogbean · 12/11/2025 17:54

DrBlackbird · 10/11/2025 09:47

Public interest case? What is that and where are the published guidelines on it? Without offering evidence, this term sounds like something out of newspaper speak.

However, academic research is held to higher standards than journalism.

You may want to argue that data "should" be scraped without the platform’s consent, but all research ethics guidelines - for reputable researchers and universities - informed consent is one key aspect in using data, even internet data. Even the association of internet researchers highlights the ethical imperative in obtaining consent when a direct quote is used.

If you are an academic researcher, then I very much hope that you revisit your understanding of research ethics, your university’s ethical guidelines, and the UK Research Integrity Office. ‘Want to’ does not override legal and ethical research requirements involving real people.

https://ukrio.org/wp-content/uploads/UKRIO-Code-of-Practice-for-Research.pdf#page17

https://aoir.org/reports/ethics3.pdf

Edited to add: it would not be up to site owners to say whether it could be studied or not. AFAIK there's no legal or necessarily ethical impediment to that

Your opinion is not in alignment with most reputable guidelines on ethical and legal use of internet data. There are some exemptions but you are making too sweeping of an opinion here. Informed consent of those whose data is being used - including whole direct quotes - is still a key ethical issue. I’m shocked that you are arguing otherwise. There are good reasons for being aware of and sensitive to potentially dangerous outcomes in using such data. Hence the need for guidelines.

Edited

I wasn't going to keep responding on this thread because I know the p.o.v. I'm advocating for isn't popular here, and I've never wanted to derail or disrupt threads.

I just want to say this and will duly shut up, IMV private or other ownership interests on the public internet shouldn't be able to gatekeep public content from public scrutiny. Think of all the sites on the public internet that should be accessible to scrutiny.

Creating T&Cs will maybe protect a public site from ripping for commercial use if the site has an appetite for lawsuits but they won't IMV protect from public interest journalism or academic research where approved protocols laid down for research are properly in place and complied with. I am familiar with those but I no longer work in HEIs.

Research protocols provide guidelines, risk assessments and hurdles for ethical approval. If researchers in universities don't conduct studies / handle data properly and as agreed, they will / should be internally sanctioned. If the university hasn't taken enough care to ensure compliance, it can be prosecuted. It depends on the case as to what the consequences might be. Because of the principles of public interest and academic freedom in democracies, the parameters of study aren't simply decided by institutions / publications / people who might be studied.

The new protocol element is the way the public internet, specifically social media, combines citizens and publication. So protocols for data gathering fall somewhere between studying publications / media and studying groups of people. If you set up a private group hosted on the internet, which no-one can see the content of, unless signed up with whatever data the organisers require, you have an expectation of privacy, just as you would if you met in person. But if the content of a site is publicly visible to consume, contributors are not private in the same way. Contributors are the monetised content, and their position isn't like private people communicating just with one another. The dilemma is how researchers can follow protocols of defaulting to anonymity for a general public, in this case site users. This is the debate.

My view is that if data is handled well, further anonymising of users is possible. The biggest risk for breaching anonymity will always be the content that users themselves publicly post. Sites can offer various aids to anonymity and erasure of content (but the content may still be permanent) and it's their responsibility not to allow users' private data they hold to be accessible. The arguments about the status of user-contributors can go on, but a site can't be both a public resource, a monetised public entertainment or a public influence, and also private.

It's not possible to acquire individual consent from user-content contributors, an unidentifiable quantity of people contributing over time, as it is from people involved in old-fashioned field studies. You could argue this means you can't study sites, but I disagree and I don't think I'm out of line with ethics panels on that. IMV site or platform-owners consent or lack of consent should not decide whether it can be scrutinised or studied.

The technologies which enable the public internet and social media aren't in the hands of its users, or even their governments, so it's important to be able to scrutinise web content independently. Public web content isn't accessible to study as easily as printed or other physical record just by the nature of the medium. It can take a lot longer and far more bodies to data gather from an ever-moving, multifaceted source of information. The biggest obstacle to getting the data therefore is time / funding. This is why researchers use ways of speeding up data collection like software. Using software to access private / hidden data is hacking and is illegal. Using software tools to retrieve and catalogue public content isn't but what you do with it is variously restricted. Academic researchers have access to commercial made-for-purpose software which requires user licenses. This removes some independent control of data from the researchers and is potentially problematic for that reason. Use of self-built open source software tools to do the same job shouldn't impede data gathering IMV, it's not ethically different from doing it by hand - the tools are not the problem so much as what is done with the data - but that doesn't mean university managements won't self-censor as they are increasingly nervous of being subject to litigation, being increasingly commercial themselves and fearful of funding cuts.

Most barriers to research come down to money / resources as well as confidence in being supported by governments and the law - the allocation of funding and validation of institutions, or threat of its withdrawal, can have a freezing effect on research. Universities, media companies, journalists and even governments are increasingly subject to the power of global internet technology barons to withdraw their technologies. For these reasons, I think the principle of public scrutiny of the public internet is a prime concern.

Aside from bad or illegal use of data, the issue with sites being studied isn't the production of papers like the one that generated this thread but the lack of motivation for people researching in unis right now to produce better, challenging studies from a wide variety of angles. Maybe working academics on MN can do that, hopefully unhampered by 'T&Cs'.

DustyWindowsills · 11/11/2025 23:36

SabrinaThwaite · 11/11/2025 23:17

Long tailed, blue, coal or just great?

I recently saw a T-shirt with a variety of said birds and the slogan ‘nice tits’.

ETA or were you going for the WC Fields / Mae West classic ‘My Little Chickadee’?

Edited

Sadly, I didn't specify. Perhaps that was my big error.

I'm rather fond of Sarah Edmonds' range of bird-themed mugs and tea towels, featuring tits, cocks, shags and floaters. Not that I'm childish or anything!

SabrinaThwaite · 11/11/2025 23:17

DustyWindowsills · 11/11/2025 09:46

Is there a rationale for deletion, e.g. forbidden words picked up by an algorithm, or is it sufficient for somebody to report a post that offends them? I had a post deleted in which I referred to our most charming visitor as a small passerine bird of the Paridae family. I also called him "he".

Long tailed, blue, coal or just great?

I recently saw a T-shirt with a variety of said birds and the slogan ‘nice tits’.

ETA or were you going for the WC Fields / Mae West classic ‘My Little Chickadee’?

DustyWindowsills · 11/11/2025 10:53

DeanElderberry · 11/11/2025 10:20

However much you try to cloak your offence in Linnaean Latin, calling other posters chickadees (even small ones) is Just Not On.

I am a bad person. 😞

DeanElderberry · 11/11/2025 10:21

Because the chickadees don't like it.

DeanElderberry · 11/11/2025 10:20

DustyWindowsills · 11/11/2025 09:46

Is there a rationale for deletion, e.g. forbidden words picked up by an algorithm, or is it sufficient for somebody to report a post that offends them? I had a post deleted in which I referred to our most charming visitor as a small passerine bird of the Paridae family. I also called him "he".

However much you try to cloak your offence in Linnaean Latin, calling other posters chickadees (even small ones) is Just Not On.

DustyWindowsills · 11/11/2025 09:46

DeanElderberry · 11/11/2025 09:19

Did you use the variant on eff orf that gets described as a 'death threat' when they run to the mods? When the BBC wanted to make their Archers boards unusable they hired a company who used a range of disruptive strategies, and that accusation was part of the package. I spotted it somewhere on MN recently and got all nostalgic but knew it was doomed.

Is there a rationale for deletion, e.g. forbidden words picked up by an algorithm, or is it sufficient for somebody to report a post that offends them? I had a post deleted in which I referred to our most charming visitor as a small passerine bird of the Paridae family. I also called him "he".

DeanElderberry · 11/11/2025 09:19

SinnerBoy · 10/11/2025 21:44

Oh dear, I was a bit mean to a scraper and had a post deleted. I hope MN have a good legal team on this case and make them delete everything. It's not as if they actually need the data, their conclusions were reached beforehand.

Did you use the variant on eff orf that gets described as a 'death threat' when they run to the mods? When the BBC wanted to make their Archers boards unusable they hired a company who used a range of disruptive strategies, and that accusation was part of the package. I spotted it somewhere on MN recently and got all nostalgic but knew it was doomed.

haXXor · 10/11/2025 23:17

IwantToRetire · 10/11/2025 22:03

Its the usual criticism that someone having a belief that sex work is exploitation therefore hates women who are (forced) in prostitution. The same as saying if you dont belief if being able to identify into a gender (sex) then you hate people who do.

Without cross polluting this thread with another one, it is as daft as saying because I am a vegitarian I hate meat eaters.

How is this research.

All they saying is that some people (especially those pesky women) are allowed to voice an opinion, and it is hateful because we the researchers (who set the bench mark) say it is impossible to have an opinion and not turn that into a hate throught.

Which probably says more about how they live their lifes.

I wonder where the funding money comes from.

What a waste.

Worse than that, it's akin to accusing someone vegetarian or vegan of hating farm animals.

Hating meat eaters is the equivalent of hating punters and pimps.

haXXor · 10/11/2025 23:13

ArabellaSaurus · 10/11/2025 19:56

'... intersex variations and fertility challenges represented as “disorders” of “normal” functioning rather than legitimate identities.'

This seems grossly offensive.

Bearing in mind that the salt-wasting form of Congenital Adrenal Hyperplasia literally kills the children who are born with it, and most DSDs have some negative impact on the people who have them, I'd say that "disorder of normal functioning" is an entirely legitimate term to use. I don't consider my migraines an "identity".

I didn't have children by choice, yet even I can see that declaring infertility to be an "identity" is hurtful and completely lacking empathy towards women who cannot concieve or gestate the children that they desperately want.

IwantToRetire · 10/11/2025 22:03

DuesToTheDirt · 10/11/2025 19:28

@IwantToRetire “biological sex,” conceptualized as reproductive anatomy and karyotype, inherently determines the “correct” sex and gender a person should possess.

They seem to have completely ignored the many GC people who don't have a gender at all. I also keep thinking about this claim:

GC users’ rhetoric consistently disparaged, targeted, and demonstrated enmity towards sex workers, who were referred to as “prostitutes” (nOvarit = 509; nFSG = 371) and accused of proliferating the “porn addicted” and “pornsick” (nOvarit = 1,350; nFSG = 482) conditions under which gender ideology allegedly proliferates. Denigration of sex workers is an established component of GC ideology."

I'm not familiar with Ovarit, and I don't understand how they came to this conclusion, but on this board, and in GC conversations elsewhere I've never come across such a thing, never mind as "an established component of GC ideology." It would come down to blaming women for men's failings, which is what men do, and is certainly not practised by the women of FWR.

Perhaps I didn't read enough of the study - since what I did read was garbage - but I can't actually see how they came to any of their conclusions.

Its the usual criticism that someone having a belief that sex work is exploitation therefore hates women who are (forced) in prostitution. The same as saying if you dont belief if being able to identify into a gender (sex) then you hate people who do.

Without cross polluting this thread with another one, it is as daft as saying because I am a vegitarian I hate meat eaters.

How is this research.

All they saying is that some people (especially those pesky women) are allowed to voice an opinion, and it is hateful because we the researchers (who set the bench mark) say it is impossible to have an opinion and not turn that into a hate throught.

Which probably says more about how they live their lifes.

I wonder where the funding money comes from.

What a waste.

SabrinaThwaite · 10/11/2025 21:54

PermanentTemporary · 10/11/2025 10:16

Well, Ovarit has closed. And MN FWR is not what it was since the Supreme Court ruling. I’m not really interested in ‘telling people to fuck off’ in principle, or endlessly rehashing recent fights.

I’m going to disagree - I’m all in favour of telling these faux researchers who try to force team women who stand up for women’s rights with neo Nazis to FRO.

Still, looking at the editorial board of BATS and the team behind CATS, I’m thinking this ‘paper’ is going to have a fairly small circulation, and will be mostly read by people who define themselves by their pronouns.

SinnerBoy · 10/11/2025 21:44

Oh dear, I was a bit mean to a scraper and had a post deleted. I hope MN have a good legal team on this case and make them delete everything. It's not as if they actually need the data, their conclusions were reached beforehand.

Magpiecomplex · 10/11/2025 21:35

I may be being unduly nice, but my reading of "intersex variations and fertility challenges" at least suggests they haven't conflated PCOS with DSDs. I have PCOS, I'm definitely a woman.

moto748e · 10/11/2025 21:24

ArabellaSaurus · 10/11/2025 19:56

'... intersex variations and fertility challenges represented as “disorders” of “normal” functioning rather than legitimate identities.'

This seems grossly offensive.

As if having a DSD in the first place is not bad enough, imagine if you kept coming across shit like this. It must be infuriating.

BunfightBetty · 10/11/2025 20:59

spannasaurus · 10/11/2025 20:15

. intersex variations and fertility challenges represented as “disorders” of “normal” functioning rather than legitimate identities.'

When did infertility become an identity?

If only I’d known infertility was something I could just identify my way out of. I could have saved thousands on IVF.

ArabellaSaurus · 10/11/2025 20:57

Exactly.

OP posts: