Really curious about the tool they used to quantify "toxicity/disruptive" comments. My initial suspicion would be that political commentary, regardless of human-perceived toxicity, might be biased toward "toxic" by an automated sentiment analysis.
In short: I am suspicious that automated tooling exists to reliably distinguish between toxic and non-toxic political discourse.