The Seven Days
In "Seven Days", we analyzed 1,575 Arabic-language headlines and their body text, from five platforms (arabic.rt.com, alarabiya.net, skynewsarabia.com, bbc.com/arabic, aljazeera.net), covering the Iran war's coverage between February 25 and March 6, 2026. The project was produced in collaboration with Arabi Facts Hub (AFH) and distributed via Muwatin.
Stories
Project Main Information
-
The articles were collected through direct scraping, using open-source Python libraries, of publicly available material from the five platforms (arabic.rt.com, alarabiya.net, skynewsarabia.com, bbc.com/arabic, aljazeera.net).
-
The analysis relied on Association Rule Mining (ARM), implemented through the FP-Growth algorithm, applied separately to headlines and body text (min_support=0.005, min_confidence=0.05). Only rules with a lift value greater than 1 were retained. Platform-level weighting was applied to counteract source imbalance across the five outlets.
-
Yes, the dataset is accessible to all journalists, researchers, and activists interested in the topic, however, we prefer that personnel who is interested in accessing the dataset, to take a workshop with us, to learn what they can about ARM analysis and FP-Growth, before start mining in our datasets. It's not mandatory, but we strongly encourage it. If you are interested, please reach out via email: info@anmat.media or using this form: https://anmat-media.github.io/anmat-workshop/
-
The datasets is available on Kaggle, you will be able to access it from here, or from Anmat account on Kaggle

