MIT Libraries logoDSpace@MIT

MIT
View Item 
  • DSpace@MIT Home
  • MIT Open Access Articles
  • MIT Open Access Articles
  • View Item
  • DSpace@MIT Home
  • MIT Open Access Articles
  • MIT Open Access Articles
  • View Item
JavaScript is disabled for your browser. Some features of this site may not work without it.

Media Cloud: Massive Open Source Collection of Global News on the Open Web

Author(s)
Roberts, Hal; Bhargava, Rahul; Valiukas, Linas; Jen, Dennis; Malik, Momin M; Bishop, Cindy Sherman; Ndulue, Emily B; Dave, Aashka; Clark, Justin; Etling, Bruce; Faris, Robert; Shah, Anushka; Rubinovitz, Jasmin; Hope, Alexis; D'Ignazio, Catherine; Bermejo, Fernando; Benkler, Yochai; Zuckerman, Ethan; ... Show more Show less
Thumbnail
DownloadPublished version (1.274Mb)
Publisher Policy

Publisher Policy

Article is made available in accordance with the publisher's policy and may be subject to US copyright law. Please refer to the publisher's site for terms of use.

Terms of use
Article is made available in accordance with the publisher's policy and may be subject to US copyright law. Please refer to the publisher's site for terms of use.
Metadata
Show full item record
Abstract
We present the first full description of Media Cloud, an open source platform based on crawling hyperlink structure in operation for over 10 years, that for many uses will be the best way to collect data for studying the media ecosystem on the open web. We document the key choices behind what data Media Cloud collects and stores, how it processes and organizes these data, and its open API access as well as user-facing tools. We also highlight the strengths and limitations of the Media Cloud collection strategy compared to relevant alternatives. We give an overview two sample datasets generated using Media Cloud and discuss how researchers can use the platform to create their own datasets.
Description
2021 Fifteenth International AAAI Conference on Web and Social Media
Date issued
2021
URI
https://hdl.handle.net/1721.1/155956
Department
Massachusetts Institute of Technology. Media Laboratory; Massachusetts Institute of Technology. Department of Urban Studies and Planning
Journal
Proceedings of the International AAAI Conference on Web and Social Media
Publisher
Association for the Advancement of Artificial Intelligence
Citation
Roberts, H., Bhargava, R., Valiukas, L., Jen, D., Malik, M. M., Bishop, C. S., Ndulue, E. B., Dave, A., Clark, J., Etling, B., Faris, R., Shah, A., Rubinovitz, J., Hope, A., D’Ignazio, C., Bermejo, F., Benkler, Y., & Zuckerman, E. (2021). Media Cloud: Massive Open Source Collection of Global News on the Open Web. Proceedings of the International AAAI Conference on Web and Social Media, 15(1), 1034-1045.
Version: Final published version

Collections
  • MIT Open Access Articles

Browse

All of DSpaceCommunities & CollectionsBy Issue DateAuthorsTitlesSubjectsThis CollectionBy Issue DateAuthorsTitlesSubjects

My Account

Login

Statistics

OA StatisticsStatistics by CountryStatistics by Department
MIT Libraries
PrivacyPermissionsAccessibilityContact us
MIT
Content created by the MIT Libraries, CC BY-NC unless otherwise noted. Notify us about copyright concerns.