Saved in:
Bibliographic Details
Main Authors: Dinneen, Jesse David, Julien, Charles-Antoine
Format: Preprint
Published: 2024
Subjects:
Online Access:https://arxiv.org/abs/2402.06421
Tags: Add Tag
No Tags, Be the first to tag this record!
_version_ 1866909100151930880
author Dinneen, Jesse David
Julien, Charles-Antoine
author_facet Dinneen, Jesse David
Julien, Charles-Antoine
contents Thoughtfully designing services and rigorously testing software to support personal information management (PIM) requires understanding the relevant collections, but relatively little is known about what people keep in their file collections, especially personal collections. Complementing recent work on the structure of 348 file collections, we examine those collections' contents, how much content is duplicated, and how collections used for personal matters differ from those used for study and work. Though all collections contain many images, some intuitively common file types are surprisingly scarce. Personal collections contain more audio than others, knowledge workers' collections contain more text documents but far fewer folders, and IT collections exhibit unusual traits. Collection duplication is correlated to collections' structural traits, but surprisingly, not to collection age. We discuss our findings in light of prior works and provide implications for various kinds of information research.
format Preprint
id arxiv_https___arxiv_org_abs_2402_06421
institution arXiv
publishDate 2024
record_format arxiv
spellingShingle What's in People's Digital File Collections?
Dinneen, Jesse David
Julien, Charles-Antoine
Human-Computer Interaction
Information Retrieval
Thoughtfully designing services and rigorously testing software to support personal information management (PIM) requires understanding the relevant collections, but relatively little is known about what people keep in their file collections, especially personal collections. Complementing recent work on the structure of 348 file collections, we examine those collections' contents, how much content is duplicated, and how collections used for personal matters differ from those used for study and work. Though all collections contain many images, some intuitively common file types are surprisingly scarce. Personal collections contain more audio than others, knowledge workers' collections contain more text documents but far fewer folders, and IT collections exhibit unusual traits. Collection duplication is correlated to collections' structural traits, but surprisingly, not to collection age. We discuss our findings in light of prior works and provide implications for various kinds of information research.
title What's in People's Digital File Collections?
topic Human-Computer Interaction
Information Retrieval
url https://arxiv.org/abs/2402.06421