Abstract
Background: In the context of rising digital inequality, concerns about how privacy impacts vulnerable populations have become increasingly relevant. However, privacy research continues to focus largely on technical, legal or policy issues, often failing to consider the perspectives of marginalised populations.
Objectives: This study explores how vulnerable populations have been represented in digital privacy research over the past two decades. It seeks to identify thematic trends, publication patterns and key contributors across scholarly outputs.
Method: A bibliometric analysis was conducted using data extracted from Scopus and Web of Science, yielding a combined dataset of 4760 articles published between 2005 and 2025. Bibliometrix (R package) was used to analyse publication trends, author productivity, geographical distribution of research, author and indexed keywords and the most relevant sources.
Results: Findings reveal that children, health-related populations and consumers dominate the discourse, while groups such as refugees, persons with disabilities, the elderly and racial or ethnic minorities are notably underrepresented.
Conclusion: Digital privacy research has evolved without sufficiently integrating the perspectives and needs of marginalised communities. These risks reinforce existing inequalities in the digital ecosystem.
Contribution: This study provides a bibliometric review to critically assess the visibility of vulnerable populations in digital privacy research. Notably, South Africa was the only African country to appear in the list of the top 20 most productive countries in this domain, ranking 18th overall. This limited representation of the African continent underscores the urgent need for research focusing on the privacy concerns of vulnerable groups in Africa.
Keywords: privacy; information privacy; digital privacy; vulnerable populations; bibliometric analysis; underrepresented groups; digital inequality; academic visibility.
Introduction
In today’s digitally interconnected world, privacy has become a paramount concern that transcends technical boundaries, touching upon profound social and ethical dimensions. Vulnerable populations, including children, the elderly, refugees, persons with disabilities and ethnic, gender or religious minorities, face disproportionate privacy risks because of limited digital agency (McDonald & Forte 2022) and unequal access to secure technologies (Scanlan 2022:725–735). In this study, we adopt the definition of vulnerable individuals articulated by McDonald and Forte (2022), as it aligns with our objective to inclusively capture a broad range of marginalised populations. According to their definition, vulnerable individuals are those who, because of their race, class, gender or sexual identity, religion or other intersectional characteristics or circumstances, are more susceptible to privacy violations that may result in emotional, financial or physical harm or neglect (McDonald & Forte 2022). Since this is a bibliometric analysis that aims to map the breadth of existing literature, it is essential to adopt a definition that is inclusive and capable of capturing as many vulnerable groups as possible. This definition offers a flexible but focused lens to assess the scope, gaps and representation necessary for a bibliometric analysis. The focus on vulnerable populations is critical because the consequences of privacy violations are not equally distributed across society. For marginalised groups, breaches of privacy can result in heightened exposure to surveillance, discrimination, exploitation and social exclusion. For example, inadequate privacy protections may expose refugees to security risks, enable profiling of minority communities, or compromise the dignity and autonomy of individuals with disabilities. Failure to adequately address these risks in research and practice may reinforce existing social inequalities and contribute to the marginalisation of already disadvantaged groups within digital environments. Digital privacy is distinguished from broader notions of information and data privacy by its specific focus on the protection of personal information within digitally networked environments (Broeckelmann 2026). While information privacy traditionally concerns the control of personal data and data privacy often relates to regulatory and technical safeguards, digital privacy encompasses the complexities introduced by online platforms, including continuous data generation, algorithmic profiling, large-scale data aggregation and persistent digital footprints. These characteristics make digital privacy particularly relevant in contemporary research, as individuals increasingly engage with digital systems that shape their visibility, identity and exposure to risk (Ahmadon et al. 2025:69–77). While scholarly interest in digital privacy has grown, privacy itself is not a universally distributed good. Rather, it is shaped by social hierarchies, access to resources and institutional power, often privileging those with greater socio-economic and political capital (Taylor 2017). Although existing research has explored specific concerns, such as children’s online safety (Livingstone & Helsper 2007:671–696; Livingstone & Smith 2014:635–654), healthcare data protection (Botes 2025; Shin et al. 2024) and the surveillance of displaced populations (Steinbrink et al. 2023:245), the broader landscape of privacy scholarship concerning vulnerable groups remains understudied and fragmented across disciplines and time (Sannon & Forte 2022). Although bibliometric analyses have been widely used to map the evolution of research domains such as privacy concerns (Benedetto & Cucchi 2022), cybersecurity (Coman et al. 2025:37–69) and information systems (Radu & Popescul 2024:13–29), these studies have primarily focused on publication trends, collaboration networks and thematic developments without explicitly examining how vulnerable populations are represented within the literature. As a result, there is limited understanding of whose privacy concerns are prioritised in scholarly discourse, highlighting a critical gap that this study seeks to address. Bibliometric analysis offers a powerful approach to this challenge, enabling systematic mapping of academic publications to identify prevailing themes, collaboration networks and keyword co-occurrences (Passas 2024:1014–1025). Such analysis can reveal not only the dominant topics in digital privacy research, but also whose privacy concerns are made visible or rendered invisible. This knowledge perspective is vital for understanding the ways in which academic research both shapes and responds to the social inequalities faced by marginalised populations. While various vulnerable groups have received scholarly attention in privacy research, others remain underrepresented (McDonald & Forte 2022; Sannon & Forte 2022), despite their increasing engagement with digital platforms. Many communities now participate in activities such as remote communication, online financial interactions and digitally mediated support services, which are contexts that raise unique and often overlooked privacy challenges. Yet, the distinct socio-cultural and ethical dimensions of privacy affecting some of these groups are rarely addressed in mainstream privacy discourse (Wang & Metzger 2024), leaving important blind spots in the literature. This paper seeks to address such gaps through a 20-year bibliometric analysis of 4760 English-language journal articles and conference papers sourced from Scopus and Web of Science, covering the period from 2005 to 2025. Using the Bibliometrix package in R, the study examines publication trends, author productivity, geographical distribution of research, author and indexed keywords and the most relevant sources. Particular attention is paid to assessing how various vulnerable populations are represented or underrepresented within the evolving landscape of privacy scholarship. By quantifying scholarly engagement with different vulnerable groups, this study exposes patterns of attention and neglect and raises critical questions about epistemic justice in digital privacy research. Ultimately, it advocates for a more inclusive, context-sensitive research agenda that reflects the diverse realities and needs of marginalised populations in an increasingly digital world.
To guide this investigation, the study is structured around the following research questions:
- Which vulnerable populations have received the most academic attention in privacy-related research?
- Which vulnerable populations are underrepresented in privacy-related research?
- How has research on digital privacy and vulnerability evolved between 2005 and 2025?
- What are the most influential journals, authors and themes shaping this domain?
Literature review
The rapid advancement and widespread adoption of digital technologies ranging from social media platforms and mobile applications (Sekalala et al. 2020:9) to biometric surveillance (Drozdowski et al. 2020:89) have introduced multifaceted challenges to individual and collective privacy. Issues surrounding data protection, informed consent and ethical data use have become increasingly urgent, particularly for vulnerable populations (Kreutzer et al. 2025:3). These groups face disproportionate privacy risks. Existing research has highlighted children (Jang & Ko 2023:1; Srivastava, Wilska & Nyrhinen 2023:235; Sun et al. 2023:1), the elderly (Tomczyk et al. 2023:5), refugees (Georgiou, Baillie & Shah 2023:4; Muhawe 2025:370) and persons with disabilities (Buthelezi et al. 2024:1), as especially susceptible within the digital privacy landscape. Vulnerability in the context of digital privacy extends beyond demographic categorisation to reflect a condition shaped by structural, technological and socio-cultural factors. It encompasses not only who is at risk, but also how and why certain individuals or groups experience heightened exposure to privacy harms within digital environments (Strathmann 2025). Factors such as limited digital literacy, restricted access to secure technologies, socio-economic inequality and marginalised social positioning can reduce individuals’ ability to control personal data, provide informed consent or resist surveillance practices. In this sense, vulnerability is relational and context-dependent, emerging from the interaction between individuals and the digital systems they engage with (Silvennoinen & Rantanen 2023:1–11). Understanding vulnerability in this way is essential for analysing how privacy risks are distributed unevenly and why certain groups remain disproportionately affected within the broader digital ecosystem. Despite this growing attention, the literature remains fragmented and unevenly distributed. Much scholarship is embedded within legal (Rhoen 2016:2; Tzanou 2013:1), healthcare (Sivan & Zukarnain 2021:1; Wang et al. 2022:1) or child protection domains (Krasznay, Rácz-Nagy & Dóra 2020:149; Siibak & Mascheroni 2021; Singh & Power 2021:99), often neglecting the wider socio-cultural, communal and spiritual dimensions of vulnerability. Furthermore, research frequently adopts normative or technical perspectives without critically examining which groups are systematically excluded from the academic discourse itself, which is a key issue raised by recent debates on digital ethics and data justice (Cinnamon 2020:214–233; Draude, Hornung & Klumbytė 2022).
Bibliometric studies on privacy and vulnerable groups
Bibliometric analysis has become a valuable quantitative tool for systematically mapping the development and structure of academic fields. By examining publication trends, citation networks, keyword co-occurrences and author collaborations, bibliometrics provides replicable insights into thematic priorities and research dynamics (Aria & Cuccurullo 2017:962). Several recent studies have applied bibliometric methods to adjacent domains, including cybersecurity (Khurana et al. 2024:202; Li, Zhang & Li 2025), data breaches (Hamid & Huda 2025:1), data justice (Draude et al. 2022), digital ethics (Guenduez, Walker & Demircioglu 2025:1) and ethical artificial intelligence (Saheb & Saheb 2024:1). However, few have focused specifically on privacy issues affecting vulnerable populations, and even fewer have critically interrogated how the academic visibility of these groups reflects broader power relations in digital policy, technology development and governance. A review of recent bibliometric analyses in the privacy domain further highlights this gap. For instance, Benedetto and Cucchi (2022) traced the historical evolution of privacy concepts across five distinct eras of technological development. The study emphasises how disciplinary shifts mirror broader trends in information technology. Their work in management and information systems identified dominant themes and theoretical frameworks in privacy research but noted the lack of attention to consumer perspectives and diverse contexts. Similarly, Rauf et al. (2024) analysed privacy issues in big data applications such as finance, e-learning and healthcare. The study revealed dominant concerns around blockchain, security and technological adoption, yet overlooked how these issues affect vulnerable or marginalised users. Other studies (Ali, Zaaba & Singh 2024:863; Shu & Liu 2021:721) explored consumer privacy, digital trust and computer privacy trends across disciplines and geographies, but none addressed the socio-cultural or ethical implications of privacy for excluded populations. Van Dijk et al. (2023:81) provided one of the most comprehensive structural mappings of the privacy research field using network and topic modelling approaches, but their findings primarily emphasise dominant academic clusters (e.g. cloud computing, location privacy) while marginalising medical or ethical communities and failing to analyse which social groups remain invisible in the literature. Collectively, these studies underscore the utility of bibliometric methods in identifying key research directions, but also reveal a shared limitation: the lack of attention to epistemic inclusion and the representation of vulnerable populations in digital privacy scholarship. This shortcoming further justifies the present study’s aim to systematically analyse how different marginalised groups, particularly those affected by intersecting social, political and technological vulnerabilities, are represented (or overlooked) in privacy-related academic discourse. One notable exception in the existing body of work is the systematic literature review conducted by Sannon and Forte (2022), which explicitly focused on the intersection of privacy and marginalisation. Their study examined publications from 2010 to 2020 across human-computer interaction (HCI), communication and privacy-focused venues and proposed the privacy responses and costs framework to articulate the complex trade-offs and harms experienced by marginalised individuals. While the review made critical contributions by surfacing thematic gaps and methodological shortcomings such as the underrepresentation of race and the lack of reporting on research practices, it was limited in both temporal scope and analytical breadth (Sannon & Forte 2022). The review spanned only a decade and did not use bibliometric tools to map author networks, keyword trends, geographic distribution or institutional collaborations. Furthermore, its disciplinary scope was primarily situated within HCI and communication, omitting other fields like law, policy and information systems where critical privacy discourse also occurs. This limitation highlights the need for a broader and more comprehensive bibliometric analysis that not only quantifies patterns across a longer time horizon, but also evaluates how vulnerable populations, including those defined by intersecting identities, are made visible (or invisible) across the academic ecosystem of privacy research. The present study addresses this gap by analysing a 20-year period (2005–2025) and leveraging bibliometric techniques to offer a multidimensional view of the scholarly attention (or neglect) towards marginalised groups in privacy discourse.
Research methods and design
Research design
This study adopts a bibliometric analysis approach to systematically map and quantify scholarly literature on privacy research concerning vulnerable populations over a 20-year span (2005–2025). Bibliometric analysis offers a robust quantitative framework to investigate publication trends, thematic evolution and patterns of academic collaboration within a research domain (Passas 2024:1014–1025). This method is particularly well-suited for identifying knowledge gaps, emergent topics and underrepresented communities across large datasets of peer-reviewed publications (Donthu et al. 2021:285).
Data sources and search strategy
Data were retrieved from two leading academic databases known for their interdisciplinary reach and high indexing quality: Scopus and Web of Science. These platforms were selected because of their extensive coverage of peer-reviewed journal articles and conference proceedings, which are the primary vehicles of scholarly communication in the field. A structured search query was formulated to capture publications that address privacy and data protection in relation to vulnerable populations. The query combined core privacy terms such as ‘privacy’, ‘data protection’ and ‘information privacy’ with descriptors of at-risk groups. The full query is shown below. The search query was adapted to the syntax requirements of each database. For Scopus, the search was conducted using the TITLE-ABS-KEY field as follows:
TITLE-ABS-KEY ((‘privacy’ OR ‘data protection’ OR ‘information privacy’) AND (‘vulnerable population*’ OR ‘marginalised group*’ OR ‘at-risk group*’ OR ‘underserved population*’ OR ‘disadvantaged group*’ OR ‘high-risk group*’ OR ‘susceptible population*’ OR ‘children’ OR ‘older persons’ OR ‘elderly’ OR ‘refugees’ OR ‘migrants’ OR ‘disabled persons’ OR ‘persons with disabilities’ OR ‘chronic illness’ OR ‘mental illness’ OR ‘physical disability*’ OR ‘pregnant*’ OR ‘drug user*’ OR ‘economically disadvantaged’ OR ‘homeless’ OR ‘domestic violence’ OR ‘incarcerated individuals’ OR ‘transgender’ OR ‘ethnic minority*’ OR ‘religious minority*’ OR ‘minority community*’))
For Web of Science, the search was conducted using the Topic field (TS), which includes title, abstract and keywords:
TS = ((‘privacy’ OR ‘data protection’ OR ‘information privacy’) AND (‘vulnerable population*’ OR ‘marginalised group*’ OR ‘at-risk group*’ OR ‘underserved population*’ OR ‘disadvantaged group*’ OR ‘high-risk group*’ OR ‘susceptible population*’ OR ‘children’ OR ‘older persons’ OR ‘elderly’ OR ‘refugees’ OR ‘migrants’ OR ‘disabled persons’ OR ‘persons with disabilities’ OR ‘chronic illness’ OR ‘mental illness’ OR ‘physical disability*’ OR ‘pregnant*’ OR ‘drug user*’ OR ‘economically disadvantaged’ OR ‘homeless’ OR ‘domestic violence’ OR ‘incarcerated individuals’ OR ‘transgender’ OR ‘ethnic minority*’ OR ‘religious minority*’ OR ‘minority community*’))
Boolean operators and truncation symbols were used to improve sensitivity while maintaining topical relevance. The search was restricted to English-language publications between 2005 and 2025 and limited to journal articles and conference papers. The Scopus search yielded 4051 articles, and Web of Science yielded 4149, resulting in an initial combined dataset of 8200 records. After the removal of 3440 duplicates, a final dataset of 4760 unique articles was retained for analysis.
Data extraction and cleaning
Bibliographic records were exported from Scopus and Web of Science in BibTeX format and merged using the mergeDbSources() function from the Bibliometrix R package (Aria & Cuccurullo 2017:962). The exported metadata included fields such as:
- Author names
- Document titles
- Abstracts
- Author and indexed keywords
- Source titles
- Affiliations
- Citation counts
- Publication years
- Digital Object Identifiers (DOIs)
Duplicates were identified and removed using combinations of DOIs, titles and author fields.
Data analysis
The final dataset was analysed using the Bibliometrix package in R (Aria & Cuccurullo 2017:962). The analytical procedures were carefully designed to address the four research questions outlined in the introduction.
Research Question 1: Which vulnerable populations have received the most academic attention in privacy-related research?
The study conducted:
- Keyword co-occurrence analysis, to identify which populations (e.g. children, refugees, elderly) appeared most frequently in author keywords and indexed terms.
- Descriptive analysis, to quantify the frequency and distribution of topics across time.
Research Question 2: Which vulnerable populations are underrepresented in privacy-related research?
To address this question, the study employed:
- Content mapping, to identify and extract references to a broad range of vulnerable population descriptors across the literature.
- Keyword co-occurrence and frequency analysis, to evaluate the visibility of various vulnerable groups, determine which are most and least represented, and examine how attention to these groups has shifted over time.
Research Question 3: How has research on digital privacy and vulnerability evolved between 2005 and 2025?
The study conducted:
- Temporal trend analysis, which traced shifts in research focus over time.
- Descriptive analysis, including publication counts, annual growth rates and citation trends, to capture the evolution of scholarly engagement with the topic.
Research Question 4: What are the most influential journals, authors and themes shaping this domain?
The study we applied:
- Descriptive metrics, including top-cited articles, prolific authors and most active journals.
To enhance interpretation, the analysis was supported by a series of visualisations, including publication trend graphs, geographic heatmaps of author affiliations, keyword co-occurrence networks and thematic evolution plots. These tools provided both a macro- and micro-level view of how privacy research concerning vulnerable populations has developed over the last two decades.
Ethical considerations
Ethical clearance to conduct this study was obtained from the Nelson Mandela University Research Ethics Committee (Ref. No. 1418).
Results
This section presents the findings from a bibliometric analysis of 4760 unique publications focused on privacy and vulnerable populations, spanning the years 2005–2025. The results are structured around key analytical areas, each aligned with the study’s research questions, including temporal trends, influential contributors, geographical distribution, thematic patterns and the visibility of vulnerable populations in scholarly discourse.
Annual publication trends (addresses RQ3: Evolution over time)
Figure 1 illustrates the evolution of scholarly output on privacy and vulnerable populations between 2005 and 2025. The annual volume of publications shows a steady increase over the two decades, rising from 39 publications in 2005 to over 500 in 2024. This upward trend reflects growing academic interest in the intersection of privacy and social vulnerability. The dataset reveals a compound annual growth rate of 11.71%, highlighting an acceleration in scholarly activity, particularly after 2015. A significant spike in publication output occurred during the COVID-19 pandemic period (2020–2023), suggesting that global crises tend to amplify privacy concerns, especially for at-risk and vulnerable populations. This growth pattern indicates a shift in framing privacy not only as a technological issue, but also increasingly as a social and ethical imperative, particularly in contexts involving marginalised or high-risk communities.
Most productive authors (addresses RQ4: Influential authors)
The authorship analysis revealed several scholars with substantial contributions to the literature on privacy and vulnerable populations as shown in Figure 2. Wang Y. was the most prolific author, with 46 publications, followed closely by Zhang Y. with 44. Other leading contributors included Cheng Y. (33), Liu Y. (29) and Liu J. (27).
Geographical distribution of research (addresses RQ4: Global research patterns)
Figure 3 presents a heatmap of corresponding author affiliations by country. Most research originates from high-income countries, with the United States (2557 publications), United Kingdom (896), Canada (569), Australia (523) and China (489) accounting for the largest shares of the literature. Collectively, these countries represent a dominant share of global scholarship on privacy and vulnerable populations. Other notable contributors include Germany (363), India (326), the Netherlands (272), Italy (237), Denmark (223), Spain (219) and Japan (200). A second tier of contributors, each with fewer than 200 publications, includes Sweden (190), Norway (170), France (161), Belgium (155), Switzerland (136), South Africa (120), Taiwan (105) and Finland (103). Importantly, South Africa is the only African country represented among the top 20 contributors, highlighting the broader underrepresentation of the African continent in this domain. This reflects a persistent geographical imbalance in global privacy research and points to the need for greater investment in context-specific studies on privacy and vulnerability across African nations. The lack of contributions from most of the continent suggests that regional privacy concerns, particularly in relation to digitally marginalised populations, may be overlooked in the current global discourse.
This asymmetry in global research output reinforces ongoing concerns about epistemic injustice, where knowledge production is concentrated in the Global North, potentially neglecting the lived experiences, regulatory challenges and cultural dimensions of privacy in the Global South, especially Africa and Latin America. Addressing this gap is essential for fostering more inclusive, equitable and socially responsive privacy research agendas.
Thematic trends and keyword analysis (addresses RQ1: Most studied vulnerable groups; RQ4: Themes)
To identify dominant research themes and conceptual patterns, a keyword co-occurrence analysis was performed using author-assigned keywords and indexed keywords. Figure 4 presents a combined visualisation of the top 20 author keywords and top 20 indexed keywords side by side, highlighting the most frequently occurring vulnerable group terms.
 |
FIGURE 4: Comparison of the top 20 vulnerable group keywords appearing in (a) author keywords and (b) indexed keywords. |
|
The most common keywords across both sets include:
- Children (Author Keywords: 376; Indexed Keywords: 531)
- Pregnancy (100; 416)
- Adolescents (172; 162)
- Women (53; 150)
- Gender (45; 103)
- Disabled Persons (n/a; 77)
- Transgender (66; 63)
- Elderly (154; 52)
- Disability (34; 41)
- Older Adults (83; 39)
- Refugees (54; 36), among others.
Notably, keywords are reported as they appear in the literature. For instance, ‘disabled persons’ is commonly used, even though more respectful and accurate terminology such as ‘people living with disabilities’ is preferable.
Keyword analysis via word cloud (supports RQ4: Thematic overview)
To provide a broad thematic overview of the research landscape, a word cloud was generated based on the combined frequency of keywords extracted from both author keywords and indexed keywords across the dataset. This visual representation emphasises the most frequently occurring terms, with larger words indicating higher frequency (Figure 5). The word cloud offers readers an accessible snapshot of dominant topics and recurring concepts within the literature, capturing prevalent research themes without the need for an exhaustive manual review. Key terms that appear prominently include ‘privacy’, ‘female’, ‘human’, ‘male’, ‘humans’, ‘adult’, ‘aged’, ‘child’, ‘article’ and ‘adolescent’, among others. These reflect significant thematic clusters around demographic groups and core privacy concerns. By presenting the entire spectrum of keywords, this section aims to give a clear and comprehensive picture of the vocabulary shaping the field.
Most relevant sources (addresses RQ4: Influential journals)
An analysis of source productivity revealed that a small number of journals dominate scholarly output in privacy-related research on vulnerable populations. Table 1 presents the 10 most prolific journals based on the number of articles published between 2005 and 2025. BMJ Open emerged as the most productive outlet with 91 articles, followed by PLOS ONE (65 articles) and the Journal of Medical Internet Research (56 articles). Other leading sources include Sensors (52), BMC Public Health (35), BMC Health Services Research (34) and IEEE Access (34). These journals reflect a strong presence of interdisciplinary and health-focused publications, emphasising the prominence of healthcare, public health and digital surveillance themes within the domain. Notably, many of these journals publish open-access content, suggesting an orientation towards accessibility and broad dissemination of findings in privacy scholarship. The concentration of publications within a handful of sources also signals possible gatekeeping effects or disciplinary clustering, which may shape whose perspectives, and which vulnerable populations receive scholarly attention.
Least represented vulnerable groups (addresses RQ2: Underrepresented populations)
Figure 6, based on both author keywords and indexed keywords, consistently shows that key vulnerable groups rarely appear as focal points in privacy research. Among the 10 least represented author keywords, several received no mentions at all, including drug users, economically disadvantaged, minority communities and poor. Other groups like incarcerated individuals, religious minorities, disabled persons, physical disability and persons with disabilities were mentioned only one to three times. Similarly, the 10 least frequent indexed keywords revealed a comparable pattern. Groups such as economically disadvantaged, incarcerated individuals, minorities, minority communities and religious minorities had zero mentions, while ethnic minorities, LGBT, older persons, indigenous and poor appeared only once or twice.
 |
FIGURE 6: Least represented vulnerable groups; (a) author keywords and (b) indexed keywords. |
|
These findings suggest a persistent gap in privacy scholarship concerning intersectional and socially vulnerable populations. Despite these groups facing heightened risks of surveillance, exploitation or data misuse, they remain largely absent from the dominant discourses shaping digital privacy frameworks. The consistent omission of groups such as the economically disadvantaged, incarcerated individuals, minority communities and people living with disabilities raises concerns about epistemic invisibility where the needs, experiences and rights of certain populations are not just underserved but effectively erased from scholarly and policy debates. This underrepresentation not only limits the inclusivity of academic knowledge production, but may also hinder the development of privacy protections that are sensitive to structural inequalities. Moving forward, there is an urgent need for digital privacy research agendas to intentionally engage with these marginalised communities, ensuring that emerging policies, tools and frameworks reflect the full diversity of social vulnerability.
Discussion
Key findings
This study offers a comprehensive bibliometric review of privacy-related research concerning vulnerable populations. It reveals consistent growth in scholarly output over the past two decades, with marked increases after 2015. The analysis further highlights geographic disparities in publication volume, a dominant thematic focus on healthcare and children and the persistent underrepresentation of many marginalised groups such as persons with disabilities, religious minorities and economically disadvantaged populations. South Africa emerges as the only African country among the top 20 contributors, ranking 18th overall.
Interpretation of findings
The increasing volume of publications and citations signals growing academic and societal interest in privacy, especially amid technological expansion and global crises like COVID-19. However, the thematic concentration on biomedical ethics and child protection, while important, reveals a narrow conceptual scope. Many vulnerable groups remain peripheral or invisible within mainstream privacy scholarship. These findings are consistent with previous literature on epistemic injustice, where knowledge production often centres Western priorities, neglecting context-specific issues in the Global South. The notable underrepresentation of African nations, with only South Africa featured in the top 20, reinforces concerns about global research imbalances. This suggests that policy and scholarly frameworks often lack sensitivity to regionally specific privacy risks such as those tied to local cultural norms, resource limitations or sociopolitical contexts. Without inclusive representation, dominant privacy frameworks may inadequately reflect the lived experiences of many at-risk communities. These findings both align with and extend existing research in the field. Similar to prior bibliometric and review studies (e.g. Benedetto & Cucchi 2022; Sannon & Forte 2022), this study confirms the dominance of themes related to healthcare, children and data protection within privacy scholarship. However, a key difference lies in the explicit focus on the representation of vulnerable populations, where this study reveals a more pronounced imbalance in scholarly attention than previously reported. While earlier studies have highlighted thematic concentrations, they have not systematically examined which social groups are rendered visible or invisible within the literature. A particularly notable and somewhat unexpected finding is the near absence of certain highly vulnerable groups such as economically disadvantaged communities, incarcerated individuals and religious minorities in both author and indexed keywords. Given the well-documented exposure of these groups to surveillance and data exploitation, their limited representation suggests a disconnect between real-world vulnerability and academic focus. This gap points to an urgent need for future research to better align scholarly inquiry with the lived experiences of marginalised populations.
Strengths and limitations
A major strength of this study lies in its use of a dual-source dataset combining Scopus and Web of Science records, which enhances the comprehensiveness of the bibliometric review. The mixed-method approach combining trend analysis, keyword co-occurrence and visual mapping enables a nuanced understanding of both research volume and conceptual priorities. However, the study is not without limitations. The analysis relied on author-supplied and indexed keywords, which may not fully capture all relevant vulnerable populations, especially when terminology is inconsistent or outdated (e.g. ‘disabled persons’ vs. ‘persons living with disabilities’). Additionally, while rigorous efforts were made to clean and standardise country names and affiliations, variations in metadata quality may have affected the accuracy of geographic attributions. Finally, the absence of full-text analysis means deeper qualitative insights into how vulnerable populations are discussed were outside the study’s scope.
Implications and recommendations
This study underscores the need for more inclusive privacy research that meaningfully incorporates the concerns of underrepresented and marginalised groups. Future studies should expand beyond biomedical and child-centric paradigms to consider a wider array of vulnerabilities, including disability, indigeneity, incarceration and economic disadvantage. Researchers should also adopt inclusive terminology and frameworks that reflect the lived realities of diverse populations, especially in the Global South. From a policy perspective, funding bodies and academic institutions must support scholarship that decentralises Western epistemologies and amplifies region-specific privacy concerns. For practice, developers and regulators should be urged to co-design digital technologies and policies with input from those communities most at risk of privacy violations.
Conclusion
This bibliometric study offers a systematic mapping of digital privacy research focused on vulnerable populations over a 20-year period. Addressing the first research question: which vulnerable populations have received the most academic attention; the analysis found that children, health-related populations (e.g. pregnant women, adolescents and the elderly) and consumers dominate the literature. These groups consistently appeared in both author and indexed keywords, suggesting a strong biomedical and child protection focus in privacy research. In response to the second research question: which populations remain underrepresented; the study revealed limited visibility of marginalised groups such as persons with disabilities, incarcerated individuals, religious minorities, economically disadvantaged communities and ethnic minorities. Their consistent absence from top keyword rankings points to conceptual blind spots and epistemic exclusion in how privacy concerns are framed and studied. With regard to the third question: how research on digital privacy and vulnerability has evolved between 2005 and 2025; the results demonstrate steady growth in scholarly output, particularly after 2015. This upward trend reflects rising global concern with digital rights and personal data, further amplified by events such as the COVID-19 pandemic. However, this momentum has not been matched by an equally expansive engagement with diverse vulnerable populations. Finally, in addressing the fourth question: what journals, authors and themes have shaped this domain; the analysis identified a small group of prolific authors and a concentration of publications in journals aligned with healthcare, ethics and digital technology. The majority of influential contributions originate from high-income countries, particularly the United States, United Kingdom and Canada, with South Africa being the only African country among the top 20 contributors, highlighting a stark geographic imbalance. In conclusion, while privacy research is expanding in volume and impact, it continues to overlook key vulnerable populations and regions. There is a pressing need for more inclusive, interdisciplinary and regionally grounded scholarship, particularly in Africa and other underrepresented contexts. Future research should aim to close these gaps through qualitative and mixed methods approaches, expanded data sources and intentional collaboration with scholars and communities from the Global South.
Acknowledgements
The authors gratefully acknowledge the support of the National Research Foundation (NRF) of South Africa, through the Black Academics Advancement Programme (BAAP), for funding the lead author’s PhD studies. The views expressed in this article are those of the authors and do not necessarily represent those of the NRF. Portions of the manuscript preparation and data analysis were supported using R programming language, specifically bibliometric and visualisation packages including bibliometrix, ggplot2 and wordcloud. These were used under the supervision of the lead author, who assumes full responsibility for the integrity and accuracy of the content.
Competing interest
The author reported that they received funding from National Research Foundation (NRF), which may be affected by the research reported in the enclosed publication. The author has disclosed those interests fully and has implemented an approved plan for managing any potential conflicts arising from their involvement. The terms of these funding arrangements have been reviewed and approved by the affiliated University in accordance with its policy on objectivity in research.
CRediT authorship contribution
Phumezo Ntlatywa: Conceptualisation, Methodology, Project administration, Visualisation, Writing – original draft. Darelle van Greunen: Supervision. All authors reviewed the article, contributed to the discussion of results, approved the final version for submission and publication and take responsibility for the integrity of its findings.
Funding information
This research was funded by the National Research Foundation (NRF) of South Africa, under the Black Academics Advancement Programme (BAAP). The funder had no role in the study design, data collection and analysis or the writing and submission of the manuscript.
Data availability
The bibliometric data supporting the findings of this study were sourced from Scopus and Web of Science databases. Cleaned datasets and R scripts used for the analysis are available from the corresponding author, Phumezo Ntlatywa, upon reasonable request.
Disclaimer
The views and opinions expressed in this article are those of the authors and are the product of professional research. They do not necessarily reflect the official policy or position of any affiliated institution, funder, agency or that of the publisher. The authors are responsible for this article’s results, findings and content.
References
Ahmadon, M.A., Napp, N., Rao, S., Silva, C., Lizar, M., Gorog, C. et al., 2025, ‘Digital privacy: Trends, challenges, and the future’, IT Professional 27(3), 69–77. https://doi.org/10.1109/MITP.2025.3546433
Ali, A.S., Zaaba, Z.F. & Singh, M.M., 2024, ‘The rise of “security and privacy”: Bibliometric analysis of computer privacy research’, International Journal of Information Security 23(2), 863–885. https://doi.org/10.1007/s10207-023-00761-4
Aria, M. & Cuccurullo, C., 2017, ‘bibliometrix: An R-tool for comprehensive science mapping analysis’, Journal of Informetrics 11(4), 959–975. https://doi.org/10.1016/j.joi.2017.08.007
Benedetto, E.D. & Cucchi, A., 2022, ‘A bibliometric analysis of privacy concerns’, in 2022 3rd International Conference on Next Generation Computing Applications (NextComp), pp. 1–7.
Botes, M., 2025, ‘Regulatory challenges of digital health: The case of mental health applications and personal data in South Africa’, Frontiers in Pharmacology 16, 1498600. https://doi.org/10.3389/fphar.2025.1498600
Broeckelmann, R., 2026, What is digital privacy? Medium, South Dakota.
Buthelezi, S.P., Zondo, N.M., Nxumalo, L.T.M. & Vilakazi, M., 2024, ‘Determining the digital divide among people with disabilities in KwaZulu-Natal’, South African Journal of Information Management 26(1), a1820. https://doi.org/10.4102/SAJIM.v26i1.1820
Cinnamon, J., 2020, ‘Data inequalities and why they matter for development’, Information Technology for Development 26(2), 214–233. https://doi.org/10.1080/02681102.2019.1650244
Coman, E., Coman, C., Alexandrescu, M.B. & Bilți, R.-S., 2025, ‘Mapping the frontiers of cybersecurity and data protection: Insights from a bibliometric study’, Electronics 14(19), 3769. https://doi.org/10.3390/electronics14193769
Donthu, N., Kumar, S., Mukherjee, D., Pandey, N. & Lim, W.M., 2021, ‘How to conduct a bibliometric analysis: An overview and guidelines’, Journal of Business Research 133, 285–296. https://doi.org/10.1016/j.jbusres.2021.04.070
Draude, C., Hornung, G. & Klumbytė, G., 2022, ‘Mapping data justice as a multidimensional concept through feminist and legal perspectives’, in A. Hepp, J. Jarke & L. Kramp (eds.), New perspectives in critical data studies: The ambivalences of data power, pp. 187–216, Springer International Publishing, Cham.
Drozdowski, P., Rathgeb, C., Dantcheva, A., Damer, N. & Busch, C., 2020, ‘Demographic bias in biometrics: A survey on an emerging challenge’, IEEE Transactions on Technology and Society 1(2), 89–103. https://doi.org/10.1109/TTS.2020.2992344
Georgiou, T., Baillie, L. & Shah, R., 2023, Investigating concerns of security and privacy among Rohingya refugees in Malaysia.
Guenduez, A.A., Walker, N. & Demircioglu, M.A., 2025, ‘Digital ethics: Global trends and divergent paths’, Government Information Quarterly 42(3), 102050. https://doi.org/10.1016/j.giq.2025.102050
Hamid, S. & Huda, M.N., 2025, ‘Mapping the landscape of government data breaches: A bibliometric analysis of literature from 2006 to 2023’, Social Sciences & Humanities Open 11, 101234. https://doi.org/10.1016/j.ssaho.2024.101234
Jang, Y. & Ko, B., 2023, ‘Online safety for children and youth under the 4Cs framework – A focus on digital policies in Australia, Canada, and the UK’, Children 10(8), 1415. https://doi.org/10.3390/children10081415
Khurana, P., Narula, S., Tiwari, N., Kapoor, R. & Arora, M., 2024, ‘Mapping the cybersecurity research: A comprehensive bibliometric analysis’, International Journal of Experimental Research and Review 46, 202–211. https://doi.org/10.52756/ijerr.2024.v46.016
Krasznay, C., Rácz-Nagy, J. & Dóra, L., 2020, ‘Privacy challenges in children’s online presence – From the developers’ perspective’, Central and Eastern European eDem and eGov Days 338, 149–158. https://doi.org/10.24989/ocg.338.12
Kreutzer, T., Orbinski, J., Appel, L., An, A., Marston, J., Boone, E. et al., 2025, ‘Ethical implications related to processing of personal data and artificial intelligence in humanitarian crises: A scoping review’, BMC Medical Ethics 26(1), 49. https://doi.org/10.1186/s12910-025-01189-2
Li, J., Zhang, Y. & Li, S., 2025, ‘Mapping the landscape of cybersecurity research: A bibliometric analysis’, International Journal of Legal Discourse 10(1), 99–119. https://doi.org/10.1515/ijld-2025-2006
Livingstone, S. & Helsper, E., 2007, ‘Gradations in digital inclusion: Children, young people and the digital divide’, New Media & Society 9(4), 671–696. https://doi.org/10.1177/1461444807080335
Livingstone, S. & Smith, P.K., 2014, ‘Annual research review: Harms experienced by child users of online and mobile technologies: The nature, prevalence and management of sexual and aggressive risks in the digital age’, Journal of Child Psychology and Psychiatry 55(6), 635–654. https://doi.org/10.1111/jcpp.12197
McDonald, N. & Forte, A., 2022, ‘Privacy and vulnerable populations’, in B.P. Knijnenburg, X. Page, P. Wisniewski, H.R. Lipford, N. Proferes & J. Romano (eds.), Modern socio-technical perspectives on privacy, pp. 337–363, Springer International Publishing, Cham.
Muhawe, C., 2025, The (in)visible immigrant’s privacy, Georgetown Law Technology Review, Georgetown University Law Center, Washington, DC.
Passas, I., 2024, ‘Bibliometric analysis: The main steps’, Encyclopedia 4(2), 1014–1025. https://doi.org/10.3390/encyclopedia4020065
Radu, L.-D. & Popescul, D., 2024, ‘Green information systems – A bibliometric analysis of the literature from 2000 to 2023’, Electronics 13(7), 1329. https://doi.org/10.3390/electronics13071329
Rauf, A., Tariq, U., Tang, H. & Shishir, M.A., 2024, ‘Bibliometric analysis: Research trends of privacy in big data and its applications’, in 2024 7th International Conference on Data Science and Information Technology (DSIT), IEEE, Nanjing, China, December 20–22, 2024, pp. 1–6.
Rhoen, M., 2016, ‘Beyond consent: Improving data protection through consumer protection law’, Internet Policy Review 5(1), 1–15. https://doi.org/10.14763/2016.1.404
Saheb, T. & Saheb, T., 2024, ‘Mapping ethical artificial intelligence policy landscape: A mixed method analysis’, Science and Engineering Ethics 30(2), 9. https://doi.org/10.1007/s11948-024-00472-6
Sannon, S. & Forte, A., 2022, ‘Privacy research with marginalized groups: What we know, what’s needed, and what’s next’, Proceedings of the ACM on Human-Computer Interaction 6(CSCW2), 455, 1–33. https://doi.org/10.1145/3555556
Scanlan, M., 2022, ‘Reassessing the disability divide: Unequal access as the world is pushed online’, Universal Access in the Information Society 21(3), 725–735. https://doi.org/10.1007/s10209-021-00803-5
Sekalala, S., Dagron, S., Forman, L. & Meier, B.M., 2020, ‘Analyzing the human rights impact of increased digital public health surveillance during the COVID-19 crisis’, Health and Human Rights 22(2), 7–20.
Shin, H., Ryu, K., Kim, J.-Y. & Lee, S., 2024, ‘Application of privacy protection technology to healthcare big data’, Digital Health 10, 20552076241282242. https://doi.org/10.1177/20552076241282242
Shu, S. & Liu, Y., 2021, ‘Looking back to move forward: A bibliometric analysis of consumer privacy research’, Journal of Theoretical and Applied Electronic Commerce Research 16(4), 727–747. https://doi.org/10.3390/jtaer16040042
Siibak, A. & Mascheroni, G., 2021, ‘Children’s data and privacy in the digital age’, in Leibniz-Institut Für Medienforschung, Hans-Bredow-Institut (HBI) (ed.), CO:RE short report series on key topics, p. 12, Leibniz Institute for Media Research | Hans-Bredow-Institut (HBI), Hamburg.
Silvennoinen, P. & Rantanen, T., 2023, ‘Digital agency of vulnerable people as experienced by rehabilitation professionals’, Technology in Society, 72, 102173. https://doi.org/10.1016/j.techsoc.2022.102173
Singh, A. & Power, T., 2021, ‘Understanding the privacy rights of the African child in the digital era’, African Human Rights Law Journal 21(1), 99–125. https://doi.org/10.17159/1996-2096/2021/v21n1a6
Sivan, R. & Zukarnain, Z.A., 2021, ‘Security and privacy in cloud-based e-health system’, Symmetry 13(5), 742. https://doi.org/10.3390/sym13050742
Srivastava, S., Wilska, T.-A. & Nyrhinen, J., 2023, ‘Children as social actors negotiating their privacy in the digital commercial context’, Childhood 30(3), 235–252. https://doi.org/10.1177/09075682231186486
Steinbrink, E., Biselli, T., Linsner, S., Herbert, F. & Reuter, C., 2023, ‘Privacy perception and behavior in safety-critical environments’, in N. Gerber, A. Stöver & K. Marky (eds.), Human factors in privacy research, pp. 237–251, Springer International Publishing, Cham.
Strathmann, C., 2025, ‘Privacy for all: Empowering vulnerable groups with diversity-oriented online protection’, in N. Yamashita, V. Evers, M. Burnett & J.P. Bigham (eds.), Proceedings of the Extended Abstracts of the CHI Conference on Human Factors in Computing Systems (CHI EA’25), Yokohama, Japan, 26 April–01 May 2025, pp. 1–4, Association for Computing Machinery, New York, NY.
Sun, R., Xue, M., Tyson, G., Wang, S., Camtepe, S. & Nepal, S., 2023, ‘Not seen, not heard in the digital world! Measuring privacy practices in children’s apps’, in Y. Ding, J. Tang, J. Sequeda, L. Aroyo, C. Castillo & G.-J. Houben (eds.), Proceedings of the ACM Web Conference 2023 (WWW ’23), Austin, Texas, USA, 30 April–04 May 2023, pp. 2166–2177, Association for Computing Machinery, New York, NY.
Taylor, L., 2017, ‘What is data justice? The case for connecting digital rights and freedoms globally’, Big Data & Society 4(2), 2053951717736335. https://doi.org/10.1177/2053951717736335
Tomczyk, Ł., Mascia, M.L., Gierszewski, D. & Walker, C., 2023, ‘Barriers to digital inclusion among older people: A intergenerational reflection on the need to develop digital competences for the group with the highest level of digital exclusion’, Innoeduca. International Journal of Technology and Educational Innovation 9(1), 5–26. https://doi.org/10.24310/innoeduca.2023.v9i1.16433
Tzanou, M., 2013, ‘Data protection as a fundamental right next to privacy? “Reconstructing” a not so new right’, International Data Privacy Law 3(2), 88–99. https://doi.org/10.1093/idpl/ipt004
Van Dijk, F., Gadellaa, J., Van Toledo, C., Spruit, M., Brinkkemper, S. & Brinkhuis, M., 2023, ‘Uncovering the structures of privacy research using bibliometric network analysis and topic modelling’, Organizational Cybersecurity Journal: Practice, Process and People 3(2), 81–99. https://doi.org/10.1108/OCJ-11-2021-0034
Wang, C., Zhang, J., Lassi, N. & Zhang, X., 2022, ‘Privacy protection in using artificial intelligence for healthcare: Chinese regulation in comparative perspective’, Healthcare 10(10), 1878. https://doi.org/10.3390/healthcare10101878
Wang, L.H. & Metzger, M.J., 2024, ‘The online privacy divide: Testing resource and identity explanations for racial/ethnic differences in privacy concerns and privacy management behaviors on social media’, Communication Research. https://doi.org/10.1177/00936502241273157
|