Skip to content

Commit 0c56973

Browse files
authored
Merge pull request #9 from openwashdata/dev
Release 0.1.0: remove list-columns from datasets
2 parents a370e98 + c89cb8f commit 0c56973

78 files changed

Lines changed: 14477 additions & 1877 deletions

File tree

Some content is hidden

Large Commits have some content hidden by default. Use the searchbox below for content that may be hidden.

‎CITATION.cff‎

Lines changed: 2 additions & 2 deletions
Original file line numberDiff line numberDiff line change
@@ -9,7 +9,7 @@ type: software
99
license: CC-BY-4.0
1010
title: 'washopenresearch: Dataset about open research data information in Water, Sanitation,
1111
and Hygiene'
12-
version: 0.0.1
12+
version: 0.1.0
1313
doi: 10.5281/zenodo.11185699
1414
abstract: The goal of washopenresearch is to provide an overview of open research
1515
data related to Water Sanitation and Hygiene (WASH). The package provides access
@@ -32,7 +32,7 @@ authors:
3232
orcid: https://orcid.org/0000-0003-2196-5015
3333
repository-code: https://github.com/openwashdata/washopenresearch
3434
url: https://github.com/openwashdata/washopenresearch
35-
date-released: '2024-05-13'
35+
date-released: '2026-07-07'
3636
contact:
3737
- family-names: Zhong
3838
given-names: Mian

‎DESCRIPTION‎

Lines changed: 4 additions & 4 deletions
Original file line numberDiff line numberDiff line change
@@ -1,6 +1,6 @@
11
Package: washopenresearch
22
Title: Dataset about open research data information in Water, Sanitation, and Hygiene
3-
Version: 0.0.1
3+
Version: 0.1.0
44
Authors@R: c(
55
person("Mian", "Zhong", , "mzhong@ethz.ch", role = c("aut", "cre"),
66
comment = c(ORCID = "0009-0009-4546-7214")),
@@ -13,11 +13,11 @@ Description: The goal of washopenresearch is to provide an overview of open rese
1313
License: CC BY 4.0
1414
Encoding: UTF-8
1515
Roxygen: list(markdown = TRUE)
16-
RoxygenNote: 7.3.1
1716
Depends:
18-
R (>= 2.10)
17+
R (>= 3.5)
1918
LazyData: true
2019
Config/Needs/website: rmarkdown
21-
Date: 2024-05-13
20+
Date: 2026-07-07
2221
URL: https://github.com/openwashdata/washopenresearch
2322
BugReports: https://github.com/openwashdata/washopenresearch/issues
23+
Config/roxygen2/version: 8.0.0

‎NEWS.md‎

Lines changed: 17 additions & 0 deletions
Original file line numberDiff line numberDiff line change
@@ -0,0 +1,17 @@
1+
# washopenresearch 0.1.0
2+
3+
## Breaking changes
4+
5+
- `washdev` and `uncnewsletter` no longer contain list-columns (#8). The multi-value variables `supp_file_type`, `supp_url`, `das_repo_url`, and `keywords` are now character columns in which multiple values are separated by `"; "`. Flat-file exports such as `write.csv()` now work directly on both datasets. Code that used `tidyr::unnest()` or `unlist()` on these columns should split the strings instead, for example with `tidyr::separate_rows(supp_file_type, sep = "; ")` or `stringr::str_split(keywords, "; ")`.
6+
- `Depends` was raised from R (>= 2.10) to R (>= 3.5), required by the serialization format of the regenerated data files.
7+
8+
## Minor improvements and fixes
9+
10+
- The CSV and XLSX exports in `inst/extdata/` now show multi-value cells as `"; "`-separated strings instead of R code literals such as `c("pdf", "docx")`.
11+
- The variable descriptions in the package documentation and the data dictionary describe the new delimited format and the correct variable types.
12+
- The article "Missed Opportunity: where is WASH research data gone?" uses `tidyr::separate_rows()` in place of `tidyr::unnest()` to expand `supp_file_type`.
13+
- The word cloud figure in the README has alt text.
14+
15+
# washopenresearch 0.0.1
16+
17+
- Initial release with the `washdev` and `uncnewsletter` datasets on data availability statements in WASH research publications.

‎R/uncnewsletter.R‎

Lines changed: 4 additions & 4 deletions
Original file line numberDiff line numberDiff line change
@@ -13,8 +13,8 @@
1313
#' \item{published_year}{Year of publication}
1414
#' \item{is_supp}{Whether the paper has supplementary materials}
1515
#' \item{num_supp}{Number of supplementary material files}
16-
#' \item{supp_file_type}{File type of the supplementary materials}
17-
#' \item{supp_url}{Website url of the supplementary materials}
16+
#' \item{supp_file_type}{File types of the supplementary materials, separated by "; " when there are multiple}
17+
#' \item{supp_url}{Website urls of the supplementary materials, separated by "; " when there are multiple}
1818
#' \item{num_authors}{Number of the authors}
1919
#' \item{first_author_name}{Name of the first author}
2020
#' \item{first_author_affiliation}{Academic affiliation of the first author}
@@ -29,7 +29,7 @@
2929
#' \item{has_das}{Whether the paper has a data availability statement}
3030
#' \item{das}{Original data availability statement of the paper. NA if it does not have a data availability statement.}
3131
#' \item{das_type}{Type of the data availability statement including in paper(data in full paper scope like supplementary material or appendix or main content) on request(data available on request to the authors) available in online repository(data is shared in a public online repository) not shareable(data is not shareable). NA if it does not have a data availability statement.}
32-
#' \item{das_repo_url}{Website url of the data if the relevant data of the paper is shared on a public repository}
33-
#' \item{keywords}{List of keywords of the paper}
32+
#' \item{das_repo_url}{Website urls of the data if the relevant data of the paper is shared on a public repository, separated by "; " when there are multiple}
33+
#' \item{keywords}{Keywords of the paper, separated by "; "}
3434
#' }
3535
"uncnewsletter"

‎R/washdev.R‎

Lines changed: 4 additions & 4 deletions
Original file line numberDiff line numberDiff line change
@@ -12,8 +12,8 @@
1212
#' \item{published_year}{Year of publication}
1313
#' \item{is_supp}{Whether the paper has supplementary materials}
1414
#' \item{num_supp}{Number of supplementary material files}
15-
#' \item{supp_file_type}{File type of the supplementary materials}
16-
#' \item{supp_url}{Website url of the supplementary materials}
15+
#' \item{supp_file_type}{File types of the supplementary materials, separated by "; " when there are multiple}
16+
#' \item{supp_url}{Website urls of the supplementary materials, separated by "; " when there are multiple}
1717
#' \item{num_authors}{Number of the authors}
1818
#' \item{first_author_name}{Name of the first author}
1919
#' \item{first_author_affiliation}{Academic affiliation of the first author}
@@ -28,7 +28,7 @@
2828
#' \item{has_das}{Whether the paper has a data availability statement}
2929
#' \item{das}{Original data availability statement of the paper. NA if it does not have a data availability statement.}
3030
#' \item{das_type}{Type of the data availability statement including in paper(data in full paper scope like supplementary material or appendix or main content) on request(data available on request to the authors) available in online repository(data is shared in a public online repository) not shareable(data is not shareable). NA if it does not have a data availability statement.}
31-
#' \item{das_repo_url}{Website url of the data if the relevant data of the paper is shared on a public repository}
32-
#' \item{keywords}{List of keywords of the paper}
31+
#' \item{das_repo_url}{Website urls of the data if the relevant data of the paper is shared on a public repository, separated by "; " when there are multiple}
32+
#' \item{keywords}{Keywords of the paper, separated by "; "}
3333
#' }
3434
"washdev"

‎README.Rmd‎

Lines changed: 2 additions & 1 deletion
Original file line numberDiff line numberDiff line change
@@ -46,7 +46,7 @@ contains two datasets from the following sources:
4646
Water
4747
News](https://waterinstitute.unc.edu/our-work/nc-water-news-newsletter)
4848

49-
![](man/figures/washdev_wordcloud.png){width="515"}
49+
![Word cloud of the most frequent keywords in articles of the Journal of Water, Sanitation and Hygiene for Development, with water, sanitation, and hygiene appearing largest](man/figures/washdev_wordcloud.png){width="515"}
5050

5151
## Installation
5252

@@ -174,6 +174,7 @@ calculate their frequency to be used.
174174

175175
```{r washdev_keyword_frequency, echo=TRUE}
176176
keywords_freq <- washdev$keywords |>
177+
str_split("; ") |>
177178
unlist() |>
178179
str_to_lower() |>
179180
table() |>

0 commit comments

Comments
 (0)