emoji_adoption_lag() compares the date an emoji was first used in your
corpus with the date Unicode released it, giving a per-glyph adoption lag in
days.
Value
A tibble with one row per emoji, most frequent first: emoji,
name, n, version, release_date, first_seen and lag_days.
lag_days is NA when the release date of the version is unknown.
Details
A lag is only as good as the corpus window: an emoji released before your
data begins will look adopted on day one, so read the lag together with n
and the span of your data. Negative lags mean the corpus contains a glyph
before its official release date – usually a vendor shipping early, or a
timestamp problem worth investigating.
Occurrences whose time is missing or unparseable are dropped.
Examples
df <- data.frame(
when = as.Date(c("2021-01-01", "2022-06-01")),
text = c("\U0001f600", "\U0001f97a")
)
emoji_adoption_lag(df, text, when)
#> # A tibble: 2 × 7
#> emoji name n version release_date first_seen lag_days
#> <chr> <chr> <int> <chr> <date> <date> <int>
#> 1 😀 grinning face 1 1.0 2015-06-09 2021-01-01 2033
#> 2 🥺 pleading face 1 11.0 2018-06-05 2022-06-01 1457