The Information and Communication sector has undoubtedly played a pivotal role in changing the way people live nowadays. Almost every area of our lives is affected by the presence and the use of the new information and communication technologies. In this regard, many researchers' attention has been attracted by the influence or the significant impact of these technologies on economic growth and development. Although the history of South Africa has had some drawbacks that could constitute a big obstacle to the emergence of a successful economic environment, the actual status of the country regarding its economy and the role that it plays in Africa towards the rest of the African countries is a vital example of an emerging economic force in Africa. This paper examines the crucial role that ICT has played and is still playing in the South African economy growth and more specifically the significance of the economic effects of the software industry. It makes use of the framework used by Heavin et al. (2003) to investigate the Irish software industry in order to analyze the impact of endogenous factors -- national, enterprise and individual -- on the software industry and its implicatio
An epidemiological model is developed for the spread of COVID-19 in South Africa. A variant of the classical compartmental SEIR model, called the SEIQRDP model, is used. As South Africa is still in the early phases of the global COVID-19 pandemic with the confirmed infectious cases not having peaked, the SEIQRDP model is first parameterized on data for Germany, Italy, and South Korea - countries for which the number of infectious cases are well past their peaks. Good fits are achieved with reasonable predictions of where the number of COVID-19 confirmed cases, deaths, and recovered cases will end up and by when. South African data for the period from 23 March to 8 May 2020 is then used to obtain SEIQRDP model parameters. It is found that the model fits the initial disease progression well, but that the long-term predictive capability of the model is rather poor. The South African SEIQRDP model is subsequently recalculated with the basic reproduction number constrained to reported values. The resulting model fits the data well, and long-term predictions appear to be reasonable. The South African SEIQRDP model predicts that the peak in the number of confirmed infectious individuals w
To curb the spread of COVID-19, many governments around the world have implemented tiered lockdowns with varying degrees of stringency. Lockdown levels are typically increased when the disease spreads and reduced when the disease abates. A predictive control approach is used to develop optimized lockdown strategies for curbing the spread of COVID-19. The strategies are then applied to South African data. The South African case is of interest as the South African government has defined five distinct levels of lockdown, which serves as a discrete control input. An epidemiological model for the spread of COVID-19 in South Africa was previously developed, and is used in conjunction with a hybrid model predictive controller to optimize lockdown management under different policy scenarios. Scenarios considered include how to flatten the curve to a level that the healthcare system can cope with, how to balance lives and livelihoods, and what impact the compliance of the population to the lockdown measures has on the spread of COVID-19. The main purpose of this paper is to show what the optimal lockdown level should be given the policy that is in place, as determined by the closed-loop fee
Using the Scopus dataset (1996-2007) a grand matrix of aggregated journal-journal citations was constructed. This matrix can be compared in terms of the network structures with the matrix contained in the Journal Citation Reports (JCR) of the Institute of Scientific Information (ISI). Since the Scopus database contains a larger number of journals and covers also the humanities, one would expect richer maps. However, the matrix is in this case sparser than in the case of the ISI data. This is due to (i) the larger number of journals covered by Scopus and (ii) the historical record of citations older than ten years contained in the ISI database. When the data is highly structured, as in the case of large journals, the maps are comparable, although one may have to vary a threshold (because of the differences in densities). In the case of interdisciplinary journals and journals in the social sciences and humanities, the new database does not add a lot to what is possible with the ISI databases.
From November 26 to December 12, 2022, Shane Wood and Kenneth Cecire, QuarkNet staff members under the University of Notre Dame, traveled to South Africa as Lecturers in the African School of Fundamental Physics and Applications (ASP) 2022, held at Nelson Mandela University in Gqeberha (Port Elizabeth) within the same calendar period. ASP is held every other year in a different African country for two or three weeks for African graduate and advanced undergraduate physics students to expose them to cutting-edge physics content and analysis techniques that may not be as available in their home institutions. Since 2016, there has been an outreach component consisting of the High School Teachers and Learners Programs. Cecire and Wood were facilitators of these programs and also acted as lecturers for two regular ASP classes. (We will use the term "student" to refer to these university students and "learners" to refer to high school students, following the practice of ASP.) They played a very active role and were quite busy during their two-week involvement. The mission was successful in terms of reaching teachers, learners, and students with new and exciting ideas and in terms of build
This report provides the first comprehensive analysis of postdoctoral research fellows (postdocs) in South African public universities. It combines an analysis of existing data with the analysis of primary data collected in the form of a survey of institutions on the postdocs they host, a bibliometric study of the research output of postdocs, and an individual survey of postdocs. The number of postdocs has been increasing steadily from 2016 to 2022 and varies across universities, with larger research-intensive universities hosting more postdocs. In terms of demographics, the proportion of black African postdocs has increased; the proportion of female postdocs has remained lower than that of males; there is an increasing proportion of older postdocs; and more than 60 percent of postdocs are foreign-born. The bibliometric analysis of the publication output of postdocs shows that it increased substantially from 2005 to 2022. Some main results of the individual survey are that a postdoc position is taken primarily to enhance prospects for employment in a permanent academic position. However, securing such positions is reported as challenging, which is supported by results that one in e
Using "Analyze Results" at the Web of Science, one can directly generate overlays onto global journal maps of science. The maps are based on the 10,000+ journals contained in the Journal Citation Reports (JCR) of the Science and Social Science Citation Indices (2011). The disciplinary diversity of the retrieval is measured in terms of Rao-Stirling's "quadratic entropy." Since this indicator of interdisciplinarity is normalized between zero and one, the interdisciplinarity can be compared among document sets and across years, cited or citing. The colors used for the overlays are based on Blondel et al.'s (2008) community-finding algorithms operating on the relations journals included in JCRs. The results can be exported from VOSViewer with different options such as proportional labels, heat maps, or cluster density maps. The maps can also be web-started and/or animated (e.g., using PowerPoint). The "citing" dimension of the aggregated journal-journal citation matrix was found to provide a more comprehensive description than the matrix based on the cited archive. The relations between local and global maps and their different functions in studying the sciences in terms of journal lit
We compare the network of aggregated journal-journal citation relations provided by the Journal Citation Reports (JCR) 2012 of the Science and Social Science Citation Indexes (SCI and SSCI) with similar data based on Scopus 2012. First, global maps were developed for the two sets separately; sets of documents can then be compared using overlays to both maps. Using fuzzy-string matching and ISSN numbers, we were able to match 10,524 journal names between the two sets; that is, 96.4% of the 10,936 journals contained in JCR or 51.2% of the 20,554 journals covered by Scopus. Network analysis was then pursued on the set of journals shared between the two databases and the two sets of unique journals. Citations among the shared journals are more comprehensively covered in JCR than Scopus, so the network in JCR is denser and more connected than in Scopus. The ranking of shared journals in terms of indegree (that is, numbers of citing journals) or total citations is similar in both databases overall (Spearman's \r{ho} > 0.97), but some individual journals rank very differently. Journals that are unique to Scopus seem to be less important--they are citing shared journals rather than bein
This article discusses a three-year study (2020 - 2022) of dominant misconceptions (DMs) for a large cohort of first-year physics course students at the University of Johannesburg (UJ), South Africa. Our study considered pre-test scores on the Force Concept Inventory using a graphical method, where we found statistical differences between the mean DM scores for the 2020 cohort, as compared to the 2021 and 2022 cohort; possibly due to the onset of COVID lockdowns. We also compared our data from South Africa with cohorts based in Spain and the Kingdom of Saudi Arabia, where the method of DMs was also applied. From this comparison, we found some differences in the preconception knowledge of the cohorts. Furthermore, we included an analysis of DMs through the `gender lens' for the South African cohort, finding no statistically significant difference between the means for DM scores of students who identify as male or female. Finally, given the diverse language backgrounds and levels of matriculation preparation for university level physics courses, we have also shown how quickly responding to student misconceptions can be efficiently addressed using the method of DMs.
It has long been known that large grounded conducting networks on the surface of Earth are affected by solar activity and geomagnetic storms. Power networks are such extensive grounded conductors and are susceptible to geomagnetically induced currents (GICs). GICs at any specific node in a power network are assumed to be linearly related to the horizontal vector components of an induced plane-wave geoelectric field by a pair of network parameters. These network parameters are not easily measured in the network, but may be estimated empirically. In this work, we present a new approach of using an ensemble of network parameters estimates. The ensembles include a huge number of parameter pair estimates calculated from simultaneously solving pairs of time instances of the governing GIC equation. Each individual estimate is not the true state of the system, but a possible state. Taking the ensemble as a whole though gives the most probable parameter estimate. The most probable parameter estimate for both network parameters, as defined by their respective ensembles, is used directly in the modelling of GICs. The ensembles themselves however allow for further analysis into the nature of G
Literature has shown that countries such as Brazil and India have successfully implemented electronic voting systems and other countries are at various piloting stages to address many challenges associated with manual paper based system such ascosts of physical ballot paper and other overheads, electoral delays, distribution of electoral materials, and general lack of confidence in the electoral process. It is in this context that this study explores how South African can leverage the opportunities that e-voting presents. Manual voting is often tedious, non-secure, and time-consuming, which leads us to think about using electronic facilities to make the process more efficient. This study proposes that the adoption of electronic voting technologies could perhaps mitigate some of these issues and challengesin the process improving the electoral process. The study used an on-line questionnaire which was administered to a broader group of voters and an in-depth semi-structured interview with the Independent Electoral Commission officials. The analysis is based on thematic analysis and diffusion of innovations theory is adopted as a theoretical lens of analysis. The findings reveal that
This paper introduces two multilingual government themed corpora in various South African languages. The corpora were collected by gathering the South African Government newspaper (Vuk'uzenzele), as well as South African government speeches (ZA-gov-multilingual), that are translated into all 11 South African official languages. The corpora can be used for a myriad of downstream NLP tasks. The corpora were created to allow researchers to study the language used in South African government publications, with a focus on understanding how South African government officials communicate with their constituents. In this paper we highlight the process of gathering, cleaning and making available the corpora. We create parallel sentence corpora for Neural Machine Translation (NMT) tasks using Language-Agnostic Sentence Representations (LASER) embeddings. With these aligned sentences we then provide NMT benchmarks for 9 indigenous languages by fine-tuning a massively multilingual pre-trained language model.
Using three years of the Journal Citation Reports (2011, 2012, and 2013), indicators of transitions in 2012 (between 2011 and 2013) are studied using methodologies based on entropy statistics. Changes can be indicated at the level of journals using the margin totals of entropy production along the row or column vectors, but also at the level of links among journals by importing the transition matrices into network analysis and visualization programs (and using community-finding algorithms). Seventy-four journals are flagged in terms of discontinuous changes in their citations; but 3,114 journals are involved in "hot" links. Most of these links are embedded in a main component; 78 clusters (containing 172 journals) are flagged as potential "hot spots" emerging at the network level. An additional finding is that PLoS ONE introduced a new communication dynamics into the database. The limitations of the methodology are elaborated using an example. The results of the study indicate where developments in the citation dynamics can be considered as significantly unexpected. This can be used as heuristic information; but what a "hot spot" in terms of the entropy statistics of aggregated cit
With the constant spread of misinformation on social media networks, a need has arisen to continuously assess the veracity of digital content. This need has inspired numerous research efforts on the development of misinformation detection (MD) models. However, many models do not use all information available to them and existing research contains a lack of relevant datasets to train the models, specifically within the South African social media environment. The aim of this paper is to investigate the transferability of knowledge of a MD model between different contextual environments. This research contributes a multimodal MD model capable of functioning in the South African social media environment, as well as introduces a South African misinformation dataset. The model makes use of multiple sources of information for misinformation detection, namely: textual and visual elements. It uses bidirectional encoder representations from transformers (BERT) as the textual encoder and a residual network (ResNet) as the visual encoder. The model is trained and evaluated on the Fakeddit dataset and a South African misinformation dataset. Results show that using South African samples in the t
Very few social media studies have been done on South African user-generated content during the COVID-19 pandemic and even fewer using hand-labelling over automated methods. Vaccination is a major tool in the fight against the pandemic, but vaccine hesitancy jeopardizes any public health effort. In this study, sentiment analysis on South African tweets related to vaccine hesitancy was performed, with the aim of training AI-mediated classification models and assessing their reliability in categorizing UGC. A dataset of 30000 tweets from South Africa were extracted and hand-labelled into one of three sentiment classes: positive, negative, neutral. The machine learning models used were LSTM, bi-LSTM, SVM, BERT-base-cased and the RoBERTa-base models, whereby their hyperparameters were carefully chosen and tuned using the WandB platform. We used two different approaches when we pre-processed our data for comparison: one was semantics-based, while the other was corpus-based. The pre-processing of the tweets in our dataset was performed using both methods, respectively. All models were found to have low F1-scores within a range of 45$\%$-55$\%$, except for BERT and RoBERTa which both achi
We present the first systematic exploration of earth tides-seismicity correlation in northwestern South America, with a special emphasis in Colombia. For this purpose, we use a dataset of ~167,000 earthquakes, gathered by the Colombian Seismological Network between 1993 and 2017. Most of the events are intermediate-depth earthquakes from the Bucaramanga seismic nest and the Cauca seismic cluster. For this purpose, we implemented a novel approach for the calculation of tidal phases that considers the relative positions of the Earth-Moon-Sun system at the time of the events. After applying the standard Schuster test to the whole dataset and to several earthquake samples (classified by time, location, magnitude and depth), we found strong correlation anomalies with the diurnal and monthly components of the tide (global log(p) values around -7.0 for the diurnal constituent and -12.1 for the monthly constituent), especially for the intermediate depth events. These anomalies suggest that around 16% of the deep earthquakes in Colombia may be triggered by tides, especially when the monthly phase is between 350$^\circ$-10$^\circ$. We attribute our positive results, which favor the tidal-tri
Global expansion of the Event Horizon Telescope (EHT) will see the strategic addition of antennas at new geographical locations, transforming the sensitivity and imaging fidelity of the $λ\sim 1$\,mm EHT array. A possible South African EHT station would leverage a strong geographical advantage, local infrastructure, and radio astronomy expertise, and have strong synergies with the Africa Millimetre Telescope in Namibia. We assessed three South African candidate millimetre sites using climatological simulations and antenna sensitivity estimates, and found at least two promising sites. These sites are comparable to some existing EHT stations during the typical April EHT observing window and outperform them during most of the year, especially the southern hemisphere winter. The results suggest that a strategically placed South African EHT station will have a sizable, positive impact on next-generation EHT objectives and the resulting black hole imaging science.
Rankings of scholarly journals based on citation data are often met with skepticism by the scientific community. Part of the skepticism is due to disparity between the common perception of journals' prestige and their ranking based on citation counts. A more serious concern is the inappropriate use of journal rankings to evaluate the scientific influence of authors. This paper focuses on analysis of the table of cross-citations among a selection of Statistics journals. Data are collected from the Web of Science database published by Thomson Reuters. Our results suggest that modelling the exchange of citations between journals is useful to highlight the most prestigious journals, but also that journal citation data are characterized by considerable heterogeneity, which needs to be properly summarized. Inferential conclusions require care in order to avoid potential over-interpretation of insignificant differences between journal ratings. Comparison with published ratings of institutions from the UK's Research Assessment Exercise shows strong correlation at aggregate level between assessed research quality and journal citation `export scores' within the discipline of Statistics.
A number of journal classification systems have been developed in bibliometrics since the launch of the Citation Indices by the Institute of Scientific Information (ISI) in the 1960s. These systems are used to normalize citation counts with respect to field-specific citation patterns. The best known system is the so-called "Web-of-Science Subject Categories" (WCs). In other systems papers are classified by algorithmic solutions. Using the Journal Citation Reports 2014 of the Science Citation Index and the Social Science Citation Index (n of journals = 11,149), we examine options for developing a new system based on journal classifications into subject categories using aggregated journal-journal citation data. Combining routines in VOSviewer and Pajek, a tree-like classification is developed. At each level one can generate a map of science for all the journals subsumed under a category. Nine major fields are distinguished at the top level. Further decomposition of the social sciences is pursued for the sake of example with a focus on journals in information science (LIS) and science studies (STS). The new classification system improves on alternative options by avoiding the problem
Publication patterns of 79 forest scientists awarded major international forestry prizes during 1990-2010 were compared with the journal classification and ranking promoted as part of the 'Excellence in Research for Australia' (ERA) by the Australian Research Council. The data revealed that these scientists exhibited an elite publication performance during the decade before and two decades following their first major award. An analysis of their 1703 articles in 431 journals revealed substantial differences between the journal choices of these elite scientists and the ERA classification and ranking of journals. Implications from these findings are that additional cross-classifications should be added for many journals, and there should be an adjustment to the ranking of several journals relevant to the ERA Field of Research classified as 0705 Forestry Sciences.