Matching Items (26)
Filtering by

Clear all filters

141461-Thumbnail Image.png
Description
In the digital humanities, there is a constant need to turn images and PDF files into plain text to apply analyses such as topic modelling, named entity recognition, and other techniques. However, although there exist different solutions to extract text embedded in PDF files or run OCR on images, they

In the digital humanities, there is a constant need to turn images and PDF files into plain text to apply analyses such as topic modelling, named entity recognition, and other techniques. However, although there exist different solutions to extract text embedded in PDF files or run OCR on images, they typically require additional training (for example, scholars have to learn how to use the command line) or are difficult to automate without programming skills. The Giles Ecosystem is a distributed system based on Apache Kafka that allows users to upload documents for text and image extraction. The system components are implemented using Java and the Spring Framework and are available under an Open Source license on GitHub (https://github.com/diging/).
ContributorsLessios-Damerow, Julia (Contributor) / Peirson, Erick (Contributor) / Laubichler, Manfred (Contributor) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2017-09-28
Description

On-going efforts to understand the dynamics of coupled social-ecological (or more broadly, coupled infrastructure) systems and common pool resources have led to the generation of numerous datasets based on a large number of case studies. This data has facilitated the identification of important factors and fundamental principles which increase our

On-going efforts to understand the dynamics of coupled social-ecological (or more broadly, coupled infrastructure) systems and common pool resources have led to the generation of numerous datasets based on a large number of case studies. This data has facilitated the identification of important factors and fundamental principles which increase our understanding of such complex systems. However, the data at our disposal are often not easily comparable, have limited scope and scale, and are based on disparate underlying frameworks inhibiting synthesis, meta-analysis, and the validation of findings. Research efforts are further hampered when case inclusion criteria, variable definitions, coding schema, and inter-coder reliability testing are not made explicit in the presentation of research and shared among the research community. This paper first outlines challenges experienced by researchers engaged in a large-scale coding project; then highlights valuable lessons learned; and finally discusses opportunities for further research on comparative case study analysis focusing on social-ecological systems and common pool resources. Includes supplemental materials and appendices published in the International Journal of the Commons 2016 Special Issue. Volume 10 - Issue 2 - 2016.

ContributorsRatajczyk, Elicia (Author) / Brady, Ute (Author) / Baggio, Jacopo (Author) / Barnett, Allain J. (Author) / Perez Ibarra, Irene (Author) / Rollins, Nathan (Author) / Rubinos, Cathy (Author) / Shin, Hoon Cheol (Author) / Yu, David (Author) / Aggarwal, Rimjhim (Author) / Anderies, John (Author) / Janssen, Marco (Author) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2016-09-09
128166-Thumbnail Image.png
Description

At the end of the dark ages, anatomy was taught as though everything that could be known was known. Scholars learned about what had been discovered rather than how to make discoveries. This was true even though the body (and the rest of biology) was very poorly understood. The renaissance

At the end of the dark ages, anatomy was taught as though everything that could be known was known. Scholars learned about what had been discovered rather than how to make discoveries. This was true even though the body (and the rest of biology) was very poorly understood. The renaissance eventually brought a revolution in how scholars (and graduate students) were trained and worked. This revolution never occurred in K-12 or university education such that we now teach young students in much the way that scholars were taught in the dark ages, we teach them what is already known rather than the process of knowing. Citizen science offers a way to change K-12 and university education and, in doing so, complete the renaissance. Here we offer an example of such an approach and call for change in the way students are taught science, change that is more possible than it has ever been and is, nonetheless, five hundred years delayed.

Created2016-03-01
127872-Thumbnail Image.png
Description

Background: Modern advances in sequencing technology have enabled the census of microbial members of many natural ecosystems. Recently, attention is increasingly being paid to the microbial residents of human-made, built ecosystems, both private (homes) and public (subways, office buildings, and hospitals). Here, we report results of the characterization of the microbial

Background: Modern advances in sequencing technology have enabled the census of microbial members of many natural ecosystems. Recently, attention is increasingly being paid to the microbial residents of human-made, built ecosystems, both private (homes) and public (subways, office buildings, and hospitals). Here, we report results of the characterization of the microbial ecology of a singular built environment, the International Space Station (ISS). This ISS sampling involved the collection and microbial analysis (via 16S rRNA gene PCR) of 15 surfaces sampled by swabs onboard the ISS. This sampling was a component of Project MERCCURI (Microbial Ecology Research Combining Citizen and University Researchers on ISS). Learning more about the microbial inhabitants of the “buildings” in which we travel through space will take on increasing importance, as plans for human exploration continue, with the possibility of colonization of other planets and moons.

Results: Sterile swabs were used to sample 15 surfaces onboard the ISS. The sites sampled were designed to be analogous to samples collected for (1) the Wildlife of Our Homes project and (2) a study of cell phones and shoes that were concurrently being collected for another component of Project MERCCURI. Sequencing of the 16S rRNA genes amplified from DNA extracted from each swab was used to produce a census of the microbes present on each surface sampled. We compared the microbes found on the ISS swabs to those from both homes on Earth and data from the Human Microbiome Project.

Conclusions: While significantly different from homes on Earth and the Human Microbiome Project samples analyzed here, the microbial community composition on the ISS was more similar to home surfaces than to the human microbiome samples. The ISS surfaces are OTU-rich with 1,036–4,294 operational taxonomic units (OTUs per sample). There was no discernible biogeography of microbes on the 15 ISS surfaces, although this may be a reflection of the small sample size we were able to obtain.

ContributorsLang, Jenna M. (Author) / Coil, David A. (Author) / Neches, Russell Y. (Author) / Brown, Wendy E. (Author) / Cavalier, Darlene (Author) / Severance, Mark (Author) / Hampton-Marcell, Jarrad T. (Author) / Gilbert, Jack A. (Author) / Eisen, Jonathan A. (Author) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2017-12-05
128562-Thumbnail Image.png
Description

We find that the flow of attention on the Web forms a directed, tree-like structure implying the time-sensitive browsing behavior of users. Using the data of a news sharing website, we construct clickstream networks in which nodes are news stories and edges represent the consecutive clicks between two stories. To

We find that the flow of attention on the Web forms a directed, tree-like structure implying the time-sensitive browsing behavior of users. Using the data of a news sharing website, we construct clickstream networks in which nodes are news stories and edges represent the consecutive clicks between two stories. To identify the flow direction of clickstreams, we define the “flow distance” of nodes (Li), which measures the average number of steps a random walker takes to reach the ith node. It is observed that Li is related with the clicks (Ci) to news stories and the age (Ti) of stories. Putting these three variables together help us understand the rise and decay of news stories from a network perspective. We also find that the studied clickstream networks preserve a stable structure over time, leading to the scaling between users and clicks. The universal scaling behavior is confirmed by the 1,000 Web forums. We suggest that the tree-like, stable structure of clickstream networks reveals the time-sensitive preference of users in online browsing. To test our assumption, we discuss three models on individual browsing behavior, and compare the simulation results with empirical data.

ContributorsWang, Cheng-Jun (Author) / Wu, Lingfei (Author) / Zhang, Jiang (Author) / Janssen, Marco (Author) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2016-09-28
128559-Thumbnail Image.png
Description

Many adaptive systems sit near a tipping or critical point. For systems near a critical point small changes to component behaviour can induce large-scale changes in aggregate structure and function. Criticality can be adaptive when the environment is changing, but entails reduced robustness through sensitivity. This tradeoff can be resolved

Many adaptive systems sit near a tipping or critical point. For systems near a critical point small changes to component behaviour can induce large-scale changes in aggregate structure and function. Criticality can be adaptive when the environment is changing, but entails reduced robustness through sensitivity. This tradeoff can be resolved when criticality can be tuned. We address the control of finite measures of criticality using data on fight sizes from an animal society model system (Macaca nemestrina, n=48). We find that a heterogeneous, socially organized system, like homogeneous, spatial systems (flocks and schools), sits near a critical point; the contributions individuals make to collective phenomena can be quantified; there is heterogeneity in these contributions; and distance from the critical point (DFC) can be controlled through biologically plausible mechanisms exploiting heterogeneity. We propose two alternative hypotheses for why a system decreases the distance from the critical point.

ContributorsDaniels, Bryan (Author) / Krakauer, David (Author) / Flack, Jessica (Author) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2017-02-10
128540-Thumbnail Image.png
Description

Maximally random jammed (MRJ) particle packings can be viewed as prototypical glasses in that they are maximally disordered while simultaneously being mechanically rigid. The prediction of the MRJ packing density ϕMRJ, among other packing properties of frictionless particles, still poses many theoretical challenges, even for congruent spheres or disks. Using

Maximally random jammed (MRJ) particle packings can be viewed as prototypical glasses in that they are maximally disordered while simultaneously being mechanically rigid. The prediction of the MRJ packing density ϕMRJ, among other packing properties of frictionless particles, still poses many theoretical challenges, even for congruent spheres or disks. Using the geometric-structure approach, we derive for the first time a highly accurate formula for MRJ densities for a very wide class of two-dimensional frictionless packings, namely, binary convex superdisks, with shapes that continuously interpolate between circles and squares. By incorporating specific attributes of MRJ states and a novel organizing principle, our formula yields predictions of ϕMRJ that are in excellent agreement with corresponding computer-simulation estimates in almost the entire α-x plane with semi-axis ratio α and small-particle relative number concentration x. Importantly, in the monodisperse circle limit, the predicted ϕMRJ = 0.834 agrees very well with the very recently numerically discovered MRJ density of 0.827, which distinguishes it from high-density “random-close packing” polycrystalline states and hence provides a stringent test on the theory. Similarly, for non-circular monodisperse superdisks, we predict MRJ states with densities that are appreciably smaller than is conventionally thought to be achievable by standard packing protocols.

ContributorsTian, Jianxiang (Author) / Xu, Yaopengxiao (Author) / Jiao, Yang (Author) / Torquato, Salvatore (Author) / Ira A. Fulton Schools of Engineering (Contributor)
Created2015-11-16
129577-Thumbnail Image.png
Description

Numerous recent investigations have been devoted to the determination of the equilibrium phase behavior and packing characteristics of hard nonspherical particles, including ellipsoids, superballs, and polyhedra, to name but just a few shapes. Systems of hard nonspherical particles exhibit a variety of stable phases with different degrees of translational and

Numerous recent investigations have been devoted to the determination of the equilibrium phase behavior and packing characteristics of hard nonspherical particles, including ellipsoids, superballs, and polyhedra, to name but just a few shapes. Systems of hard nonspherical particles exhibit a variety of stable phases with different degrees of translational and orientational order, including isotropic liquid, solid crystal, rotator and a variety of liquid crystal phases. In this paper, we employ a Monte Carlo implementation of the adaptive-shrinking-cell (ASC) numerical scheme and free-energy calculations to ascertain with high precision the equilibrium phase behavior of systems of congruent Archimedean truncated tetrahedra over the entire range of possible densities up to the maximal nearly space-filling density. In particular, we find that the system undergoes two first-order phase transitions as the density increases: first a liquid–solid transition and then a solid–solid transition. The isotropic liquid phase coexists with the Conway–Torquato (CT) crystal phase at intermediate densities, verifying the result of a previous qualitative study [J. Chem. Phys. 2011, 135, 151101]. The freezing- and melting-point packing fractions for this transition are respectively ϕF = 0.496 ± 0.006 and ϕM = 0.591 ± 0.005.

At higher densities, we find that the CT phase undergoes another first-order phase transition to one associated with the densest-known crystal, with coexistence densities in the range ϕ ∈ [0.780 ± 0.002, 0.802 ± 0.003]. We find no evidence for stable rotator (or plastic) or nematic phases. We also generate the maximally random jammed (MRJ) packings of truncated tetrahedra, which may be regarded to be the glassy end state of a rapid compression of the liquid. Specifically, we systematically study the structural characteristics of the MRJ packings, including the centroidal pair correlation function, structure factor and orientational pair correlation function. We find that such MRJ packings are hyperuniform with an average packing fraction of 0.770, which is considerably larger than the corresponding value for identical spheres (≈ 0.64). We conclude with some simple observations concerning what types of phase transitions might be expected in general hard-particle systems based on the particle shape and which would be good glass formers.

ContributorsChen, Duyu (Author) / Jiao, Yang (Author) / Torquato, Salvatore (Author) / Ira A. Fulton Schools of Engineering (Contributor)
Created2014-07-17
129572-Thumbnail Image.png
Description

X-ray tomography has provided a non-destructive means for microstructure characterization in three and four dimensions. A stochastic procedure to accurately reconstruct material microstructure from limited-angle X-ray tomographic projections is presented and its utility is demonstrated by reconstructing a variety of distinct heterogeneous materials and elucidating the information content of different

X-ray tomography has provided a non-destructive means for microstructure characterization in three and four dimensions. A stochastic procedure to accurately reconstruct material microstructure from limited-angle X-ray tomographic projections is presented and its utility is demonstrated by reconstructing a variety of distinct heterogeneous materials and elucidating the information content of different projection data sets. A small number of projections (e.g. 20–40) are necessary for accurate reconstructions via the stochastic procedure, indicating its high efficiency in using limited structural information.

ContributorsLi, Hechao (Author) / Chawla, Nikhilesh (Author) / Jiao, Yang (Author) / Ira A. Fulton Schools of Engineering (Contributor)
Created2014-09-01
128744-Thumbnail Image.png
Description

Sequential affect dynamics generated during the interaction of intimate dyads, such as married couples, are associated with a cascade of effects - some good and some bad - on each partner, close family members, and other social contacts. Although the effects are well documented, the probabilistic structures associated with micro-social

Sequential affect dynamics generated during the interaction of intimate dyads, such as married couples, are associated with a cascade of effects - some good and some bad - on each partner, close family members, and other social contacts. Although the effects are well documented, the probabilistic structures associated with micro-social processes connected to the varied outcomes remain enigmatic. Using extant data we developed a method of classifying and subsequently generating couple dynamics using a Hierarchical Dirichlet Process Hidden semi-Markov Model (HDP-HSMM). Our findings indicate that several key aspects of existing models of marital interaction are inadequate: affect state emissions and their durations, along with the expected variability differences between distressed and nondistressed couples are present but highly nuanced; and most surprisingly, heterogeneity among highly satisfied couples necessitate that they be divided into subgroups. We review how this unsupervised learning technique generates plausible dyadic sequences that are sensitive to relationship quality and provide a natural mechanism for computational models of behavioral and affective micro-social processes.

ContributorsGriffin, William (Author) / Li, Xun (Author) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2016-05-17