Matching Items (34)
141461-Thumbnail Image.png
Description
In the digital humanities, there is a constant need to turn images and PDF files into plain text to apply analyses such as topic modelling, named entity recognition, and other techniques. However, although there exist different solutions to extract text embedded in PDF files or run OCR on images, they

In the digital humanities, there is a constant need to turn images and PDF files into plain text to apply analyses such as topic modelling, named entity recognition, and other techniques. However, although there exist different solutions to extract text embedded in PDF files or run OCR on images, they typically require additional training (for example, scholars have to learn how to use the command line) or are difficult to automate without programming skills. The Giles Ecosystem is a distributed system based on Apache Kafka that allows users to upload documents for text and image extraction. The system components are implemented using Java and the Spring Framework and are available under an Open Source license on GitHub (https://github.com/diging/).
ContributorsLessios-Damerow, Julia (Contributor) / Peirson, Erick (Contributor) / Laubichler, Manfred (Contributor) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2017-09-28
130413-Thumbnail Image.png
Description
Because collective cognition emerges from local signaling among group members, deciphering communication systems is crucial to understanding the underlying mechanisms. Alarm signals are widespread in the social insects and can elicit a variety of behavioral responses to danger, but the functional plasticity of these signals has not been well studied.

Because collective cognition emerges from local signaling among group members, deciphering communication systems is crucial to understanding the underlying mechanisms. Alarm signals are widespread in the social insects and can elicit a variety of behavioral responses to danger, but the functional plasticity of these signals has not been well studied. Here we report an alarm pheromone in the ant Temnothorax rugatulus that elicits two different behaviors depending on context. When an ant was tethered inside an unfamiliar nest site and unable to move freely, she released a pheromone from her mandibular gland that signaled other ants to reject this nest as a potential new home, presumably to avoid potential danger. When the same pheromone was presented near the ants' home nest, they were instead attracted to it, presumably to respond to a threat to the colony. We used coupled gas chromatography/mass spectrometry to identify candidate compounds from the mandibular gland and tested each one in a nest choice bioassay. We found that 2,5-dimethylpyrazine was sufficient to induce rejection of a marked new nest and also to attract ants when released at the home nest. This is the first detailed investigation of chemical communication in the leptothoracine ants. We discuss the possibility that this pheromone's deterrent function can improve an emigrating colony's nest site selection performance.
Created2014-09-01
134136-Thumbnail Image.png
Description
Biomarkers are the cornerstone of modern-day medicine. They are defined as any biological substance in or outside the body that gives insight to the body's condition. Doctors and researchers can measure specific biomarkers to diagnose and treat patients, such as the concentration of hemoglobin Alc and its connection to diabetes.

Biomarkers are the cornerstone of modern-day medicine. They are defined as any biological substance in or outside the body that gives insight to the body's condition. Doctors and researchers can measure specific biomarkers to diagnose and treat patients, such as the concentration of hemoglobin Alc and its connection to diabetes. There are a variety of methods, or assays, to detect biomarkers, but the most common assay is enzyme-linked immunosorbent assay (ELISA). A new-generation assay termed mass spectrometric immunoassay (MSIA) can measure proteoforms, the different chemical variations of proteins, and their relative abundance. ELISA on the other hand measures the overall concentration of protein in the sample. Measuring each of the proteoforms of a protein is important because only one or two variations could be biologically significant and/or cause diseases. However, running MSIA is expensive. For this reason, an alternative plate-based MSIA technique was tested for its ability to detect the proteoforms of a protein called apolipoprotein C-III (ApoC-III). This technique combines the protein capturing procedure of ELISA to isolate the protein with detection in a mass spectrometer. A larger amount of ApoC-III present in the body indicates a considerable risk for coronary heart disease. The precision of the assay is determined on the coefficient of variation (CV). A CV value is the ratio of standard deviation in relation to the mean, represented as a percentage. The smaller the percentage, the less variation the assay has, and therefore the more ability it has to detect subtle changes in the biomarker. An accepted CV would be less than 10% for single-day tests (intra-day) and less than 15% for multi-day tests (inter-day). The plate-based MSIA was started by first coating a 96-well round bottom plate with 2.5 micrograms of ApoC-III antibody. Next, a series of steps were conducted: a buffer wash, then the sample incubation, followed by another buffer wash and two consecutive water washes. After the final wash, the wells were filled with a MALDI matrix, then spotted onto a gold plate to dry. The dry gold target was then placed into a MALDI-TOF mass spectrometer to produce mass spectra for each spot. The mass spectra were calibrated and the area underneath each of the four peaks representing the ApoC-III proteoforms was exported as an Excel file. The intra-day CV values were found by dividing the standard deviation by the average relative abundance of each peak. After repeating the same procedure for three more days, the inter-day CVs were found using the same method. After completing the experiment, the CV values were all within the acceptable guidelines. Therefore, the plate-based MSIA is a viable alternative for finding proteoforms than the more expensive MSIA tips. To further validate this, additional tests will need to be conducted with different proteins and number of samples to determine assay flexibility.
ContributorsTieu, Luc (Author) / Borges, Chad (Thesis director) / Nedelkov, Dobrin (Committee member) / Harrington Bioengineering Program (Contributor) / Barrett, The Honors College (Contributor)
Created2017-12
129567-Thumbnail Image.png
Description

Human protein diversity arises as a result of alternative splicing, single nucleotide polymorphisms (SNPs) and posttranslational modifications. Because of these processes, each protein can exists as multiple variants in vivo. Tailored strategies are needed to study these protein variants and understand their role in health and disease. In this work

Human protein diversity arises as a result of alternative splicing, single nucleotide polymorphisms (SNPs) and posttranslational modifications. Because of these processes, each protein can exists as multiple variants in vivo. Tailored strategies are needed to study these protein variants and understand their role in health and disease. In this work we utilized quantitative mass spectrometric immunoassays to determine the protein variants concentration of beta-2-microglobulin, cystatin C, retinol binding protein, and transthyretin, in a population of 500 healthy individuals. Additionally, we determined the longitudinal concentration changes for the protein variants from four individuals over a 6 month period. Along with the native forms of the four proteins, 13 posttranslationally modified variants and 7 SNP-derived variants were detected and their concentration determined. Correlations of the variants concentration with geographical origin, gender, and age of the individuals were also examined. This work represents an important step toward building a catalog of protein variants concentrations and examining their longitudinal changes.

ContributorsTrenchevska, Olgica (Author) / Phillips, David A. (Author) / Nelson, Randall (Author) / Nedelkov, Dobrin (Author) / Biodesign Institute (Contributor)
Created2014-06-23
129259-Thumbnail Image.png
Description

What's a profession without a code of ethics? Being a legitimate profession almost requires drafting a code and, at least nominally, making members follow it. Codes of ethics (henceforth “codes”) exist for a number of reasons, many of which can vary widely from profession to profession - but above all

What's a profession without a code of ethics? Being a legitimate profession almost requires drafting a code and, at least nominally, making members follow it. Codes of ethics (henceforth “codes”) exist for a number of reasons, many of which can vary widely from profession to profession - but above all they are a form of codified self-regulation. While codes can be beneficial, it argues that when we scratch below the surface, there are many problems at their root. In terms of efficacy, codes can serve as a form of ethical window dressing, rather than effective rules for behavior. But even more that, codes can degrade the meaning behind being a good person who acts ethically for the right reasons.

Created2013-11-30
128778-Thumbnail Image.png
Description

Online communities are becoming increasingly important as platforms for large-scale human cooperation. These communities allow users seeking and sharing professional skills to solve problems collaboratively. To investigate how users cooperate to complete a large number of knowledge-producing tasks, we analyze Stack Exchange, one of the largest question and answer systems

Online communities are becoming increasingly important as platforms for large-scale human cooperation. These communities allow users seeking and sharing professional skills to solve problems collaboratively. To investigate how users cooperate to complete a large number of knowledge-producing tasks, we analyze Stack Exchange, one of the largest question and answer systems in the world. We construct attention networks to model the growth of 110 communities in the Stack Exchange system and quantify individual answering strategies using the linking dynamics on attention networks. We identify two answering strategies. Strategy A aims at performing maintenance by doing simple tasks, whereas strategy B aims at investing time in doing challenging tasks. Both strategies are important: empirical evidence shows that strategy A decreases the median waiting time for answers and strategy B increases the acceptance rate of answers. In investigating the strategic persistence of users, we find that users tends to stick on the same strategy over time in a community, but switch from one strategy to the other across communities. This finding reveals the different sets of knowledge and skills between users. A balance between the population of users taking A and B strategies that approximates 2:1, is found to be optimal to the sustainable growth of communities.

ContributorsWu, Lingfei (Author) / Baggio, Jacopo (Author) / Janssen, Marco (Author) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2016-03-02
128687-Thumbnail Image.png
Description

Proteins can exist as multiple proteoforms in vivo, as a result of alternative splicing and single-nucleotide polymorphisms (SNPs), as well as posttranslational processing. To address their clinical significance in a context of diagnostic information, proteoforms require a more in-depth analysis. Mass spectrometric immunoassays (MSIA) have been devised for studying structural

Proteins can exist as multiple proteoforms in vivo, as a result of alternative splicing and single-nucleotide polymorphisms (SNPs), as well as posttranslational processing. To address their clinical significance in a context of diagnostic information, proteoforms require a more in-depth analysis. Mass spectrometric immunoassays (MSIA) have been devised for studying structural diversity in human proteins. MSIA enables protein profiling in a simple and high-throughput manner, by combining the selectivity of targeted immunoassays, with the specificity of mass spectrometric detection. MSIA has been used for qualitative and quantitative analysis of single and multiple proteoforms, distinguishing between normal fluctuations and changes related to clinical conditions. This mini review offers an overview of the development and application of mass spectrometric immunoassays for clinical and population proteomics studies. Provided are examples of some recent developments, and also discussed are the trends and challenges in mass spectrometry-based immunoassays for the next-phase of clinical applications.

ContributorsTrenchevska, Olgica (Author) / Nelson, Randall (Author) / Nedelkov, Dobrin (Author) / Biodesign Institute (Contributor)
Created2016-03-17
128933-Thumbnail Image.png
Description

Introduction: Apolipoprotein C-III (apoC-III) regulates triglyceride (TG) metabolism. In plasma, apoC-III exists in non-sialylated (apoC-III0a without glycosylation and apoC-III[subscript 0b] with glycosylation), monosialylated (apoC-III1) or disialylated (apoC-III2) proteoforms. Our aim was to clarify the relationship between apoC-III sialylation proteoforms with fasting plasma TG concentrations.

Methods: In 204 non-diabetic adolescent participants, the

Introduction: Apolipoprotein C-III (apoC-III) regulates triglyceride (TG) metabolism. In plasma, apoC-III exists in non-sialylated (apoC-III0a without glycosylation and apoC-III[subscript 0b] with glycosylation), monosialylated (apoC-III1) or disialylated (apoC-III2) proteoforms. Our aim was to clarify the relationship between apoC-III sialylation proteoforms with fasting plasma TG concentrations.

Methods: In 204 non-diabetic adolescent participants, the relative abundance of apoC-III plasma proteoforms was measured using mass spectrometric immunoassay.

Results: Compared with the healthy weight subgroup (n = 16), the ratios of apoC-III0a, apoC-III0b, and apoC-III1 to apoC-III2 were significantly greater in overweight (n = 33) and obese participants (n = 155). These ratios were positively correlated with BMI z-scores and negatively correlated with measures of insulin sensitivity (S[subscript i]). The relationship of apoC-III1 / apoC-III2 with Si persisted after adjusting for BMI (p = 0.02). Fasting TG was correlated with the ratio of apoC-III0a / apoC-III2 (r = 0.47, p<0.001), apoC-III0b / apoC-III2 (r = 0.41, p<0.001), apoC-III1 / apoC-III2 (r = 0.43, p<0.001). By examining apoC-III concentrations, the association of apoC-III proteoforms with TG was driven by apoC-III0a (r = 0.57, p<0.001), apoC-III0b (r = 0.56. p<0.001) and apoC-III1 (r = 0.67, p<0.001), but not apoC-III2 (r = 0.006, p = 0.9) concentrations, indicating that apoC-III relationship with plasma TG differed in apoC-III2 compared with the other proteoforms.

Conclusion: We conclude that apoC-III0a, apoC-III0b, and apoC-III1, but not apoC-III2 appear to be under metabolic control and associate with fasting plasma TG. Measurement of apoC-III proteoforms can offer insights into the biology of TG metabolism in obesity.

ContributorsYassine, Hussein N. (Author) / Trenchevska, Olgica (Author) / Ramrakhiani, Ambika (Author) / Parekh, Aarushi (Author) / Koska, Juraj (Author) / Walker, Ryan W. (Author) / Billheimer, Dean (Author) / Reaven, Peter D. (Author) / Yen, Frances T. (Author) / Nelson, Randall (Author) / Goran, Michael I. (Author) / Nedelkov, Dobrin (Author) / Biodesign Institute (Contributor)
Created2015-12-03
Description

On-going efforts to understand the dynamics of coupled social-ecological (or more broadly, coupled infrastructure) systems and common pool resources have led to the generation of numerous datasets based on a large number of case studies. This data has facilitated the identification of important factors and fundamental principles which increase our

On-going efforts to understand the dynamics of coupled social-ecological (or more broadly, coupled infrastructure) systems and common pool resources have led to the generation of numerous datasets based on a large number of case studies. This data has facilitated the identification of important factors and fundamental principles which increase our understanding of such complex systems. However, the data at our disposal are often not easily comparable, have limited scope and scale, and are based on disparate underlying frameworks inhibiting synthesis, meta-analysis, and the validation of findings. Research efforts are further hampered when case inclusion criteria, variable definitions, coding schema, and inter-coder reliability testing are not made explicit in the presentation of research and shared among the research community. This paper first outlines challenges experienced by researchers engaged in a large-scale coding project; then highlights valuable lessons learned; and finally discusses opportunities for further research on comparative case study analysis focusing on social-ecological systems and common pool resources. Includes supplemental materials and appendices published in the International Journal of the Commons 2016 Special Issue. Volume 10 - Issue 2 - 2016.

ContributorsRatajczyk, Elicia (Author) / Brady, Ute (Author) / Baggio, Jacopo (Author) / Barnett, Allain J. (Author) / Perez Ibarra, Irene (Author) / Rollins, Nathan (Author) / Rubinos, Cathy (Author) / Shin, Hoon Cheol (Author) / Yu, David (Author) / Aggarwal, Rimjhim (Author) / Anderies, John (Author) / Janssen, Marco (Author) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2016-09-09
129155-Thumbnail Image.png
Description

The impetus for discovery and evaluation of protein biomarkers has been accelerated by recent development of advanced technologies for rapid and broad proteome analyses. Mass spectrometry (MS)-based protein assays hold great potential for in vitro biomarker studies. Described here is the development of a multiplex mass spectrometric immunoassay (MSIA) for

The impetus for discovery and evaluation of protein biomarkers has been accelerated by recent development of advanced technologies for rapid and broad proteome analyses. Mass spectrometry (MS)-based protein assays hold great potential for in vitro biomarker studies. Described here is the development of a multiplex mass spectrometric immunoassay (MSIA) for quantification of apolipoprotein C-I (apoC-I), apolipoprotein C-II (apoC-II), apolipoprotein C-III (apoC-III) and their proteoforms. The multiplex MSIA assay was fast (∼40 min) and high-throughput (96 samples at a time). The assay was applied to a small cohort of human plasma samples, revealing the existence of multiple proteoforms for each apolipoprotein C. The quantitative aspect of the assay enabled determination of the concentration for each proteoform individually. Low-abundance proteoforms, such as fucosylated apoC-III, were detected in less than 20% of the samples. The distribution of apoC-III proteoforms varied among samples with similar total apoC-III concentrations. The multiplex analysis of the three apolipoproteins C and their proteoforms using quantitative MSIA represents a significant step forward toward better understanding of their physiological roles in health and disease.

ContributorsTrenchevska, Olgica (Author) / Schaab, Matthew (Author) / Nelson, Randall (Author) / Nedelkov, Dobrin (Author) / Biodesign Institute (Contributor)
Created2015-06-15