This growing collection consists of scholarly works authored by ASU-affiliated faculty, staff, and community members, and it contains many open access articles. ASU-affiliated authors are encouraged to Share Your Work in KEEP.

Displaying 1 - 10 of 54
Filtering by

Clear all filters

141461-Thumbnail Image.png
Description
In the digital humanities, there is a constant need to turn images and PDF files into plain text to apply analyses such as topic modelling, named entity recognition, and other techniques. However, although there exist different solutions to extract text embedded in PDF files or run OCR on images, they

In the digital humanities, there is a constant need to turn images and PDF files into plain text to apply analyses such as topic modelling, named entity recognition, and other techniques. However, although there exist different solutions to extract text embedded in PDF files or run OCR on images, they typically require additional training (for example, scholars have to learn how to use the command line) or are difficult to automate without programming skills. The Giles Ecosystem is a distributed system based on Apache Kafka that allows users to upload documents for text and image extraction. The system components are implemented using Java and the Spring Framework and are available under an Open Source license on GitHub (https://github.com/diging/).
ContributorsLessios-Damerow, Julia (Contributor) / Peirson, Erick (Contributor) / Laubichler, Manfred (Contributor) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2017-09-28
129552-Thumbnail Image.png
Description

S-cysteinylated albumin and methionine-oxidized apolipoprotein A-I (apoA-I) have been posed as candidate markers of diseases associated with oxidative stress. Here, a dilute-and-shoot form of LC–electrospray ionization–MS requiring half a microliter of blood plasma was employed to simultaneously quantify the relative abundance of these oxidized proteoforms in samples stored at −80

S-cysteinylated albumin and methionine-oxidized apolipoprotein A-I (apoA-I) have been posed as candidate markers of diseases associated with oxidative stress. Here, a dilute-and-shoot form of LC–electrospray ionization–MS requiring half a microliter of blood plasma was employed to simultaneously quantify the relative abundance of these oxidized proteoforms in samples stored at −80 °C, −20 °C, and room temperature and exposed to multiple freeze-thaw cycles and other adverse conditions in order to assess the possibility that protein oxidation may occur as a result of poor sample storage or handling. Samples from a healthy donor and a participant with poorly controlled type 2 diabetes started at the same low level of protein oxidation and behaved similarly; significant increases in albumin oxidation via S-cysteinylation were found to occur within hours at room temperature and days at −20 °C. Methionine oxidation of apoA-I took place on a longer time scale, setting in after albumin oxidation reached a plateau. Freeze–thaw cycles had a minimal effect on protein oxidation. In matched collections, protein oxidation in serum was the same as that in plasma. Albumin and apoA-I oxidation were not affected by sample headspace or the degree to which vials were sealed. ApoA-I, however, was unexpectedly found to oxidize faster in samples with lower surface-area-to-volume ratios. An initial survey of samples from patients with inflammatory conditions normally associated with elevated oxidative stress-including acute myocardial infarction and prostate cancer—demonstrated a lack of detectable apoA-I oxidation. Albumin S-cysteinylation in these samples was consistent with known but relatively brief exposures to temperatures above −30 °C (the freezing point of blood plasma). Given their properties and ease of analysis, these oxidized proteoforms, once fully validated, may represent the first markers of blood plasma specimen integrity based on direct measurement of oxidative molecular damage that can occur under suboptimal storage conditions.

ContributorsBorges, Chad (Author) / Rehder, Douglas (Author) / Jensen, Sally (Author) / Schaab, Matthew (Author) / Sherma, Nisha (Author) / Yassine, Hussein (Author) / Nikolova, Boriana (Author) / Breburda, Christian (Author) / Department of Chemistry and Biochemistry (Contributor)
Created2014-07-01
129567-Thumbnail Image.png
Description

Human protein diversity arises as a result of alternative splicing, single nucleotide polymorphisms (SNPs) and posttranslational modifications. Because of these processes, each protein can exists as multiple variants in vivo. Tailored strategies are needed to study these protein variants and understand their role in health and disease. In this work

Human protein diversity arises as a result of alternative splicing, single nucleotide polymorphisms (SNPs) and posttranslational modifications. Because of these processes, each protein can exists as multiple variants in vivo. Tailored strategies are needed to study these protein variants and understand their role in health and disease. In this work we utilized quantitative mass spectrometric immunoassays to determine the protein variants concentration of beta-2-microglobulin, cystatin C, retinol binding protein, and transthyretin, in a population of 500 healthy individuals. Additionally, we determined the longitudinal concentration changes for the protein variants from four individuals over a 6 month period. Along with the native forms of the four proteins, 13 posttranslationally modified variants and 7 SNP-derived variants were detected and their concentration determined. Correlations of the variants concentration with geographical origin, gender, and age of the individuals were also examined. This work represents an important step toward building a catalog of protein variants concentrations and examining their longitudinal changes.

ContributorsTrenchevska, Olgica (Author) / Phillips, David A. (Author) / Nelson, Randall (Author) / Nedelkov, Dobrin (Author) / Biodesign Institute (Contributor)
Created2014-06-23
129259-Thumbnail Image.png
Description

What's a profession without a code of ethics? Being a legitimate profession almost requires drafting a code and, at least nominally, making members follow it. Codes of ethics (henceforth “codes”) exist for a number of reasons, many of which can vary widely from profession to profession - but above all

What's a profession without a code of ethics? Being a legitimate profession almost requires drafting a code and, at least nominally, making members follow it. Codes of ethics (henceforth “codes”) exist for a number of reasons, many of which can vary widely from profession to profession - but above all they are a form of codified self-regulation. While codes can be beneficial, it argues that when we scratch below the surface, there are many problems at their root. In terms of efficacy, codes can serve as a form of ethical window dressing, rather than effective rules for behavior. But even more that, codes can degrade the meaning behind being a good person who acts ethically for the right reasons.

Created2013-11-30
128778-Thumbnail Image.png
Description

Online communities are becoming increasingly important as platforms for large-scale human cooperation. These communities allow users seeking and sharing professional skills to solve problems collaboratively. To investigate how users cooperate to complete a large number of knowledge-producing tasks, we analyze Stack Exchange, one of the largest question and answer systems

Online communities are becoming increasingly important as platforms for large-scale human cooperation. These communities allow users seeking and sharing professional skills to solve problems collaboratively. To investigate how users cooperate to complete a large number of knowledge-producing tasks, we analyze Stack Exchange, one of the largest question and answer systems in the world. We construct attention networks to model the growth of 110 communities in the Stack Exchange system and quantify individual answering strategies using the linking dynamics on attention networks. We identify two answering strategies. Strategy A aims at performing maintenance by doing simple tasks, whereas strategy B aims at investing time in doing challenging tasks. Both strategies are important: empirical evidence shows that strategy A decreases the median waiting time for answers and strategy B increases the acceptance rate of answers. In investigating the strategic persistence of users, we find that users tends to stick on the same strategy over time in a community, but switch from one strategy to the other across communities. This finding reveals the different sets of knowledge and skills between users. A balance between the population of users taking A and B strategies that approximates 2:1, is found to be optimal to the sustainable growth of communities.

ContributorsWu, Lingfei (Author) / Baggio, Jacopo (Author) / Janssen, Marco (Author) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2016-03-02
128687-Thumbnail Image.png
Description

Proteins can exist as multiple proteoforms in vivo, as a result of alternative splicing and single-nucleotide polymorphisms (SNPs), as well as posttranslational processing. To address their clinical significance in a context of diagnostic information, proteoforms require a more in-depth analysis. Mass spectrometric immunoassays (MSIA) have been devised for studying structural

Proteins can exist as multiple proteoforms in vivo, as a result of alternative splicing and single-nucleotide polymorphisms (SNPs), as well as posttranslational processing. To address their clinical significance in a context of diagnostic information, proteoforms require a more in-depth analysis. Mass spectrometric immunoassays (MSIA) have been devised for studying structural diversity in human proteins. MSIA enables protein profiling in a simple and high-throughput manner, by combining the selectivity of targeted immunoassays, with the specificity of mass spectrometric detection. MSIA has been used for qualitative and quantitative analysis of single and multiple proteoforms, distinguishing between normal fluctuations and changes related to clinical conditions. This mini review offers an overview of the development and application of mass spectrometric immunoassays for clinical and population proteomics studies. Provided are examples of some recent developments, and also discussed are the trends and challenges in mass spectrometry-based immunoassays for the next-phase of clinical applications.

ContributorsTrenchevska, Olgica (Author) / Nelson, Randall (Author) / Nedelkov, Dobrin (Author) / Biodesign Institute (Contributor)
Created2016-03-17
128975-Thumbnail Image.png
Description

Background: Cysteine sulfenic acid (Cys-SOH) plays important roles in the redox regulation of numerous proteins. As a relatively unstable posttranslational protein modification it is difficult to quantify the degree to which any particular protein is modified by Cys-SOH within a complex biological environment. The goal of these studies was to move

Background: Cysteine sulfenic acid (Cys-SOH) plays important roles in the redox regulation of numerous proteins. As a relatively unstable posttranslational protein modification it is difficult to quantify the degree to which any particular protein is modified by Cys-SOH within a complex biological environment. The goal of these studies was to move a step beyond detection and into the relative quantification of Cys-SOH within specific proteins found in a complex biological setting--namely, human plasma.

Results: This report describes the possibilities and limitations of performing such analyses based on the use of thionitrobenzoic acid and dimedone-based probes which are commonly employed to trap Cys-SOH. Results obtained by electrospray ionization-based mass spectrometric immunoassay reveal the optimal type of probe for such analyses as well as the reproducible relative quantification of Cys-SOH within albumin and transthyretin extracted from human plasma--the latter as a protein previously unknown to be modified by Cys-SOH.

Conclusions: The relative quantification of Cys-SOH within specific proteins in a complex biological setting can be accomplished, but several analytical precautions related to trapping, detecting, and quantifying Cys-SOH must be taken into account prior to pursuing its study in such matrices.

ContributorsRehder, Douglas (Author) / Borges, Chad (Author) / Biodesign Institute (Contributor)
Created2010-07-01
128933-Thumbnail Image.png
Description

Introduction: Apolipoprotein C-III (apoC-III) regulates triglyceride (TG) metabolism. In plasma, apoC-III exists in non-sialylated (apoC-III0a without glycosylation and apoC-III[subscript 0b] with glycosylation), monosialylated (apoC-III1) or disialylated (apoC-III2) proteoforms. Our aim was to clarify the relationship between apoC-III sialylation proteoforms with fasting plasma TG concentrations.

Methods: In 204 non-diabetic adolescent participants, the

Introduction: Apolipoprotein C-III (apoC-III) regulates triglyceride (TG) metabolism. In plasma, apoC-III exists in non-sialylated (apoC-III0a without glycosylation and apoC-III[subscript 0b] with glycosylation), monosialylated (apoC-III1) or disialylated (apoC-III2) proteoforms. Our aim was to clarify the relationship between apoC-III sialylation proteoforms with fasting plasma TG concentrations.

Methods: In 204 non-diabetic adolescent participants, the relative abundance of apoC-III plasma proteoforms was measured using mass spectrometric immunoassay.

Results: Compared with the healthy weight subgroup (n = 16), the ratios of apoC-III0a, apoC-III0b, and apoC-III1 to apoC-III2 were significantly greater in overweight (n = 33) and obese participants (n = 155). These ratios were positively correlated with BMI z-scores and negatively correlated with measures of insulin sensitivity (S[subscript i]). The relationship of apoC-III1 / apoC-III2 with Si persisted after adjusting for BMI (p = 0.02). Fasting TG was correlated with the ratio of apoC-III0a / apoC-III2 (r = 0.47, p<0.001), apoC-III0b / apoC-III2 (r = 0.41, p<0.001), apoC-III1 / apoC-III2 (r = 0.43, p<0.001). By examining apoC-III concentrations, the association of apoC-III proteoforms with TG was driven by apoC-III0a (r = 0.57, p<0.001), apoC-III0b (r = 0.56. p<0.001) and apoC-III1 (r = 0.67, p<0.001), but not apoC-III2 (r = 0.006, p = 0.9) concentrations, indicating that apoC-III relationship with plasma TG differed in apoC-III2 compared with the other proteoforms.

Conclusion: We conclude that apoC-III0a, apoC-III0b, and apoC-III1, but not apoC-III2 appear to be under metabolic control and associate with fasting plasma TG. Measurement of apoC-III proteoforms can offer insights into the biology of TG metabolism in obesity.

ContributorsYassine, Hussein N. (Author) / Trenchevska, Olgica (Author) / Ramrakhiani, Ambika (Author) / Parekh, Aarushi (Author) / Koska, Juraj (Author) / Walker, Ryan W. (Author) / Billheimer, Dean (Author) / Reaven, Peter D. (Author) / Yen, Frances T. (Author) / Nelson, Randall (Author) / Goran, Michael I. (Author) / Nedelkov, Dobrin (Author) / Biodesign Institute (Contributor)
Created2015-12-03
Description

On-going efforts to understand the dynamics of coupled social-ecological (or more broadly, coupled infrastructure) systems and common pool resources have led to the generation of numerous datasets based on a large number of case studies. This data has facilitated the identification of important factors and fundamental principles which increase our

On-going efforts to understand the dynamics of coupled social-ecological (or more broadly, coupled infrastructure) systems and common pool resources have led to the generation of numerous datasets based on a large number of case studies. This data has facilitated the identification of important factors and fundamental principles which increase our understanding of such complex systems. However, the data at our disposal are often not easily comparable, have limited scope and scale, and are based on disparate underlying frameworks inhibiting synthesis, meta-analysis, and the validation of findings. Research efforts are further hampered when case inclusion criteria, variable definitions, coding schema, and inter-coder reliability testing are not made explicit in the presentation of research and shared among the research community. This paper first outlines challenges experienced by researchers engaged in a large-scale coding project; then highlights valuable lessons learned; and finally discusses opportunities for further research on comparative case study analysis focusing on social-ecological systems and common pool resources. Includes supplemental materials and appendices published in the International Journal of the Commons 2016 Special Issue. Volume 10 - Issue 2 - 2016.

ContributorsRatajczyk, Elicia (Author) / Brady, Ute (Author) / Baggio, Jacopo (Author) / Barnett, Allain J. (Author) / Perez Ibarra, Irene (Author) / Rollins, Nathan (Author) / Rubinos, Cathy (Author) / Shin, Hoon Cheol (Author) / Yu, David (Author) / Aggarwal, Rimjhim (Author) / Anderies, John (Author) / Janssen, Marco (Author) / ASU-SFI Center for Biosocial Complex Systems (Contributor)
Created2016-09-09
129155-Thumbnail Image.png
Description

The impetus for discovery and evaluation of protein biomarkers has been accelerated by recent development of advanced technologies for rapid and broad proteome analyses. Mass spectrometry (MS)-based protein assays hold great potential for in vitro biomarker studies. Described here is the development of a multiplex mass spectrometric immunoassay (MSIA) for

The impetus for discovery and evaluation of protein biomarkers has been accelerated by recent development of advanced technologies for rapid and broad proteome analyses. Mass spectrometry (MS)-based protein assays hold great potential for in vitro biomarker studies. Described here is the development of a multiplex mass spectrometric immunoassay (MSIA) for quantification of apolipoprotein C-I (apoC-I), apolipoprotein C-II (apoC-II), apolipoprotein C-III (apoC-III) and their proteoforms. The multiplex MSIA assay was fast (∼40 min) and high-throughput (96 samples at a time). The assay was applied to a small cohort of human plasma samples, revealing the existence of multiple proteoforms for each apolipoprotein C. The quantitative aspect of the assay enabled determination of the concentration for each proteoform individually. Low-abundance proteoforms, such as fucosylated apoC-III, were detected in less than 20% of the samples. The distribution of apoC-III proteoforms varied among samples with similar total apoC-III concentrations. The multiplex analysis of the three apolipoproteins C and their proteoforms using quantitative MSIA represents a significant step forward toward better understanding of their physiological roles in health and disease.

ContributorsTrenchevska, Olgica (Author) / Schaab, Matthew (Author) / Nelson, Randall (Author) / Nedelkov, Dobrin (Author) / Biodesign Institute (Contributor)
Created2015-06-15