HMMRATAC: a Hidden Markov ModeleR for ATAC-seq

Evan D Tarbell; Tao Liu

doi:10.1093/nar/gkz533

HMMRATAC: a Hidden Markov ModeleR for ATAC-seq

Nucleic Acids Research ◽

10.1093/nar/gkz533 ◽

2019 ◽

Vol 47 (16) ◽

pp. e91-e91 ◽

Cited By ~ 17

Author(s):

Evan D Tarbell ◽

Tao Liu

Keyword(s):

Machine Learning ◽

Hidden Markov ◽

Current Data ◽

Supervised Machine Learning ◽

Learning Approach ◽

Analysis Tool ◽

Peak Calling ◽

Entire Genome ◽

Machine Learning Approach ◽

Accessible Chromatin

Abstract ATAC-seq has been widely adopted to identify accessible chromatin regions across the genome. However, current data analysis still utilizes approaches initially designed for ChIP-seq or DNase-seq, without considering the transposase digested DNA fragments that contain additional nucleosome positioning information. We present the first dedicated ATAC-seq analysis tool, a semi-supervised machine learning approach named HMMRATAC. HMMRATAC splits a single ATAC-seq dataset into nucleosome-free and nucleosome-enriched signals, learns the unique chromatin structure around accessible regions, and then predicts accessible regions across the entire genome. We show that HMMRATAC outperforms the popular peak-calling algorithms on published human ATAC-seq datasets. We find that single-end sequenced or size-selected ATAC-seq datasets result in a loss of sensitivity compared to paired-end datasets without size-selection.

Download Full-text

HMMRATAC: a Hidden Markov ModeleR for ATAC-seq

10.1101/306621 ◽

2018 ◽

Cited By ~ 1

Author(s):

Evan D. Tarbell ◽

Tao Liu

Keyword(s):

Machine Learning ◽

Hidden Markov ◽

Current Data ◽

Supervised Machine Learning ◽

Learning Approach ◽

Analysis Tool ◽

Peak Calling ◽

Entire Genome ◽

Machine Learning Approach ◽

Accessible Chromatin

ABSTRACTATAC-seq has been widely adopted to identify accessible chromatin regions across the genome. However, current data analysis still utilizes approaches initially designed for ChIP-seq or DNase-seq, without considering the transposase digested DNA fragments that contain additional nucleosome positioning information. We present the first dedicated ATAC-seq analysis tool, a semi-supervised machine learning approach named HMMRATAC. HMMRATAC splits a single ATAC-seq dataset into nucleosome-free and nucleosome-enriched signals, learns the unique chromatin structure around accessible regions, and then predicts accessible regions across the entire genome. We show that HMMRATAC outperforms the popular peak-calling algorithms on published human ATAC-seq datasets. We find that single-end sequenced or size-selected ATAC-seq datasets result in a loss of sensitivity compared to paired-end datasets without size-selection.

Download Full-text

Mol2vec: Unsupervised Machine Learning Approach with Chemical Intuition

10.26434/chemrxiv.5513581.v1 ◽

2017 ◽

Author(s):

Sabrina Jaeger ◽

Simone Fulle ◽

Samo Turk

Keyword(s):

Machine Learning ◽

Language Processing ◽

Supervised Machine Learning ◽

Learning Approach ◽

Learning Approaches ◽

Unsupervised Machine Learning ◽

Feature Representations ◽

Machine Learning Approach ◽

The Individual ◽

Vector Representations

Inspired by natural language processing techniques we here introduce Mol2vec which is an unsupervised machine learning approach to learn vector representations of molecular substructures. Similarly, to the Word2vec models where vectors of closely related words are in close proximity in the vector space, Mol2vec learns vector representations of molecular substructures that are pointing in similar directions for chemically related substructures. Compounds can finally be encoded as vectors by summing up vectors of the individual substructures and, for instance, feed into supervised machine learning approaches to predict compound properties. The underlying substructure vector embeddings are obtained by training an unsupervised machine learning approach on a so-called corpus of compounds that consists of all available chemical matter. The resulting Mol2vec model is pre-trained once, yields dense vector representations and overcomes drawbacks of common compound feature representations such as sparseness and bit collisions. The prediction capabilities are demonstrated on several compound property and bioactivity data sets and compared with results obtained for Morgan fingerprints as reference compound representation. Mol2vec can be easily combined with ProtVec, which employs the same Word2vec concept on protein sequences, resulting in a proteochemometric approach that is alignment independent and can be thus also easily used for proteins with low sequence similarities.

Download Full-text

Coreference resolution of Korean anaphoric zero objects: Towards a supervised machine learning approach

International Journal of Computer Science and Information Technology for Education ◽

10.21742/ijcsite.2016.1.01 ◽

2016 ◽

Vol 1 (1) ◽

pp. 1-6

Author(s):

Euhee Kim ◽

◽

Myung-Kwan Park ◽

Keyword(s):

Machine Learning ◽

Supervised Machine Learning ◽

Learning Approach ◽

Coreference Resolution ◽

Machine Learning Approach

Download Full-text

A Supervised Machine Learning Approach for the Credibility Assessment of User-Generated Content

Wireless Personal Communications ◽

10.1007/s11277-021-08136-5 ◽

2021 ◽

Author(s):

Praphula Kumar Jain ◽

Rajendra Pamula ◽

Sarfraj Ansari

Keyword(s):

Machine Learning ◽

Supervised Machine Learning ◽

User Generated Content ◽

Learning Approach ◽

Credibility Assessment ◽

Machine Learning Approach

Download Full-text

Fast and robust supervised machine learning approach for classification and prediction of Parkinson’s disease onset

Computer Methods in Biomechanics and Biomedical Engineering Imaging & Visualization ◽

10.1080/21681163.2021.1941262 ◽

2021 ◽

pp. 1-17

Author(s):

Lavanya Madhuri Bollipo ◽

Kadambari K V

Keyword(s):

Machine Learning ◽

Parkinson’S Disease ◽

Parkinson's Disease ◽

Disease Onset ◽

Supervised Machine Learning ◽

Learning Approach ◽

Machine Learning Approach

Download Full-text

Supervised Machine Learning Approach For The Prediction of Breast Cancer

2020 International Conference on System, Computation, Automation and Networking (ICSCAN) ◽

10.1109/icscan49426.2020.9262403 ◽

2020 ◽

Author(s):

Tarun Jain ◽

Vivek Kumar Verma ◽

Mahek Agarwal ◽

Anju Yadav ◽

Ashish Jain

Keyword(s):

Breast Cancer ◽

Machine Learning ◽

Supervised Machine Learning ◽

Learning Approach ◽

Machine Learning Approach

Download Full-text

A supervised machine learning approach to author disambiguation in the Web of Science

Journal of Informetrics ◽

10.1016/j.joi.2021.101166 ◽

2021 ◽

Vol 15 (3) ◽

pp. 101166

Author(s):

Andreas Rehs

Keyword(s):

Machine Learning ◽

Web Of Science ◽

Supervised Machine Learning ◽

Learning Approach ◽

Machine Learning Approach ◽

Author Disambiguation ◽

The Web

Download Full-text

Supervised machine-learning approach for the optimal arrangement of active hotspots in three-dimensional integrated circuits

IEEE Transactions on Components Packaging and Manufacturing Technology ◽

10.1109/tcpmt.2021.3109662 ◽

2021 ◽

pp. 1-1

Author(s):

Srikanth Rangarajan ◽

Leila Choobineh ◽

Bahgat Sammakia

Keyword(s):

Machine Learning ◽

Integrated Circuits ◽

Three Dimensional ◽

Supervised Machine Learning ◽

Learning Approach ◽

Optimal Arrangement ◽

Machine Learning Approach

Download Full-text

Supervised Machine Learning Approach for Subjectivity/Objectivity Classification of Social Data

Information Systems - Lecture Notes in Business Information Processing ◽

10.1007/978-3-030-44322-1_15 ◽

2020 ◽

pp. 193-205

Author(s):

Rim Chiha ◽

Mounir Ben Ayed

Keyword(s):

Machine Learning ◽

Supervised Machine Learning ◽

Learning Approach ◽

Social Data ◽

Machine Learning Approach

Download Full-text

A Supervised Machine Learning Approach to Fake News Identification

Intelligent Data Communication Technologies and Internet of Things - Lecture Notes on Data Engineering and Communications Technologies ◽

10.1007/978-3-030-34080-3_22 ◽

2019 ◽

pp. 197-204

Author(s):

Anisha Datta ◽

Shukrity Si

Keyword(s):

Machine Learning ◽

Supervised Machine Learning ◽

Learning Approach ◽

Fake News ◽

Machine Learning Approach

Download Full-text