| | SLO | ENG | Cookies and privacy

Bigger font | Smaller font

Show document Help

Title:Iskanje vezavnih mest in rangiranje genov na podlagi CLIP in RNAseq podatkov
Authors:ID Haberman, Nejc (Author)
ID Štiglic, Gregor (Mentor) More about this mentor... New window
ID Ule, Jernej (Comentor)
Files:.pdf MAG_Haberman_Nejc_2013.pdf (2,11 MB)
MD5: 0FC45A6CB12690C810C6F92802DEC20B
 
Language:Slovenian
Work type:Master's thesis/paper
Typology:2.09 - Master's Thesis
Organization:FZV - Faculty of Health Sciences
Abstract:Proteini, ki se vežejo na RNK (ang. RNA-binding proteins), imajo pomembno vlogo pri regulaciji posttrankripcijskih procesov in so ključni regulatorji genske ekspresije. Za preučevanje vezavnih proteinov na RNK v živih celicah sta najprimernejši metodi CLIP (UV cross-linking and immunoprecipitation) in iCLIP (individual-nucleotide resolution CLIP). RNA-seq je postala standardizirana metoda za merjenje genske ekspresije celotnega transkriptoma. V naši raziskavi smo preučevali najprimernejšo normalizacijo CLIP-podatkov z RNA-seq podatki. Glavni problem je, da posamezna CLIP/iCLIP-metoda določi več vezavnih mest v visoko izraženih delih molekule RNK, ki posledično prekrijejo nizko izražena vezavna mesta. Zato je potreba po RNA-seq podatkih, da lahko z njimi normaliziramo vezavna mesta CLIP/iCLIP- podatkov. V ta namen smo načrtovali in implementirali različne metode normalizacij CLIP- in RNA-seq podatkov in jih primerjali z do sedaj znanimi rezultati za protein LIN28A. V našo raziskavo smo vključili programsko orodje Piranha, ki že uporablja metodo normalizacije z regresijskim modelom ZTNBR (zero-truncated negative binomial distributions regression), ki pa se ni izkazala za primerno rešitev našega problema. Izdelali smo hibridno metodo, ki izboljša objavljeno metodo, ki uporablja preprosto normalizacijo vsote RNA-seq odčitkov pri proteinu LIN28A. Hibridna metoda uporablja statistični model ZTNB (zero-truncated negative binomial distributions) za identifikacijo signifikantnih vezavnih mest in je del programskega orodja Piranha. Za vse opisane metode normalizacij smo razvili cevovod programskih skript in bioinformatskih orodij za primerjalno analizo metod. Prav tako smo razvili programski cevovod za preprocesiranje iCLIP-podatkov, ki jih lahko uporabimo v omenjenih normalizacijah. Vse metode bodo prosto dostopne na spletu in se bodo lahko uporabljale v prihodnjih raziskavah za analizo vezavnih proteinov na RNK s CLIP/iCLIP- in ostalimi sorodnimi metodami sekveniranja.
Keywords:normalizacija, vezavni proteini na RNK, CLIP, iCLIP, RNA-seq, LIN28A, bioinformatska orodja, genska obogatitev
Place of publishing:Maribor
Publisher:[N. Haberman]
Year of publishing:2013
PID:20.500.12556/DKUM-40731 New window
UDC:577
COBISS.SI-ID:1916836 New window
NUK URN:URN:SI:UM:DK:VTLNPYNQ
Publication date in DKUM:06.08.2013
Views:2776
Downloads:247
Metadata:XML DC-XML DC-RDF
Categories:FZV
:
Copy citation
  
Average score:(0 votes)
Your score:Voting is allowed only for logged in users.
Share:Bookmark and Share



Hover the mouse pointer over a document title to show the abstract or click on the title to get all document metadata.

Secondary language

Language:English
Title:Binding site identification and gene ranking based on CLIP and RNAseq data
Abstract:RNA binding proteins (RBPs) are key players in post-transcriptional processes and have a major role in the regulation of gene expression. CLIP (UV cross-linking and immunoprecipitation) and iCLIP (individual-nucleotide resolution CLIP) are the most appropriate methods to study protein-RNA interactions in living cells. For gene expression measurements across the transcriptome, RNA-seq has become the standard method of choice. In our study, we searched for the most appropriate method to normalize CLIP data and integrating it with matched RNA-seq data. The main need for normalization stems from the fact that the number of reads detected by CLIP/iCLIP methods depends on the expression level of each mRNA. Therefore, RNA-seq data is needed to normalize the CLIP/iCLIP data. We developed different types of algorithms for data normalization, and tested them against a published method and data sets for LIN28A protein using CLIP and RNA-seq data. In our study we included Piranha software with integrated CLIP and RNA-seq normalization using a ZTNBR (zero-truncated negative binomial distributions regression) model and we have found that ZTNBR is not suitable for the described problem. We devised a hybrid method which extends the previous simplistic normalization method by using a statistical ZTNB (zero-truncated negative binomial distributions) model for significant binding site identification and which also is a part of Piranha software. For all normalization methods we developed a pipeline of scripts and bioinformatics tools for comparative analysis of the different types of normalizations. We also developed a pre-processing pipeline for the iCLIP data which can be used for the same type of normalizations. All the methods will be freely available online and we imagine their frequent use in future studies of RNA binding proteins exploiting CLIP and related sequencing techniques.
Keywords:normalization, RNA binding proteins, CLIP, iCLIP, RNA-seq, LIN28A, bioinformatic tools, gene enrichment


Comments

Leave comment

You must log in to leave a comment.

Comments (0)
0 - 0 / 0
 
There are no comments!

Back
Logos of partners University of Maribor University of Ljubljana University of Primorska University of Nova Gorica