run_metadata
2 rows where devstage_curation = "Pharyngula" and tissue_curation_coarse = "Liver and Biliary System"
This data as json, CSV (advanced)
| Link | rowid ▼ | run.accession | experiment.accession | sample.accession | study.accession | bioproject | study.title | study.alias | study.type | study.abstract | study.attributes | study.PMIDs | sample.description | sample.title | sample.alias | sample.centername | sample.attributes | GEOsample.title | GEOsample.dataprocessing | GEOsample.source | GEOsample.treatmentprotocol | GEOsample.extractprotocol | GEOsample.growthprotocol | GEOsample.characteristics | GEOsample.accession | experiment.title | experiment.alias | experiment.library_name | experiment.design_description | experiment.library_construction_protocol | experiment.attributes | experiment.library_strategy | experiment.library_source | experiment.library_selection | experiment.library_layout | experiment.platform | experiment.instrument_model | experiment.spot_descriptor | experiment.study_ref | run.title | run.attributes | run.filename | run.semantic_name | run.total_bases | run.total_spots | run.alias | run.read_lengths | run.base_counts | run.r1_length | run.r2_length | run.r3_length | run.r4_length | run.Acount | run.Ccount | run.Gcount | run.Tcount | run.Ncount | run.experiment | run.pool_member | submission.accession | submission.srasource | submission.bioprojectsource | seqdetective.n_mates | seqdetective.mapping_rate.mate1 | seqdetective.mapping_rate.mate2 | seqdetective.nofeature_rate.mate1 | seqdetective.nofeature_rate.mate2 | seqdetective.sparsity.mate1 | seqdetective.sparsity.mate2 | seqdetective.pos_strand_rate.mate1 | seqdetective.pos_strand_rate.mate2 | seqdetective.readlen.mate1 | seqdetective.readlen.mate2 | seqdetective.judgement.mate1 | seqdetective.judgement.mate2 | seqdetective.judgement.reason | platform_family | instrument_generation | read_bias | selection_class | prep_kit | sc_or_bulk | tech_class | technology | tech_variant | submission.bioprojectsource.country | earliest_date | devstage_curation | devstage_curation_coarse | tissue_curation | tissue_curation_coarse |
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| 53015 | 53015 | SRR9662024 | SRX6422900 | SRS5079690 | SRP213938 | PRJNA553572 | A map of cis regulatory elements and 3D genome structures in zebrafish | GSE134055 | Other | The zebrafish has been widely used for the study of human disease and development as 70% of the protein coding genes are conserved between the two species. Annotation of functional control elements of the zebrafish genome however has lagged behind that of other model systems such as mouse and Drosophila. Based on multi omics approaches taken in the ENCODE and Roadmap Epigenomics projects we performed RNA seq ATAC seq ChIP seq and Hi C experiments in ten adult and two embryonic tissues to generate a comprehensive map of transcriptomes and regulatory elements in the zebrafish Tuebingen reference strain. Overall we have identified 235 596 cis regulatory elements which potentially shape the tissue specific and developmental stage specific gene expression in zebrafish. A comparison of zebrafish human and mouse regulatory elements allowed us to identify both evolutionarily conserved and species specific regulatory sequences. Furthermore through the analysis of Hi C data in zebrafish brain and muscle we observed different levels of 3D genome organization including compartment topological associating domains TADs and chromatin loops in zebrafish. A subset of TADs are deeply conserved between zebrafish and human. This work provides an additional epigenomic anchor for the functional annotation of vertebrate genomes and the study of evolutionally conserved elements of 3D genome organization. Overall design: 13 tissues from adult and embryonic stage were examined using ChIP Seq H3K27ac and H3K4me3 RNA Seq 11 of them were examined using ATAC seq WGBS and ChIP seq H3K9me3 and H3K9me2 and one scATAC seq in brain. Additionally we performed HiC experiments in adult muscle and brain. Please note that for the samples GSM4661977 GSM4662088 [1] each processed data generated from both replicates is linked to the corresponding *rep1 sample records [2] the input sample used for each ChIP sample is indicated in the description field in the corresponding input sample records. | pubmed:33239788;pubmed:35649578 | YueLab RNA Seq Liver rep2 | GSM3934892 | source name:Tissue|strain:Tuebingen|tissue:Liver | YueLab RNA Seq Liver rep2 | RNA seq reads were aligned to zv10 genome assembly using STAR; ChIP seq and ATAC seq reads were aligned to zv10 genome assembly using BWA HiC reads were aligned to zv10 genome assembly using Bowtie2 The TPM value of gene expression was caculated using RSEM ChIP seq and ATAC seq peaks were called using MACS2 with the following setting: ChIP seq q value <10e 2 p value<10e 5 Change>1 FC>2. ATAC seq: q value<10e 2 and p value<10e 5 HiC matrix was generated using HiC Pro Genome build: zv10 Supplementary files format and content: tab delimited text files include TPM values for each Sample; the narrowPeak files included the peaks for each Sample; The .hic file were the matrix of Hi C for each Sample.**All replicates were merged | Tissue | For each RNA seq experiment the same tissues combined from at least two Tuebingen fish were used as one replicate. For embryonic trunk ten 1 dpf fish were dechorionated with pronase and trunk were cut off for RNA seq. For embryonic neuron green cells from TgHuc:Kaede cells were sorted by FACS and approxinately 20 000 cells were used for one replicate. The tissue RNA was extracted from Trizol® according to the protocol Invitrogen. The cDNA libraries were performed using SureSelect Strand Specific RNA Library Preparation Kit Agilent according to the manufacturer’s protocol. Briefly polyA RNA was purified from 1000 ng of total RNA using oligo dT beads Invitrogen. Extracted RNA was first fragmented then followed by reverse transcription end repair adenylation adaptor ligation and subsequent PCR amplification. The final product was checked by size distribution and concentration using BioAnalyzer High Sensitivity DNA Kit Agilent and Kapa Library Quantification Kit Kapa Biosystems and then followed by pair end 2X 50 bp high throughput sequencing using HiSeq 2500 Illumina. | Embryonic and ault Tuebingen zebrafish were raised under standard laboratory conditions | strain:Tuebingen|tissue:Liver | GSM3934892 | GSM3934892: YueLab RNA Seq Liver rep2; Danio rerio; RNA Seq | GSM3934892 | 1 | For each RNA seq experiment the same tissues combined from at least two Tuebingen fish were used as one replicate. For embryonic trunk ten 1 dpf fish were dechorionated with pronase and trunk were cut off for RNA seq. For embryonic neuron green cells from TgHuc:Kaede cells were sorted by FACS and approxinately 20 000 cells were used for one replicate. The tissue RNA was extracted from Trizol® according to the protocol Invitrogen. The cDNA libraries were performed using SureSelect Strand Specific RNA Library Preparation Kit Agilent according to the manufacturer's protocol. Briefly polyA RNA was purified from 1000 ng of total RNA using oligo dT beads Invitrogen. Extracted RNA was first fragmented then followed by reverse transcription end repair adenylation adaptor ligation and subsequent PCR amplification. The final product was checked by size distribution and concentration using BioAnalyzer High Sensitivity DNA Kit Agilent and Kapa Library Quantification Kit Kapa Biosystems and then followed by pair end 2X 50 bp high throughput sequencing using HiSeq 2500 Illumina. | GEO Accession:GSM3934892 | RNA-Seq | TRANSCRIPTOMIC | cDNA | PAIRED | ILLUMINA | Illumina Genome Analyzer | SRP213938 | YueLab-RNA-Seq-Liver-rep2_1.fastq.gz YueLab-RNA-Seq-Liver-rep2_2.fastq.gz | fastq fastq | 2453892473.0 | 18494458.0 | GSM3934892 r1 | 0:66.33 1:66.36 | A:639899749;C:572026186;G:561420643;T:680465509;N:80386 | 66 | 66 | 639899749 | 572026186 | 561420643 | 680465509 | 80386 | SRX6422900 | SRS5079690 | SRA919194 | GEO | Feng Yue, Department of Biochemistry and Molecular Genetics, Northwestern University Feinberg School of Medicine | 2 | 0.95181 | 0.96045 | 0.10335 | 0.09874 | 0.7665 | 0.76238 | 0.5468 | 0.55172 | 67 | 65 | B | B | biological fallback assumption | illumina | early_illumina | unknown | poly_a | unknown | bulk | unknown | unknown | United States | 2019-07-09 | Pharyngula | Embryo | Liver | Liver and Biliary System | |||||||||||
| 53016 | 53016 | SRR9662023 | SRX6422899 | SRS5079689 | SRP213938 | PRJNA553572 | A map of cis regulatory elements and 3D genome structures in zebrafish | GSE134055 | Other | The zebrafish has been widely used for the study of human disease and development as 70% of the protein coding genes are conserved between the two species. Annotation of functional control elements of the zebrafish genome however has lagged behind that of other model systems such as mouse and Drosophila. Based on multi omics approaches taken in the ENCODE and Roadmap Epigenomics projects we performed RNA seq ATAC seq ChIP seq and Hi C experiments in ten adult and two embryonic tissues to generate a comprehensive map of transcriptomes and regulatory elements in the zebrafish Tuebingen reference strain. Overall we have identified 235 596 cis regulatory elements which potentially shape the tissue specific and developmental stage specific gene expression in zebrafish. A comparison of zebrafish human and mouse regulatory elements allowed us to identify both evolutionarily conserved and species specific regulatory sequences. Furthermore through the analysis of Hi C data in zebrafish brain and muscle we observed different levels of 3D genome organization including compartment topological associating domains TADs and chromatin loops in zebrafish. A subset of TADs are deeply conserved between zebrafish and human. This work provides an additional epigenomic anchor for the functional annotation of vertebrate genomes and the study of evolutionally conserved elements of 3D genome organization. Overall design: 13 tissues from adult and embryonic stage were examined using ChIP Seq H3K27ac and H3K4me3 RNA Seq 11 of them were examined using ATAC seq WGBS and ChIP seq H3K9me3 and H3K9me2 and one scATAC seq in brain. Additionally we performed HiC experiments in adult muscle and brain. Please note that for the samples GSM4661977 GSM4662088 [1] each processed data generated from both replicates is linked to the corresponding *rep1 sample records [2] the input sample used for each ChIP sample is indicated in the description field in the corresponding input sample records. | pubmed:33239788;pubmed:35649578 | YueLab RNA Seq Liver rep1 | GSM3934891 | source name:Tissue|strain:Tuebingen|tissue:Liver | YueLab RNA Seq Liver rep1 | RNA seq reads were aligned to zv10 genome assembly using STAR; ChIP seq and ATAC seq reads were aligned to zv10 genome assembly using BWA HiC reads were aligned to zv10 genome assembly using Bowtie2 The TPM value of gene expression was caculated using RSEM ChIP seq and ATAC seq peaks were called using MACS2 with the following setting: ChIP seq q value <10e 2 p value<10e 5 Change>1 FC>2. ATAC seq: q value<10e 2 and p value<10e 5 HiC matrix was generated using HiC Pro Genome build: zv10 Supplementary files format and content: tab delimited text files include TPM values for each Sample; the narrowPeak files included the peaks for each Sample; The .hic file were the matrix of Hi C for each Sample.**All replicates were merged | Tissue | For each RNA seq experiment the same tissues combined from at least two Tuebingen fish were used as one replicate. For embryonic trunk ten 1 dpf fish were dechorionated with pronase and trunk were cut off for RNA seq. For embryonic neuron green cells from TgHuc:Kaede cells were sorted by FACS and approxinately 20 000 cells were used for one replicate. The tissue RNA was extracted from Trizol® according to the protocol Invitrogen. The cDNA libraries were performed using SureSelect Strand Specific RNA Library Preparation Kit Agilent according to the manufacturer’s protocol. Briefly polyA RNA was purified from 1000 ng of total RNA using oligo dT beads Invitrogen. Extracted RNA was first fragmented then followed by reverse transcription end repair adenylation adaptor ligation and subsequent PCR amplification. The final product was checked by size distribution and concentration using BioAnalyzer High Sensitivity DNA Kit Agilent and Kapa Library Quantification Kit Kapa Biosystems and then followed by pair end 2X 50 bp high throughput sequencing using HiSeq 2500 Illumina. | Embryonic and ault Tuebingen zebrafish were raised under standard laboratory conditions | strain:Tuebingen|tissue:Liver | GSM3934891 | GSM3934891: YueLab RNA Seq Liver rep1; Danio rerio; RNA Seq | GSM3934891 | 1 | For each RNA seq experiment the same tissues combined from at least two Tuebingen fish were used as one replicate. For embryonic trunk ten 1 dpf fish were dechorionated with pronase and trunk were cut off for RNA seq. For embryonic neuron green cells from TgHuc:Kaede cells were sorted by FACS and approxinately 20 000 cells were used for one replicate. The tissue RNA was extracted from Trizol® according to the protocol Invitrogen. The cDNA libraries were performed using SureSelect Strand Specific RNA Library Preparation Kit Agilent according to the manufacturer's protocol. Briefly polyA RNA was purified from 1000 ng of total RNA using oligo dT beads Invitrogen. Extracted RNA was first fragmented then followed by reverse transcription end repair adenylation adaptor ligation and subsequent PCR amplification. The final product was checked by size distribution and concentration using BioAnalyzer High Sensitivity DNA Kit Agilent and Kapa Library Quantification Kit Kapa Biosystems and then followed by pair end 2X 50 bp high throughput sequencing using HiSeq 2500 Illumina. | GEO Accession:GSM3934891 | RNA-Seq | TRANSCRIPTOMIC | cDNA | PAIRED | ILLUMINA | Illumina Genome Analyzer | SRP213938 | YueLab-RNA-Seq-Liver-rep1_1.fastq.gz YueLab-RNA-Seq-Liver-rep1_2.fastq.gz | fastq fastq | 3947397329.0 | 32675197.0 | GSM3934891 r1 | 0:60.48 1:60.33 | A:1018232733;C:908244909;G:926976675;T:1093882633;N:60379 | 60 | 60 | 1018232733 | 908244909 | 926976675 | 1093882633 | 60379 | SRX6422899 | SRS5079689 | SRA919194 | GEO | Feng Yue, Department of Biochemistry and Molecular Genetics, Northwestern University Feinberg School of Medicine | 2 | 0.97167 | 0.9794 | 0.0735 | 0.06958 | 0.79279 | 0.79454 | 0.54083 | 0.54983 | 60 | 60 | B | B | biological fallback assumption | illumina | early_illumina | unknown | poly_a | unknown | bulk | unknown | unknown | United States | 2019-07-09 | Pharyngula | Embryo | Liver | Liver and Biliary System |
Advanced export
JSON shape: default, array, newline-delimited
CREATE TABLE run_metadata("run.accession" VARCHAR, "experiment.accession" VARCHAR, "sample.accession" VARCHAR, "study.accession" VARCHAR, bioproject VARCHAR, "study.title" VARCHAR, "study.alias" VARCHAR, "study.type" VARCHAR, "study.abstract" VARCHAR, "study.attributes" VARCHAR, "study.PMIDs" VARCHAR, "sample.description" VARCHAR, "sample.title" VARCHAR, "sample.alias" VARCHAR, "sample.centername" VARCHAR, "sample.attributes" VARCHAR, "GEOsample.title" VARCHAR, "GEOsample.dataprocessing" VARCHAR, "GEOsample.source" VARCHAR, "GEOsample.treatmentprotocol" VARCHAR, "GEOsample.extractprotocol" VARCHAR, "GEOsample.growthprotocol" VARCHAR, "GEOsample.characteristics" VARCHAR, "GEOsample.accession" VARCHAR, "experiment.title" VARCHAR, "experiment.alias" VARCHAR, "experiment.library_name" VARCHAR, "experiment.design_description" VARCHAR, "experiment.library_construction_protocol" VARCHAR, "experiment.attributes" VARCHAR, "experiment.library_strategy" VARCHAR, "experiment.library_source" VARCHAR, "experiment.library_selection" VARCHAR, "experiment.library_layout" VARCHAR, "experiment.platform" VARCHAR, "experiment.instrument_model" VARCHAR, "experiment.spot_descriptor" VARCHAR, "experiment.study_ref" VARCHAR, "run.title" VARCHAR, "run.attributes" VARCHAR, "run.filename" VARCHAR, "run.semantic_name" VARCHAR, "run.total_bases" DOUBLE, "run.total_spots" DOUBLE, "run.alias" VARCHAR, "run.read_lengths" VARCHAR, "run.base_counts" VARCHAR, "run.r1_length" BIGINT, "run.r2_length" BIGINT, "run.r3_length" BIGINT, "run.r4_length" BIGINT, "run.Acount" BIGINT, "run.Ccount" BIGINT, "run.Gcount" BIGINT, "run.Tcount" BIGINT, "run.Ncount" BIGINT, "run.experiment" VARCHAR, "run.pool_member" VARCHAR, "submission.accession" VARCHAR, "submission.srasource" VARCHAR, "submission.bioprojectsource" VARCHAR, "seqdetective.n_mates" BIGINT, "seqdetective.mapping_rate.mate1" DOUBLE, "seqdetective.mapping_rate.mate2" DOUBLE, "seqdetective.nofeature_rate.mate1" DOUBLE, "seqdetective.nofeature_rate.mate2" DOUBLE, "seqdetective.sparsity.mate1" DOUBLE, "seqdetective.sparsity.mate2" DOUBLE, "seqdetective.pos_strand_rate.mate1" DOUBLE, "seqdetective.pos_strand_rate.mate2" DOUBLE, "seqdetective.readlen.mate1" BIGINT, "seqdetective.readlen.mate2" BIGINT, "seqdetective.judgement.mate1" VARCHAR, "seqdetective.judgement.mate2" VARCHAR, "seqdetective.judgement.reason" VARCHAR, platform_family VARCHAR, instrument_generation VARCHAR, read_bias VARCHAR, selection_class VARCHAR, prep_kit VARCHAR, sc_or_bulk VARCHAR, tech_class VARCHAR, technology VARCHAR, tech_variant VARCHAR, "submission.bioprojectsource.country" VARCHAR, earliest_date DATE, devstage_curation VARCHAR, devstage_curation_coarse VARCHAR, tissue_curation VARCHAR, tissue_curation_coarse VARCHAR);;
CREATE INDEX idx_run_bioproject ON run_metadata(bioproject);;
CREATE INDEX idx_run_run_accession ON run_metadata("run.accession");;