Os01g0173100

From RiceWiki
Revision as of 01:57, 8 June 2014 by Hermione (talk | contribs) (Evolution)
Jump to: navigation, search

Please input one-sentence summary here.

Annotated Information

Function

Please input function information here.

Expression

Please input expression information here.

Evolution

To understand the evolutionary relationships within the Alba protein superfamily, a phylogenetic tree was constructed. The unrooted phylogram revealed presence of two separate lineages, the first group consisting of animal Alba proteins, exemplified by the human RNase/MRP subunits Rpp20/Pop7, and the second containing archaeal and plant Alba proteins. OsAlba1 (A2WL83, O. sativa ssp. indica) is identical to Q94E63 (O. sativa ssp. japonica) and I1NKP6 (O. glaberrima). It clustered together with monocot (Oryza brachyantha, Zea mays, and Hordeum vulgare) Alba proteins, suggesting a closer proximity to these as compared to dicot Alba proteins.

Multiple sequence alignment (MSA) of OsAlba1 with other known Alba sequences from diverse sources including Archaea, Arabidopsis, yeast, human, Plasmodium and Trypanosoma revealed a high degree of conservation of amino acid in the regions of 59–72 and 101–122 in MSA, corresponding to 46–59 and 71–92, respectively, in OsAlba1. MSA showed presence of a conserved glycine residue at position 77 within Alba domain. Ten homologs of Alba protein have thus far been reported in UniProt database from indica rice, nine of which correspond to specific known genes with known locus ID on chromosomes. OsAlba1 is located on chromosome 1, albeit being a complete sequence of 152 aa in contrast to A2WL83 with 124 aa (UniProtKB/TrEMBL). The accession number of chromosome 1 of Oryza sativa ssp indica is GI 57015219 in NCBI, while that of japonica is GI 57015276. Available sequence of OsAlba1 in indica (position 4200197–4201583) corresponds to position 3805619–3807005 of japonica, indicating its different position on chromosome 1. The Alba sequences were retrieved from UniProt and a phylogenetic analysis of the Alba proteins from O. sativa ssp. indica was carried out. The phylogram appeared to be divided into two groups based on the percentage of protein sequence covered by Alba domain. The Alba domain for Group I had sequence coverage of less than 30%, while Group II comprised those with higher sequence coverage, when compared with the complete protein sequence .MSA of the phylogenetic tree is shown in Supplementary. All the ten sequences, harboring only a single domain i.e., Alba, are uncharacterized.

Labs working on this gene

Please input related labs here.

References

Please input cited references here.

Structured Information

Gene Name

Os01g0173100

Description

Alba, DNA/RNA-binding protein family protein

Version

NM_001048692.1 GI:115434797 GeneID:4327120

Length

1684 bp

Definition

Oryza sativa Japonica Group Os01g0173100, complete gene.

Source

Oryza sativa Japonica Group

 ORGANISM  Oryza sativa Japonica Group
           Eukaryota; Viridiplantae; Streptophyta; Embryophyta; Tracheophyta;
           Spermatophyta; Magnoliophyta; Liliopsida; Poales; Poaceae; BEP
           clade; Ehrhartoideae; Oryzeae; Oryza.
Chromosome

Chromosome 1

Location

Chromosome 1:3741778..3743461

Sequence Coding Region

3741949..3742044,3742170..3742234,3742377..3742439,3743057..3743105,3743202..3743387

Expression

GEO Profiles:Os01g0173100

Genome Context

<gbrowseImage1> name=NC_008394:3741778..3743461 source=RiceChromosome01 preset=GeneLocation </gbrowseImage1>

Gene Structure

<gbrowseImage2> name=NC_008394:3741778..3743461 source=RiceChromosome01 preset=GeneLocation </gbrowseImage2>

Coding Sequence

<cdnaseq>atggcggtggaggagatcaccgagggggtgaggaacctggccgtggagggggagcccgcggcggcggcggcggcggcgggaggtggtggtgagggggcgcagaggagggcggccgggagcagcagcaaccgcatccaggtgtccaacaccaagaagccactcttcttctatgtcaacctcgccaagaggtacatgcagcagcacggcgatgtcgagctctccgcgctcgggatggccattgcaacagttgtaactgttgcggagattcttaagaataacgggtttgctgttgaaaagaagattagaacatctacggtggaaataaacgatgaatcgagagttcgcccgctccaaaaggctaagattgagatagtgttagaaaagagcgagaaatttgatgagctgatggctgccgcagcggaagagagggaagctgcggaagctgaggagcaggcctga</cdnaseq>

Protein Sequence

<aaseq>MAVEEITEGVRNLAVEGEPAAAAAAAGGGGEGAQRRAAGSSSNR IQVSNTKKPLFFYVNLAKRYMQQHGDVELSALGMAIATVVTVAEILKNNGFAVEKKIR TSTVEINDESRVRPLQKAKIEIVLEKSEKFDELMAAAAEEREAAEAEEQA</aaseq>

Gene Sequence

<dnaseqindica>1418..1513#1228..1292#1023..1085#357..405#75..260#gcacaacccaccaccaccaccaccaccgatctgcggagagcgagctaccaccaccgagcggcgtagcttgagcgatggcggtggaggagatcaccgagggggtgaggaacctggccgtggagggggagcccgcggcggcggcggcggcggcgggaggtggtggtgagggggcgcagaggagggcggccgggagcagcagcaaccgcatccaggtgtccaacaccaagaagccactcttcttctatgtcaacctcgccaaggtttgtcactctacccgcctcgcctcctctccgtgccctagccgcgagcagcagcggtttttgttgactgattggattggtttcgtttttatgcagaggtacatgcagcagcacggcgatgtcgagctctccgcgctcgggatgggtgcgtctcgctctcttctccctgttcatccctggtctcgatttggcgttaattcggggggttttagggtaataggagtgactcttgttgagttgattcgacctagaagttattaggttgtgggagaatgcgaacgatgcaagaagttactgtttgcatgtgcaattgctctgttgtttcatgattgacttctggattcccaatctatttttatcaaggtggagcatttttctgtttcgcgtttgtgtgctaggtaatcttagtattgtgttctaggatatgccattttgcgaagcagtgtatagttcaatagttggtcttgaagacttaaacgggtggtctttgtttcgtactcgcattgatatgtacataaccactcataaaatgagtcatgacaccctttcaccactcttcattacctttgcttttggtcaagtaaaactatttactgttcttaaaaaaaaaagaagctgtttataaatgtttaactcttagatgagtgctgactgtcagcccatttggttggtcggtgtagttcgatatttcctttcttcaagatgtggttttatctctactgttgaaggctgtacaaaccatattgtttgatttgtttgcagccattgcaacagttgtaactgttgcggagattcttaagaataacgggtttgctgttgaaaagagtaagaaattctttccttgcaggattgtactgcttgttttgtctcattatctgttttaacattactgttgtgggcatgcagtaatccacaatattcattgttctctactctgaattttccattttcaccttcgcttgtgtagagattagaacatctacggtggaaataaacgatgaatcgagagttcgcccgctccaaaaggctaaggtcaggctttagttgccaaagccacttatttctctcttgttaaaaaaaaaaagcctgttcttaacatgagctttctttctttctttttctaaacctgaactgattgagagctttgtttcctacagattgagatagtgttagaaaagagcgagaaatttgatgagctgatggctgccgcagcggaagagagggaagctgcggaagctgaggagcaggcctgataaggactaggagcgatgacctttgttgtgtgctttgttcctgttgtttaatactgcgtgctttagagatggcatgccttcttgtttaagttgtgtttttcttttatggcagtgtgtgtaaactcagtgactgatgataatctgtttgcaattgagttgagacagcttctc</dnaseqindica>

External Link(s)

NCBI Gene:Os01g0173100, RefSeq:Os01g0173100