Basic Information

Gene Symbol
-
Assembly
GCA_905332935.1
Location
HG995196.1:15948051-15979082[+]

Transcription Factor Domain

TF Family
TF_bZIP
Domain
bZIP domain
PFAM
AnimalTFDB
TF Group
Basic Domians group
Description
bZIP proteins are homo- or heterodimers that contain highly basic DNA binding regions adjacent to regions of α-helix that fold together as coiled coils
Hmmscan Out
# of c-Evalue i-Evalue score bias hmm coord from hmm coord to ali coord from ali coord to env coord from env coord to acc
1 4 0.077 33 4.8 0.0 26 46 99 119 97 126 0.87
2 4 0.077 33 4.8 0.0 26 46 865 885 863 892 0.87
3 4 0.077 33 4.8 0.0 26 46 1430 1450 1428 1457 0.87
4 4 0.077 33 4.8 0.0 26 46 1981 2001 1979 2008 0.87

Sequence Information

Coding Sequence
ATGCTTCGTGTGTTTCGCATGCGAGTGCGATCCGACGACGAAGAGTCGGTCGCCTCGTCGCTGGCCTCGTCGGTTTCGAGAGGCAAGAAGCGGCGAATCGTGTCGACGGTCGCGAGCGCGTCGGAGGAGCTCGCGATCGACGTGAGGACTTGCAGTCCGGCGGACGCTCACGTCGAGTTTACGAGGCAAGCGGCACGAATCATATACGTGTCGTCGTCATCGTCAAACATGAAGGGGACGCATATAAAGATCCTCAGGGACGCGTCCAAGGTAATGCAGGAGCTGTGGACGACGAAGATCGCGGAAATGGAAGCGCGGCTGTTCGCGCTCGAGAAAGAGAACGCAGCCCTTCGCGAAGGGACGTCCAGAATGGCAGTTTGCGCCAACGTGTGCGCGCAGTGCGGCGGCCCGGCTTCCGAGCCCGGCCGTCCTCCGCGAGAGCGCAACAACGACGGTGCGACCGAAGATGACCTGGAAAGGAGGTTTCGGCAGCTAGAAGCCTCCCTGAAGAGAATCGAGGAGCGTCTCGGGGAACGACAACCGTGCGTCCCAGAGCCCCGACCCGTCCCGGAACCCCGACGCGTCGCGGATCCCCCCGCCACTCGCCGCGCCACCCCGACAACTGCGCCCCCGAGGGAGCAGGCGGATGGCGAATGGCGAGTGGTCGAAAGGAGGAGGAACAGGAGAAGAACGGCCGCCGCCGCCGCGGTCGGCAAAGAAGTATCGACAACATCAAGAAGAGGGGCGAACGCGGCCCCACCGCCGTCGACGAGGAGTACAGGCGCGCCAAAACCGAGCGTCGCCGCCACCGCCACGAGGAAGGGGCCGATACCGGCACCGAGGACAATGCCGCCGCCCCCCCGGTCACCGCAGACATCCGCGGTGACCGTAACGTTGAAAGAAGGATCCAGCATGTCGTACGCCGACGTGCTGGCAACGGCCCGAAGGACGATCCCGCTCGCGGAGATCGGCGTGGGGAAACTCGGAATGAAAAAGAGTATGACGGGCGGCATTATAATCGAAGTGCCCGGGGACAAGGACAGGGAAAAAGCGGAGCGGCTAGCGACGCGTCTGTCCGAGGTGCTAGGCCCGGCCACGGCAAAGGTCGCAGCCCCGACGAAGACCGCGGAGCTGAGAGTATCCGGGCTAGATATCTCGGTTACCAAAGAGGAGTTGCGGCAAGCATTGGCCGCAGCGGCGGGTTGCAGAGCCGCCGATATACGAACGGGAGACATCGGCATCGCCAGAGACGGCCTCGCAATGGCCTGGATCAAATGCCCGGAGGTCGCAGCCTGGAAGTTGGCGCAGGCAAGAAGAGTGCCTATGGGGTGGTCGAAGGCAAAGATCAGGGCCATCCCGAAGAAACCCCTGCAGTGCTACAAATGCCTGCAGTACGGTCACGTGCGAGCCACCTGCACGTCCACCGTGAACCGAGAGCACCTGTGCTTCAGGTGCGGCGGAGCCGGACACCGCGCCAAAGGCTGTCCCGCCTCCGCACCCAAGTGCCTCCTGTGCGAGTCGCTCGGGGTGCCCGCCAACCACAAGATGGGCGGAGACGCGTTTTGCCGATCCGGTGCGGAGAATTTCGGAGAAGTTGGTGGGCCCATGTCCGCCAAGACCACTCACCTCACCTTCCCGTGGGGCTCCCCGTCACCCTGCACCGGCTTTGACCGGGACAGGGTGCGAAAGAGAGGACAAGGACGACGCCGGACAAAGAAGAGGCGCGAGCCGCTCCCGCCAGCCGCCACCGCCGCCGAGATGGCGACTGAAGGAACGCGACAAAGAGATGCTACGGGCGGGGGCCATAGTCGCCGCCTGGAGCTGGGACGCCCGACGGATCACCGAGTTGGGAAGCGTGGACGAGGAGGCGGAAAGCCTCCGAAGATACATGAGCGCGGCGTGCGACGCCTCGATGCCGCGCTCCGTTCCACCGCGCGGTTGCCGCCGCGGAGTGTACTGGTGGACGCCGGAGATCGCCGAGATGCGAGCCAACTGCATCCGGGCACGAAGAGGATTTCTGAGGGCCCGCCGTCGCAGGCTGACGCGCGACGAAGAGGAGATCTCTCGCCGCTACGAGGCCTACAGGGAAGCGAGACGCACCCTCCAGCGTGCGATCAAGACGGCGAAAGCTCGGTGCTGGACGGAGCTCGTGGAGACGGTCGAATCCGACCCGTGGGGACGCCCCTACAAGAGGTCCCGGAAGTAGCCCCAAGAAGAGTGACCAGGGGCTCGTTGAAGGCCGCGGAGAACGTGTTCCTTTCTCCGCCTATGGTGATTGCCGACACACCGTCGATCCCAACCGCAGGTGTGTTTCGCATGCGAGTGCGATCCGACGACGAAGAGTCGGTCGCCTCGTCGCTGGCCTCGTCGGTTTCGAGAGGCAAGAAGCGGCGAATCGTGTCGACGGTCGCGAGCGCGTCGGAGGAGCTCGCGATCGACGTGAGGACTTGCAGTCCGGCGGACGCTCACGTCGAGTTTACGAGGCAAGCGGCACGAATCATATACGTGTCGTCGTCATCGTCAAACATGAAGGGGACGCATATAAAGATCCTCAGGGACGCGTCCAAGGTAATGCAGGAGCTGTGGACGACGAAGATCGCGGAAATGGAAGCGCGGCTGTTCGCGCTCGAGAAAGAGAACGCAGCCCTTCGCGAAGGGACGTCCAGAATGGCAGTTTGCGCCAACGTGTGCGCGCAGTGCGGCGGCCCGGCTTCCGAGCCCGGCCGTCCTCCGCGAGAGCGCAACAACGACGGTGCGACCGAAGATGACCTGGAAAGGAGGTTTCGGCAGCTAGAAGCCTCCCTGAAGAGAATCGAGGAGCGTCTCGGGGAACGACAACCGTGCGTCCCAGAGCCCCGACCCGTCCCGGAACCCCGACGCGTCGCGGATCCCCCCGCCACTCGCCGCGCCACCCCGACAGCTGCGCCCCCGAGGGAGCAGGCGGATGGCGAATGGCGAGTGGTCGAAAGGAGGAGGAACAGGAGAAGAACGGCCGCCGCCGCCGCGGTCGGCAAAGAAGTATCGACAACATCAAGAAGAGGGGCGAACGCGGCCCCACCGCCGTCGACGAGGAGTACAGGCGCGCCAAAACCGAGCGTCGCCGCCACCGCCACGAGGAAGGGGCCGATACCGGCACCGAGGACAATGCCGCCGCCCCCCCGGTCACCGCAGACATCCGCGGTGACCGTAACGTTGAAAGAAGGATCCAGCATGTCGTACGCCGACGTGCTGGCAACGGCCCGAAGGACGATCCCGCTCGCGGAGATCGGCGTGGGGAAACTCGGAATGAAAAAGAGTATGACGGGCGGCATTATAATCGAAGTGCCCGGGGACAAGGACAGAGAAAAAGCGGAGCGGCTAGCGACGCGTCTGTCCGAGGTGCTAGGCCCGGCCACGGCAAAGGTCGCAGCCCCGACGAAGACCGCGGAGCTGAGAGTATCCGGGCTAGATATCTCGGTTACCAAAGAGGAGTTGCGGCAAGCATTGGCCGCAGCGGCGGGTTGCAGAGCCGCCGATATACGAACGGGAGACATCGGCATCGCCAGAGACGGCCTCGCAATGGCCTGGATCAAGTGCCCGGAGGTCGCAGCCTGGAAGTTGGCGCAGGCAAGAAGAGTGCCTATGGGGTGGTCGAAGGCAAAGATCAGGGCCATCCCAAAGAAACCCCTGCAGTGCTACAAATGCCTGCAGTACGGTCACGTGCGAGCCACCTGCACGTCCACCGTGAACCGAGAGCACCTGTGCTTCAGGTGCGGCGGAGCCGGACACCGCGCCAAAGGCTGTCCCGCCTCCGCACCCAAGTGCCTCCTGTGCGAGTCGCTCGGGGTGCCCGCCAACCACAAGATGGGCGGAGACGCGTACGCTAGAGCGGACCAAAGGACAACAGAGGTCCCGGAAGTAGCCCCAAGAAGAGTGACCAGGGGCTCGTTGAAGGCCGCGGAGAACGTGTTCCTTTCTCCGCCTATGGTGATTGCCGACACACCGTCGATCCCAACCGCAGGTGTGTTTCGCATGCGAGTGCGATCCGACGACGAAGAGTCGGTCGCCTCGTCGCTGGCCTCGTCGGTTTCGAGAGGCAAGAAGCGGCGAATCGTGTCGACGGTCGCGAGCGCGTCGGAGGAGCTCGCGATCGACGTGAGGACTTGCAGTCCGGCGGACGCTCACGTCGAGTTTACGAGGCAAGCGGCACGAATCATATACGTGTCGTCGTCATCGTCAAACATGAAGGGGACGCATATAAAGATCCTCAGGGACGCGTCCAAGGTAATGCAGGAGCTGTGGACGACGAAGATCGCGGAAATGGAAGCGCGGCTGTTCGCGCTCGAGAAAGAGAACGCAGCCCTTCGCGAAGGGACGTCCAGAATGGCAGTTTGCGCCAACGTGTGCGCGCAGTGCGGCGGCCCGGCTTCCGAGCCCGGCCGTCCTCCGCGAGAGCGCAACAACGACGGTGCGACCGAAGATGACCTGGAAAGGAGGTTTCGGCAGCTCGAAGCCTCCCTGAAGAGAATCGAGGAGCGTCTAGGGGAACGACAACCGCGCGTCCCAGAGCCCCGACCCGTCCCGGAACCCCGACGCGTCGCGGATCCCCCCGCCACTCGCCGCGCCACCCCGACAACTGCGACCCCGAGGGAGCAGGCGGGTGGCGAATGGCGAGTGGTCGAAAGGAGGAGGAACAGGAGAAGAACGGCCGCCGCCGCCGCGGTCGGCGAAGAAGCATCGACAACAACAAGAAGAGGGGCGAACGCGGCCCCACCGCCGTCGACGAGGAGTACAGGCGCGCCAAAACCGAGCGTCGCCGCCACCGCCACGAGGAAGGGGCCGATACCGGCACCGAGGATAATGCCGCCGCCCCCTCGGTCACCGCAGACATCCGCGGTGACCGTAACGTTGAAAGAAGGATCCAGCATGTCGTACGCCGACGTGCTGGCAACGGCCCGAAGGACGATCCCGCTCGCGGAGATCGGCGTGGGGAAACTCGGAATGAAAAAGAGTATGACGGGCGGCATTATAATCGAAGTGCCCGGGGACAAGGACAGAGAAAAAGCGGAGCGGCTAGCGACGCGTCTGTCCGAGGTGCTAGGCCCGGCCACGGCAAAGGTCGCAGCCCCGACGAAGACCGCGGAGCTGAGAGTATCCGGGCTAGATATCTCGGTTACCAAAGAGGAGTTGCGGCAAGCATTGGCCGCAGCGGCGGGTTGCAGAGCCGCCGATATACGAACGGGAGACATCGGCATCGCCAGAGACGGCCTCGCAATGGCCTGGATCAAGTGCCCGGAGGTCGCAGCCTGGAAGTTGGCGCAGGCAAGAAGAGTGCCTATGGGGTGGTCGAAGGCAAAGATCAGGGCCATCCCGAAGAAACCCCTGCAGTGCTACAAATGCCTGCAGTACGGTCACGTGCGAGCCACCTGCACGTCCACCGTGAACCGAGAGCACCTGTGCTTCAGGTGCGGCGGAGCCGGACACCGCGCCAAAGGCTGTCCCGCCTCCGCACCCAAGTGCCTCCTGTGCGAGTCGCTCGGGGTGCCCGCCAACCACAAGATGGGCGGAGACGCGTCCCCAAGAAGAGTGACCAGGGGCTCGTTGAAGGCCGCGGAGAACGTGTTCCTTTCTCCGCCTATGGTGATTGCCGACACACCGTCGATCCCAACCGCAGGTGTGTTTCGCATGCGAGTGCGATCCGACGACGAAGAGTCGGTCGCCTCGTCGCTGGCCTCGTCGGTTTCGAGAGGCAAGAAGCGGCGAATCGTGTCGACGGTCGCGAGCGCGTCGGAGGAGCTCGCGATCGACGTGAGGACTTGCAGTCCGGCGGACGCTCACGTCGAGTTTACGAGGCAAGCGGCACGAATCATATACGTGTCGTCGTCATCGTCAAACATGAAGGGGACGCATATAAAGATCCTCAGGGACGCGTCCAAGGTAATGCAGGAGCTGTGGACGACGAAGATCGCGGAAATGGAAGCGCGGCTGTTCGCGCTCGAGAAAGAGAACGCAGCCCTTCGCGAAGGGACGTCCAGAATGGCAGTTTGCGCCAACGTGTGCGCGCAGTGCGGCGGCCCGGCTTCCGAGCCCGGCCGTCCTCCGCGAGAGCGCAACAACGACGGTGCGACCGAAGAGGACCTGGAAAGGAGATTTCGGCAGCTCGAAGCCTCCCTGAAGAGAATCGAGGAGCGTCTCGGGGAACGACAACCGCGCGTCCCAGAGCCCCGACCCGTCCCGGAACCCCGACGCGTCGCGGATCCCCCCGCCACTCGCCGCGCCACCCCGACAACTGCGCCCCCGAGGGAGCAGGCGGATGGCGAATGGCGAGTGGTCGAAAGGAGGAGGAACAGGAGAAGAACGGCCGCCGCCGCCGCGGTCGGCGAAGAAGGATCGACAACATCAAGAAGAGGGGCGAACGCGGCCCCACCGCCGTCGACGAGGAGTACAGGCGCGCCAAAACCGAGCGTCGCCGCCACCGCCACGAGGAAGGGGCCGATACCGGCACCGAGGACAATGCCGCTGCCCCCTCGGTCACCGCAGACATCCGCGGTGACCGTAACGTTGAAAGAAGGATCCAGCATGTCGTACGCCGACGTGCTGGCAACGGCCCGAAGGACGATCCCGCTCGCGGAGATCGGCGTGGGGAAACTCGGAATGAAAAAGAGTATGACGGGCGGTATTATAATCGAAGTGCCCGGGGACAAGGACAGAGAAAAAGCGGAGCGGCTAGCGACGCGTCTGTCCGAGGTGCTAGGCCCGGCCACGGCAAAGGTCGCAGCCCCGACGAAGACCGCGGAGCTGAGAGTATCCGGGCTAGATATCTCGGTTACCAAAGAGGAGTTGCGGCAAGCATTGGCCGCAGCGGCGGGTTGCAGAGCCGCCGATATACGAACGGGAGACATCGGCATCGCCAGAGACGGCCTCGCAATGGCCTGGATCAAGTGCCCGGAGGTCGCAGCCTGGAAGTTGGCGCAGGCAAGAAGAGTGCCTATGGGGTGGTCGAAGGCAAAGATCAGGGCCATCCCGAAGAAACCCCTGCAGTGCTACAAATGCCTGCAGTACGGTCACGTGCGAGCCACCTGCACGTCCACCGTGAACCGAGAGCACCTGTGCTTCAGGTGCGGCGGAGCCGGACACCGCGCCAAAGGCTGTCCCGCCTCCGCACCCAAGTGCCTCCTGTGCGAGTCGCTCGGGGTGCCCGCCAACCACAAGATGGGCGGAGACGCGTATCGTCGTCGTCGTCGTCGCGACGCAGAAGAGACACCGCGCGCCGCAGGGTTCGCGGCGCTGCCGCCGCCGCCGCCACCCCAAGAGCGTCGGACCAAGATCGAGGCGCTAGGGGGAGTGGTCCCCCCTAGCGACCAAACCAACAAGTGA
Protein Sequence
MLRVFRMRVRSDDEESVASSLASSVSRGKKRRIVSTVASASEELAIDVRTCSPADAHVEFTRQAARIIYVSSSSSNMKGTHIKILRDASKVMQELWTTKIAEMEARLFALEKENAALREGTSRMAVCANVCAQCGGPASEPGRPPRERNNDGATEDDLERRFRQLEASLKRIEERLGERQPCVPEPRPVPEPRRVADPPATRRATPTTAPPREQADGEWRVVERRRNRRRTAAAAAVGKEVSTTSRRGANAAPPPSTRSTGAPKPSVAATATRKGPIPAPRTMPPPPRSPQTSAVTVTLKEGSSMSYADVLATARRTIPLAEIGVGKLGMKKSMTGGIIIEVPGDKDREKAERLATRLSEVLGPATAKVAAPTKTAELRVSGLDISVTKEELRQALAAAAGCRAADIRTGDIGIARDGLAMAWIKCPEVAAWKLAQARRVPMGWSKAKIRAIPKKPLQCYKCLQYGHVRATCTSTVNREHLCFRCGGAGHRAKGCPASAPKCLLCESLGVPANHKMGGDAFCRSGAENFGEVGGPMSAKTTHLTFPWGSPSPCTGFDRDRVRKRGQGRRRTKKRREPLPPAATAAEMATEGTRQRDATGGGHSRRLELGRPTDHRVGKRGRGGGKPPKIHERGVRRLDAALRSTARLPPRSVLVDAGDRRDASQLHPGTKRISEGPPSQADARRRGDLSPLRGLQGSETHPPACDQDGESSVLDGARGDGRIRPVGTPLQEVPEVAPRRVTRGSLKAAENVFLSPPMVIADTPSIPTAGVFRMRVRSDDEESVASSLASSVSRGKKRRIVSTVASASEELAIDVRTCSPADAHVEFTRQAARIIYVSSSSSNMKGTHIKILRDASKVMQELWTTKIAEMEARLFALEKENAALREGTSRMAVCANVCAQCGGPASEPGRPPRERNNDGATEDDLERRFRQLEASLKRIEERLGERQPCVPEPRPVPEPRRVADPPATRRATPTAAPPREQADGEWRVVERRRNRRRTAAAAAVGKEVSTTSRRGANAAPPPSTRSTGAPKPSVAATATRKGPIPAPRTMPPPPRSPQTSAVTVTLKEGSSMSYADVLATARRTIPLAEIGVGKLGMKKSMTGGIIIEVPGDKDREKAERLATRLSEVLGPATAKVAAPTKTAELRVSGLDISVTKEELRQALAAAAGCRAADIRTGDIGIARDGLAMAWIKCPEVAAWKLAQARRVPMGWSKAKIRAIPKKPLQCYKCLQYGHVRATCTSTVNREHLCFRCGGAGHRAKGCPASAPKCLLCESLGVPANHKMGGDAYARADQRTTEVPEVAPRRVTRGSLKAAENVFLSPPMVIADTPSIPTAGVFRMRVRSDDEESVASSLASSVSRGKKRRIVSTVASASEELAIDVRTCSPADAHVEFTRQAARIIYVSSSSSNMKGTHIKILRDASKVMQELWTTKIAEMEARLFALEKENAALREGTSRMAVCANVCAQCGGPASEPGRPPRERNNDGATEDDLERRFRQLEASLKRIEERLGERQPRVPEPRPVPEPRRVADPPATRRATPTTATPREQAGGEWRVVERRRNRRRTAAAAAVGEEASTTTRRGANAAPPPSTRSTGAPKPSVAATATRKGPIPAPRIMPPPPRSPQTSAVTVTLKEGSSMSYADVLATARRTIPLAEIGVGKLGMKKSMTGGIIIEVPGDKDREKAERLATRLSEVLGPATAKVAAPTKTAELRVSGLDISVTKEELRQALAAAAGCRAADIRTGDIGIARDGLAMAWIKCPEVAAWKLAQARRVPMGWSKAKIRAIPKKPLQCYKCLQYGHVRATCTSTVNREHLCFRCGGAGHRAKGCPASAPKCLLCESLGVPANHKMGGDASPRRVTRGSLKAAENVFLSPPMVIADTPSIPTAGVFRMRVRSDDEESVASSLASSVSRGKKRRIVSTVASASEELAIDVRTCSPADAHVEFTRQAARIIYVSSSSSNMKGTHIKILRDASKVMQELWTTKIAEMEARLFALEKENAALREGTSRMAVCANVCAQCGGPASEPGRPPRERNNDGATEEDLERRFRQLEASLKRIEERLGERQPRVPEPRPVPEPRRVADPPATRRATPTTAPPREQADGEWRVVERRRNRRRTAAAAAVGEEGSTTSRRGANAAPPPSTRSTGAPKPSVAATATRKGPIPAPRTMPLPPRSPQTSAVTVTLKEGSSMSYADVLATARRTIPLAEIGVGKLGMKKSMTGGIIIEVPGDKDREKAERLATRLSEVLGPATAKVAAPTKTAELRVSGLDISVTKEELRQALAAAAGCRAADIRTGDIGIARDGLAMAWIKCPEVAAWKLAQARRVPMGWSKAKIRAIPKKPLQCYKCLQYGHVRATCTSTVNREHLCFRCGGAGHRAKGCPASAPKCLLCESLGVPANHKMGGDAYRRRRRRDAEETPRAAGFAALPPPPPPQERRTKIEALGGVVPPSDQTNK

Similar Transcription Factors

Sequence clustering based on sequence similarity using MMseqs2

100% Identity
iTF_00220268;
90% Identity
iTF_00220293;
80% Identity
iTF_00220293;