Skip Header

You are using a version of browser that may not display all the features of this website. Please consider upgrading your browser.
Entry version 168 (16 Oct 2019)
Sequence version 2 (01 Mar 2005)
Previous versions | rss
Help videoAdd a publicationFeedback
Protein

Pulmonary surfactant-associated protein A1

Gene

SFTPA1

Organism
Homo sapiens (Human)
Status
Reviewed-Annotation score:

Annotation score:5 out of 5

<p>The annotation score provides a heuristic measure of the annotation content of a UniProtKB entry or proteome. This score <strong>cannot</strong> be used as a measure of the accuracy of the annotation as we cannot define the ‘correct annotation’ for any given protein.<p><a href='/help/annotation_score' target='_top'>More...</a></p>
-Experimental evidence at protein leveli <p>This indicates the type of evidence that supports the existence of the protein. Note that the ‘protein existence’ evidence does not give information on the accuracy or correctness of the sequence(s) displayed.<p><a href='/help/protein_existence' target='_top'>More...</a></p>

<p>This section provides any useful information about the protein, mostly biological knowledge.<p><a href='/help/function_section' target='_top'>More...</a></p>Functioni

In presence of calcium ions, it binds to surfactant phospholipids and contributes to lower the surface tension at the air-liquid interface in the alveoli of the mammalian lung and is essential for normal respiration. Enhances the expression of MYO18A/SP-R210 on alveolar macrophages (By similarity).By similarity
(Microbial infection) Recognition of M.tuberculosis by dendritic cells may occur partially via this molecule (PubMed:17158455, PubMed:21203928). Can recognize, bind, and opsonize pathogens to enhance their elimination by alveolar macrophages (PubMed:21123169).3 Publications

Miscellaneous

Pulmonary surfactant consists of 90% lipid and 10% protein. There are 4 surfactant-associated proteins: 2 collagenous, carbohydrate-binding glycoproteins (SP-A and SP-D) and 2 small hydrophobic proteins (SP-B and SP-C).

<p>The <a href="http://www.geneontology.org/">Gene Ontology (GO)</a> project provides a set of hierarchical controlled vocabulary split into 3 categories:<p><a href='/help/gene_ontology' target='_top'>More...</a></p>GO - Molecular functioni

GO - Biological processi

<p>UniProtKB Keywords constitute a <a href="http://www.uniprot.org/keywords">controlled vocabulary</a> with a hierarchical structure. Keywords summarise the content of a UniProtKB entry and facilitate the search for proteins of interest.<p><a href='/help/keywords' target='_top'>More...</a></p>Keywordsi

Biological processGaseous exchange
LigandCalcium, Lectin

Enzyme and pathway databases

Reactome - a knowledgebase of biological pathways and processes

More...
Reactomei
R-HSA-166016 Toll Like Receptor 4 (TLR4) Cascade
R-HSA-168179 Toll Like Receptor TLR1:TLR2 Cascade
R-HSA-391160 Signal regulatory protein family interactions
R-HSA-5683826 Surfactant metabolism
R-HSA-5686938 Regulation of TLR by endogenous ligand
R-HSA-5688849 Defective CSF2RB causes pulmonary surfactant metabolism dysfunction 5 (SMDP5)
R-HSA-5688890 Defective CSF2RA causes pulmonary surfactant metabolism dysfunction 4 (SMDP4)

<p>This section provides information about the protein and gene name(s) and synonym(s) and about the organism that is the source of the protein sequence.<p><a href='/help/names_and_taxonomy_section' target='_top'>More...</a></p>Names & Taxonomyi

<p>This subsection of the <a href="http://www.uniprot.org/help/names_and_taxonomy_section">Names and taxonomy</a> section provides an exhaustive list of all names of the protein, from commonly used to obsolete, to allow unambiguous identification of a protein.<p><a href='/help/protein_names' target='_top'>More...</a></p>Protein namesi
Recommended name:
Pulmonary surfactant-associated protein A1
Short name:
PSP-A
Short name:
PSPA
Short name:
SP-A
Short name:
SP-A1
Alternative name(s):
35 kDa pulmonary surfactant-associated protein
Alveolar proteinosis protein
Collectin-4
<p>This subsection of the <a href="http://www.uniprot.org/help/names_and_taxonomy_section">Names and taxonomy</a> section indicates the name(s) of the gene(s) that code for the protein sequence(s) described in the entry. Four distinct tokens exist: ‘Name’, ‘Synonyms’, ‘Ordered locus names’ and ‘ORF names’.<p><a href='/help/gene_name' target='_top'>More...</a></p>Gene namesi
Name:SFTPA1
Synonyms:COLEC4, PSAP, SFTP1, SFTPA, SFTPA1B
<p>This subsection of the <a href="http://www.uniprot.org/help/names_and_taxonomy_section">Names and taxonomy</a> section provides information on the name(s) of the organism that is the source of the protein sequence.<p><a href='/help/organism-name' target='_top'>More...</a></p>OrganismiHomo sapiens (Human)
<p>This subsection of the <a href="http://www.uniprot.org/help/names_and_taxonomy_section">Names and taxonomy</a> section shows the unique identifier assigned by the NCBI to the source organism of the protein. This is known as the ‘taxonomic identifier’ or ‘taxid’.<p><a href='/help/taxonomic_identifier' target='_top'>More...</a></p>Taxonomic identifieri9606 [NCBI]
<p>This subsection of the <a href="http://www.uniprot.org/help/names_and_taxonomy_section">Names and taxonomy</a> section contains the taxonomic hierarchical classification lineage of the source organism. It lists the nodes as they appear top-down in the taxonomic tree, with the more general grouping listed first.<p><a href='/help/taxonomic_lineage' target='_top'>More...</a></p>Taxonomic lineageiEukaryotaMetazoaChordataCraniataVertebrataEuteleostomiMammaliaEutheriaEuarchontogliresPrimatesHaplorrhiniCatarrhiniHominidaeHomo
<p>This subsection of the <a href="http://www.uniprot.org/help/names_and_taxonomy_section">Names and taxonomy</a> section is present for entries that are part of a <a href="http://www.uniprot.org/proteomes">proteome</a>, i.e. of a set of proteins thought to be expressed by organisms whose genomes have been completely sequenced.<p><a href='/help/proteomes_manual' target='_top'>More...</a></p>Proteomesi
  • UP000005640 <p>A UniProt <a href="http://www.uniprot.org/manual/proteomes_manual">proteome</a> can consist of several components. <br></br>The component name refers to the genomic component encoding a set of proteins.<p><a href='/help/proteome_component' target='_top'>More...</a></p> Componenti: Chromosome 10

Organism-specific databases

Human Gene Nomenclature Database

More...
HGNCi
HGNC:10798 SFTPA1

Online Mendelian Inheritance in Man (OMIM)

More...
MIMi
178630 gene

neXtProt; the human protein knowledge platform

More...
neXtProti
NX_Q8IWL2

<p>This section provides information on the location and the topology of the mature protein in the cell.<p><a href='/help/subcellular_location_section' target='_top'>More...</a></p>Subcellular locationi

Extracellular region or secreted Cytosol Plasma membrane Cytoskeleton Lysosome Endosome Peroxisome ER Golgi apparatus Nucleus Mitochondrion Manual annotation Automatic computational assertionGraphics by Christian Stolte & Seán O’Donoghue; Source: COMPARTMENTS

Keywords - Cellular componenti

Extracellular matrix, Secreted, Surface film

<p>This section provides information on the disease(s) and phenotype(s) associated with a protein.<p><a href='/help/pathology_and_biotech_section' target='_top'>More...</a></p>Pathology & Biotechi

<p>This subsection of the ‘Pathology and Biotech’ section provides information on the disease(s) associated with genetic variations in a given protein. The information is extracted from the scientific literature and diseases that are also described in the <a href="http://www.ncbi.nlm.nih.gov/sites/entrez?db=omim">OMIM</a> database are represented with a <a href="http://www.uniprot.org/diseases">controlled vocabulary</a> in the following way:<p><a href='/help/involvement_in_disease' target='_top'>More...</a></p>Involvement in diseasei

Pulmonary fibrosis, idiopathic (IPF)
Disease susceptibility is associated with variations affecting the gene represented in this entry.
Disease descriptionA lung disease characterized by shortness of breath, radiographically evident diffuse pulmonary infiltrates, and varying degrees of inflammation and fibrosis on biopsy. In some cases, the disorder can be rapidly progressive and characterized by sequential acute lung injury with subsequent scarring and end-stage lung disease.
Related information in OMIM
Respiratory distress syndrome in premature infants (RDS)1 Publication
Disease susceptibility may be associated with variations affecting the gene represented in this entry. The association between SFTPA1 alleles and respiratory distress syndrome in premature infants is dependent on a variation Ile to Thr at position 131 in SFTPB.
Disease descriptionA lung disease affecting usually premature newborn infants. It is characterized by deficient gas exchange, diffuse atelectasis, high-permeability lung edema and fibrin-rich alveolar deposits called 'hyaline membranes'.
Related information in OMIM

Organism-specific databases

DisGeNET

More...
DisGeNETi
653509

MalaCards human disease database

More...
MalaCardsi
SFTPA1
MIMi178500 phenotype
267450 phenotype

Open Targets

More...
OpenTargetsi
ENSG00000122852

Orphanet; a database dedicated to information on rare diseases and orphan drugs

More...
Orphaneti
2032 Idiopathic pulmonary fibrosis

The Pharmacogenetics and Pharmacogenomics Knowledge Base

More...
PharmGKBi
PA35710

Miscellaneous databases

Pharos NIH Druggable Genome Knowledgebase

More...
Pharosi
Q8IWL2

Chemistry databases

Drug and drug target database

More...
DrugBanki
DB03814 2-(N-Morpholino)-Ethanesulfonic Acid

Polymorphism and mutation databases

BioMuta curated single-nucleotide variation and disease association database

More...
BioMutai
SFTPA1

Domain mapping of disease mutations (DMDM)

More...
DMDMi
60416440

<p>This section describes post-translational modifications (PTMs) and/or processing events.<p><a href='/help/ptm_processing_section' target='_top'>More...</a></p>PTM / Processingi

Molecule processing

Feature keyPosition(s)DescriptionActionsGraphical viewLength
<p>This subsection of the ‘PTM / Processing’ section denotes the presence of an N-terminal signal peptide.<p><a href='/help/signal' target='_top'>More...</a></p>Signal peptidei1 – 20Add BLAST20
<p>This subsection of the ‘PTM / Processing’ section describes the extent of a polypeptide chain in the mature protein following processing.<p><a href='/help/chain' target='_top'>More...</a></p>ChainiPRO_000001745721 – 248Pulmonary surfactant-associated protein A1Add BLAST228

Amino acid modifications

Feature keyPosition(s)DescriptionActionsGraphical viewLength
<p>This subsection of the PTM / Processing":/help/ptm_processing_section section describes the positions of cysteine residues participating in disulfide bonds.<p><a href='/help/disulfid' target='_top'>More...</a></p>Disulfide bondi26Interchain1 Publication
<p>This subsection of the ‘PTM / Processing’ section specifies the position and type of each modified residue excluding <a href="http://www.uniprot.org/manual/lipid">lipids</a>, <a href="http://www.uniprot.org/manual/carbohyd">glycans</a> and <a href="http://www.uniprot.org/manual/crosslnk">protein cross-links</a>.<p><a href='/help/mod_res' target='_top'>More...</a></p>Modified residuei304-hydroxyprolineBy similarity1
Modified residuei334-hydroxyprolineBy similarity1
Modified residuei364-hydroxyprolineBy similarity1
Modified residuei424-hydroxyprolineBy similarity1
Modified residuei544-hydroxyprolineBy similarity1
Modified residuei574-hydroxyprolineBy similarity1
Modified residuei634-hydroxyprolineBy similarity1
Modified residuei674-hydroxyprolineBy similarity1
Modified residuei704-hydroxyprolineBy similarity1
Disulfide bondi155 ↔ 246PROSITE-ProRule annotation1 Publication
<p>This subsection of the <a href="http://www.uniprot.org/help/ptm_processing_section">PTM / Processing</a> section specifies the position and type of each covalently attached glycan group (mono-, di-, or polysaccharide).<p><a href='/help/carbohyd' target='_top'>More...</a></p>Glycosylationi207N-linked (GlcNAc...) asparagineCurated1
Disulfide bondi224 ↔ 238PROSITE-ProRule annotation1 Publication

<p>This subsection of the <a href="http://www.uniprot.org/help/ptm_processing_section">PTM/processing</a> section describes post-translational modifications (PTMs). This subsection <strong>complements</strong> the information provided at the sequence level or describes modifications for which <strong>position-specific data is not yet available</strong>.<p><a href='/help/post-translational_modification' target='_top'>More...</a></p>Post-translational modificationi

N-acetylated.1 Publication

Keywords - PTMi

Acetylation, Disulfide bond, Glycoprotein, Hydroxylation

Proteomic databases

The CPTAC Assay portal

More...
CPTACi
CPTAC-1214

MassIVE - Mass Spectrometry Interactive Virtual Environment

More...
MassIVEi
Q8IWL2

PaxDb, a database of protein abundance averages across all three domains of life

More...
PaxDbi
Q8IWL2

PeptideAtlas

More...
PeptideAtlasi
Q8IWL2

PRoteomics IDEntifications database

More...
PRIDEi
Q8IWL2

ProteomicsDB: a multi-organism proteome resource

More...
ProteomicsDBi
33955
70869 [Q8IWL2-1]

PTM databases

iPTMnet integrated resource for PTMs in systems biology context

More...
iPTMneti
Q8IWL2

Comprehensive resource for the study of protein post-translational modifications (PTMs) in human, mouse and rat.

More...
PhosphoSitePlusi
Q8IWL2

<p>This section provides information on the expression of a gene at the mRNA or protein level in cells or in tissues of multicellular organisms.<p><a href='/help/expression_section' target='_top'>More...</a></p>Expressioni

Gene expression databases

Bgee dataBase for Gene Expression Evolution

More...
Bgeei
ENSG00000122852 Expressed in 68 organ(s), highest expression level in right lung

ExpressionAtlas, Differential and Baseline Expression

More...
ExpressionAtlasi
Q8IWL2 baseline and differential

Genevisible search portal to normalized and curated expression data from Genevestigator

More...
Genevisiblei
Q8IWL2 HS

Organism-specific databases

Human Protein Atlas

More...
HPAi
CAB016793
HPA042638
HPA045752
HPA049368

<p>This section provides information on the quaternary structure of a protein and on interaction(s) with other proteins or protein complexes.<p><a href='/help/interaction_section' target='_top'>More...</a></p>Interactioni

<p>This subsection of the <a href="http://www.uniprot.org/help/interaction_section">'Interaction'</a> section provides information about the protein quaternary structure and interaction(s) with other proteins or protein complexes (with the exception of physiological receptor-ligand interactions which are annotated in the <a href="http://www.uniprot.org/help/function_section">'Function'</a> section).<p><a href='/help/subunit_structure' target='_top'>More...</a></p>Subunit structurei

Oligomeric complex of 6 set of homotrimers.

1 Publication

(Microbial infection) Binds M.bovis cell surface protein Apa via its glycosylated sites; probably also recognizes other bacterial moieties.

1 Publication

(Microbial infection) Binds to the S.aureus extracellular adherence protein, Eap, thereby enhancing phagocytosis and killing of S.aureus by alveolar macrophages.

1 Publication

<p>This subsection of the '<a href="http://www.uniprot.org/help/interaction_section%27">Interaction</a> section provides information about binary protein-protein interactions. The data presented in this section are a quality-filtered subset of binary interactions automatically derived from the <a href="http://www.ebi.ac.uk/intact/">IntAct database</a>. It is updated on a monthly basis. Each binary interaction is displayed on a separate line.<p><a href='/help/binary_interactions' target='_top'>More...</a></p>Binary interactionsi

Protein-protein interaction databases

The Biological General Repository for Interaction Datasets (BioGrid)

More...
BioGridi
575839, 4 interactors

Protein interaction database and analysis system

More...
IntActi
Q8IWL2, 2 interactors

STRING: functional protein association networks

More...
STRINGi
9606.ENSP00000397082

<p>This section provides information on the tertiary and secondary structure of a protein.<p><a href='/help/structure_section' target='_top'>More...</a></p>Structurei

3D structure databases

SWISS-MODEL Repository - a database of annotated 3D protein structure models

More...
SMRi
Q8IWL2

Database of comparative protein structure models

More...
ModBasei
Search...

<p>This section provides information on sequence similarities with other proteins and the domain(s) present in a protein.<p><a href='/help/family_and_domains_section' target='_top'>More...</a></p>Family & Domainsi

Domains and Repeats

Feature keyPosition(s)DescriptionActionsGraphical viewLength
<p>This subsection of the <a href="http://www.uniprot.org/help/family_and_domains_section">Family and Domains</a> section describes the position and type of a domain, which is defined as a specific combination of secondary structures organized into a characteristic three-dimensional structure or fold.<p><a href='/help/domain' target='_top'>More...</a></p>Domaini28 – 100Collagen-likeAdd BLAST73
Domaini132 – 248C-type lectinPROSITE-ProRule annotationAdd BLAST117

<p>This subsection of the ‘Family and domains’ section provides information about the sequence similarity with other proteins.<p><a href='/help/sequence_similarities' target='_top'>More...</a></p>Sequence similaritiesi

Belongs to the SFTPA family.Curated

Keywords - Domaini

Collagen, Signal

Phylogenomic databases

evolutionary genealogy of genes: Non-supervised Orthologous Groups

More...
eggNOGi
KOG4297 Eukaryota
ENOG410XPJ1 LUCA

Ensembl GeneTree

More...
GeneTreei
ENSGT00940000156653

InParanoid: Eukaryotic Ortholog Groups

More...
InParanoidi
Q8IWL2

KEGG Orthology (KO)

More...
KOi
K10067

Identification of Orthologs from Complete Genome Data

More...
OMAi
ATQEACT

Database of Orthologous Groups

More...
OrthoDBi
1172460at2759

Database for complete collections of gene phylogenies

More...
PhylomeDBi
Q8IWL2

TreeFam database of animal gene trees

More...
TreeFami
TF330481

Family and domain databases

Conserved Domains Database

More...
CDDi
cd03591 CLECT_collectin_like, 1 hit

Gene3D Structural and Functional Annotation of Protein Families

More...
Gene3Di
3.10.100.10, 1 hit

Integrated resource of protein families, domains and functional sites

More...
InterProi
View protein in InterPro
IPR001304 C-type_lectin-like
IPR016186 C-type_lectin-like/link_sf
IPR018378 C-type_lectin_CS
IPR033990 Collectin_CTLD
IPR016187 CTDL_fold

Pfam protein domain database

More...
Pfami
View protein in Pfam
PF00059 Lectin_C, 1 hit

Simple Modular Architecture Research Tool; a protein domain database

More...
SMARTi
View protein in SMART
SM00034 CLECT, 1 hit

Superfamily database of structural and functional annotation

More...
SUPFAMi
SSF56436 SSF56436, 1 hit

PROSITE; a protein domain and family database

More...
PROSITEi
View protein in PROSITE
PS00615 C_TYPE_LECTIN_1, 1 hit
PS50041 C_TYPE_LECTIN_2, 1 hit

<p>This section displays by default the canonical protein sequence and upon request all isoforms described in the entry. It also includes information pertinent to the sequence(s), including <a href="http://www.uniprot.org/help/sequence_length">length</a> and <a href="http://www.uniprot.org/help/sequences">molecular weight</a>. The information is filed in different subsections. The current subsections and their content are listed below:<p><a href='/help/sequences_section' target='_top'>More...</a></p>Sequences (2+)i

<p>This subsection of the <a href="http://www.uniprot.org/help/sequences_section">Sequence</a> section indicates if the <a href="http://www.uniprot.org/help/canonical_and_isoforms">canonical sequence</a> displayed by default in the entry is complete or not.<p><a href='/help/sequence_status' target='_top'>More...</a></p>Sequence statusi: Complete.

<p>This subsection of the <a href="http://www.uniprot.org/help/sequences_section">Sequence</a> section indicates if the <a href="http://www.uniprot.org/help/canonical_and_isoforms">canonical sequence</a> displayed by default in the entry is in its mature form or if it represents the precursor.<p><a href='/help/sequence_processing' target='_top'>More...</a></p>Sequence processingi: The displayed sequence is further processed into a mature form.

This entry describes 2 <p>This subsection of the ‘Sequence’ section lists the alternative protein sequences (isoforms) that can be generated from the same gene by a single or by the combination of up to four biological events (alternative promoter usage, alternative splicing, alternative initiation and ribosomal frameshifting). Additionally, this section gives relevant information on each alternative protein isoform.<p><a href='/help/alternative_products' target='_top'>More...</a></p> isoformsi produced by alternative splicing. AlignAdd to basket

This entry has 2 described isoforms and 1 potential isoform that is computationally mapped.Show allAlign All

Isoform 1 (identifier: Q8IWL2-1) [UniParc]FASTAAdd to basket

This isoform has been chosen as the <div> <p><b>What is the canonical sequence?</b><p><a href='/help/canonical_and_isoforms' target='_top'>More...</a></p>canonicali sequence. All positional information in this entry refers to it. This is also the sequence that appears in the downloadable versions of the entry.

« Hide
        10         20         30         40         50
MWLCPLALNL ILMAASGAVC EVKDVCVGSP GIPGTPGSHG LPGRDGRDGL
60 70 80 90 100
KGDPGPPGPM GPPGEMPCPP GNDGLPGAPG IPGECGEKGE PGERGPPGLP
110 120 130 140 150
AHLDEELQAT LHDFRHQILQ TRGALSLQGS IMTVGEKVFS SNGQSITFDA
160 170 180 190 200
IQEACARAGG RIAVPRNPEE NEAIASFVKK YNTYAYVGLT EGPSPGDFRY
210 220 230 240
SDGTPVNYTN WYRGEPAGRG KEQCVEMYTD GQWNDRNCLY SRLTICEF
Length:248
Mass (Da):26,242
Last modified:March 1, 2005 - v2
<p>The checksum is a form of redundancy check that is calculated from the sequence. It is useful for tracking sequence updates.</p> <p>It should be noted that while, in theory, two different sequences could have the same checksum value, the likelihood that this would happen is extremely low.</p> <p>However UniProtKB may contain entries with identical sequences in case of multiple genes (paralogs).</p> <p>The checksum is computed as the sequence 64-bit Cyclic Redundancy Check value (CRC64) using the generator polynomial: x<sup>64</sup> + x<sup>4</sup> + x<sup>3</sup> + x + 1. The algorithm is described in the ISO 3309 standard. </p> <p class="publication">Press W.H., Flannery B.P., Teukolsky S.A. and Vetterling W.T.<br /> <strong>Cyclic redundancy and other checksums</strong><br /> <a href="http://www.nrbook.com/b/bookcpdf.php">Numerical recipes in C 2nd ed., pp896-902, Cambridge University Press (1993)</a>)</p> Checksum:iAFFFCF38B87BE081
GO
Isoform 2 (identifier: Q8IWL2-2) [UniParc]FASTAAdd to basket

The sequence of this isoform differs from the canonical sequence as follows:
     1-1: M → MRPCQVPGAATGPRAM

Show »
Length:263
Mass (Da):27,736
Checksum:i39F4A2E09F3F0AC3
GO

<p>In eukaryotic reference proteomes, unreviewed entries that are likely to belong to the same gene are computationally mapped, based on gene identifiers from Ensembl, EnsemblGenomes and model organism databases.<p><a href='/help/gene_centric_isoform_mapping' target='_top'>More...</a></p>Computationally mapped potential isoform sequencesi

There is 1 potential isoform mapped to this entry.BLASTAlignShow allAdd to basket
EntryEntry nameProtein names
Gene namesLengthAnnotation
A0A0C4DG36A0A0C4DG36_HUMAN
Pulmonary surfactant-associated pro...
SFTPA1
158Annotation score:

Annotation score:1 out of 5

<p>The annotation score provides a heuristic measure of the annotation content of a UniProtKB entry or proteome. This score <strong>cannot</strong> be used as a measure of the accuracy of the annotation as we cannot define the ‘correct annotation’ for any given protein.<p><a href='/help/annotation_score' target='_top'>More...</a></p>

Experimental Info

Feature keyPosition(s)DescriptionActionsGraphical viewLength
<p>This subsection of the ‘Sequence’ section reports difference(s) between the canonical sequence (displayed by default in the entry) and the different sequence submissions merged in the entry. These various submissions may originate from different sequencing projects, different types of experiments, or different biological samples. Sequence conflicts are usually of unknown origin.<p><a href='/help/conflict' target='_top'>More...</a></p>Sequence conflicti45D → H in AAA36510 (PubMed:2995821).Curated1
Sequence conflicti54P → L in AAA36510 (PubMed:2995821).Curated1
Sequence conflicti100P → R in AAA36510 (PubMed:2995821).Curated1

<p>This subsection of the ‘Sequence’ section provides information on polymorphic variants. If the variant is associated with a disease state, the description of the latter can be found in the <a href="http://www.uniprot.org/manual/involvement_in_disease">'Involvement in disease'</a> subsection.<p><a href='/help/polymorphism' target='_top'>More...</a></p>Polymorphismi

At least 5 allelic variants of SFTPA1 are known: 6A, 6A2, 6A3, 6A4 and 6A5. The sequence shown is that of allele 6A3.

Natural variant

Feature keyPosition(s)DescriptionActionsGraphical viewLength
<p>This subsection of the ‘Sequence’ section describes natural variant(s) of the protein sequence.<p><a href='/help/variant' target='_top'>More...</a></p>Natural variantiVAR_0635175P → L1 PublicationCorresponds to variant dbSNP:rs72659389Ensembl.1
Natural variantiVAR_0041849N → T2 PublicationsCorresponds to variant dbSNP:rs139899873Ensembl.1
Natural variantiVAR_02129219V → A in allele 6A and allele 6A(5). 4 PublicationsCorresponds to variant dbSNP:rs1059047Ensembl.1
Natural variantiVAR_01223150L → V in allele 6A(2). 3 PublicationsCorresponds to variant dbSNP:rs1136450Ensembl.1
Natural variantiVAR_012232219R → W Associated with susceptibility to idiopathic pulmonary fibrosis in smokers; allele 6A(4) and allele 6A(5). 4 PublicationsCorresponds to variant dbSNP:rs4253527Ensembl.1
Natural variantiVAR_012233223Q → K. Corresponds to variant dbSNP:rs1965708EnsemblClinVar.1

Alternative sequence

Feature keyPosition(s)DescriptionActionsGraphical viewLength
<p>This subsection of the ‘Sequence’ section describes the sequence of naturally occurring alternative protein isoform(s). The changes in the amino acid sequence may be due to alternative splicing, alternative promoter usage, alternative initiation, or ribosomal frameshifting.<p><a href='/help/var_seq' target='_top'>More...</a></p>Alternative sequenceiVSP_0468021M → MRPCQVPGAATGPRAM in isoform 2. Curated1

Sequence databases

Select the link destinations:

EMBL nucleotide sequence database

More...
EMBLi

GenBank nucleotide sequence database

More...
GenBanki

DNA Data Bank of Japan; a nucleotide sequence database

More...
DDBJi
Links Updated
M30838 Genomic DNA Translation: AAA36510.1
M13686 mRNA Translation: AAA60211.1
HQ021433 mRNA Translation: ADO27676.1
HQ021434 mRNA Translation: ADO27677.1
HQ021435 mRNA Translation: ADO27678.1
HQ021436 mRNA Translation: ADO27679.1
HQ021437 mRNA Translation: ADO27680.1
HQ021438 mRNA Translation: ADO27681.1
HQ021439 mRNA Translation: ADO27682.1
HQ021440 mRNA Translation: ADO27683.1
HQ021441 mRNA Translation: ADO27684.1
HQ021442 mRNA Translation: ADO27685.1
AK290703 mRNA Translation: BAF83392.1
AY198391 Genomic DNA Translation: AAO13486.1
BX248123 Genomic DNA No translation available.
CH471083 Genomic DNA Translation: EAW54657.1
CH471083 Genomic DNA Translation: EAW54651.1
CH471083 Genomic DNA Translation: EAW54658.1
BC029913 mRNA Translation: AAH29913.1
BC111570 mRNA Translation: AAI11571.1
BC171875 mRNA Translation: AAI71875.1

The Consensus CDS (CCDS) project

More...
CCDSi
CCDS44444.2 [Q8IWL2-2]
CCDS44445.1 [Q8IWL2-1]

Protein sequence database of the Protein Information Resource

More...
PIRi
A24622 LNHUPS
A25720 LNHUP6

NCBI Reference Sequences

More...
RefSeqi
NP_001087239.2, NM_001093770.2 [Q8IWL2-2]
NP_001158116.1, NM_001164644.1 [Q8IWL2-1]
NP_001158119.1, NM_001164647.1 [Q8IWL2-1]
NP_005402.3, NM_005411.4 [Q8IWL2-1]
XP_005270119.1, XM_005270062.4 [Q8IWL2-1]
XP_006718016.1, XM_006717953.2 [Q8IWL2-2]

Genome annotation databases

Ensembl eukaryotic genome annotation project

More...
Ensembli
ENST00000398636; ENSP00000381633; ENSG00000122852 [Q8IWL2-1]
ENST00000419470; ENSP00000397082; ENSG00000122852 [Q8IWL2-2]
ENST00000428376; ENSP00000411102; ENSG00000122852 [Q8IWL2-1]

Database of genes from NCBI RefSeq genomes

More...
GeneIDi
653509

KEGG: Kyoto Encyclopedia of Genes and Genomes

More...
KEGGi
hsa:653509

UCSC genome browser

More...
UCSCi
uc001kap.4 human [Q8IWL2-1]

Keywords - Coding sequence diversityi

Alternative splicing, Polymorphism

<p>This section provides links to proteins that are similar to the protein sequence(s) described in this entry at different levels of sequence identity thresholds (100%, 90% and 50%) based on their membership in UniProt Reference Clusters (<a href="http://www.uniprot.org/help/uniref">UniRef</a>).<p><a href='/help/similar_proteins_section' target='_top'>More...</a></p>Similar proteinsi

<p>This section is used to point to information related to entries and found in data collections other than UniProtKB.<p><a href='/help/cross_references_section' target='_top'>More...</a></p>Cross-referencesi

<p>This subsection of the <a href="http://www.uniprot.org/manual/cross_references_section">Cross-references</a> section provides links to various web resources that are relevant for a specific protein.<p><a href='/help/web_resource' target='_top'>More...</a></p>Web resourcesi

SeattleSNPs
Functional Glycomics Gateway - Glycan Binding

Pulmonary surfactant protein SP-A1

Sequence databases

Select the link destinations:
EMBLi
GenBanki
DDBJi
Links Updated
M30838 Genomic DNA Translation: AAA36510.1
M13686 mRNA Translation: AAA60211.1
HQ021433 mRNA Translation: ADO27676.1
HQ021434 mRNA Translation: ADO27677.1
HQ021435 mRNA Translation: ADO27678.1
HQ021436 mRNA Translation: ADO27679.1
HQ021437 mRNA Translation: ADO27680.1
HQ021438 mRNA Translation: ADO27681.1
HQ021439 mRNA Translation: ADO27682.1
HQ021440 mRNA Translation: ADO27683.1
HQ021441 mRNA Translation: ADO27684.1
HQ021442 mRNA Translation: ADO27685.1
AK290703 mRNA Translation: BAF83392.1
AY198391 Genomic DNA Translation: AAO13486.1
BX248123 Genomic DNA No translation available.
CH471083 Genomic DNA Translation: EAW54657.1
CH471083 Genomic DNA Translation: EAW54651.1
CH471083 Genomic DNA Translation: EAW54658.1
BC029913 mRNA Translation: AAH29913.1
BC111570 mRNA Translation: AAI11571.1
BC171875 mRNA Translation: AAI71875.1
CCDSiCCDS44444.2 [Q8IWL2-2]
CCDS44445.1 [Q8IWL2-1]
PIRiA24622 LNHUPS
A25720 LNHUP6
RefSeqiNP_001087239.2, NM_001093770.2 [Q8IWL2-2]
NP_001158116.1, NM_001164644.1 [Q8IWL2-1]
NP_001158119.1, NM_001164647.1 [Q8IWL2-1]
NP_005402.3, NM_005411.4 [Q8IWL2-1]
XP_005270119.1, XM_005270062.4 [Q8IWL2-1]
XP_006718016.1, XM_006717953.2 [Q8IWL2-2]

3D structure databases

SMRiQ8IWL2
ModBaseiSearch...

Protein-protein interaction databases

BioGridi575839, 4 interactors
IntActiQ8IWL2, 2 interactors
STRINGi9606.ENSP00000397082

Chemistry databases

DrugBankiDB03814 2-(N-Morpholino)-Ethanesulfonic Acid

PTM databases

iPTMnetiQ8IWL2
PhosphoSitePlusiQ8IWL2

Polymorphism and mutation databases

BioMutaiSFTPA1
DMDMi60416440

Proteomic databases

CPTACiCPTAC-1214
MassIVEiQ8IWL2
PaxDbiQ8IWL2
PeptideAtlasiQ8IWL2
PRIDEiQ8IWL2
ProteomicsDBi33955
70869 [Q8IWL2-1]

Genome annotation databases

EnsembliENST00000398636; ENSP00000381633; ENSG00000122852 [Q8IWL2-1]
ENST00000419470; ENSP00000397082; ENSG00000122852 [Q8IWL2-2]
ENST00000428376; ENSP00000411102; ENSG00000122852 [Q8IWL2-1]
GeneIDi653509
KEGGihsa:653509
UCSCiuc001kap.4 human [Q8IWL2-1]

Organism-specific databases

Comparative Toxicogenomics Database

More...
CTDi
653509
DisGeNETi653509

GeneCards: human genes, protein and diseases

More...
GeneCardsi
SFTPA1
HGNCiHGNC:10798 SFTPA1
HPAiCAB016793
HPA042638
HPA045752
HPA049368
MalaCardsiSFTPA1
MIMi178500 phenotype
178630 gene
267450 phenotype
neXtProtiNX_Q8IWL2
OpenTargetsiENSG00000122852
Orphaneti2032 Idiopathic pulmonary fibrosis
PharmGKBiPA35710

GenAtlas: human gene database

More...
GenAtlasi
Search...

Phylogenomic databases

eggNOGiKOG4297 Eukaryota
ENOG410XPJ1 LUCA
GeneTreeiENSGT00940000156653
InParanoidiQ8IWL2
KOiK10067
OMAiATQEACT
OrthoDBi1172460at2759
PhylomeDBiQ8IWL2
TreeFamiTF330481

Enzyme and pathway databases

ReactomeiR-HSA-166016 Toll Like Receptor 4 (TLR4) Cascade
R-HSA-168179 Toll Like Receptor TLR1:TLR2 Cascade
R-HSA-391160 Signal regulatory protein family interactions
R-HSA-5683826 Surfactant metabolism
R-HSA-5686938 Regulation of TLR by endogenous ligand
R-HSA-5688849 Defective CSF2RB causes pulmonary surfactant metabolism dysfunction 5 (SMDP5)
R-HSA-5688890 Defective CSF2RA causes pulmonary surfactant metabolism dysfunction 4 (SMDP4)

Miscellaneous databases

The Gene Wiki collection of pages on human genes and proteins

More...
GeneWikii
Pulmonary_surfactant-associated_protein_A1

Database of phenotypes from RNA interference screens in Drosophila and Homo sapiens

More...
GenomeRNAii
653509
PharosiQ8IWL2

Protein Ontology

More...
PROi
PR:Q8IWL2

The Stanford Online Universal Resource for Clones and ESTs

More...
SOURCEi
Search...

Gene expression databases

BgeeiENSG00000122852 Expressed in 68 organ(s), highest expression level in right lung
ExpressionAtlasiQ8IWL2 baseline and differential
GenevisibleiQ8IWL2 HS

Family and domain databases

CDDicd03591 CLECT_collectin_like, 1 hit
Gene3Di3.10.100.10, 1 hit
InterProiView protein in InterPro
IPR001304 C-type_lectin-like
IPR016186 C-type_lectin-like/link_sf
IPR018378 C-type_lectin_CS
IPR033990 Collectin_CTLD
IPR016187 CTDL_fold
PfamiView protein in Pfam
PF00059 Lectin_C, 1 hit
SMARTiView protein in SMART
SM00034 CLECT, 1 hit
SUPFAMiSSF56436 SSF56436, 1 hit
PROSITEiView protein in PROSITE
PS00615 C_TYPE_LECTIN_1, 1 hit
PS50041 C_TYPE_LECTIN_2, 1 hit

ProtoNet; Automatic hierarchical classification of proteins

More...
ProtoNeti
Search...

MobiDB: a database of protein disorder and mobility annotations

More...
MobiDBi
Search...

<p>This section provides general information on the entry.<p><a href='/help/entry_information_section' target='_top'>More...</a></p>Entry informationi

<p>This subsection of the ‘Entry information’ section provides a mnemonic identifier for a UniProtKB entry, but it is not a stable identifier. Each reviewed entry is assigned a unique entry name upon integration into UniProtKB/Swiss-Prot.<p><a href='/help/entry_name' target='_top'>More...</a></p>Entry nameiSFTA1_HUMAN
<p>This subsection of the ‘Entry information’ section provides one or more accession number(s). These are stable identifiers and should be used to cite UniProtKB entries. Upon integration into UniProtKB, each entry is assigned a unique accession number, which is called ‘Primary (citable) accession number’.<p><a href='/help/accession_numbers' target='_top'>More...</a></p>AccessioniPrimary (citable) accession number: Q8IWL2
Secondary accession number(s): A8K3T8
, B7ZW50, E3VLD8, E3VLD9, E3VLE0, E3VLE1, G5E9J3, P07714, Q14DV4, Q5RIR5, Q5RIR7, Q6PIT0, Q8TC19
<p>This subsection of the ‘Entry information’ section shows the date of integration of the entry into UniProtKB, the date of the last sequence update and the date of the last annotation modification (‘Last modified’). The version number for both the entry and the <a href="http://www.uniprot.org/help/canonical_and_isoforms">canonical sequence</a> are also displayed.<p><a href='/help/entry_history' target='_top'>More...</a></p>Entry historyiIntegrated into UniProtKB/Swiss-Prot: April 1, 1988
Last sequence update: March 1, 2005
Last modified: October 16, 2019
This is version 168 of the entry and version 2 of the sequence. See complete history.
<p>This subsection of the ‘Entry information’ section indicates whether the entry has been manually annotated and reviewed by UniProtKB curators or not, in other words, if the entry belongs to the Swiss-Prot section of UniProtKB (<strong>reviewed</strong>) or to the computer-annotated TrEMBL section (<strong>unreviewed</strong>).<p><a href='/help/entry_status' target='_top'>More...</a></p>Entry statusiReviewed (UniProtKB/Swiss-Prot)
Annotation programChordata Protein Annotation Program
DisclaimerAny medical or genetic information present in this entry is provided for research, educational and informational purposes only. It is not in any way intended to be used as a substitute for professional medical advice, diagnosis, treatment or care.

<p>This section contains any relevant information that doesn’t fit in any other defined sections<p><a href='/help/miscellaneous_section' target='_top'>More...</a></p>Miscellaneousi

Keywords - Technical termi

Complete proteome, Reference proteome

Documents

  1. SIMILARITY comments
    Index of protein domains and families
  2. Human polymorphisms and disease mutations
    Index of human polymorphisms and disease mutations
  3. MIM cross-references
    Online Mendelian Inheritance in Man (MIM) cross-references in UniProtKB/Swiss-Prot
  4. Human chromosome 10
    Human chromosome 10: entries, gene names and cross-references to MIM
  5. Human entries with polymorphisms or disease mutations
    List of human entries with polymorphisms or disease mutations
UniProt is an ELIXIR core data resource
Main funding by: National Institutes of Health

We'd like to inform you that we have updated our Privacy Notice to comply with Europe’s new General Data Protection Regulation (GDPR) that applies since 25 May 2018.

Do not show this banner again