human reference protein sequences (Biotechnology Information)
Structured Review

Human Reference Protein Sequences, supplied by Biotechnology Information, used in various techniques. Bioz Stars score: 86/100, based on 1 PubMed citations. ZERO BIAS - scores, article reviews, protocol conditions and more
https://www.bioz.com/product/human+reference+sequences/pmc12412403-112-12-20?v=Biotechnology+Information
Average 86 stars, based on 1 article reviews
Images
1) Product Images from "DeepHVI: A multimodal deep learning framework for predicting human-virus protein-protein interactions using protein language models"
Article Title: DeepHVI: A multimodal deep learning framework for predicting human-virus protein-protein interactions using protein language models
Journal: Biosafety and Health
doi: 10.1016/j.bsheal.2025.07.005
Figure Legend Snippet: Schematic overview of training and inference workflows. A) Training phase: Positive and negative sample pairs are processed through the embedding module to generate four modality-specific embedding representations, human/virus sequence embedding is encoded by ESM-2 to represent amino acid sequence features, human/virus chemical embedding is extracted from AAindex profiles to represent protein biochemical properties. The cross-fusion module computes loss via contrastive learning and executes forward propagation. Model parameters are iteratively updated using the Adam optimizer, with final weights preserved for inference. B) Binary task inference. Pre-trained weights are loaded to compute task-specific losses, enabling downstream classification prediction. C) Conditional generative inference. A sequence decoder module translates fused modality embeddings into human protein sequences, with outputs ranked to return the top five highest-confidence matches.
Techniques Used: Virus, Sequencing
Figure Legend Snippet: Cosine similarity analysis of generated human protein sequences and viral protein sequences. A) Cosine similarity between generated and ground-truth human proteins, illustrating the distribution of cosine similarity values between human protein sequences. B) Cosine similarity between generated and ground-truth viral proteins, depicting the distribution for viral protein sequences, which exhibits slightly lower and more variable similarity scores. In both cases, the distributions are sharply peaked around 0.8, indicating a generally strong semantic alignment across samples. Human and viral protein sequences from the test set were analyzed using DeepHVI to generate a density distribution of similarity scores between reconstructed sequences and human interactors. Abbreviations: Std, standard deviation; Min, minimum; Max, maximum.
Techniques Used: Generated, Standard Deviation