Conserved heavy/light contacts and germline preferences revealed by a large-scale analysis of natively paired human antibody sequences and structural data.
Pawel Dudzic, Dawid Chomicz, Weronika Bielska, Igor Jaszczyszyn, Michał Zieliński, Bartosz Janusz, Sonia Wróbel, Marguerite-Marie Le Pannérer, Andrew Philips, Prabakaran Ponraj, Sandeep Kumar, Konrad Krawczyk
{"title":"Conserved heavy/light contacts and germline preferences revealed by a large-scale analysis of natively paired human antibody sequences and structural data.","authors":"Pawel Dudzic, Dawid Chomicz, Weronika Bielska, Igor Jaszczyszyn, Michał Zieliński, Bartosz Janusz, Sonia Wróbel, Marguerite-Marie Le Pannérer, Andrew Philips, Prabakaran Ponraj, Sandeep Kumar, Konrad Krawczyk","doi":"10.1038/s42003-025-08388-y","DOIUrl":null,"url":null,"abstract":"<p><p>Understanding the pairing preferences and structural interactions between antibody heavy and light chains can enhance our ability to design more effective and specific therapeutic antibodies. Insights from natural antibody repertoires and conserved contact sites help reduce autoreactivity and improve drug safety and efficacy. Current databases represent only a limited portion of the estimated diversity of unique paired antibody molecules. To address this, we introduce PairedAbNGS, a novel database with paired heavy/light antibody chains. To our knowledge, this is the largest resource for paired natural antibody sequences with 58 bioprojects and over 14 million assembled productive sequences. Using this dataset, we investigated heavy and light chain variable (V) gene pairing preferences and found significant biases beyond gene usage frequencies, possibly due to receptor editing favoring less autoreactive combinations. Analyzing the available antibody structures from the Protein Data Bank, we studied conserved contact residues between heavy and light chains, particularly interactions between the CDR3 region of one chain and the FWR2 region of the opposite chain. Examination of amino acid pairs at key contact sites revealed significant deviations of amino acids distributions compared to random pairings, in the heavy chain's CDR3 region contacting the opposite chain, indicating specific interactions might be crucial for proper chain pairing. This observation is further reinforced by preferential IGHV-IGLJ and IGLV-IGHJ pairing preferences. We hope that both our resources and the findings would contribute to improving the engineering of biological drugs. We make the database accessible at https://naturalantibody.com/paired-ab-ngs as a valuable tool for biological and machine-learning applications.</p>","PeriodicalId":10552,"journal":{"name":"Communications Biology","volume":"8 1","pages":"1110"},"PeriodicalIF":5.1000,"publicationDate":"2025-07-26","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"https://www.ncbi.nlm.nih.gov/pmc/articles/PMC12297541/pdf/","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Communications Biology","FirstCategoryId":"99","ListUrlMain":"https://doi.org/10.1038/s42003-025-08388-y","RegionNum":1,"RegionCategory":"生物学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"","PubModel":"","JCR":"Q1","JCRName":"BIOLOGY","Score":null,"Total":0}
引用次数: 0
Abstract
Understanding the pairing preferences and structural interactions between antibody heavy and light chains can enhance our ability to design more effective and specific therapeutic antibodies. Insights from natural antibody repertoires and conserved contact sites help reduce autoreactivity and improve drug safety and efficacy. Current databases represent only a limited portion of the estimated diversity of unique paired antibody molecules. To address this, we introduce PairedAbNGS, a novel database with paired heavy/light antibody chains. To our knowledge, this is the largest resource for paired natural antibody sequences with 58 bioprojects and over 14 million assembled productive sequences. Using this dataset, we investigated heavy and light chain variable (V) gene pairing preferences and found significant biases beyond gene usage frequencies, possibly due to receptor editing favoring less autoreactive combinations. Analyzing the available antibody structures from the Protein Data Bank, we studied conserved contact residues between heavy and light chains, particularly interactions between the CDR3 region of one chain and the FWR2 region of the opposite chain. Examination of amino acid pairs at key contact sites revealed significant deviations of amino acids distributions compared to random pairings, in the heavy chain's CDR3 region contacting the opposite chain, indicating specific interactions might be crucial for proper chain pairing. This observation is further reinforced by preferential IGHV-IGLJ and IGLV-IGHJ pairing preferences. We hope that both our resources and the findings would contribute to improving the engineering of biological drugs. We make the database accessible at https://naturalantibody.com/paired-ab-ngs as a valuable tool for biological and machine-learning applications.
期刊介绍:
Communications Biology is an open access journal from Nature Research publishing high-quality research, reviews and commentary in all areas of the biological sciences. Research papers published by the journal represent significant advances bringing new biological insight to a specialized area of research.