{"id":2135,"date":"2026-07-10T13:00:00","date_gmt":"2026-07-10T13:00:00","guid":{"rendered":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/"},"modified":"2026-07-10T13:59:11","modified_gmt":"2026-07-10T13:59:11","slug":"accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit","status":"publish","type":"post","link":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/","title":{"rendered":"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit"},"content":{"rendered":"<p><br \/>\n<\/p>\n<div>\n<p class=\"wp-block-paragraph\">Biomolecular construction prediction and co-folding with fashions like OpenFold3 at the moment are mainstream, large-scale workloads powering drug discovery and protein design. More and more, they\u2019re pushed end-to-end by AI brokers. For an agent to run that pipeline effectively, each step must be quick and scalable: A number of Sequence Alignment (MSA) era, co-folding inference, serving, and multi-GPU scale-out. A bottleneck anyplace limits total throughput.\u00a0<\/p>\n<p class=\"wp-block-paragraph\">Velocity and memory-efficiency are important for key drug discovery workflows similar to digital screening and prediction of enormous molecular assemblies. In digital screening, hundreds of thousands to billions of compounds are screened in opposition to one or a couple of protein targets. Whereas co-folding fashions usually give the most effective predicted constructions, they are often costly to run, making them impractical for digital screening functions. That&#8217;s the place NVIDIA acceleration turns into key, making potential the deployment of OpenFold3 and associated strategies on the scale of enormous compound libraries.<\/p>\n<p class=\"wp-block-paragraph\">Velocity can be necessary for predicting giant molecular assemblies involving a number of proteins and hundreds of amino acid residues, as co-folding mannequin runtime scales cubically with the variety of residues. A fair greater problem, nonetheless, is reminiscence use, as single GPU reminiscence may be restricted, putting a tough ceiling on the scale of complexes that may be predicted in a single shot. Strategies to scale back reminiscence necessities, and that distribute prediction duties throughout a number of GPUs, would allow qualitatively new functions which might be merely unfeasible at present.<\/p>\n<p class=\"wp-block-paragraph\">NVIDIA has constructed instruments to speed up and enhance the effectivity of every step of the construction prediction and co-folding workflow. NVIDIA BioNeMo Agent Toolkit offers brokers seamless entry to the instruments they should speed up biology and chemistry workflows. On this put up, we break down the accelerations for every stage on NVIDIA B300 and H100 GPUs, then present how these phases may be executed by an agent (see Determine 1, beneath).<\/p>\n<div class=\"wp-block-image\">\n<figure data-wp-context=\"{&quot;imageId&quot;:&quot;6a50fa7d13702&quot;}\" data-wp-interactive=\"core\/image\" data-wp-key=\"6a50fa7d13702\" class=\"aligncenter size-full wp-lightbox-container\"><img decoding=\"async\" width=\"1382\" height=\"980\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on--pointerdown=\"actions.preloadImage\" data-wp-on--pointerenter=\"actions.preloadImageWithDelay\" data-wp-on--pointerleave=\"actions.cancelPreload\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5.webp\" alt=\"A flowchart showing the BioNeMo Agent Toolkit pipeline for protein structure prediction, from input sequences through GPU MSA Generation, Co-Folding with OpenFold 3, and Multi-GPU Scale-out to a predicted structure.\" class=\"wp-image-119698\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5.webp 1382w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-162x115.png 162w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-300x213.png 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-768x545.png 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-625x443.png 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-645x457.png 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-423x300.png 423w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-127x90.png 127w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-362x257.png 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-155x110.png 155w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-1024x726.png 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-762x540.png 762w\" sizes=\"(max-width: 1382px) 100vw, 1382px\"\/><img loading=\"lazy\" decoding=\"async\" width=\"1382\" height=\"980\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on--pointerdown=\"actions.preloadImage\" data-wp-on--pointerenter=\"actions.preloadImageWithDelay\" data-wp-on--pointerleave=\"actions.cancelPreload\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5.webp\" alt=\"A flowchart showing the BioNeMo Agent Toolkit pipeline for protein structure prediction, from input sequences through GPU MSA Generation, Co-Folding with OpenFold 3, and Multi-GPU Scale-out to a predicted structure.\" class=\"lazyload wp-image-119698\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5.webp 1382w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-162x115.png 162w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-300x213.png 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-768x545.png 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-625x443.png 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-645x457.png 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-423x300.png 423w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-127x90.png 127w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-362x257.png 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-155x110.png 155w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-1024x726.png 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image3-5-762x540.png 762w\" data-sizes=\"(max-width: 1382px) 100vw, 1382px\"\/><figcaption class=\"wp-element-caption\">Determine 1. The NVIDIA-accelerated structure-prediction workflow: GPU MSA, cuEquivariance, optimized inference, and Fold-CP, all orchestrated by the BioNeMo Agent Toolkit<\/figcaption><\/figure>\n<\/div>\n<h2 id=\"remove_the_msa_bottleneck_with_gpu_msa\" class=\"wp-block-heading\">Take away the MSA bottleneck with GPU MSA<\/h2>\n<p class=\"wp-block-paragraph\">For co-folding fashions, constructing the MSA has historically been a CPU-bound step that may dominate wall-clock time. MMseqs2-GPU strikes homology search onto NVIDIA GPUs, lowering this bottleneck whereas scaling with sequence size on each NVIDIA Hopper and NVIDIA Blackwell architectures. <\/p>\n<p class=\"wp-block-paragraph\">The most recent GPU accelerated model provides Hopper and Blackwell particular optimizations, together with environment friendly assist for larger-than-GPU-memory database search on NVIDIA Grace programs and extra speedups from improved Blackwell DPX directions obtainable from CUDA 13.2. These GPU contributions have been upstreamed again into the principle MMseqs2 repository so the entire group can profit from the accelerations.<\/p>\n<p class=\"wp-block-paragraph\">The MSA Search NIM makes use of MMseqs2-GPU, whose Nature Strategies paper reviews as much as 177\u00d7 quicker alignment than CPU JackHMMER on a single L40S. In our benchmarking, the stage scales easily previous 10k tokens on H100 and B300 GPUs (see Determine 2, beneath). The MSA Search NIM may be referred to as instantly, self-hosted or wrapped as a software in an agentic workflow.\u00a0<\/p>\n<figure class=\"wp-block-table\">Notice: We present learn how to name the NIM API endpoint hosted on construct.nvidia.com, however for large-scale workloads, the self-hosted NIM is the popular beneficial endpoint. construct.nvidia.com endpoints aren&#8217;t constructed to deal with giant quantity requests and can time-out.<\/figure>\n<div class=\"wp-block-syntaxhighlighter-code \">\n# Add the MSA Search NIM talent to your agent (browse all: add &#8211;list)<br \/>\nnpx abilities add NVIDIA-BioNeMo\/bionemo-agent-toolkit &#8211;skill msa-search-nim &#8211;agent claude-code<br \/>\n# Use hosted API on construct.nvidia.com (nothing to obtain)<br \/>\n# It&#8217;s also possible to use the talent to obtain the NIM container and supply self-hosted API endpoint. We do not cowl that on this tutorial.<br \/>\nexport NVIDIA_API_KEY=<\/p>\n<p># Simply immediate the agent:<br \/>\n# It&#8217;s a must to obtain a pattern goal.fasta file. You may immediate the agent to obtain it for you or level to an already present file.<br \/>\n&#8220;Construct an MSA for the sequence in goal.fasta with the MSA Search NIM.&#8221;\n<\/p><\/div>\n<div class=\"wp-block-image\">\n<figure data-wp-context=\"{&quot;imageId&quot;:&quot;6a50fa7d148b8&quot;}\" data-wp-interactive=\"core\/image\" data-wp-key=\"6a50fa7d148b8\" class=\"aligncenter size-full wp-lightbox-container\"><img decoding=\"async\" width=\"1999\" height=\"1238\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on--pointerdown=\"actions.preloadImage\" data-wp-on--pointerenter=\"actions.preloadImageWithDelay\" data-wp-on--pointerleave=\"actions.cancelPreload\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2.webp\" alt=\"A line chart comparing OpenFold3 MSA stage latency by sequence length between H100 (463s) and B300 (429s) GPUs at ~10k tokens, showing near-linear scaling with B300 slightly faster.\" class=\"wp-image-119700\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-179x111.png 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-300x186.png 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-768x476.png 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-625x387.png 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-1536x951.png 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-645x399.png 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-484x300.png 484w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-145x90.png 145w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-362x224.png 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-178x110.png 178w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-1024x634.png 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-872x540.png 872w\" sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><img loading=\"lazy\" decoding=\"async\" width=\"1999\" height=\"1238\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on--pointerdown=\"actions.preloadImage\" data-wp-on--pointerenter=\"actions.preloadImageWithDelay\" data-wp-on--pointerleave=\"actions.cancelPreload\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2.webp\" alt=\"A line chart comparing OpenFold3 MSA stage latency by sequence length between H100 (463s) and B300 (429s) GPUs at ~10k tokens, showing near-linear scaling with B300 slightly faster.\" class=\"lazyload wp-image-119700\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-179x111.png 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-300x186.png 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-768x476.png 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-625x387.png 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-1536x951.png 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-645x399.png 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-484x300.png 484w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-145x90.png 145w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-362x224.png 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-178x110.png 178w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-1024x634.png 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image6-2-872x540.png 872w\" data-sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><figcaption class=\"wp-element-caption\">Determine 2. MSA-stage latency scales close to linearly to 10k+ tokens on H100 and B300 (429 s vs 463 s at ~10k tokens)<\/figcaption><\/figure>\n<\/div>\n<h2 id=\"fold_at_sota_speed_with_cuequivariance_and_the_openfold3_nim\" class=\"wp-block-heading\">Fold at SOTA velocity with cuEquivariance and the OpenFold3 NIM<\/h2>\n<p class=\"wp-block-paragraph\">cuEquivariance is a CUDA-X library of geometric studying primitives for atomistic modeling and it gives accelerated variations of the Triangle Consideration, Triangle Multiplication and Consideration Pair Bias kernels that dominate co-folding. On B300 it cuts latency as much as ~3\u00d7 (see Desk 1, beneath).\u00a0<\/p>\n<figure class=\"wp-block-table\">Sequence lengthPyTorch (OSS)cuEquivarianceSpeedup1,024 (H100)79.4 s41.6 s1.9\u00d71,024 (B300)49.1 s27.3 s1.8\u00d71,536 (B300)149.0 s56.2 s2.7\u00d72,048 (B300)300.1 s97.6 s3.1\u00d7<figcaption class=\"wp-element-caption\">Desk 1. Comparability of OpenFold3 forward-pass latency with and with out cuEquivariance<\/figcaption><\/figure>\n<p class=\"wp-block-paragraph\">cuEquivariance kernels are built-in instantly into the OSS fashions like OpenFold3 (offered as an non-compulsory dependency), OpenFold2, RosettaFold3, Protenix and Boltz. <\/p>\n<p class=\"wp-block-paragraph\">As a result of the accelerations are upstreamed into these OSS fashions, a researcher will get the speedups mechanically just by working the mannequin they already use on an NVIDIA GPU. CuEquivariance kernels additionally extends most sequence size to ~5.9k tokens whereas PyTorch runs out of reminiscence past ~1.5k\u20132.5k tokens.<\/p>\n<div class=\"wp-block-image\">\n<figure data-wp-context=\"{&quot;imageId&quot;:&quot;6a50fa7d1597b&quot;}\" data-wp-interactive=\"core\/image\" data-wp-key=\"6a50fa7d1597b\" class=\"aligncenter size-full wp-lightbox-container\"><img decoding=\"async\" width=\"2000\" height=\"2000\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on--pointerdown=\"actions.preloadImage\" data-wp-on--pointerenter=\"actions.preloadImageWithDelay\" data-wp-on--pointerleave=\"actions.cancelPreload\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4.webp\" alt=\"A line chart comparing OpenFold3 inference latency between cuEquivariance and PyTorch OSS on H100 and B300 GPUs, showing cuEquivariance extends maximum sequence length to ~5.9k tokens while PyTorch runs out of memory beyond ~1.5k\u20132.5k tokens.\" class=\"wp-image-119702\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4.webp 2000w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-115x115.png 115w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-300x300.png 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-768x768.png 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-625x625.png 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-1536x1536.png 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-645x645.png 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-90x90.png 90w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-32x32.png 32w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-50x50.png 50w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-64x64.png 64w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-96x96.png 96w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-128x128.png 128w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-150x150.png 150w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-362x362.png 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-110x110.png 110w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-1024x1024.png 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-540x540.png 540w\" sizes=\"(max-width: 2000px) 100vw, 2000px\"\/><img loading=\"lazy\" decoding=\"async\" width=\"2000\" height=\"2000\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on--pointerdown=\"actions.preloadImage\" data-wp-on--pointerenter=\"actions.preloadImageWithDelay\" data-wp-on--pointerleave=\"actions.cancelPreload\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4.webp\" alt=\"A line chart comparing OpenFold3 inference latency between cuEquivariance and PyTorch OSS on H100 and B300 GPUs, showing cuEquivariance extends maximum sequence length to ~5.9k tokens while PyTorch runs out of memory beyond ~1.5k\u20132.5k tokens.\" class=\"lazyload wp-image-119702\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4.webp 2000w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-115x115.png 115w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-300x300.png 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-768x768.png 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-625x625.png 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-1536x1536.png 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-645x645.png 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-90x90.png 90w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-32x32.png 32w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-50x50.png 50w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-64x64.png 64w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-96x96.png 96w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-128x128.png 128w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-150x150.png 150w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-362x362.png 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-110x110.png 110w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-1024x1024.png 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image2-4-540x540.png 540w\" data-sizes=\"(max-width: 2000px) 100vw, 2000px\"\/><figcaption class=\"wp-element-caption\">Determine 3. cuEquivariance lowers latency and extends most sequence size over PyTorch on H100 and B300<\/figcaption><\/figure>\n<\/div>\n<p class=\"wp-block-paragraph\">On high of cuEquivariance, the OpenFold3 NIM applies additional inference optimizations that compound the achieve (see Determine 4, beneath), attaining sequence lengths of as much as ~6,400 on a single B300. These extra accelerations are delivered by the NIM; for SOTA out of the field, builders can name the NIM endpoint instantly or compose it into an agentic workflow.<\/p>\n<div class=\"wp-block-image\">\n<figure data-wp-context=\"{&quot;imageId&quot;:&quot;6a50fa7d162b5&quot;}\" data-wp-interactive=\"core\/image\" data-wp-key=\"6a50fa7d162b5\" class=\"aligncenter size-full wp-lightbox-container\"><img decoding=\"async\" width=\"1999\" height=\"668\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on--pointerdown=\"actions.preloadImage\" data-wp-on--pointerenter=\"actions.preloadImageWithDelay\" data-wp-on--pointerleave=\"actions.cancelPreload\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5.webp\" alt=\"Side-by-side line charts showing OpenFold3 forward-pass inference latency, with the fully optimized NIM 3.25x faster than OSS on H100 and 1.44x faster on B300, supporting sequences up to 6,400 tokens.\" class=\"wp-image-119704\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-179x60.png 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-300x100.png 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-768x257.png 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-625x209.png 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-1536x513.png 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-645x216.png 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-500x167.png 500w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-160x53.png 160w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-362x121.png 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-329x110.png 329w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-1024x342.png 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-960x321.png 960w\" sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><img loading=\"lazy\" decoding=\"async\" width=\"1999\" height=\"668\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on--pointerdown=\"actions.preloadImage\" data-wp-on--pointerenter=\"actions.preloadImageWithDelay\" data-wp-on--pointerleave=\"actions.cancelPreload\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5.webp\" alt=\"Side-by-side line charts showing OpenFold3 forward-pass inference latency, with the fully optimized NIM 3.25x faster than OSS on H100 and 1.44x faster on B300, supporting sequences up to 6,400 tokens.\" class=\"lazyload wp-image-119704\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-179x60.png 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-300x100.png 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-768x257.png 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-625x209.png 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-1536x513.png 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-645x216.png 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-500x167.png 500w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-160x53.png 160w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-362x121.png 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-329x110.png 329w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-1024x342.png 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/image4-5-960x321.png 960w\" data-sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><figcaption class=\"wp-element-caption\">Determine 4. The totally optimized OpenFold3 NIM on H100 and B300, attaining over 3x and 4x quicker, respectively, folding as much as 6,400 tokens on a single B300\u00a0<\/figcaption><\/figure>\n<\/div>\n<div class=\"wp-block-syntaxhighlighter-code \">\n# Add the OpenFold3 NIM talent to your agent<br \/>\nnpx abilities add NVIDIA-BioNeMo\/bionemo-agent-toolkit &#8211;skill openfold3-nim &#8211;agent claude-code<\/p>\n<p># Hosted API on construct.nvidia.com<br \/>\n# It&#8217;s also possible to use the talent to obtain the NIM container and supply self-hosted API endpoint. We do not cowl that on this tutorial.<br \/>\nexport NVIDIA_API_KEY=<\/p>\n<p># Immediate the agent:<br \/>\n# It&#8217;s a must to obtain a pattern goal.fasta file. You may immediate the agent to obtain it for you or level to an already present file.<br \/>\n&#8220;Fold goal.fasta with OpenFold3 utilizing the MSA from the earlier step; return the ranked constructions with confidence scores.&#8221;\n<\/p><\/div>\n<h2 id=\"scale_beyond_one_gpu_with_fold-cp\" class=\"wp-block-heading\">Scale past one GPU with Fold-CP<\/h2>\n<p class=\"wp-block-paragraph\">Single-GPU reminiscence has traditionally capped co-folding fashions at a couple of thousand residues. On NVIDIA B300 (Blackwell Extremely), the bigger HBM and Blackwell-generation effectivity mixed with the cuEquivariance and superior inference optimizations above push that ceiling considerably larger with no mannequin adjustments.\u00a0<\/p>\n<p class=\"wp-block-paragraph\">In lots of instances, being single-device certain is probably not ample. Fold-CP introduces a brand new parallelization method such that per-device reminiscence requirement scales as O(N\u00b2\/P) the place N is token depend and P is the variety of GPUs, reaching 32,000 tokens on 64 B300 with the Boltz-2 mannequin\u2014a couple of 12\u00d7 leap over the single-GPU restrict.\u00a0<\/p>\n<p class=\"wp-block-paragraph\">To strive Fold-CP, merely level your agent to the Boltz-CP codebase and ask it to run multi-GPU inference.\u00a0<\/p>\n<div class=\"wp-block-syntaxhighlighter-code \">\n# Context-parallel inference throughout 4 GPUs with Fold-CP (boltz-cp)<br \/>\ngit clone https:\/\/github.com\/NVIDIA-Digital-Bio\/boltz-cp &amp;&amp; cd boltz-cp<br \/>\n# Set up dependencies after which run the command beneath<br \/>\ntorchrun &#8211;nnodes 1 &#8211;nproc_per_node 4<br \/>\n  src\/boltz\/distributed\/important.py predict \/path\/to\/preprocessed_data<br \/>\n  &#8211;out_dir .\/predictions<br \/>\n  &#8211;size_dp 1 &#8211;size_cp 4<br \/>\n  &#8211;recycling_steps 3 &#8211;sampling_steps 200 &#8211;diffusion_samples 5\n<\/div>\n<h2 id=\"accelerating_the_end-to-end_co-folding_pipeline\" class=\"wp-block-heading\">Accelerating the end-to-end co-folding pipeline<\/h2>\n<p class=\"wp-block-paragraph\">Construction prediction efficiency is now an end-to-end programs downside. For OpenFold3, the sensible workflow spans MSA era, co-folding inference, deployment, and the reminiscence limits that decide how giant a organic meeting may be modeled. <\/p>\n<p class=\"wp-block-paragraph\">NVIDIA accelerates every layer of that workflow: MSA Search NIM speeds homology search by 177x, cuEquivariance and OpenFold3 NIM decrease inference latency by as much as 4x on Blackwell GPUs, and Fold-CP reveals how context parallelism can lengthen co-folding past a single GPU to 32,000-token complexes on 64 NVIDIA B300 GPUs.\u00a0<\/p>\n<p class=\"wp-block-paragraph\">Collectively, these instruments make construction prediction quicker, extra scalable, and simpler to compose into agentic discovery workflows, serving to researchers transfer from mannequin predictions to bigger, extra helpful organic programs.<\/p>\n<h2 id=\"what_these_accelerations_unlock\" class=\"wp-block-heading\">What these accelerations unlock<\/h2>\n<p class=\"wp-block-paragraph\">These enhancements open up lessons of structural biology issues that have been beforehand out of attain. In digital screening, quicker co-folding inference implies that structure-based strategies, which have traditionally been reserved for the ultimate phases of a drug discovery marketing campaign, can now be utilized at far earlier phases and at a lot bigger library scales, enhancing the standard and variety of hits that advance by the pipeline. <\/p>\n<p class=\"wp-block-paragraph\">For giant biomolecular assemblies, the mix of prolonged single-GPU capability on B300 and the context-parallel Fold-CP framework shifts what&#8217;s modelable: complexes on the scale of the ribosome, the spliceosome, or giant signaling assemblies have been structurally intractable for co-folding fashions, and these accelerations start to alter that. To make this concrete, folding a posh of ~10,000 residues, roughly the dimensions of the bacterial ribosome, would have been prohibitively costly or just out of attain on a single GPU; on B300 with Fold-CP, such predictions change into tractable throughout a multi-GPU node. <\/p>\n<p class=\"wp-block-paragraph\">These are qualitative shifts within the questions that structural biology can ask computationally, not merely enhancements in throughput. Making these instruments accessible by open-source integrations and agentic APIs will speed up the tempo at which computational predictions translate into organic perception and, finally, into new medicines.<\/p>\n<h2 id=\"getting_started\u00a0\" class=\"wp-block-heading\">Getting began\u00a0<\/h2>\n<p class=\"wp-block-paragraph\">Strive the accelerated OpenFold3 workflow with NVIDIA BioNeMo Agent Toolkit, beginning with these instruments:<\/p>\n<h4 class=\"wp-block-heading\">Acknowledgments\u00a0<\/h4>\n<p class=\"wp-block-paragraph\">We\u2019d prefer to thank our broader NVIDIA staff for growing the benchmarks and instruments:\u00a0Franco Pellegrini, Lalit Vaidya, Duc Tran, Tien Pham, Maximilian Stadler, Alejandro Chacon, Quan Vu, Simon Chu, Brian Roland, Dejun Lin, Joseph Chang, Hoa La, Jonathan Mitchell, Vishanth Iyer, Timur Rvachov, Christian Dallago, Christian Hundt. We\u2019d additionally prefer to thank the broader OpenFold and OMSF groups for our collaboration and their contributions.<\/p>\n<\/div>\n<p><br \/>\n<br \/><a href=\"https:\/\/developer.nvidia.com\/blog\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/\">Source link <\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>Biomolecular construction prediction and co-folding with fashions like OpenFold3 at the moment are mainstream, large-scale workloads powering drug discovery and protein design. More and more, they\u2019re pushed end-to-end by AI brokers. For an agent to run that pipeline effectively, each step must be quick and scalable: A number of Sequence Alignment (MSA) era, co-folding inference, [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":2137,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"fifu_image_url":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp","fifu_image_alt":"","jnews-multi-image_gallery":[],"jnews_single_post":[],"jnews_primary_category":[],"jnews_override_bookmark_settings":[],"jnews_social_meta":[],"jnews_override_counter":[],"footnotes":""},"categories":[3],"tags":[1913,457,1423,2653,1972,81,750,2654],"class_list":["post-2135","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-ai-platforms-apps","tag-accelerating","tag-agent","tag-bionemo","tag-cofolding","tag-endtoend","tag-nvidia","tag-performance","tag-toolkit"],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v27.7 - https:\/\/yoast.com\/product\/yoast-seo-wordpress\/ -->\n<title>Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit - Future News 24<\/title>\n<meta name=\"description\" content=\"Biomolecular structure prediction and co&#x2d;folding with models like OpenFold3 are now mainstream, large&#x2d;scale workloads powering drug discovery and protein design.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit - Future News 24\" \/>\n<meta property=\"og:description\" content=\"Biomolecular structure prediction and co&#x2d;folding with models like OpenFold3 are now mainstream, large&#x2d;scale workloads powering drug discovery and protein design.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/\" \/>\n<meta property=\"og:site_name\" content=\"Future News 24\" \/>\n<meta property=\"article:published_time\" content=\"2026-07-10T13:00:00+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2026-07-10T13:59:11+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp\" \/>\n<meta name=\"author\" content=\"Future News 24\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:image\" content=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Future News 24\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"8 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/#article\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/\"},\"author\":{\"name\":\"Future News 24\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#\\\/schema\\\/person\\\/cecad1bde21cfc357cf70128144d6c83\"},\"headline\":\"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit\",\"datePublished\":\"2026-07-10T13:00:00+00:00\",\"dateModified\":\"2026-07-10T13:59:11+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/\"},\"wordCount\":1637,\"commentCount\":0,\"publisher\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#organization\"},\"image\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/developer-blogs.nvidia.com\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp\",\"keywords\":[\"Accelerating\",\"Agent\",\"BioNeMo\",\"CoFolding\",\"EndtoEnd\",\"NVIDIA\",\"performance\",\"Toolkit\"],\"articleSection\":[\"AI Platforms &amp; Apps\"],\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"CommentAction\",\"name\":\"Comment\",\"target\":[\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/#respond\"]}]},{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/\",\"url\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/\",\"name\":\"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit - Future News 24\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/#primaryimage\"},\"image\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/developer-blogs.nvidia.com\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp\",\"datePublished\":\"2026-07-10T13:00:00+00:00\",\"dateModified\":\"2026-07-10T13:59:11+00:00\",\"description\":\"Biomolecular structure prediction and co&#x2d;folding with models like OpenFold3 are now mainstream, large&#x2d;scale workloads powering drug discovery and protein design.\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/#primaryimage\",\"url\":\"https:\\\/\\\/developer-blogs.nvidia.com\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp\",\"contentUrl\":\"https:\\\/\\\/developer-blogs.nvidia.com\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/07\\\/10\\\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/futurenews24.com\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#website\",\"url\":\"https:\\\/\\\/futurenews24.com\\\/\",\"name\":\"Future News 24\",\"description\":\"The Smart Hub for AI and Next-Gen Innovation\",\"publisher\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/futurenews24.com\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Organization\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#organization\",\"name\":\"Future News 24\",\"url\":\"https:\\\/\\\/futurenews24.com\\\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#\\\/schema\\\/logo\\\/image\\\/\",\"url\":\"https:\\\/\\\/futurenews24.com\\\/wp-content\\\/uploads\\\/2026\\\/06\\\/fn24-favicon.png\",\"contentUrl\":\"https:\\\/\\\/futurenews24.com\\\/wp-content\\\/uploads\\\/2026\\\/06\\\/fn24-favicon.png\",\"width\":250,\"height\":250,\"caption\":\"Future News 24\"},\"image\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#\\\/schema\\\/logo\\\/image\\\/\"}},{\"@type\":\"Person\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#\\\/schema\\\/person\\\/cecad1bde21cfc357cf70128144d6c83\",\"name\":\"Future News 24\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g\",\"url\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g\",\"contentUrl\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g\",\"caption\":\"Future News 24\"},\"sameAs\":[\"https:\\\/\\\/futurenews24.com\"],\"url\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/author\\\/mridulpahuja20\\\/\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit - Future News 24","description":"Biomolecular structure prediction and co&#x2d;folding with models like OpenFold3 are now mainstream, large&#x2d;scale workloads powering drug discovery and protein design.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/","og_locale":"en_US","og_type":"article","og_title":"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit - Future News 24","og_description":"Biomolecular structure prediction and co&#x2d;folding with models like OpenFold3 are now mainstream, large&#x2d;scale workloads powering drug discovery and protein design.","og_url":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/","og_site_name":"Future News 24","article_published_time":"2026-07-10T13:00:00+00:00","article_modified_time":"2026-07-10T13:59:11+00:00","og_image":[{"url":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp","type":"","width":"","height":""}],"author":"Future News 24","twitter_card":"summary_large_image","twitter_image":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp","twitter_misc":{"Written by":"Future News 24","Est. reading time":"8 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/#article","isPartOf":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/"},"author":{"name":"Future News 24","@id":"https:\/\/futurenews24.com\/#\/schema\/person\/cecad1bde21cfc357cf70128144d6c83"},"headline":"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit","datePublished":"2026-07-10T13:00:00+00:00","dateModified":"2026-07-10T13:59:11+00:00","mainEntityOfPage":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/"},"wordCount":1637,"commentCount":0,"publisher":{"@id":"https:\/\/futurenews24.com\/#organization"},"image":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/#primaryimage"},"thumbnailUrl":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp","keywords":["Accelerating","Agent","BioNeMo","CoFolding","EndtoEnd","NVIDIA","performance","Toolkit"],"articleSection":["AI Platforms &amp; Apps"],"inLanguage":"en-US","potentialAction":[{"@type":"CommentAction","name":"Comment","target":["https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/#respond"]}]},{"@type":"WebPage","@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/","url":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/","name":"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit - Future News 24","isPartOf":{"@id":"https:\/\/futurenews24.com\/#website"},"primaryImageOfPage":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/#primaryimage"},"image":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/#primaryimage"},"thumbnailUrl":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp","datePublished":"2026-07-10T13:00:00+00:00","dateModified":"2026-07-10T13:59:11+00:00","description":"Biomolecular structure prediction and co&#x2d;folding with models like OpenFold3 are now mainstream, large&#x2d;scale workloads powering drug discovery and protein design.","breadcrumb":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/#primaryimage","url":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp","contentUrl":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/07\/hc-social-openfold-3-blog-1920x1080-5405363-1.webp"},{"@type":"BreadcrumbList","@id":"https:\/\/futurenews24.com\/index.php\/2026\/07\/10\/accelerating-end-to-end-co-folding-performance-with-nvidia-bionemo-agent-toolkit\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/futurenews24.com\/"},{"@type":"ListItem","position":2,"name":"Accelerating Finish-to-Finish Co-Folding Efficiency with NVIDIA BioNeMo Agent Toolkit"}]},{"@type":"WebSite","@id":"https:\/\/futurenews24.com\/#website","url":"https:\/\/futurenews24.com\/","name":"Future News 24","description":"The Smart Hub for AI and Next-Gen Innovation","publisher":{"@id":"https:\/\/futurenews24.com\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/futurenews24.com\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/futurenews24.com\/#organization","name":"Future News 24","url":"https:\/\/futurenews24.com\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/futurenews24.com\/#\/schema\/logo\/image\/","url":"https:\/\/futurenews24.com\/wp-content\/uploads\/2026\/06\/fn24-favicon.png","contentUrl":"https:\/\/futurenews24.com\/wp-content\/uploads\/2026\/06\/fn24-favicon.png","width":250,"height":250,"caption":"Future News 24"},"image":{"@id":"https:\/\/futurenews24.com\/#\/schema\/logo\/image\/"}},{"@type":"Person","@id":"https:\/\/futurenews24.com\/#\/schema\/person\/cecad1bde21cfc357cf70128144d6c83","name":"Future News 24","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/secure.gravatar.com\/avatar\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g","url":"https:\/\/secure.gravatar.com\/avatar\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g","caption":"Future News 24"},"sameAs":["https:\/\/futurenews24.com"],"url":"https:\/\/futurenews24.com\/index.php\/author\/mridulpahuja20\/"}]}},"_links":{"self":[{"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/posts\/2135","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/comments?post=2135"}],"version-history":[{"count":1,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/posts\/2135\/revisions"}],"predecessor-version":[{"id":2136,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/posts\/2135\/revisions\/2136"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/media\/2137"}],"wp:attachment":[{"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/media?parent=2135"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/categories?post=2135"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/tags?post=2135"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}