{"id":954,"date":"2026-06-12T21:12:00","date_gmt":"2026-06-12T21:12:00","guid":{"rendered":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/"},"modified":"2026-06-13T17:59:29","modified_gmt":"2026-06-13T17:59:29","slug":"nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark","status":"publish","type":"post","link":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/","title":{"rendered":"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark"},"content":{"rendered":"<p><br \/>\n<\/p>\n<div>\n<p>AI brokers have essentially modified the complexity of inference workloads. Till now, the business has struggled to outline a regular for measuring how inference methods carry out underneath these circumstances. Synthetic Evaluation AgentPerf (AA-AgentPerf) gives the business\u2019s first multi-vendor open benchmarks profiling trajectories which might be consultant of real-world AI agent coding duties.\u00a0<\/p>\n<p>This put up explains how AA-AgentPerf units a brand new normal for measuring agentic workload efficiency, and the way NVIDIA excessive co-design helps ship as much as 20x higher agentic coding efficiency than earlier generations.<\/p>\n<h2 id=\"what_is_aa-agentperf\" class=\"wp-block-heading\">What&#8217;s AA-AgentPerf?<\/h2>\n<p>AA-AgentPerf is a {hardware} benchmark created by Synthetic Evaluation that measures the variety of concurrent AI brokers an inference system can assist whereas assembly predefined, model-specific efficiency service degree goal (SLO) tiers. An SLO is outlined as a selected threshold of output token velocity and time-to-first-token (TTFT). The benchmark outcomes are normalized per accelerator and per megawatt to allow comparability throughout {hardware} configurations.<\/p>\n<div class=\"wp-block-image\">\n<figure data-wp-context=\"{&quot;imageId&quot;:&quot;6a2d9a5c2588b&quot;}\" data-wp-interactive=\"core\/image\" data-wp-key=\"6a2d9a5c2588b\" class=\"aligncenter size-full wp-lightbox-container\"><img decoding=\"async\" width=\"1999\" height=\"1339\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures.webp\" alt=\"A diagram showing multiple AI agents (labeled Agent 1 through Agent N) feeding requests simultaneously into a central AI workflow\u2014consisting of LLM calls and tool use\u2014with three efficiency metrics measured at the bottom: per kW, per $\/hr, and per accelerator. &#10;\" class=\"wp-image-118656\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-172x115.jpg 172w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-300x201.jpg 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-768x514.jpg 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-625x419.jpg 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-1536x1029.jpg 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-645x432.jpg 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-448x300.jpg 448w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-134x90.jpg 134w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-362x242.jpg 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-164x110.jpg 164w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-1024x686.jpg 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-806x540.jpg 806w\" sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><img loading=\"lazy\" decoding=\"async\" width=\"1999\" height=\"1339\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures.webp\" alt=\"A diagram showing multiple AI agents (labeled Agent 1 through Agent N) feeding requests simultaneously into a central AI workflow\u2014consisting of LLM calls and tool use\u2014with three efficiency metrics measured at the bottom: per kW, per $\/hr, and per accelerator. &#10;\" class=\"lazyload wp-image-118656\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-172x115.jpg 172w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-300x201.jpg 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-768x514.jpg 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-625x419.jpg 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-1536x1029.jpg 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-645x432.jpg 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-448x300.jpg 448w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-134x90.jpg 134w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-362x242.jpg 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-164x110.jpg 164w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-1024x686.jpg 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/aa-agentperf-measures-806x540.jpg 806w\" data-sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><figcaption class=\"wp-element-caption\">Determine 1. The AA-AgentPerf {hardware} benchmark measures the throughput and effectivity of operating a number of AI brokers in parallel\u00a0<\/figcaption><\/figure>\n<\/div>\n<h2 id=\"measuring_representative_agentic_coding_performance\" class=\"wp-block-heading\">Measuring consultant agentic coding efficiency<\/h2>\n<p>Agentic workloads are distinctive as a result of LLM-driven selections typically produce non-deterministic sequences of requests and gear calls. Probably the most tough a part of measuring agent efficiency is to precisely seize this non-determinism in a consultant agent trajectory\u2014the entire sequence of actions, selections, and observations made by an agent because it traverses via a job from starting to finish (Determine 2).\u00a0<\/p>\n<div class=\"wp-block-image\">\n<figure data-wp-context=\"{&quot;imageId&quot;:&quot;6a2d9a5c2694f&quot;}\" data-wp-interactive=\"core\/image\" data-wp-key=\"6a2d9a5c2694f\" class=\"aligncenter size-full wp-lightbox-container\"><img decoding=\"async\" width=\"1920\" height=\"434\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory.gif\" alt=\"A simple diagram showing a left-to-right pipeline: a gray box labeled \u201cRequest\u201d points to a horizontal sequence of green boxes labeled \u201cLLM Call,\u201d \u201cTool Call,\u201d \u201cLLM Call,\u201d and \u201cTool Call,\u201d followed by ellipsis and a final output window icon. At the top, an arrow labeled \u201cAgent\u2019s Trajectory\u201d indicates the direction of the process.&#10;\" class=\"wp-image-118592\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory.gif 1920w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-179x40.gif 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-300x68.gif 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-768x174.gif 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-625x141.gif 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-1536x347.gif 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-645x146.gif 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-500x113.gif 500w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-160x36.gif 160w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-362x82.gif 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-487x110.gif 487w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-1024x231.gif 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-960x217.gif 960w\" sizes=\"(max-width: 1920px) 100vw, 1920px\"\/><img loading=\"lazy\" decoding=\"async\" width=\"1920\" height=\"434\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory.gif\" alt=\"A simple diagram showing a left-to-right pipeline: a gray box labeled \u201cRequest\u201d points to a horizontal sequence of green boxes labeled \u201cLLM Call,\u201d \u201cTool Call,\u201d \u201cLLM Call,\u201d and \u201cTool Call,\u201d followed by ellipsis and a final output window icon. At the top, an arrow labeled \u201cAgent\u2019s Trajectory\u201d indicates the direction of the process.&#10;\" class=\"lazyload wp-image-118592\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory.gif 1920w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-179x40.gif 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-300x68.gif 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-768x174.gif 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-625x141.gif 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-1536x347.gif 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-645x146.gif 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-500x113.gif 500w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-160x36.gif 160w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-362x82.gif 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-487x110.gif 487w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-1024x231.gif 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-trajectory-960x217.gif 960w\" data-sizes=\"(max-width: 1920px) 100vw, 1920px\"\/><figcaption class=\"wp-element-caption\">Determine 2. An agent\u2019s trajectory from person request to closing reply<\/figcaption><\/figure>\n<\/div>\n<p>AA-AgentPerf captures this by measuring GPU efficiency throughout prerecorded agentic coding trajectories with interleaved reasoning and gear use, whereas simulating interturn latency with a consultant baseline for CPU tool-call efficiency. These trajectories are constructed round fixing points in public code repositories throughout a number of use-cases,12+ programming languages, and response from frontier fashions. Along with rigorous definition of the trajectories, the Synthetic Evaluation group additionally:<\/p>\n<p>Leveraged consultant cached, enter, and output sequence lengths for requests, starting from 5K to 131K with a imply of roughly 27K.<\/p>\n<p>Mapped instrument calls to consultant CPU-side duties in agentic coding workflows and simulated instrument calls throughout a distribution with a one-second median delay time. The identical CPU tool-call baseline was then utilized throughout all methods examined.<\/p>\n<p>Retains the test-set personal to forestall benchmark-targeted optimization.<\/p>\n<h2 id=\"aa-agentperf_testing_and_measurement_methodology\" class=\"wp-block-heading\">AA-AgentPerf testing and measurement methodology<\/h2>\n<p>The AA-AgentPerf harness measures the variety of concurrent brokers an inference system can assist whereas assembly SLO necessities (Determine 3). At launch, this benchmark focuses on testing DeepSeek-V4-Professional throughout a number of SLO tiers derived from Synthetic Evaluation serverless API benchmarking information. This ensures that the benchmarks replicate quality-of-service ranges noticed in manufacturing suppliers at this time.\u00a0<\/p>\n<div class=\"wp-block-image\">\n<figure data-wp-context=\"{&quot;imageId&quot;:&quot;6a2d9a5c27c15&quot;}\" data-wp-interactive=\"core\/image\" data-wp-key=\"6a2d9a5c27c15\" class=\"aligncenter size-full wp-lightbox-container\"><img decoding=\"async\" width=\"1999\" height=\"1276\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency.webp\" alt=\"A scatter plot titled \u201cSLO Thresholds Define Max Concurrency\u201d shows output speed on the vertical axis and max concurrent users on the horizontal axis. Several bright green dots slope downward from left to right, transitioning to gray dots beyond a highlighted green point on the curve. Dashed horizontal and vertical lines from that green point mark the SLO threshold and the maximum number of concurrent users that still meet the target SLO.&#10;\" class=\"wp-image-118594\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-179x115.jpg 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-300x191.jpg 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-768x490.jpg 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-625x399.jpg 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-1536x980.jpg 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-645x412.jpg 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-470x300.jpg 470w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-141x90.jpg 141w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-362x231.jpg 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-172x110.jpg 172w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-1024x654.jpg 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-846x540.jpg 846w\" sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><img loading=\"lazy\" decoding=\"async\" width=\"1999\" height=\"1276\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency.webp\" alt=\"A scatter plot titled \u201cSLO Thresholds Define Max Concurrency\u201d shows output speed on the vertical axis and max concurrent users on the horizontal axis. Several bright green dots slope downward from left to right, transitioning to gray dots beyond a highlighted green point on the curve. Dashed horizontal and vertical lines from that green point mark the SLO threshold and the maximum number of concurrent users that still meet the target SLO.&#10;\" class=\"lazyload wp-image-118594\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-179x115.jpg 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-300x191.jpg 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-768x490.jpg 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-625x399.jpg 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-1536x980.jpg 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-645x412.jpg 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-470x300.jpg 470w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-141x90.jpg 141w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-362x231.jpg 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-172x110.jpg 172w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-1024x654.jpg 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/slo-thresholds-define-max-concurrency-846x540.jpg 846w\" data-sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><figcaption class=\"wp-element-caption\">Determine 3. SLO thresholds cap what number of customers will be served at goal velocity<\/figcaption><\/figure>\n<\/div>\n<p>Throughout a benchmarking run, AA-AgentPerf sends GPUs 1000&#8217;s of concurrent requests drawn from its prerecorded agent trajectory dataset. To make sure unbiased outcomes for every run, dynamic prefixes are added at first of each trajectory section. Strict SLO thresholds are enforced all through the trajectory, and the best concurrency degree that satisfies these necessities is recorded because the official benchmark outcome for a given SLO (Determine 3). This course of is then repeated throughout a number of SLO tiers to seize totally different person expertise targets (Desk 1).<\/p>\n<figure class=\"wp-block-table aligncenter\">ModelSLO tierP25 output velocity (tokens\/second)P95 TTFT (seconds)DeepSeek-V4-ProSLO #13010SLO #21005SLO #33003<figcaption class=\"wp-element-caption\">Desk 1. SLO tiers and TTFT necessities for AA-AgentPerf DeepSeek-V4-PRO exams<\/figcaption><\/figure>\n<h2 id=\"how_to_interpret_aa-agentperf_results\" class=\"wp-block-heading\">Find out how to interpret AA-AgentPerf outcomes<\/h2>\n<p>The core AA-AgentPerf metric is runtime energy per megawatt\u2014a sensible normalization for representing information middle scale efficiency. Desk 2 outlines learn how to leverage the reported efficiency to estimate what number of agentic periods could possibly be supported for a given energy funds.\u00a0<\/p>\n<figure class=\"wp-block-table aligncenter\">BenchmarkValue of metricNVIDIA GB300 NVL72NVIDIA H200Concurrent brokers per MWEnergy effectivity: What number of lively brokers a system can assist for a given energy budget61.4K2.6KConcurrent brokers per GPUHardware effectivity: How a lot serving capability is achieved per GPU57.51.4<figcaption class=\"wp-element-caption\">Desk 2. Find out how to leverage the metrics reported by AgentPerf to help in capability planning for information facilities aiming to assist agentic functions at scale. Numbers replicate AA-AgentPerf outcomes for SLO=30 configurations<\/figcaption><\/figure>\n<p>On launch day, NVIDIA GB300 NVL72 delivers as much as 20x extra concurrent brokers per megawatt than the earlier era, NVIDIA H200 (Determine 4).<\/p>\n<div class=\"wp-block-image\">\n<figure data-wp-context=\"{&quot;imageId&quot;:&quot;6a2d9a5c28e9f&quot;}\" data-wp-interactive=\"core\/image\" data-wp-key=\"6a2d9a5c28e9f\" class=\"aligncenter size-full wp-lightbox-container\"><img decoding=\"async\" width=\"1999\" height=\"1109\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1.webp\" alt=\"Bar chart comparing concurrent agents per megawatt for GB300 vs H200 at 20 and 60 tokens per second, showing GB300 providing about 20x higher capacity than H200.&#10;\" class=\"wp-image-118603\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-179x99.jpg 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-300x166.jpg 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-768x426.jpg 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-625x347.jpg 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-1536x852.jpg 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-645x358.jpg 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-500x277.jpg 500w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-160x90.jpg 160w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-362x201.jpg 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-198x110.jpg 198w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-1024x568.jpg 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-960x533.jpg 960w\" sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><img loading=\"lazy\" decoding=\"async\" width=\"1999\" height=\"1109\" data-wp-class--hide=\"state.isContentHidden\" data-wp-class--show=\"state.isContentVisible\" data-wp-init=\"callbacks.setButtonStyles\" data-wp-on--click=\"actions.showLightbox\" data-wp-on--load=\"callbacks.setButtonStyles\" data-wp-on-window--resize=\"callbacks.setButtonStyles\" src=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1.webp\" alt=\"Bar chart comparing concurrent agents per megawatt for GB300 vs H200 at 20 and 60 tokens per second, showing GB300 providing about 20x higher capacity than H200.&#10;\" class=\"lazyload wp-image-118603\" srcset=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1.webp 1999w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-179x99.jpg 179w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-300x166.jpg 300w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-768x426.jpg 768w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-625x347.jpg 625w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-1536x852.jpg 1536w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-645x358.jpg 645w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-500x277.jpg 500w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-160x90.jpg 160w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-362x201.jpg 362w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-198x110.jpg 198w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-1024x568.jpg 1024w, https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/nvidia-gb300-nvl72-agentic-coding-performance-1-960x533.jpg 960w\" data-sizes=\"(max-width: 1999px) 100vw, 1999px\"\/><figcaption class=\"wp-element-caption\">Determine 4. NVIDIA GB300 NVL72 helps way more concurrent coding brokers per megawatt than H200 at each 20 and 60 tokens per second service degree targets, reaching roughly 20x larger agent capability<\/figcaption><\/figure>\n<\/div>\n<p>This efficiency highlights how GB300 NVL72 is ready to ship throughout large-scale agentic coding workloads, from routing long-lived periods effectively to preserving combination of consultants (MoEs) and GPUs totally utilized throughout many concurrent agent periods..<\/p>\n<p>SGLang, TensorRT LLM, or vLLM: Agent runtimes apply optimizations akin to WideEP and DeepEP to unfold MoE skilled execution throughout the complete NVL72 area, maximizing efficient batch sizes and scaling successfully to 1000&#8217;s of brokers.<\/p>\n<p>DeepGEMM and Mega MoE optimizations: MXFP4\/MXFP8 kernels and fused MoE overlap NVLink communication with tensor core compute to spice up token throughput for reasoning and code era.<\/p>\n<p>NVIDIA NVLink scale-up area: GB300 NVL72 hyperlinks 72 GPUs right into a single high-bandwidth NVLink material, so each GPU can quickly share parameters, KV cache, and intermediate outcomes\u2014essential for quick, coordinated execution of agentic coding methods.<\/p>\n<h2 id=\"looking_forward_nvidia_vera_rubin_platform\" class=\"wp-block-heading\">Wanting ahead: NVIDIA Vera Rubin platform<\/h2>\n<p>AA-AgentPerf establishes the usual for evaluating agentic inference, and the outcomes spotlight how tightly built-in {hardware} and software program can unlock step-function features in concurrency and effectivity. NVIDIA GB300 NVL72 demonstrates as much as 20x larger agentic coding efficiency.\u00a0<\/p>\n<p>The NVIDIA Vera Rubin platform is anticipated to increase these features by leveraging 50 PFLOPs of NVFP4 compute and leveraging the Vera CPU to speed up LLM instrument calls and enhance end-to-end efficiency, economics, and effectivity for agentic workflows.\u00a0<\/p>\n<p>To be taught extra about why agentic workloads place distinctive calls for on inference infrastructure and the way the NVIDIA Vera Rubin platform optimizes efficiency, see Constructing for the Rising Complexity of Agentic Methods with Excessive Co-Design.<\/p>\n<h3 id=\"acknowledgments\" class=\"wp-block-heading\">Acknowledgments<\/h3>\n<p>This work was made doable via the experience and engineering contributions of Jatin Gangani, Iman Tabrizian, Xiaoming Chen, Peiheng Hu, Taizhong Wu, Shichen Li, Manu Maheswari, and plenty of different proficient NVIDIA engineers.<\/p>\n<\/div>\n<p><br \/>\n<br \/><a href=\"https:\/\/developer.nvidia.com\/blog\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/\">Source link <\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>AI brokers have essentially modified the complexity of inference workloads. Till now, the business has struggled to outline a regular for measuring how inference methods carry out underneath these circumstances. Synthetic Evaluation AgentPerf (AA-AgentPerf) gives the business\u2019s first multi-vendor open benchmarks profiling trajectories which might be consultant of real-world AI agent coding duties.\u00a0 This put [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":956,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"fifu_image_url":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-1.webp","fifu_image_alt":"","jnews-multi-image_gallery":[],"jnews_single_post":[],"jnews_primary_category":[],"jnews_override_bookmark_settings":[],"jnews_social_meta":[],"jnews_override_counter":[],"footnotes":""},"categories":[3],"tags":[1312,15,1315,1314,1313,81,750],"class_list":["post-954","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-ai-platforms-apps","tag-achieves","tag-agentic","tag-benchmark","tag-coding","tag-leading","tag-nvidia","tag-performance"],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v27.7 - https:\/\/yoast.com\/product\/yoast-seo-wordpress\/ -->\n<title>NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark - Future News 24<\/title>\n<meta name=\"description\" content=\"AI agents have fundamentally changed the complexity of inference workloads. Until now, the industry has struggled to define a standard for measuring how&#8230;\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark - Future News 24\" \/>\n<meta property=\"og:description\" content=\"AI agents have fundamentally changed the complexity of inference workloads. Until now, the industry has struggled to define a standard for measuring how&#8230;\" \/>\n<meta property=\"og:url\" content=\"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/\" \/>\n<meta property=\"og:site_name\" content=\"Future News 24\" \/>\n<meta property=\"article:published_time\" content=\"2026-06-12T21:12:00+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2026-06-13T17:59:29+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-1.webp\" \/>\n<meta name=\"author\" content=\"Future News 24\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:image\" content=\"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-1.webp\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Future News 24\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"5 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/#article\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/\"},\"author\":{\"name\":\"Future News 24\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#\\\/schema\\\/person\\\/cecad1bde21cfc357cf70128144d6c83\"},\"headline\":\"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark\",\"datePublished\":\"2026-06-12T21:12:00+00:00\",\"dateModified\":\"2026-06-13T17:59:29+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/\"},\"wordCount\":1048,\"commentCount\":0,\"publisher\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#organization\"},\"image\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/developer-blogs.nvidia.com\\\/wp-content\\\/uploads\\\/2026\\\/06\\\/ai-agent-1.webp\",\"keywords\":[\"Achieves\",\"Agentic\",\"Benchmark\",\"Coding\",\"Leading\",\"NVIDIA\",\"performance\"],\"articleSection\":[\"AI Platforms &amp; Apps\"],\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"CommentAction\",\"name\":\"Comment\",\"target\":[\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/#respond\"]}]},{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/\",\"url\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/\",\"name\":\"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark - Future News 24\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/#primaryimage\"},\"image\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/developer-blogs.nvidia.com\\\/wp-content\\\/uploads\\\/2026\\\/06\\\/ai-agent-1.webp\",\"datePublished\":\"2026-06-12T21:12:00+00:00\",\"dateModified\":\"2026-06-13T17:59:29+00:00\",\"description\":\"AI agents have fundamentally changed the complexity of inference workloads. Until now, the industry has struggled to define a standard for measuring how&#8230;\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/#primaryimage\",\"url\":\"https:\\\/\\\/developer-blogs.nvidia.com\\\/wp-content\\\/uploads\\\/2026\\\/06\\\/ai-agent-1.webp\",\"contentUrl\":\"https:\\\/\\\/developer-blogs.nvidia.com\\\/wp-content\\\/uploads\\\/2026\\\/06\\\/ai-agent-1.webp\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/2026\\\/06\\\/12\\\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\\\/\\\/futurenews24.com\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#website\",\"url\":\"https:\\\/\\\/futurenews24.com\\\/\",\"name\":\"Future News 24\",\"description\":\"The Smart Hub for AI and Next-Gen Innovation\",\"publisher\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/futurenews24.com\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Organization\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#organization\",\"name\":\"Future News 24\",\"url\":\"https:\\\/\\\/futurenews24.com\\\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#\\\/schema\\\/logo\\\/image\\\/\",\"url\":\"https:\\\/\\\/futurenews24.com\\\/wp-content\\\/uploads\\\/2026\\\/06\\\/fn24-favicon.png\",\"contentUrl\":\"https:\\\/\\\/futurenews24.com\\\/wp-content\\\/uploads\\\/2026\\\/06\\\/fn24-favicon.png\",\"width\":250,\"height\":250,\"caption\":\"Future News 24\"},\"image\":{\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#\\\/schema\\\/logo\\\/image\\\/\"}},{\"@type\":\"Person\",\"@id\":\"https:\\\/\\\/futurenews24.com\\\/#\\\/schema\\\/person\\\/cecad1bde21cfc357cf70128144d6c83\",\"name\":\"Future News 24\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g\",\"url\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g\",\"contentUrl\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g\",\"caption\":\"Future News 24\"},\"sameAs\":[\"https:\\\/\\\/futurenews24.com\"],\"url\":\"https:\\\/\\\/futurenews24.com\\\/index.php\\\/author\\\/mridulpahuja20\\\/\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark - Future News 24","description":"AI agents have fundamentally changed the complexity of inference workloads. Until now, the industry has struggled to define a standard for measuring how&#8230;","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/","og_locale":"en_US","og_type":"article","og_title":"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark - Future News 24","og_description":"AI agents have fundamentally changed the complexity of inference workloads. Until now, the industry has struggled to define a standard for measuring how&#8230;","og_url":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/","og_site_name":"Future News 24","article_published_time":"2026-06-12T21:12:00+00:00","article_modified_time":"2026-06-13T17:59:29+00:00","og_image":[{"url":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-1.webp","type":"","width":"","height":""}],"author":"Future News 24","twitter_card":"summary_large_image","twitter_image":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-1.webp","twitter_misc":{"Written by":"Future News 24","Est. reading time":"5 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/#article","isPartOf":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/"},"author":{"name":"Future News 24","@id":"https:\/\/futurenews24.com\/#\/schema\/person\/cecad1bde21cfc357cf70128144d6c83"},"headline":"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark","datePublished":"2026-06-12T21:12:00+00:00","dateModified":"2026-06-13T17:59:29+00:00","mainEntityOfPage":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/"},"wordCount":1048,"commentCount":0,"publisher":{"@id":"https:\/\/futurenews24.com\/#organization"},"image":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/#primaryimage"},"thumbnailUrl":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-1.webp","keywords":["Achieves","Agentic","Benchmark","Coding","Leading","NVIDIA","performance"],"articleSection":["AI Platforms &amp; Apps"],"inLanguage":"en-US","potentialAction":[{"@type":"CommentAction","name":"Comment","target":["https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/#respond"]}]},{"@type":"WebPage","@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/","url":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/","name":"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark - Future News 24","isPartOf":{"@id":"https:\/\/futurenews24.com\/#website"},"primaryImageOfPage":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/#primaryimage"},"image":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/#primaryimage"},"thumbnailUrl":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-1.webp","datePublished":"2026-06-12T21:12:00+00:00","dateModified":"2026-06-13T17:59:29+00:00","description":"AI agents have fundamentally changed the complexity of inference workloads. Until now, the industry has struggled to define a standard for measuring how&#8230;","breadcrumb":{"@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/#primaryimage","url":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-1.webp","contentUrl":"https:\/\/developer-blogs.nvidia.com\/wp-content\/uploads\/2026\/06\/ai-agent-1.webp"},{"@type":"BreadcrumbList","@id":"https:\/\/futurenews24.com\/index.php\/2026\/06\/12\/nvidia-achieves-leading-agentic-coding-performance-on-first-agentic-ai-benchmark\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/futurenews24.com\/"},{"@type":"ListItem","position":2,"name":"NVIDIA Achieves Main Agentic Coding Efficiency on First Agentic AI Benchmark"}]},{"@type":"WebSite","@id":"https:\/\/futurenews24.com\/#website","url":"https:\/\/futurenews24.com\/","name":"Future News 24","description":"The Smart Hub for AI and Next-Gen Innovation","publisher":{"@id":"https:\/\/futurenews24.com\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/futurenews24.com\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/futurenews24.com\/#organization","name":"Future News 24","url":"https:\/\/futurenews24.com\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/futurenews24.com\/#\/schema\/logo\/image\/","url":"https:\/\/futurenews24.com\/wp-content\/uploads\/2026\/06\/fn24-favicon.png","contentUrl":"https:\/\/futurenews24.com\/wp-content\/uploads\/2026\/06\/fn24-favicon.png","width":250,"height":250,"caption":"Future News 24"},"image":{"@id":"https:\/\/futurenews24.com\/#\/schema\/logo\/image\/"}},{"@type":"Person","@id":"https:\/\/futurenews24.com\/#\/schema\/person\/cecad1bde21cfc357cf70128144d6c83","name":"Future News 24","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/secure.gravatar.com\/avatar\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g","url":"https:\/\/secure.gravatar.com\/avatar\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/d57f07142d73cb5503ab2446ea7bc9ef3d0a5ba378d64a6157692311e42bf097?s=96&d=mm&r=g","caption":"Future News 24"},"sameAs":["https:\/\/futurenews24.com"],"url":"https:\/\/futurenews24.com\/index.php\/author\/mridulpahuja20\/"}]}},"_links":{"self":[{"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/posts\/954","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/comments?post=954"}],"version-history":[{"count":1,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/posts\/954\/revisions"}],"predecessor-version":[{"id":955,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/posts\/954\/revisions\/955"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/media\/956"}],"wp:attachment":[{"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/media?parent=954"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/categories?post=954"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/futurenews24.com\/index.php\/wp-json\/wp\/v2\/tags?post=954"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}