{"id":2797,"date":"2026-08-12T11:42:25","date_gmt":"2026-08-12T18:42:25","guid":{"rendered":"https:\/\/devblogs.microsoft.com\/foundry\/?p=2797"},"modified":"2026-08-12T11:42:25","modified_gmt":"2026-08-12T18:42:25","slug":"azure-content-understanding-gpt-5-series-guide-model-selection-grounding-improvements-and-confidence-enhancements","status":"publish","type":"post","link":"https:\/\/devblogs.microsoft.com\/foundry\/azure-content-understanding-gpt-5-series-guide-model-selection-grounding-improvements-and-confidence-enhancements\/","title":{"rendered":"Azure Content Understanding GPT-5 Series Guide: Model Selection, Grounding Improvements, and Confidence Enhancements"},"content":{"rendered":"<p>Enterprise content is no longer just something people consume. As organizations increasingly rely on AI to extract and act on information from documents, images, audio, and video, Azure Content Understanding is expanding support for the GPT-5 series and improving grounding and confidence to deliver greater flexibility, efficiency, and quality.<\/p>\n<p>This expanded model catalog enables organizations to choose the right level of intelligence for each workload, helping reduce costs for high-volume processing while preserving access to advanced reasoning capabilities where needed. It also provides optimized pipelines tuned for each model. Preprocessing allows the models to support larger files and higher quality than the simple LLM document pipelines. Generating grounding and confidence scores enables automated validation and higher straight-through processing rates. It also provides a clear path forward as older foundation models retire, allowing customers to transition to newer generations of models without redesigning their Content Understanding workflows.<\/p>\n<h2>What\u2019s New<\/h2>\n<p>This release expands Content Understanding support to the <strong>GPT-5, GPT-5.1, GPT-5.2, GPT-5.4, and GPT-5.5<\/strong> <strong>series<\/strong> including standard, mini, and nano models across document, image, video, and speech analysis.<\/p>\n<p>Just as important, the release introduces an updated grounding and confidence scoring method that generates higher quality outputs and reduces overall cost. In our tested configurations, it consumed up to 28% fewer total inference tokens and the full-inference LLM <strong>cost decrease by up to 25%<\/strong> while <strong>improving<\/strong> <strong>confidence scores accuracy and grounding accuracy by up to 14% and 3% <\/strong>respectively (measured by AUROC and grounding exact match).<\/p>\n<h2>Choosing the Best Model for Your Task<\/h2>\n<p>Think of model selection as a mixing board, with <strong>quality on your content<\/strong> and <strong>end-to-end cost<\/strong> as the two faders to adjust. A model that excels on forms may not lead on other tasks such as video segmentation, speech classification, or image generation. As organizations increasingly leverage AI to extract information from content selecting the right model is critical for balancing accuracy, cost, latency, throughput, compliance, and regional deployment requirements.<\/p>\n<p>While we cannot benchmark every combination of input types, schema definitions, and deployment topologies, below is a set of starting points and recommendations based on our testing of the most common scenarios.<\/p>\n<p>We encourage customers to leverage the table above to choose a model short list, and then evaluate on your own data to make the decision. Meanwhile, you should consider other factors such as your budget, quality bar, regional availability, throughput target, and existing capacity.<\/p>\n<h2>Table 1: General Model Selection Guidelines<\/h2>\n<table style=\"width: 72.3971%;\" width=\"691\">\n<tbody>\n<tr>\n<td style=\"width: 19.5699%;\" width=\"166\"><strong>Modality<\/strong><\/td>\n<td style=\"width: 24.8988%;\" width=\"166\"><strong>Balanced recommendation<\/strong><\/td>\n<td style=\"width: 29.996%;\" width=\"166\"><strong>Best quality<\/strong><\/td>\n<td style=\"width: 59.3362%;\" width=\"194\"><strong>Lower-cost choice<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"width: 19.5699%;\" width=\"166\"><a href=\"#documents-and-speech-a-tight-full-size-cluster\"><strong>Document<\/strong><\/a><\/td>\n<td style=\"width: 24.8988%;\" width=\"166\">GPT-5.1 or GPT-5.2<\/td>\n<td style=\"width: 29.996%;\" width=\"166\">GPT-5.5 about +<strong>2% better quality<\/strong> at about <strong>101% higher cost<\/strong><\/td>\n<td style=\"width: 59.3362%;\" width=\"194\"><strong>GPT-5.4 Mini<\/strong> costs about <strong>50% less<\/strong> with an average &#8211;<strong>2% lower quality<\/strong> on our answer match metric than balanced GPT-5.2<\/td>\n<\/tr>\n<tr>\n<td style=\"width: 19.5699%;\" width=\"166\"><a href=\"#video-the-biggest-model-is-not-the-best-model\"><strong>Video<\/strong><\/a><\/td>\n<td style=\"width: 24.8988%;\" width=\"166\">GPT-5 or GPT-5.1<\/td>\n<td style=\"width: 29.996%;\" width=\"166\">Same as balanced for this use case<\/td>\n<td style=\"width: 59.3362%;\" width=\"194\"><strong>GPT-5 Mini<\/strong> costs about <strong>28% less<\/strong> than GPT-5 with about &#8211;<strong>7% lower overall Generation F1<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"width: 19.5699%;\" width=\"166\"><a href=\"#documents-and-speech-a-tight-full-size-cluster\"><strong>Speech<sup>1<\/sup><\/strong><\/a><\/td>\n<td style=\"width: 24.8988%;\" width=\"166\">GPT-5.1 or GPT-5.2<\/td>\n<td style=\"width: 29.996%;\" width=\"166\">GPT-5.5 about +<strong>2% better quality <\/strong>than GPT-5.2 at about <strong>101% higher cost<\/strong><\/td>\n<td style=\"width: 59.3362%;\" width=\"194\"><strong>GPT-5.4 Mini<\/strong> costs about <strong>48% less<\/strong> with about &#8211;<strong>2 Answer Match points lower quality<\/strong> than balanced GPT-5.2<\/td>\n<\/tr>\n<tr>\n<td style=\"width: 19.5699%;\" width=\"166\"><a href=\"#images-first-decide-whether-quality-means-classify-or-generate\"><strong>Image<\/strong><\/a><\/td>\n<td style=\"width: 24.8988%;\" width=\"166\">GPT-5.1<\/td>\n<td style=\"width: 29.996%;\" width=\"166\">GPT-5.5 delivers about +<strong>3% better quality<\/strong> for about <strong>130% higher cost<\/strong><\/td>\n<td style=\"width: 59.3362%;\" width=\"194\">For classification, <strong>GPT-5 Mini<\/strong> costs about <strong>52% less<\/strong> than GPT-5.1 with about &#8211;<strong>12% lower Classify F1<\/strong><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p><sup>1 <\/sup><strong>Note: <\/strong>Speech and document extraction are similar tasks, so we currently recommend the same models for both.<\/p>\n<p>Detailed model quality and cost analysis is included in the <a href=\"#_Appendix:_Detailed_Model\">Appendix<\/a>.<\/p>\n<h2>Grounding and Confidence improvements<\/h2>\n<p>In this release, we also improved grounding efficiency and refreshed the underlying confidence scoring method, so customers can get more useful evidence and ranking signals without building complex post-processing systems.<\/p>\n<h3>Grounding: Fewer Tokens Consumed<\/h3>\n<p>The updated grounding system more efficiently identifies the source for the extracted data. Across all model types, we observed <strong>20-30% fewer input tokens<\/strong> and <strong>18-28% fewer total inference tokens per document<\/strong>.<\/p>\n<p><figure id=\"attachment_2799\" aria-labelledby=\"figcaption_attachment_2799\" class=\"wp-caption alignnone\" ><a href=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/token-reduction.webp\"><img decoding=\"async\" class=\" wp-image-2799\" src=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/token-reduction-300x158.webp\" alt=\"Bar chart comparing input and output tokens per document for GPT-4.1 and GPT-5.2 with previous and updated grounding methods. Updated grounding reduces tokens by 28% for GPT-4.1 and 23% for GPT-5.2, highlighted with red bars and specific token counts.\" width=\"524\" height=\"276\" srcset=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/token-reduction-300x158.webp 300w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/token-reduction-1024x540.webp 1024w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/token-reduction-768x405.webp 768w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/token-reduction.webp 1206w\" sizes=\"(max-width: 524px) 100vw, 524px\" \/><\/a><figcaption id=\"figcaption_attachment_2799\" class=\"wp-caption-text\">Updated grounding reduced zero-shot inference tokens by 28% for GPT-4.1 and 23% for GPT-5.2<\/figcaption><\/figure><\/p>\n<p>Those reductions lowered the full-inference LLM bill by <strong>11-25%<\/strong> across the tested model-and-labeled-sample configurations and reduced P50 and P95 latency. Grounding accuracy remains similar to the previous method.<\/p>\n<h3>Confidence: Improved Accuracy and Generalizability<\/h3>\n<p>We crafted a new confidence scoring method that is applicable to a broader set of models. Measured by <strong>AUROC<\/strong>, which measures how reliable the confidence scores rank correct fields above incorrect fields across various threshold, the new confidence scoring system improved by about <strong>9% for GPT-4.1<\/strong> and <strong>14% for GPT-5.2<\/strong> against the previous method.<\/p>\n<p><figure id=\"attachment_2801\" aria-labelledby=\"figcaption_attachment_2801\" class=\"wp-caption alignnone\" ><a href=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/Conf-ranking.webp\"><img decoding=\"async\" class=\" wp-image-2801\" src=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/Conf-ranking-300x167.webp\" alt=\"Conf ranking image\" width=\"510\" height=\"284\" srcset=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/Conf-ranking-300x167.webp 300w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/Conf-ranking-768x428.webp 768w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/Conf-ranking.webp 1005w\" sizes=\"(max-width: 510px) 100vw, 510px\" \/><\/a><figcaption id=\"figcaption_attachment_2801\" class=\"wp-caption-text\">Confidence-sensitive workloads: We observed a significant confidence-model quality drop with GPT-5 Mini and GPT-5 Nano. Avoid these models when confidence quality is important to your workflow.<\/figcaption><\/figure><\/p>\n<p>A field&#8217;s baseline confidence score combines several inputs, and score distributions can differ by field type. For straight-through processing, set acceptance thresholds field by field and recalibrate them whenever you switch models.<\/p>\n<h2>Appendix: Detailed Model Quality\/Cost Analysis<\/h2>\n<p>In this Appendix, we present more details on the model quality and cost analysis conducted in-house. These results may not be representative for every production workload, but it should be helpful to the reader as a guidance for model selection.<\/p>\n<h3>Documents\/Speech: Finding the Right Balance<\/h3>\n<h3>Dataset<\/h3>\n<p>The evaluation dataset span from structured to semi-structured to unstructured documents, with a total of 71 document types. Since Content Understanding supports field extraction with and without labeled samples, we tested configurations where there are zero, one, five, and all available labeled samples, and average across them to calculate the average accuracy of a given model.<\/p>\n<p>Quality is reported as a macro average across leaf fields and analyzers, excluding container fields such as arrays and objects. The evaluated documents averaged <strong>3.35 pages<\/strong>. Note workloads with substantially longer files, different schemas, or different training-example strategies may see a different cost and quality frontier.<\/p>\n<p>We use these document results to guide the current speech recommendation because both tasks extract structured fields from source content. We are not publishing a separate speech accuracy graph.<\/p>\n<h3>Results<\/h3>\n<p>In the figure below, GPT-5.1 and GPT-5.2 stand out as well-balanced models between AI quality and cost, and they differ by about <strong>1% in Answer Match score<\/strong>. At the top of the quality range, GPT-5.5 delivers about <strong>2% more Answer Match<\/strong> than GPT-5.2, but increases estimated average cost by about <strong>101%<\/strong>.<\/p>\n<p>Compared with GPT-5.2, GPT-5.4 Mini reduces estimated cost by about <strong>48%<\/strong> while giving up about <strong>2% in Answer Match<\/strong>. We recommend customers to start with GPT-5.1\/5.2 or GPT-5.4 mini, then add GPT-5.5 when its quality improvement can justify the premium.<\/p>\n<p><figure id=\"attachment_2802\" aria-labelledby=\"figcaption_attachment_2802\" class=\"wp-caption alignnone\" ><a href=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/docs.webp\"><img decoding=\"async\" class=\" wp-image-2802\" src=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/docs-300x209.webp\" alt=\"docs image\" width=\"514\" height=\"358\" srcset=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/docs-300x209.webp 300w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/docs-1024x715.webp 1024w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/docs-768x536.webp 768w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/docs.webp 1206w\" sizes=\"(max-width: 514px) 100vw, 514px\" \/><\/a><figcaption id=\"figcaption_attachment_2802\" class=\"wp-caption-text\">GPT-5.1 and GPT-5.2 form the balanced document cost-quality frontier<\/figcaption><\/figure><\/p>\n<h3>Video: the Biggest Model is not the Best Model<\/h3>\n<h3>Dataset<\/h3>\n<p>The video evaluation combined two distinct workload shapes: a <strong>segmentation-focused, 60-minute video dataset<\/strong>, and a short whole-video dataset averaging <strong>just under one minute<\/strong>. We compute generation F1 score covering both scalar answer fields and timestamped custom segments, e.g., semantically relevant time windows such as when a logo appears on screen, the duration of a news segment, or an ad break.<\/p>\n<p>This benchmark is most relevant to workloads that extract both video-level facts and segments from long media. Short clips, different frame density, or schemas without segmentation may produce a different ranking.<\/p>\n<h3>Results<\/h3>\n<p>As shown in the figure below, GPT-5 reached <strong>89.0% overall Generation F1<\/strong>, GPT-5.1 reached <strong>88.6%<\/strong>, and GPT-5.4 reached <strong>87.5%<\/strong>. GPT-5 and GPT-5.1 form both the balanced and best-quality pair for this use case; GPT-5 also cost about <strong>18% less<\/strong> than the GPT-4.1 baseline in the release summary.<\/p>\n<p>For a lower-cost video option, we recommend customers start with <strong>GPT-5 Mini<\/strong>. It cost about <strong>28% less<\/strong> than GPT-5 on the segmentation-focused benchmark while giving up about <strong>7% overall Generation F1<\/strong>.<\/p>\n<p>GPT-5 matches baseline video segmentation quality at 18% lower benchmark cost<\/p>\n<p><figure id=\"attachment_2805\" aria-labelledby=\"figcaption_attachment_2805\" class=\"wp-caption alignnone\" ><a href=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/video.webp\"><img decoding=\"async\" class=\" wp-image-2805\" src=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/video-300x209.webp\" alt=\"video image\" width=\"508\" height=\"354\" srcset=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/video-300x209.webp 300w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/video-1024x713.webp 1024w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/video-768x535.webp 768w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/video.webp 1206w\" sizes=\"(max-width: 508px) 100vw, 508px\" \/><\/a><figcaption id=\"figcaption_attachment_2805\" class=\"wp-caption-text\">GPT-5 matches baseline video segmentation quality at 18% lower benchmark cost<\/figcaption><\/figure><\/p>\n<h3>Images: Select Different Models for Classify and Generate tasks<\/h3>\n<h3>Dataset<\/h3>\n<p>The image evaluation covered <strong>17 zero-shot datasets<\/strong> with five repeats. The set spanned classification and generative tasks across product, apparel, industrial-defect, scene, and object-oriented content.<\/p>\n<p>Classification used F1 as metrics. Generative fields used a <strong>1-7 rubric score<\/strong>, so the two quality axes answer different questions and should not be collapsed into one winner.<\/p>\n<h3>Results<\/h3>\n<p>As shown in the figure below, GPT-5.1 led classification at <strong>69.5% Classify F1<\/strong>, making it an attractive choice for that use case. For generation, GPT-5.5 led at <strong>5.0 out of 7<\/strong>, about <strong>3% higher<\/strong> than GPT-5 Mini at about <strong>375% higher cost<\/strong>.<\/p>\n<p><a href=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img.webp\"><img decoding=\"async\" class=\"alignnone wp-image-2804\" src=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-300x209.webp\" alt=\"img image\" width=\"515\" height=\"359\" srcset=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-300x209.webp 300w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-1024x715.webp 1024w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-768x536.webp 768w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img.webp 1206w\" sizes=\"(max-width: 515px) 100vw, 515px\" \/><\/a><\/p>\n<p>For a lower-cost classification option, GPT-5 Mini cost about <strong>52% less<\/strong> than GPT-5.1 while giving up about <strong>12% Classify F1<\/strong>. We recommend customers to start with GPT-5.1 for classification-led workloads, and GPT-5 Mini for generation-led workloads. If necessary, test if GPT-5.5&#8217;s quality gain could justify the premium.<\/p>\n<p><figure id=\"attachment_2803\" aria-labelledby=\"figcaption_attachment_2803\" class=\"wp-caption alignnone\" ><a href=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-2.webp\"><img decoding=\"async\" class=\" wp-image-2803\" src=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-2-300x209.webp\" alt=\"img 2 image\" width=\"512\" height=\"357\" srcset=\"https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-2-300x209.webp 300w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-2-1024x713.webp 1024w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-2-768x535.webp 768w, https:\/\/devblogs.microsoft.com\/foundry\/wp-content\/uploads\/sites\/89\/2026\/08\/img-2.webp 1206w\" sizes=\"(max-width: 512px) 100vw, 512px\" \/><\/a><figcaption id=\"figcaption_attachment_2803\" class=\"wp-caption-text\">GPT-5.1 leads image classification while costing less than the GPT-4.1 baseline<\/figcaption><\/figure><\/p>\n<h3>So what should teams do next?<\/h3>\n<p>The expanded catalog gives every channel on the mixing board a wider range. The next step is to choose two or three candidates for your workload and test which combination of quality, cost, availability, and throughput deserves the final setting.<\/p>\n<p>Keep the comparison simple: use the same analyzer, schema, representative input set, and labeled examples for every run. Change only the model deployment, then compare output quality, latency, token usage, and failure rate. Before testing, check the analyzer&#8217;s supportedModels response and confirm that each candidate is available in your region.<\/p>\n<h3>Start in Content Understanding Studio<\/h3>\n<ol>\n<li>In Content Understanding Studio, open <strong>Settings<\/strong>, add your Foundry resource, and configure its <a href=\"https:\/\/learn.microsoft.com\/en-us\/azure\/ai-services\/content-understanding\/concepts\/models-deployments?tabs=studio#option-1-set-default-deployments-at-the-resource-level\">default model deployments<\/a>. Studio can deploy required models automatically when no suitable default exists.<\/li>\n<li>Follow the <a href=\"https:\/\/learn.microsoft.com\/en-us\/azure\/ai-services\/content-understanding\/quickstart\/content-understanding-studio?tabs=portal%2Ccu-studio\">Content Understanding Studio quickstart<\/a> to select an analyzer and run it on your own representative content.<\/li>\n<li>Test each candidate against the same files. Review the extracted fields and raw response, and record the quality, latency, and usage that matter to your workload.<\/li>\n<\/ol>\n<h3>Or compare deployments through the REST API<\/h3>\n<p>For a repeatable evaluation, pass a different modelDeployments mapping in each analyze request. A request-level mapping <a href=\"https:\/\/learn.microsoft.com\/en-us\/azure\/ai-services\/content-understanding\/concepts\/models-deployments?tabs=studio#option-2-pass-model-deployments-in-each-analyze-request\">overrides the resource defaults<\/a>, so you can keep the analyzer and inputs unchanged while swapping the completion deployment.<\/p>\n<p>For example, you can call:<\/p>\n<pre class=\"prettyprint language-default\"><code class=\"language-default\">POST \/contentunderstanding\/analyzers\/myInvoice:analyze\r\n\r\n{\r\n\u00a0 \"inputs\": [\r\n\u00a0\u00a0\u00a0 {\r\n\u00a0\u00a0\u00a0\u00a0\u00a0 \"url\": \"&lt;representative-input-url&gt;\"\r\n\u00a0\u00a0\u00a0 }\r\n\u00a0 ],\r\n\u00a0 \"modelDeployments\": {\r\n\u00a0\u00a0\u00a0 \"prebuilt-analyzer-completion\": \"&lt;candidate-deployment-name&gt;\",\r\n\u00a0\u00a0\u00a0 \"prebuilt-analyzer-embedding\": \"&lt;embedding-deployment-name&gt;\"\r\n\u00a0 }\r\n}<\/code><\/pre>\n<p>Run the same request once per candidate, then compare the extracted results and the response&#8217;s usage data. Start with the balanced recommendation, add the lower-cost option, and include the best-quality model only when the remaining accuracy gap matters to the workflow.<\/p>\n<h2>Additional Links<\/h2>\n<ul>\n<li>For more on new capabilities in the latest Content Understanding Preview release &#8211; <a href=\"https:\/\/aka.ms\/cu-august-blog-new-features\">From Sync APIs to support for the GPT-5 model series and agentic<\/a><\/li>\n<li>To try out Content Understanding in CU Studio &#8211; <a href=\"https:\/\/learn.microsoft.com\/en-us\/azure\/ai-services\/content-understanding\/quickstart\/content-understanding-studio?tabs=portal%2Ccu-studio\">Quickstart Try out Content Understanding Studio or Foundry portal &#8211; Foundry Tools | Microsoft Learn<\/a><\/li>\n<li>To use the CU APIs and SDK see &#8211; <a href=\"https:\/\/learn.microsoft.com\/en-us\/azure\/ai-services\/content-understanding\/quickstart\/use-rest-api?tabs=portal%2Cdocument&amp;pivots=programming-language-rest\">Quickstart: Azure Content Understanding in Foundry Tools &#8211; Foundry Tools | Microsoft Learn<\/a><\/li>\n<\/ul>\n","protected":false},"excerpt":{"rendered":"<p>Enterprise content is no longer just something people consume. As organizations increasingly rely on AI to extract and act on information from documents, images, audio, and video, Azure Content Understanding is expanding support for the GPT-5 series and improving grounding and confidence to deliver greater flexibility, efficiency, and quality. This expanded model catalog enables organizations [&hellip;]<\/p>\n","protected":false},"author":202690,"featured_media":1563,"comment_status":"open","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"_acf_changed":false,"footnotes":""},"categories":[1],"tags":[25,32,10,144,33,121,28],"class_list":["post-2797","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-microsoft-foundry","tag-agents","tag-ai","tag-ai-agents","tag-content-understanding","tag-foundry","tag-foundry-tools","tag-whats-new"],"acf":[],"blog_post_summary":"<p>Enterprise content is no longer just something people consume. As organizations increasingly rely on AI to extract and act on information from documents, images, audio, and video, Azure Content Understanding is expanding support for the GPT-5 series and improving grounding and confidence to deliver greater flexibility, efficiency, and quality. This expanded model catalog enables organizations [&hellip;]<\/p>\n","_links":{"self":[{"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/posts\/2797","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/users\/202690"}],"replies":[{"embeddable":true,"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/comments?post=2797"}],"version-history":[{"count":2,"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/posts\/2797\/revisions"}],"predecessor-version":[{"id":2831,"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/posts\/2797\/revisions\/2831"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/media\/1563"}],"wp:attachment":[{"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/media?parent=2797"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/categories?post=2797"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/devblogs.microsoft.com\/foundry\/wp-json\/wp\/v2\/tags?post=2797"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}