{"id":52857,"date":"2026-07-23T11:30:00","date_gmt":"2026-07-23T18:30:00","guid":{"rendered":"https:\/\/azure.microsoft.com\/en-us\/blog\/?p=52857"},"modified":"2026-07-23T11:39:33","modified_gmt":"2026-07-23T18:39:33","slug":"att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd","status":"publish","type":"post","link":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/","title":{"rendered":"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD"},"content":{"rendered":"\n<h2 class=\"wp-block-heading\" id=\"first-of-its-kind-telecom-ai-deployment\">First-of-its-kind telecom AI deployment<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">Telecommunications organizations are increasingly looking to AI to help teams navigate highly specialized domains, but generic models often lack the industry-specific knowledge needed to understand telecom networks, standards, and operations. To address that gap, AT&amp;T created their Open Telco (OTel) models, the next generation of telecom-focused AI designed to bring deeper telecommunications expertise into AI systems. Building OTel2.0 required more than training a large language model, it reflected a broader issue many organizations face: how to build domain-specific AI systems at scale while balancing cost, performance, and operational complexity. Cost management quickly became a key consideration. To continue advancing telecom-focused AI, AT&amp;T needed a platform capable of supporting OTel2.0 development at an entirely new scale.<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">Where teams previously had to own and manage deployments, infrastructure, and the associated operational overhead, <a href=\"https:\/\/learn.microsoft.com\/en-us\/azure\/foundry\/concepts\/managed-compute-overview\">Foundry Managed Compute<\/a> provided a more streamlined way to access dedicated graphics processing unit (GPU) capacity. This transformation requires more than powerful models; it requires the ability to scale without compromising cost, flexibility, or performance.<\/p>\n\n\n\n<div class=\"wp-block-buttons is-content-justification-center is-layout-flex wp-container-core-buttons-is-layout-a89b3969 wp-block-buttons-is-layout-flex\">\n<div class=\"wp-block-button\"><a class=\"wp-block-button__link wp-element-button\" href=\"https:\/\/www.gsma.com\/newsroom\/press-release\/gsma-launches-open-telco-ai-to-accelerate-development-of-telco%E2%80%91grade-ai\/\" target=\"_blank\" rel=\"noopener noreferrer noreferrer noopener\">Learn how OTel2.0 scales telecom AI<strong style=\"font-family: inherit;font-size: 15.008px;font-style: inherit;letter-spacing: 0.30016px;text-transform: inherit\"><\/strong><\/a><\/div>\n<\/div>\n\n\n\n<p class=\"wp-block-paragraph\">\n  Using <a href=\"https:\/\/devblogs.microsoft.com\/foundry\/announcing-foundry-managed-compute\/\">Microsoft Foundry Managed Compute<\/a>, AT&amp;T was able to experiment across multiple open models, optimize workloads across different GPU architectures, and process massive volumes of telecom data all within a unified platform. The result was an AI development environment capable of supporting trillions of tokens while giving teams the flexibility to iterate, optimize, and innovate faster.\n<\/p>\n\n\n\n<h2 class=\"wp-block-heading\" id=\"model-choice-meets-infrastructure-flexibility\">Model choice meets infrastructure flexibility <\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">Building OTel2.0 required flexibility across both models and infrastructure. Rather than standardizing on a single model, AT&amp;T adopted a multi open-model strategy. Open models were central to AT&amp;T\u2019s approach because they provided the flexibility to work with approved telecom data, tailor the workflow for domain-specific model development, and support large-scale experimentation with greater control over cost and deployment strategy. Through <a href=\"https:\/\/azure.microsoft.com\/en-us\/products\/ai-foundry\" target=\"_blank\" rel=\"noopener noreferrer\">Microsoft Foundry<\/a>, the team deployed several models from the Hugging Face collection, including <a href=\"https:\/\/azure.microsoft.com\/en-us\/products\/phi\" target=\"_blank\" rel=\"noopener noreferrer\">Phi-4<\/a>, OSS-120B, and Gemma-4, to support different stages of development, from synthetic data generation and data preparation to reasoning-intensive workloads and broader model development efforts. Phi-4 played a significant role in this process, processing more than 700 billion tokens a month as part of the broader data preparation and training workflow for OTel2.0.<\/p>\n\n\n\n<blockquote class=\"wp-block-quote has-quote-default-font-size is-layout-flow wp-block-quote-is-layout-flow\">\n<p class=\"wp-block-paragraph\">Every company in the world needs to build its own AI, and that is only possible with open models and open source. AT&amp;T is championing this vision, building on open models like Phi-4 and Gemma, and giving OTel back to the community as a telecom AI foundation others can build upon. Microsoft Foundry makes this practical at scale, bringing the latest open models from the Hugging Face collection together with AMD and NVIDIA GPUs in one place, so teams can pick the right model and the right hardware, then deploy in hours instead of weeks.<\/p>\n<cite>\u2014Jeff Boudier, Vice President of Product, Hugging Face<\/cite><\/blockquote>\n\n\n\n<p class=\"wp-block-paragraph\"><a href=\"https:\/\/www.gsma.com\/newsroom\/press-release\/gsma-launches-open-telco-ai-to-accelerate-development-of-telco%E2%80%91grade-ai\/\" target=\"_blank\" rel=\"noreferrer noopener\">Developing OTel2.0<\/a> also required infrastructure capable of operating at telecom scale. AT&amp;T used approximately <strong>530 GPUs<\/strong> through Microsoft Foundry Managed Compute spanning multiple GPU architectures including <strong>430 AMD Instinct\u2122 MI300X GPUs<\/strong>. This heterogenous approach gave AT&amp;T more flexibility in how models were deployed and\u00a0optimized as requirements evolved.<\/p>\n\n\n\n<figure class=\"wp-block-table aligncenter is-style-stripes\"><table class=\"has-white-color has-electric-blue-400-background-color has-text-color has-background has-link-color has-fixed-layout\"><thead><tr><th class=\"has-text-align-left\" data-align=\"left\">Model<\/th><th class=\"has-text-align-left\" data-align=\"left\">Example workload<\/th><\/tr><\/thead><tbody><tr><td class=\"has-text-align-left\" data-align=\"left\">Phi-4<\/td><td class=\"has-text-align-left\" data-align=\"left\">Around 700B tokens a month for <br>data preparation and synthetic <br>data generation<\/td><\/tr><tr><td class=\"has-text-align-left\" data-align=\"left\">OSS 120B<\/td><td class=\"has-text-align-left\" data-align=\"left\">Higher-reasoning workloads <\/td><\/tr><tr><td class=\"has-text-align-left\" data-align=\"left\">Gemma 4<\/td><td class=\"has-text-align-left\" data-align=\"left\">OTel2.0 development workflows<\/td><\/tr><\/tbody><\/table><figcaption class=\"wp-element-caption\">Table 1:&nbsp;Explains what&nbsp;open source&nbsp;models were used and how<\/figcaption><\/figure>\n\n\n\n<p class=\"wp-block-paragraph\">This flexibility illustrates a broader trend across AI development. Organizations increasingly need platforms that allow them to choose the right model for the job, optimize for cost and performance, and scale workloads without rebuilding operational environments. <a href=\"https:\/\/azure.microsoft.com\/en-us\/products\/ai-foundry\" target=\"_blank\" rel=\"noopener noreferrer\">Microsoft Foundry<\/a> brings model choice, infrastructure flexibility, governance, and operational scale together in a unified platform that supports those requirements.<\/p>\n\n\n\n<div class=\"wp-block-buttons is-content-justification-center is-layout-flex wp-container-core-buttons-is-layout-a89b3969 wp-block-buttons-is-layout-flex\">\n<div class=\"wp-block-button\"><a class=\"wp-block-button__link wp-element-button\" href=\"https:\/\/azure.microsoft.com\/en-us\/products\/ai-foundry\" target=\"_blank\" rel=\"noopener noreferrer\">Start building with Microsoft Foundry<\/a><\/div>\n<\/div>\n\n\n\n<p class=\"wp-block-paragraph\">Beyond flexibility and cost, deployment speed is a critical factor for many AI initiatives. As workloads expand and new models are evaluated, the ability to access GPU capacity quickly enables teams to move from experimentation to execution faster without lengthy provisioning cycles. With Foundry Managed Compute, AT&amp;T could deploy and scale models in days rather than waiting weeks for infrastructure to become available, helping accelerate development timelines and maintain momentum across OTel2.0 development.<\/p>\n\n\n\n<h2 class=\"wp-block-heading\" id=\"optimizing-cost-without-limiting-innovation\">Optimizing cost without limiting innovation  <\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">\n  As AI workloads grow, economics become as important as model performance. For AT&amp;T,  one of the primary objectives was to lower AI model consumption costs while continuing to drive meaningful business value through AI-powered innovation. By using open models on <a href=\"https:\/\/learn.microsoft.com\/en-us\/azure\/foundry\/concepts\/managed-compute-overview\">Microsoft Foundry Managed Compute<\/a>, AT&amp;T was able to support large-scale data preparation and model development using a different economic model built around dedicated GPU infrastructure and open-model flexibility.\n<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">The impact became clear at scale. In support of OTel2.0, AT&amp;T processed approximately 1T tokens, consisting of raw documents from GSMA supplemented by synthetic data generated. Generating the data using open-source models like Phi-4, served by Microsoft&#8217;s Foundry Managed Compute, saved tens of millions of dollars versus using frontier models. This allowed teams to invest in larger-scale experimentation and development while maintaining a focus on business value and operational efficiency.<\/p>\n\n\n\n<figure class=\"wp-block-table aligncenter is-style-stripes has-body-medium-font-size\"><table class=\"has-white-color has-electric-blue-400-background-color has-text-color has-background has-link-color has-fixed-layout\"><thead><tr><th>Metric<\/th><th class=\"has-text-align-left\" data-align=\"left\">Value<\/th><\/tr><\/thead><tbody><tr><td>OTel 1.0 Downloads<\/td><td class=\"has-text-align-left\" data-align=\"left\">Over 25M<\/td><\/tr><tr><td>GPUs Used Through Foundry Managed Compute<\/td><td class=\"has-text-align-left\" data-align=\"left\">About 530<\/td><\/tr><tr><td>Tokens Processed for OTel2.0<\/td><td class=\"has-text-align-left\" data-align=\"left\">About 1T<\/td><\/tr><tr><td>Tokens Trained for OTel2.0<\/td><td class=\"has-text-align-left\" data-align=\"left\">About&nbsp;400&nbsp;B<\/td><\/tr><tr><td>Models used to train OTel<\/td><td class=\"has-text-align-left\" data-align=\"left\">Phi-4, OSS 120B, Gemma 4<\/td><\/tr><\/tbody><\/table><figcaption class=\"wp-element-caption\">Table 2: Quick facts about the OTel model family and metrics around what was used to build OTel2.0&nbsp;<\/figcaption><\/figure>\n\n\n\n<blockquote class=\"wp-block-quote has-quote-default-font-size is-layout-flow wp-block-quote-is-layout-flow\">\n<p class=\"wp-block-paragraph\">When you are processing hundreds of billions of tokens, infrastructure becomes part of the problem you solve. Foundry Managed Compute gave us access to GPU capacity at scale so our teams could focus on advancing OTel2.0 instead of managing infrastructure.<\/p>\n<cite>\u2014Mark Austin, Vice President, Data Science and AI at AT&amp;T<\/cite><\/blockquote>\n\n\n\n<p class=\"wp-block-paragraph\">\n  At this scale, infrastructure is no longer simply a <a id=\"post-52857-_Int_JNQm0EQ5\"><\/a>deployment consideration. It <a id=\"post-52857-_Int_PdbJnM1r\"><\/a>becomes a strategic component of AI development.\n<\/p>\n\n\n\n<h2 class=\"wp-block-heading\" id=\"accelerating-the-next-wave-of-production-scale-ai\">Accelerating the next wave of production-scale AI<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">OTel 2.0 demonstrates how organizations can combine open models, scalable infrastructure, and domain expertise to build production-ready AI systems. By matching different models to different workloads and optimizing infrastructure for cost and performance, AT&amp;T was able to process trillions of tokens while maintaining operational efficiency.&nbsp;<\/p>\n\n\n\n<p class=\"wp-block-paragraph\">As organizations move from AI experimentation to production deployment, they increasingly need the flexibility to choose the right models, optimize infrastructure, and scale efficiently. <a href=\"https:\/\/azure.microsoft.com\/en-us\/products\/ai-foundry\" target=\"_blank\" rel=\"noopener noreferrer\">Microsoft Foundry<\/a> and <a href=\"https:\/\/learn.microsoft.com\/en-us\/azure\/foundry\/concepts\/managed-compute-overview\" target=\"_blank\" rel=\"noopener noreferrer\">Foundry Managed Compute<\/a> help support that transition by bringing those capabilities together in a unified platform.<\/p>\n\n\n\n<h3 class=\"wp-block-heading\" id=\"learn-more\">Learn more<\/h3>\n\n\n\n<ul class=\"wp-block-list\">\n<li class=\"wp-block-list-item\"><a href=\"https:\/\/blogs.microsoft.com\/blog\/2026\/07\/20\/microsoft-expands-azure-ai-and-hpc-infrastructure-with-amd\/\" target=\"_blank\" rel=\"noopener noreferrer\">Read&nbsp;Scott&nbsp;Guthrie&#8217;s&nbsp;blog<\/a> about&nbsp;Azure AI and HPC infrastructure.<\/li>\n\n\n\n<li class=\"wp-block-list-item\">Learn more about <a href=\"https:\/\/www.gsma.com\/newsroom\/press-release\/gsma-launches-open-telco-ai-to-accelerate-development-of-telco%E2%80%91grade-ai\/\" target=\"_blank\" rel=\"noreferrer noopener\">OTel2.0<\/a>.<\/li>\n<\/ul>\n\n\n\n<p class=\"wp-block-paragraph\"><strong>Explore session topics from <a href=\"https:\/\/www.amd.com\/en\/corporate\/events\/advancing-ai.html\" target=\"_blank\" rel=\"noopener noreferrer\">AMD\u2019s Advancing AI<\/a><\/strong>:<\/p>\n\n\n\n<ul class=\"wp-block-list\">\n<li class=\"wp-block-list-item\"><a href=\"https:\/\/www.amd.com\/en\/corporate\/events\/advancing-ai\/sessions-catalog\/sovereign-ai-at-scale--enterprise-ready-ai-with-microsoft-and-amd.html\" target=\"_blank\" rel=\"noopener noreferrer\">From GPUs to CPUs: Optimizing Every AI Workload with Azure and AMD<\/a><\/li>\n\n\n\n<li class=\"wp-block-list-item\"><a href=\"https:\/\/www.amd.com\/en\/corporate\/events\/advancing-ai\/sessions-catalog\/whats-next-for-ai-infrastructure-in-the-cloud.html\" target=\"_blank\" rel=\"noopener noreferrer\">What&#8217;s Next for AI Infrastructure in the Cloud?<\/a><\/li>\n<\/ul>\n\n\n\n<div class=\"is-style-inline-centered wp-block-bloginabox-theme-promotional\">\n\t\n<div class=\"promotional promotional--has-media promotional--media-left\">\n\t<div class=\"promotional__wrapper\">\n\t\t<div class=\"promotional__content-wrapper\">\n\t\t\t<div class=\"promotional__content\">\n\t\t\t\t\n\n<h2 class=\"wp-block-heading\" id=\"powering-the-future-of-ai-on-azure\">Powering the Future of AI on Azure<\/h2>\n\n\n\n<p class=\"wp-block-paragraph\">Discover how Microsoft and AMD are expanding Azure AI and HPC infrastructure.<\/p>\n\n\n\n<div class=\"wp-block-buttons is-layout-flex wp-block-buttons-is-layout-flex\">\n<div class=\"wp-block-button\"><a data-bi-an=\"Global CTA\" data-bi-ct=\"cta link\" data-bi-id=\"cta-block\" class=\"wp-block-button__link wp-element-button\" href=\"https:\/\/blogs.microsoft.com\/blog\/2026\/07\/20\/microsoft-expands-azure-ai-and-hpc-infrastructure-with-amd\/\" target=\"_blank\" rel=\"noopener noreferrer\">Read more<\/a><\/div>\n<\/div>\n\n\t\t\t<\/div>\n\t\t<\/div>\n\t\t\t\t\t<div class=\"promotional__media-wrapper\">\n\t\t\t\t<div class=\"promotional__media\">\n\t\t\t\t\t\t\t\t\t\t\t<img decoding=\"async\" width=\"1920\" height=\"1242\" src=\"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2027\/06\/FY26AE-Azure-Composite-DeepBlue-006-H.jpg\" class=\"attachment-full size-full\" alt=\"Abstract shapes in blue and green and purple.\" loading=\"lazy\" srcset=\"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2027\/06\/FY26AE-Azure-Composite-DeepBlue-006-H.jpg 1920w, https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2027\/06\/FY26AE-Azure-Composite-DeepBlue-006-H-300x194.jpg 300w, https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2027\/06\/FY26AE-Azure-Composite-DeepBlue-006-H-1024x662.jpg 1024w, https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2027\/06\/FY26AE-Azure-Composite-DeepBlue-006-H-768x497.jpg 768w, https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2027\/06\/FY26AE-Azure-Composite-DeepBlue-006-H-1536x994.jpg 1536w\" sizes=\"auto, (max-width: 1920px) 100vw, 1920px\" \/>\t\t\t\t\t\t\t\t\t<\/div>\n\t\t\t<\/div>\n\t\t\t<\/div>\n<\/div>\n<\/div>\n","protected":false},"excerpt":{"rendered":"<p>AT&amp;T processed approximately one trillion tokens while developing OTel2.0 using Microsoft Foundry Managed Compute, open AI models, and AMD and NVIDIA GPU infrastructure. Discover how flexible model choice and scalable infrastructure are enabling production-scale telecom AI.<\/p>\n","protected":false},"author":58,"featured_media":52957,"comment_status":"open","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"ep_exclude_from_search":false,"_classifai_error":"","_classifai_text_to_speech_error":"","_alt_title":"","ms-ems-related-posts":[],"footnotes":"","azure_community_cta_settings":[]},"categories":[1454],"tags":[3423,3421,3418,3420,3422,3417],"audience":[3054,3053],"content-type":[1527],"product":[3164],"tech-community":[],"coauthors":[3091],"class_list":["post-52857","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-ai-machine-learning","tag-amd-gpus","tag-att","tag-foundry-managed-compute","tag-microsoft-foundry","tag-open-source-models","tag-otel-2-0","audience-business-decision-makers","audience-it-decision-makers","content-type-customer-stories","product-microsoft-foundry","review-flag-1680286581-56","review-flag-1-1680286581-825","review-flag-2-1680286581-601","review-flag-4-1680286581-250","review-flag-new-1680286579-546"],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v27.4 - https:\/\/yoast.com\/product\/yoast-seo-wordpress\/ -->\n<title>AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD | Microsoft Azure Blog<\/title>\n<meta name=\"description\" content=\"Discover how AT&amp;T used Microsoft Foundry Managed Compute, AMD and NVIDIA GPUs, and open AI models to process one trillion tokens, reduce AI costs, and accelerate telecom AI innovation at scale.\" \/>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD | Microsoft Azure Blog\" \/>\n<meta property=\"og:description\" content=\"Discover how AT&amp;T used Microsoft Foundry Managed Compute, AMD and NVIDIA GPUs, and open AI models to process one trillion tokens, reduce AI costs, and accelerate telecom AI innovation at scale.\" \/>\n<meta property=\"og:url\" content=\"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/\" \/>\n<meta property=\"og:site_name\" content=\"Microsoft Azure Blog\" \/>\n<meta property=\"article:publisher\" content=\"https:\/\/www.facebook.com\/microsoftazure\" \/>\n<meta property=\"article:published_time\" content=\"2026-07-23T18:30:00+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2026-07-23T18:39:33+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2026\/07\/Foundry-AT-T-OTel-2_2-2-1.jpg\" \/>\n\t<meta property=\"og:image:width\" content=\"1920\" \/>\n\t<meta property=\"og:image:height\" content=\"1080\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/jpeg\" \/>\n<meta name=\"author\" content=\"Steve Sweetman\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:image\" content=\"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2026\/07\/Foundry-AT-T-OTel-2_2-2-1.jpg\" \/>\n<meta name=\"twitter:creator\" content=\"@azure\" \/>\n<meta name=\"twitter:site\" content=\"@azure\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Steve Sweetman\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"5 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\\\/\\\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/#article\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/\"},\"author\":[{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/author\\\/steve-sweetman\\\/\",\"@type\":\"Person\",\"@name\":\"Steve Sweetman\"}],\"headline\":\"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD\",\"datePublished\":\"2026-07-23T18:30:00+00:00\",\"dateModified\":\"2026-07-23T18:39:33+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/\"},\"wordCount\":1158,\"commentCount\":0,\"publisher\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/#organization\"},\"image\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/Foundry-AT-T-OTel-2_2-2-1.jpg\",\"keywords\":[\"AMD GPUs\",\"AT&amp;T\",\"Foundry Managed Compute\",\"Microsoft Foundry\",\"Open Source models\",\"OTel 2.0\"],\"articleSection\":[\"AI + machine learning\"],\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"CommentAction\",\"name\":\"Comment\",\"target\":[\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/#respond\"]}]},{\"@type\":\"WebPage\",\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/\",\"url\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/\",\"name\":\"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD | Microsoft Azure Blog\",\"isPartOf\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/#website\"},\"primaryImageOfPage\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/#primaryimage\"},\"image\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/#primaryimage\"},\"thumbnailUrl\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/Foundry-AT-T-OTel-2_2-2-1.jpg\",\"datePublished\":\"2026-07-23T18:30:00+00:00\",\"dateModified\":\"2026-07-23T18:39:33+00:00\",\"description\":\"Discover how AT&T used Microsoft Foundry Managed Compute, AMD and NVIDIA GPUs, and open AI models to process one trillion tokens, reduce AI costs, and accelerate telecom AI innovation at scale.\",\"breadcrumb\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/\"]}]},{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/#primaryimage\",\"url\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/Foundry-AT-T-OTel-2_2-2-1.jpg\",\"contentUrl\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/wp-content\\\/uploads\\\/2026\\\/07\\\/Foundry-AT-T-OTel-2_2-2-1.jpg\",\"width\":1920,\"height\":1080,\"caption\":\"How Microsoft Foundry helped power AT&T's OTel2.0.\"},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\\\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Blog home\",\"item\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"AI + machine learning\",\"item\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/category\\\/ai-machine-learning\\\/\"},{\"@type\":\"ListItem\",\"position\":3,\"name\":\"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/#website\",\"url\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/\",\"name\":\"Microsoft Azure Blog\",\"description\":\"Get the latest Azure news, updates, and announcements from the Azure blog. From product updates to hot topics, hear from the Azure experts.\",\"publisher\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/?s={search_term_string}\"},\"query-input\":{\"@type\":\"PropertyValueSpecification\",\"valueRequired\":true,\"valueName\":\"search_term_string\"}}],\"inLanguage\":\"en-US\"},{\"@type\":\"Organization\",\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/#organization\",\"name\":\"Microsoft Azure Blog\",\"url\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/#\\\/schema\\\/logo\\\/image\\\/\",\"url\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/wp-content\\\/uploads\\\/2024\\\/06\\\/microsoft_logo.webp\",\"contentUrl\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/wp-content\\\/uploads\\\/2024\\\/06\\\/microsoft_logo.webp\",\"width\":512,\"height\":512,\"caption\":\"Microsoft Azure Blog\"},\"image\":{\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/#\\\/schema\\\/logo\\\/image\\\/\"},\"sameAs\":[\"https:\\\/\\\/www.facebook.com\\\/microsoftazure\",\"https:\\\/\\\/x.com\\\/azure\",\"https:\\\/\\\/www.instagram.com\\\/microsoftdeveloper\\\/\",\"https:\\\/\\\/www.linkedin.com\\\/company\\\/16188386\",\"https:\\\/\\\/www.youtube.com\\\/user\\\/windowsazure\"]},{\"@type\":\"Person\",\"@id\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/#\\\/schema\\\/person\\\/c505c4707a8b0733197cd32e481f1fae\",\"name\":\"Estela Virko\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/e0c80aef45905808b531834966f343f124eab0a088eaf8c5091c92e213cf58e5?s=96&d=mm&r=ge79ba33f176c2122a103010d5e208f8b\",\"url\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/e0c80aef45905808b531834966f343f124eab0a088eaf8c5091c92e213cf58e5?s=96&d=mm&r=g\",\"contentUrl\":\"https:\\\/\\\/secure.gravatar.com\\\/avatar\\\/e0c80aef45905808b531834966f343f124eab0a088eaf8c5091c92e213cf58e5?s=96&d=mm&r=g\",\"caption\":\"Estela Virko\"},\"url\":\"https:\\\/\\\/azure.microsoft.com\\\/en-us\\\/blog\\\/author\\\/estelavirko\\\/\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD | Microsoft Azure Blog","description":"Discover how AT&T used Microsoft Foundry Managed Compute, AMD and NVIDIA GPUs, and open AI models to process one trillion tokens, reduce AI costs, and accelerate telecom AI innovation at scale.","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/","og_locale":"en_US","og_type":"article","og_title":"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD | Microsoft Azure Blog","og_description":"Discover how AT&T used Microsoft Foundry Managed Compute, AMD and NVIDIA GPUs, and open AI models to process one trillion tokens, reduce AI costs, and accelerate telecom AI innovation at scale.","og_url":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/","og_site_name":"Microsoft Azure Blog","article_publisher":"https:\/\/www.facebook.com\/microsoftazure","article_published_time":"2026-07-23T18:30:00+00:00","article_modified_time":"2026-07-23T18:39:33+00:00","og_image":[{"width":1920,"height":1080,"url":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2026\/07\/Foundry-AT-T-OTel-2_2-2-1.jpg","type":"image\/jpeg"}],"author":"Steve Sweetman","twitter_card":"summary_large_image","twitter_image":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2026\/07\/Foundry-AT-T-OTel-2_2-2-1.jpg","twitter_creator":"@azure","twitter_site":"@azure","twitter_misc":{"Written by":"Steve Sweetman","Est. reading time":"5 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/#article","isPartOf":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/"},"author":[{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/author\/steve-sweetman\/","@type":"Person","@name":"Steve Sweetman"}],"headline":"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD","datePublished":"2026-07-23T18:30:00+00:00","dateModified":"2026-07-23T18:39:33+00:00","mainEntityOfPage":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/"},"wordCount":1158,"commentCount":0,"publisher":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/#organization"},"image":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/#primaryimage"},"thumbnailUrl":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2026\/07\/Foundry-AT-T-OTel-2_2-2-1.jpg","keywords":["AMD GPUs","AT&amp;T","Foundry Managed Compute","Microsoft Foundry","Open Source models","OTel 2.0"],"articleSection":["AI + machine learning"],"inLanguage":"en-US","potentialAction":[{"@type":"CommentAction","name":"Comment","target":["https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/#respond"]}]},{"@type":"WebPage","@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/","url":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/","name":"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD | Microsoft Azure Blog","isPartOf":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/#website"},"primaryImageOfPage":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/#primaryimage"},"image":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/#primaryimage"},"thumbnailUrl":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2026\/07\/Foundry-AT-T-OTel-2_2-2-1.jpg","datePublished":"2026-07-23T18:30:00+00:00","dateModified":"2026-07-23T18:39:33+00:00","description":"Discover how AT&T used Microsoft Foundry Managed Compute, AMD and NVIDIA GPUs, and open AI models to process one trillion tokens, reduce AI costs, and accelerate telecom AI innovation at scale.","breadcrumb":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/"]}]},{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/#primaryimage","url":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2026\/07\/Foundry-AT-T-OTel-2_2-2-1.jpg","contentUrl":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2026\/07\/Foundry-AT-T-OTel-2_2-2-1.jpg","width":1920,"height":1080,"caption":"How Microsoft Foundry helped power AT&T's OTel2.0."},{"@type":"BreadcrumbList","@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Blog home","item":"https:\/\/azure.microsoft.com\/en-us\/blog\/"},{"@type":"ListItem","position":2,"name":"AI + machine learning","item":"https:\/\/azure.microsoft.com\/en-us\/blog\/category\/ai-machine-learning\/"},{"@type":"ListItem","position":3,"name":"AT&amp;T and Microsoft scale trillion-token workloads with Microsoft Foundry and AMD"}]},{"@type":"WebSite","@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/#website","url":"https:\/\/azure.microsoft.com\/en-us\/blog\/","name":"Microsoft Azure Blog","description":"Get the latest Azure news, updates, and announcements from the Azure blog. From product updates to hot topics, hear from the Azure experts.","publisher":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/azure.microsoft.com\/en-us\/blog\/?s={search_term_string}"},"query-input":{"@type":"PropertyValueSpecification","valueRequired":true,"valueName":"search_term_string"}}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/#organization","name":"Microsoft Azure Blog","url":"https:\/\/azure.microsoft.com\/en-us\/blog\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/#\/schema\/logo\/image\/","url":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2024\/06\/microsoft_logo.webp","contentUrl":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-content\/uploads\/2024\/06\/microsoft_logo.webp","width":512,"height":512,"caption":"Microsoft Azure Blog"},"image":{"@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/#\/schema\/logo\/image\/"},"sameAs":["https:\/\/www.facebook.com\/microsoftazure","https:\/\/x.com\/azure","https:\/\/www.instagram.com\/microsoftdeveloper\/","https:\/\/www.linkedin.com\/company\/16188386","https:\/\/www.youtube.com\/user\/windowsazure"]},{"@type":"Person","@id":"https:\/\/azure.microsoft.com\/en-us\/blog\/#\/schema\/person\/c505c4707a8b0733197cd32e481f1fae","name":"Estela Virko","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/secure.gravatar.com\/avatar\/e0c80aef45905808b531834966f343f124eab0a088eaf8c5091c92e213cf58e5?s=96&d=mm&r=ge79ba33f176c2122a103010d5e208f8b","url":"https:\/\/secure.gravatar.com\/avatar\/e0c80aef45905808b531834966f343f124eab0a088eaf8c5091c92e213cf58e5?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/e0c80aef45905808b531834966f343f124eab0a088eaf8c5091c92e213cf58e5?s=96&d=mm&r=g","caption":"Estela Virko"},"url":"https:\/\/azure.microsoft.com\/en-us\/blog\/author\/estelavirko\/"}]}},"bloginabox_animated_featured_image":null,"bloginabox_display_generated_audio":true,"distributor_meta":false,"distributor_terms":false,"distributor_media":false,"distributor_original_site_name":"Microsoft Azure Blog","distributor_original_site_url":"https:\/\/azure.microsoft.com\/en-us\/blog","push-errors":false,"_links":{"self":[{"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/posts\/52857","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/users\/58"}],"replies":[{"embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/comments?post=52857"}],"version-history":[{"count":46,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/posts\/52857\/revisions"}],"predecessor-version":[{"id":52975,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/posts\/52857\/revisions\/52975"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/media\/52957"}],"wp:attachment":[{"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/media?parent=52857"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/categories?post=52857"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/tags?post=52857"},{"taxonomy":"audience","embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/audience?post=52857"},{"taxonomy":"content-type","embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/content-type?post=52857"},{"taxonomy":"product","embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/product?post=52857"},{"taxonomy":"tech-community","embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/tech-community?post=52857"},{"taxonomy":"author","embeddable":true,"href":"https:\/\/azure.microsoft.com\/en-us\/blog\/wp-json\/wp\/v2\/coauthors?post=52857"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}