AT&T and Microsoft leverage Foundry and AMD for trillion-token AI workloads
AT&T has developed its Open Telco (OTel) 2.0 AI models using Microsoft Foundry Managed Compute and AMD GPUs to handle trillion-token workloads. This partnership addresses the need for domain-specific AI in telecommunications by offering a scalable, cost-effective platform for training and deploying models. The solution enables flexibility in model choice and infrastructure, allowing organizations to accelerate AI development and production deployment.
- →Foundry Managed Compute for scalable GPU capacity
- →Multi open-model strategy for flexibility and cost control
- →AT&T develops domain-specific AI with Microsoft Foundry and AMD
- →Accelerated AI deployment speeds
- →Cost optimization for large-scale AI
Features (2) ›
- Foundry Managed Compute for scalable GPU capacity
Microsoft Foundry Managed Compute provides streamlined access to dedicated GPU capacity, abstracting infrastructure management for AI development. AT&T utilized approximately 530 GPUs through this service, including 430 AMD Instinct™ MI300X GPUs, to experiment, optimize, and process large volumes of telecom data.
- Multi open-model strategy for flexibility and cost control
AT&T adopted a multi open-model strategy, deploying models like Phi-4, OSS-120B, and Gemma-4 from Hugging Face to support various stages of OTel2.0 development. Phi-4 processed over 700 billion tokens monthly for data preparation, demonstrating cost savings compared to frontier models.
Enhancements (2) ›
- Accelerated AI deployment speeds
Foundry Managed Compute allows for rapid deployment and scaling of models, reducing provisioning cycles. AT&T could deploy and scale models in days, accelerating development timelines for OTel2.0 compared to traditional weeks-long infrastructure availability.
- Cost optimization for large-scale AI
By using open models on Microsoft Foundry Managed Compute, AT&T achieved significant cost savings, processing approximately 1 trillion tokens for OTel2.0. This approach saved tens of millions of dollars compared to using frontier models, allowing for greater investment in experimentation.
Notes (1) ›
- AT&T develops domain-specific AI with Microsoft Foundry and AMD
AT&T has created its Open Telco (OTel) 2.0 AI models, designed for the telecommunications industry, using Microsoft Foundry Managed Compute and AMD Instinct™ MI300X GPUs. This platform allows for handling trillion-token workloads, offering flexibility in model choice and infrastructure optimization.
https://azure.microsoft.com/en-us/blog/att-and-microsoft-scale-trillion-token-workloads-with-microsoft-foundry-and-amd/
Related releases
- Claude Opus 5 available on Azure Databricks Azure Updates ·
- Azure API Management adds AI Gateway in public preview Azure Updates ·
- Azure Firewall supports HTTP header insertion Azure Updates ·
- Microsoft Foundry GA includes GPT-5.6 models and APAC Data Zone Microsoft Azure Blog ·
- Azure Kubernetes Service 1.33 reaches end of life in 7 days endoflife.date ·
- Azure Database for PostgreSQL 11 reaches end of life in 7 days endoflife.date ·