Close Menu

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    Apple Set to Launch Its First Foldable iPhone, the Duo, in October with Advanced Features

    September 16, 2026

    Bank of England to Review Interest Rate Stability and Gilt Reduction Strategy in September Meeting

    September 15, 2026

    Endangered Australian Sea Lion Succumbs to H5N1 Bird Flu on Kangaroo Island for the First Time

    September 14, 2026
    Manchester ExaminerManchester Examiner
    • Home
    • Contact Us
    • Automotive
    • Business
    • Entertainment
    • Health
    • Lifestyle
    • Luxury
    • News
    • Sports
    • Technology
    • Travel
    Manchester ExaminerManchester Examiner
    Home » Maia 200 boosts Microsoft Azure with new AI inference silicon
    Technology

    Maia 200 boosts Microsoft Azure with new AI inference silicon

    January 28, 2026
    Facebook WhatsApp Twitter Pinterest LinkedIn Telegram Tumblr Email Reddit VKontakte

    EuroWire, SAN FRANCISCO: Microsoft on Jan. 26 introduced Maia 200, the second generation of its in-house artificial intelligence accelerator, built to run AI models in production across Azure data centres. The company said Maia 200 is designed for inference, the stage where trained models generate responses to live requests, and will be used to support a range of Microsoft AI services.

    Maia 200 boosts Microsoft Azure with new AI inference silicon
    Maia 200 brings high-bandwidth HBM3e memory to Microsoft’s Azure AI inference platform. (AI-generated image)

    Maia 200 is manufactured on TSMC’s 3-nanometer process and includes more than 140 billion transistors, Microsoft said. The chip pairs compute with a new memory system that includes 216 gigabytes of HBM3e high-bandwidth memory and about 272 megabytes of on-chip SRAM, aimed at sustaining large-scale token generation and other inference-heavy workloads.

    Microsoft said Maia 200 delivers more than 10 petaflops of performance at 4-bit precision and about 5 petaflops at 8-bit precision, formats commonly used to run modern generative AI efficiently. The company also said the system is designed around a 750-watt power envelope and is built with scalable networking so chips can be linked for larger deployments.

    The company said the new hardware has begun coming online in an Azure U.S. Central data centre in Iowa, with an additional location planned in Arizona. Microsoft described Maia 200 as its most efficient inference system deployed to date, reporting a 30% improvement in performance per dollar compared with its existing inference systems.

    AI inference focus and Azure deployment

    Microsoft said Maia 200 is intended to support AI products and services that rely on high-volume, low-latency model execution, including workloads running in Azure and Microsoft’s own applications. The company said it has designed the chip and the surrounding system as part of an end-to-end infrastructure approach that includes silicon, servers, networking and software for deploying AI models at scale.

    Alongside the chip, Microsoft announced early access to a Maia software development kit for developers and researchers working on model optimization. The company said the tooling is aimed at helping teams compile and tune models for Maia-based systems, and is structured to fit into common AI development workflows used for deploying inference in the cloud.

    Performance claims and model support

    Microsoft said Maia 200 is built to run large language models and advanced reasoning systems, and that it will be used for internal and hosted model deployments in Azure. The company has positioned the chip as a production inference accelerator, distinguishing it from training-focused systems that are typically used to build models before deployment.

    Microsoft has accelerated custom silicon work as demand has grown for compute to serve generative AI applications, where costs and availability of accelerators can affect how quickly services scale. Maia 200 follows Maia 100, which Microsoft introduced in 2023, and represents the company’s latest iteration of its dedicated AI accelerator line for datacenter inference.

    Related Posts

    Apple Set to Launch Its First Foldable iPhone, the Duo, in October with Advanced Features

    September 16, 2026

    Apple Unveils iPhone 18 Pro Series with Adjustable Aperture, Enhanced AI Features, and Longer Battery Life

    September 10, 2026

    Nvidia Moves Forward with $12.93 Billion Acquisition of Hugging Face, Expected to Complete by 2027

    September 4, 2026

    Austria’s First Military Satellite, BEACONSAT, Set to Launch in 2027, Marking a Major Space Defense Milestone

    August 24, 2026

    New Mexico Court Orders Meta to Pay $567 Million for Youth Safety Violations and Mental Health Initiatives

    August 8, 2026

    Russia and US Agree on Two-Year Timeline for Managed ISS Deorbit and Retirement

    August 6, 2026
    Editor's Pick

    Apple Set to Launch Its First Foldable iPhone, the Duo, in October with Advanced Features

    September 16, 2026

    Bank of England to Review Interest Rate Stability and Gilt Reduction Strategy in September Meeting

    September 15, 2026

    Endangered Australian Sea Lion Succumbs to H5N1 Bird Flu on Kangaroo Island for the First Time

    September 14, 2026

    Austria’s Central Bank Projects Reduced Growth and Inflation Forecasts for 2026 and Beyond

    September 14, 2026

    Gold Approaches One-Week Lows After Sharp 2 Percent Drop as Markets Reassess Rate Expectations

    September 12, 2026

    European stocks close lower following ECB rate hike across bourses

    September 12, 2026

    European Union Reports 35% Drop in Irregular Border Crossings Over Eight Months Period

    September 12, 2026

    India and Russia Set Ambitious Goal to Achieve $100 Billion in Bilateral Trade by 2030 Through Expanded Economic Collaboration

    September 12, 2026
    © 2024 Manchester Examiner | All Rights Reserved
    • Home
    • Contact Us

    Type above and press Enter to search. Press Esc to cancel.