Close Menu

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    From Bangkok to the World: Central Park Unveils “Bangkok’s World Stage” as Its Next Global Ambition

    September 10, 2026

    Aircom Appoints Khurram Chaudhry as General Manager

    September 10, 2026

    Dr Hanan H. Balkhy Launches Candidacy for WHO Director-General, Pledges “A WHO the World Can Count On”

    September 10, 2026
    Facebook X (Twitter) Instagram
    MEA News HubMEA News Hub
    • Automotive
    • Business
    • Entertainment
    • Health
    • Lifestyle
    • Luxury
    • News
    • Sports
    • Technology
    • Travel
    MEA News HubMEA News Hub
    Home » Microsoft launches Maia 200 chip for Azure AI inference
    Technology

    Microsoft launches Maia 200 chip for Azure AI inference

    January 28, 2026
    Facebook Twitter Pinterest LinkedIn Reddit WhatsApp Email

    MENA Newswire, SAN FRANCISCO: Microsoft on Jan. 26 introduced Maia 200, the second generation of its in-house artificial intelligence accelerator, built to run AI models in production across Azure data centres. The company said Maia 200 is designed for inference, the stage where trained models generate responses to live requests, and will be used to support a range of Microsoft AI services.

    Microsoft launches Maia 200 chip for Azure AI inference
    Microsoft Maia 200 targets faster AI inference in Azure data centers using custom silicon. (AI-generated image)

    Maia 200 is manufactured on TSMC’s 3-nanometer process and includes more than 140 billion transistors, Microsoft said. The chip pairs compute with a new memory system that includes 216 gigabytes of HBM3e high-bandwidth memory and about 272 megabytes of on-chip SRAM, aimed at sustaining large-scale token generation and other inference-heavy workloads.

    Microsoft said Maia 200 delivers more than 10 petaflops of performance at 4-bit precision and about 5 petaflops at 8-bit precision, formats commonly used to run modern generative AI efficiently. The company also said the system is designed around a 750-watt power envelope and is built with scalable networking so chips can be linked for larger deployments.

    The company said the new hardware has begun coming online in an Azure U.S. Central data centre in Iowa, with an additional location planned in Arizona. Microsoft described Maia 200 as its most efficient inference system deployed to date, reporting a 30% improvement in performance per dollar compared with its existing inference systems.

    AI inference focus and Azure deployment

    Microsoft said Maia 200 is intended to support AI products and services that rely on high-volume, low-latency model execution, including workloads running in Azure and Microsoft’s own applications. The company said it has designed the chip and the surrounding system as part of an end-to-end infrastructure approach that includes silicon, servers, networking and software for deploying AI models at scale.

    Alongside the chip, Microsoft announced early access to a Maia software development kit for developers and researchers working on model optimization. The company said the tooling is aimed at helping teams compile and tune models for Maia-based systems, and is structured to fit into common AI development workflows used for deploying inference in the cloud.

    Performance claims and model support

    Microsoft said Maia 200 is built to run large language models and advanced reasoning systems, and that it will be used for internal and hosted model deployments in Azure. The company has positioned the chip as a production inference accelerator, distinguishing it from training-focused systems that are typically used to build models before deployment.

    Microsoft has accelerated custom silicon work as demand has grown for compute to serve generative AI applications, where costs and availability of accelerators can affect how quickly services scale. Maia 200 follows Maia 100, which Microsoft introduced in 2023, and represents the company’s latest iteration of its dedicated AI accelerator line for datacenter inference.

    Share. Facebook Twitter Pinterest LinkedIn WhatsApp Reddit Email

    Related Posts

    US battery push confronts China’s supply chain dominance

    September 9, 2026

    Japan deploys AI to detect investment fraud earlier

    September 3, 2026

    China digital industry revenue reaches 20.71 trillion yuan

    September 1, 2026

    Microsoft and HUMAIN Expand Strategic Collaboration at LEAP 2026 with New Enterprise AI Offering and AI PC

    August 31, 2026

    HUMAIN and KORA Partner to Build Operating System at LEAP 2026

    August 31, 2026

    TestMu Conference 2026: TestMu AI’s Flagship Event Marks a New Milestone in the World of Agentic Engineering and Quality

    August 27, 2026
    Latest News
    Business

    Egypt reserves set new record at $57.2 billion in August

    September 9, 2026

    Egypt’s net international reserves rose to a record $57.2145 billion at the end of August 2026. Reserves stood at $56.2939 billion at the end of July. That means the total increased by about $920.6 million during August, or roughly 1.6%. The latest reading extends a series of monthly gains in 2026 and takes Egypt’s official foreign reserves above $57 billion for the first time.

    Ballon d’Or 2026 nominees confirmed for London ceremony

    September 9, 2026

    Panama Canal weighs tighter transit cap amid drought

    September 9, 2026

    US battery push confronts China’s supply chain dominance

    September 9, 2026

    Etihad Airways adds Saudi Red Sea route from Abu Dhabi

    September 9, 2026

    Volkswagen signs agreement to turn car plant into defense hub

    September 9, 2026

    Anak Krakatau ash hits nearly 3,000 flights in Indonesia

    September 8, 2026

    Air Arabia boosts Sharjah Bangkok routes to 28 weekly flights

    September 5, 2026
    © 2026 MEA News Hub | All Rights Reserved
    • Home
    • Contact Us

    Type above and press Enter to search. Press Esc to cancel.