CONNECT WITH US
Aug 11
Microsoft seeks expanded Maia 300 production to tap into demand for cheaper compute
Microsoft plans to unveil its new Maia 300 accelerator chip this fall, with a public announcement possible as early as next month, according to The Information. Although Microsoft has lagged Google and Amazon in scaling its in-house AI chips, it is targeting far higher production volumes for Maia 300 than for Maia 200 in hopes of attracting major clients such as Anthropic.

Nvidia is intensifying its open-source strategy through its upcoming trillion-parameter model in the Nemotron 4 family, which is meant to compete with the world's leading open-source models, according to The Information. While this may put it in the awkward position of vying against its customers, the chipmaker's advocacy for open-source AI could see the company benefit from more chip sales across the board.

At the 2026 OCP APAC Summit, debate over networking architectures in cloud AI data centers sharpened as industry players said the main bottleneck is no longer compute; beyond memory, networking has become a key constraint on how far AI infrastructure can scale.

The use of optical communications in cloud AI data-center interconnect architectures is rising sharply, and Taiwan IC design firms including MediaTek, Realtek Semiconductor, Himax Technologies, and Elan Microelectronics are moving into the market from different angles, with mass production for their product lines all pointing to 2027.
Foxconn will hold its second quarter 2026 earnings call on the afternoon of August 12, as AI racks continue to ship and July consolidated revenue topped NT$900 billion (approx. US$27.9 billion) for the first time. Investors are expected to focus on the shipment pace of AI servers and racks, plus how changes in the underlying transaction model are reshaping the company's operating structure.
Nvidia and Broadcom outlined different paths for scaling AI networks at the OCP APAC Summit in Taipei on Aug. 11, as clusters expand from rack-scale systems to entire data centers and across multiple sites.
Supermicro executives fielded a wave of questions from Wall Street analysts following the company's fourth-quarter and full-year fiscal 2026 earnings report. Much of the discussion centered on the sustainability of its gross margins, the composition of its record order backlog, and how the company is balancing rapid growth with profitability.

SpaceXAI has launched Grok Bot, an AI agent designed to take on workplace tasks and operate digital tools with limited human intervention, marking the company's latest move to push Grok beyond a conventional chatbot.

Amazon is moving ahead with a dedicated gas-fired power plant for a new data center campus in Pecos County, Texas, as part of an effort to speed AI infrastructure deployment outside the state grid. The project, called GW Ranch, is designed to support behind-the-meter self-generation and avoid the interconnection delays that have slowed other power-heavy facilities.

CoreWeave reported second-quarter fiscal 2026 revenue of US$2.6 billion, up 112% year-over-year and 24% sequentially, and raised its full-year outlook as demand for its AI cloud infrastructure continued to outstrip supply, executives said on the company's August 11 earnings call.

IBM and Together AI have announced a collaboration to deploy a dedicated Nvidia-powered inference cluster on IBM Cloud, a move that could reshape how enterprises access open-source AI services worldwide. The planned system is designed to improve performance, lower token costs, and expand global access to scalable AI infrastructure as demand for real-time applications grows.

Lumentum Holdings said it is pulling forward capacity investments across lasers, transceivers, and optical circuit switches (OCS) after fiscal fourth-quarter revenue surged 109% year-over-year to US$1.01 billion, with the company now guiding to hit its long-stated US$1.25 billion quarterly revenue target more than a quarter ahead of schedule.