GPUs and AI accelerators, data centers, power, and training infrastructure.
-
Data centers may face temporary power cuts to prevent blackouts on largest US grid
PJM Interconnection, the largest power grid in the US serving 67 million customers, will begin temporarily curtailing power to data centers of 50 megawatts or larger starting June 2027 to prevent blackouts. Data center electricity demand is projected to quadruple by 2035, and a recent auction for new generating capacity fell short of meeting demand.
-
Satya Nadella says companies that trust one AI for everything may not survive
In a July 27 CNN interview, Microsoft CEO Satya Nadella warned that companies relying entirely on one external AI provider 'will not remain a firm' because they've effectively outsourced their own thinking. He argued businesses need an AI gateway that separates their prompts from the underlying model, plus retained interaction metadata they could eventually use to train their own models.
-
Ilya Sutskever's Safe Superintelligence partners with Nvidia to scale its AI research
Ilya Sutskever's Safe Superintelligence (SSI) ended two years in stealth to announce a long-term strategic partnership with Nvidia, which Bloomberg reports involves a roughly $5 billion investment; the deal gives SSI access to Nvidia's next-generation Vera Rubin GPU platform, expected to boost SSI's compute 'by an order of magnitude.'
-
AMD and Cerebras Launch AI Inference Solution
AMD and Cerebras jointly announced a disaggregated AI inference solution at "Advancing AI 2026" on July 23, combining AMD's rackscale Helios systems with Cerebras's Wafer-Scale Engine. By pairing AMD Instinct GPUs' high throughput with the Wafer-Scale Engine's ultra-fast token generation in a single workflow, the companies say they can deliver up to 5x higher tokens per second per watt. The combined solution is set to debut through Cerebras Cloud in the second half of 2026.
-
Google reportedly working on ultra-efficient AI chip for Gemini
Google is reportedly building a next-generation server chip codenamed "Frozen v2," designed exclusively to run Gemini. Instead of a general-purpose AI accelerator, parts of Gemini's model architecture would be hardwired directly into the chip, targeting 6-10x more tokens generated per unit of power than Google's current TPUs. A 2028 launch is being targeted, and Alphabet's stock rose about 3% on the news.
-
Google just had its first negative cash flow quarter due to massive AI spending
Google's Q2 revenue hit $119.8 billion, beating expectations, but AI infrastructure spending grew even faster, pushing the company into negative free cash flow (-$5.8 billion) for the first time since going public. Google raised its 2026 capital expenditure guidance to as much as $205 billion, up from a prior $180-190 billion estimate and roughly six times what it spent back in 2022 ($22 billion). Its stock fell about 4.5% the day after the earnings report.
-
Google justifies its massive AI spending with a booming cloud business
Google Cloud revenue grew 82% year-over-year to $24.8 billion, beating Wall Street's $22.46 billion estimate, with backlog reaching $514 billion. Companywide net income surged to $112.1 billion from $28.1 billion a year earlier, and total revenue rose 24% to $119.8 billion. The company credited the growth to enterprise adoption of AI solutions and infrastructure, with Gemini app users climbing to 950 million. Annual capital expenditure is running at $180-190 billion.