GPU와 AI 전용 칩, 데이터센터, 전력과 학습 인프라에 관한 소식. GPUs and AI accelerators, data centers, power, and training infrastructure.
-
Data centers may face temporary power cuts to prevent blackouts on largest US grid
미국 최대 전력망 운영사인 PJM 인터커넥션이 정전을 막기 위해 2027년 6월부터 50메가와트 이상 대형 데이터센터에 대한 일시적 전력 공급 제한을 시행하기로 했다. 데이터센터 전력 수요가 2035년까지 지금의 4배로 늘어날 것으로 예상되는 가운데, 신규 발전 용량 확보를 위한 최근 경매도 수요를 채우지 못했다. PJM Interconnection, the largest power grid in the US serving 67 million customers, will begin temporarily curtailing power to data centers of 50 megawatts or larger starting June 2027 to prevent blackouts. Data center electricity demand is projected to quadruple by 2035, and a recent auction for new generating capacity fell short of meeting demand.
-
Satya Nadella says companies that trust one AI for everything may not survive
사티아 나델라 마이크로소프트 CEO가 7월 27일 CNN 인터뷰에서 '하나의 AI 모델에 모든 것을 맡기는 기업은 살아남지 못할 것'이라고 경고했다. 그는 '이런 통제력이 없는 기업은 사실상 자신의 사고 자체를 외주화한 것이므로 기업으로 남지 못할 것'이라며, 기업이 모델과 자사 프롬프트를 분리하는 AI 게이트웨이를 구축하고 모든 상호작용의 메타데이터를 축적해 자체 모델을 학습시킬 수 있어야 한다고 주장했다. In a July 27 CNN interview, Microsoft CEO Satya Nadella warned that companies relying entirely on one external AI provider 'will not remain a firm' because they've effectively outsourced their own thinking. He argued businesses need an AI gateway that separates their prompts from the underlying model, plus retained interaction metadata they could eventually use to train their own models.
-
Ilya Sutskever's Safe Superintelligence partners with Nvidia to scale its AI research
일리야 수츠케버가 이끄는 세이프 슈퍼인텔리전스(SSI)가 2년간의 스텔스 모드를 끝내고 엔비디아와 장기 파트너십을 맺었다고 발표했다. 블룸버그 보도로는 이번 투자 규모가 50억 달러에 달하며, SSI는 이를 통해 엔비디아의 차세대 GPU 플랫폼 '베라 루빈'에 접근해 컴퓨팅 자원을 한 자릿수 단위(order of magnitude)만큼 늘릴 수 있게 된다. Ilya Sutskever's Safe Superintelligence (SSI) ended two years in stealth to announce a long-term strategic partnership with Nvidia, which Bloomberg reports involves a roughly $5 billion investment; the deal gives SSI access to Nvidia's next-generation Vera Rubin GPU platform, expected to boost SSI's compute 'by an order of magnitude.'
-
AMD and Cerebras Launch AI Inference Solution
AMD와 세레브라스(Cerebras)가 7월 23일 'Advancing AI 2026' 행사에서 AMD의 랙스케일 시스템 '헬리오스(Helios)'와 세레브라스의 웨이퍼스케일 엔진(WSE)을 결합한 분리형(disaggregated) AI 추론 솔루션을 공동 발표했다. AMD 인스팅트(Instinct) GPU의 높은 처리량과 세레브라스 웨이퍼스케일 엔진의 초고속 토큰 생성 속도를 하나의 워크플로로 묶어, 와트당 초당 토큰 처리량(T/s/W)을 최대 5배까지 끌어올릴 수 있다고 밝혔다. 이 통합 솔루션은 세레브라스 클라우드를 통해 2026년 하반기 중 첫 출시될 예정이다. AMD and Cerebras jointly announced a disaggregated AI inference solution at "Advancing AI 2026" on July 23, combining AMD's rackscale Helios systems with Cerebras's Wafer-Scale Engine. By pairing AMD Instinct GPUs' high throughput with the Wafer-Scale Engine's ultra-fast token generation in a single workflow, the companies say they can deliver up to 5x higher tokens per second per watt. The combined solution is set to debut through Cerebras Cloud in the second half of 2026.
-
Google reportedly working on ultra-efficient AI chip for Gemini
구글이 제미나이 전용으로 설계된 차세대 서버 칩 '프로즌 v2'를 개발 중이라고 알려졌다. 범용 AI 가속기가 아니라 제미나이 모델 구조 일부를 아예 반도체 회로에 새겨 넣는 방식으로, 기존 텐서 처리 장치(TPU) 대비 전력당 토큰 생성량을 6~10배까지 끌어올리는 것이 목표다. 출시 목표 시점은 2028년으로 전해졌고, 소식이 알려지자 알파벳 주가는 약 3% 올랐다. Google is reportedly building a next-generation server chip codenamed "Frozen v2," designed exclusively to run Gemini. Instead of a general-purpose AI accelerator, parts of Gemini's model architecture would be hardwired directly into the chip, targeting 6-10x more tokens generated per unit of power than Google's current TPUs. A 2028 launch is being targeted, and Alphabet's stock rose about 3% on the news.
-
Google just had its first negative cash flow quarter due to massive AI spending
구글의 2분기 매출은 1198억 달러로 시장 예상을 웃돌았지만, AI 인프라에 쏟아붓는 돈이 그보다 더 빠르게 늘면서 회사는 상장 이후 처음으로 마이너스 잉여현금흐름(-58억 달러)을 기록했다. 구글은 올해 자본적 지출 전망치를 기존 1800억~1900억 달러에서 최대 2050억 달러로 올려 잡았는데, 이는 2022년 220억 달러의 약 6배에 달하는 규모다. 실적 발표 후 구글 주가는 하루 만에 약 4.5% 떨어졌다. Google's Q2 revenue hit $119.8 billion, beating expectations, but AI infrastructure spending grew even faster, pushing the company into negative free cash flow (-$5.8 billion) for the first time since going public. Google raised its 2026 capital expenditure guidance to as much as $205 billion, up from a prior $180-190 billion estimate and roughly six times what it spent back in 2022 ($22 billion). Its stock fell about 4.5% the day after the earnings report.
-
Google justifies its massive AI spending with a booming cloud business
Google Cloud 매출이 전년 동기 대비 82% 성장한 248억 달러를 기록해 월가 예상(224.6억 달러)을 상회했으며, 미완료 계약 잔고는 5,140억 달러에 달한다. 전사 순이익은 전년 281억 달러에서 1,121억 달러로 급증했고 전체 매출은 24% 성장한 1,198억 달러를 기록했다. 회사는 이 성장의 핵심 동력이 '기업 AI 솔루션 및 인프라 채택'이라고 설명했으며, Gemini 앱 사용자는 9억 5천만 명으로 늘었다. 연간 자본지출은 1,800억~1,900억 달러 규모다. Google Cloud revenue grew 82% year-over-year to $24.8 billion, beating Wall Street's $22.46 billion estimate, with backlog reaching $514 billion. Companywide net income surged to $112.1 billion from $28.1 billion a year earlier, and total revenue rose 24% to $119.8 billion. The company credited the growth to enterprise adoption of AI solutions and infrastructure, with Gemini app users climbing to 950 million. Annual capital expenditure is running at $180-190 billion.