Subscribe to our newsletter
Get the latest updates delivered to your inbox every day, and stay up-to-date for free 🧠📈
Share
Linkedin
Twitter
Facebook
Whatsapp
Copy Link
Intel announced fresh job cuts in its Data Center and AI Group, the division responsible for server CPUs and AI chips. The layoffs come despite the unit generating $5.05 billion in Q1 2026 revenue, up 22% year-over-year. CEO Lip-Bu Tan continues his turnaround plan as Intel shares have surged over 317% in the past year.
Chinese researchers at Zhejiang University have revealed a cyber-physical vulnerability called Bit2Watt that allows malicious cloud tenants to weaponize GPU workloads and potentially cause widespread blackouts. The theoretical attack exploits the power consumption patterns of AI datacenters, demonstrating how approximately 1,000 GPUs could create cascading failures affecting over 80% of large-scale power systems.
Alphabet reported Google Cloud revenue jumped 82% to $24.8 billion, driven by enterprise AI adoption and infrastructure demand. The results exceeded analyst expectations despite concerns over capital expenditures rising to as much as $205 billion for the year. CEO Sundar Pichai defended the spending, citing strong demand indicators and long-term deals that justify the company's AI infrastructure investments.
Welsh chipmaker IQE raised its 2026 revenue growth forecast to above 30%, up from 20%, betting on surging AI data-centre demand for its indium phosphide materials. The upgrade comes despite an 18% revenue decline to £97.3m in 2025, with first-half 2026 trading beating expectations at £64m.
A leak suggests Qualcomm's upcoming Snapdragon 4 Gen 6 will bring dedicated AI processing to budget Android phones for the first time. The chip features a Hexagon DSP and separate machine learning engine, potentially enabling live translation, smarter camera tools, and voice features on entry-level devices costing under $400.
Chinese AI lab Z.AI has completed a 1-gigawatt data centre running exclusively on domestically produced chips, marking a major milestone in China's push for AI self-sufficiency. The facility operates multiple clusters with over 10,000 chips each to train GLM models, demonstrating that Chinese labs can scale frontier AI development despite US export controls blocking access to Nvidia's most advanced processors.
China's Ministry of Commerce is consulting leading tech firms including Alibaba, ByteDance, and Huawei about implementing export controls on advanced AI models, training data, and semiconductor technologies. The proposed restrictions could ban Chinese chip designers from using TSMC and limit downloads of open-weight models, marking a significant shift in how Beijing protects its technological advancements.
San Francisco-based AI infrastructure startup Infinity has raised $15 million at a $100 million valuation to break Nvidia's software lock-in. Its AI agent Ignition automates the creation of low-level code needed to run AI models on any chip architecture, reducing what typically takes months or years to just hours or days. The company already generates millions in annual revenue from chip partnerships.
Google is building a specialized AI chip called Frozen v2 that embeds Gemini's architecture directly into hardware. The chip could deliver 6-10 times better efficiency than current TPUs, addressing severe compute shortages that forced Google Cloud to turn away customers. Deployment targets 2028, but the approach trades flexibility for performance.
Alphabet is developing Frozen v2, a server chip that embeds its Gemini AI model directly into hardware to achieve 6-10 times more tokens per unit of power than current TPUs. The move signals a strategic shift from spending more on AI infrastructure to making AI operations dramatically cheaper, with deployment targeted for 2028.
Samsung is testing a new AI chip called Gaia with laptop makers HP and Lenovo, targeting mass production in 2027. Built by the same team behind Exynos phone chips, Gaia's memory-centric design handles on-device AI tasks efficiently. The architecture could eventually address the persistent thermal and performance issues that have plagued Exynos-powered Galaxy phones for years.
AMD launched Helios, its first rack-scale AI platform designed to compete with Nvidia's dominance in data center computing. The system combines 72 Instinct MI455X GPUs and has already secured major customers including Microsoft, OpenAI, Meta, and Anthropic. With superior performance metrics and gigawatt-scale deployments planned, AMD positions itself as a serious Nvidia competitor in the AI accelerator market.
Uday Ruddaraju has been promoted to Chief Technology Officer of Compute at OpenAI, a year after joining from Elon Musk's xAI. The Indian-origin tech leader will oversee the company's mission to build the world's largest compute footprint, supporting frontier models like GPT-5.6. His team focuses on scaling compute infrastructure across distributed systems, hardware, and data centers as the AI race intensifies.
Etched, founded by three Harvard dropouts in 2022, has closed a $300 million Series C funding round at a $10.3 billion valuation—doubling its worth in just seven months. Led by Sequoia with participation from SK Hynix and Andreessen Horowitz, the AI chip startup is building specialized hardware for inference to challenge Nvidia's dominance with custom chips designed for low-voltage operation and cluster-scale memory.
General Compute landed $400 million in debt financing from Upper90 to build an AI inference cloud using SambaNova and AMD chips instead of Nvidia GPUs. The deal marks the first time inference-specific chips have been used as loan collateral, signaling a shift toward cost-efficient alternatives as AI infrastructure costs come under scrutiny.
Don’t drown in AI news. We cut through the noise - filtering, ranking and summarizing the most important AI news, breakthroughs and research daily. Follow topics that matter to you and stay ahead.