Share
Linkedin
Twitter
Facebook
Whatsapp
Copy Link
China's Ministry of Commerce is consulting leading tech firms including Alibaba, ByteDance, and Huawei about implementing export controls on advanced AI models, training data, and semiconductor technologies. The proposed restrictions could ban Chinese chip designers from using TSMC and limit downloads of open-weight models, marking a significant shift in how Beijing protects its technological advancements.
San Francisco-based AI infrastructure startup Infinity has raised $15 million at a $100 million valuation to break Nvidia's software lock-in. Its AI agent Ignition automates the creation of low-level code needed to run AI models on any chip architecture, reducing what typically takes months or years to just hours or days. The company already generates millions in annual revenue from chip partnerships.
Google is building a specialized AI chip called Frozen v2 that embeds Gemini's architecture directly into hardware. The chip could deliver 6-10 times better efficiency than current TPUs, addressing severe compute shortages that forced Google Cloud to turn away customers. Deployment targets 2028, but the approach trades flexibility for performance.
Alphabet is developing Frozen v2, a server chip that embeds its Gemini AI model directly into hardware to achieve 6-10 times more tokens per unit of power than current TPUs. The move signals a strategic shift from spending more on AI infrastructure to making AI operations dramatically cheaper, with deployment targeted for 2028.
Samsung is testing a new AI chip called Gaia with laptop makers HP and Lenovo, targeting mass production in 2027. Built by the same team behind Exynos phone chips, Gaia's memory-centric design handles on-device AI tasks efficiently. The architecture could eventually address the persistent thermal and performance issues that have plagued Exynos-powered Galaxy phones for years.
AMD launched Helios, its first rack-scale AI platform designed to compete with Nvidia's dominance in data center computing. The system combines 72 Instinct MI455X GPUs and has already secured major customers including Microsoft, OpenAI, Meta, and Anthropic. With superior performance metrics and gigawatt-scale deployments planned, AMD positions itself as a serious Nvidia competitor in the AI accelerator market.
Uday Ruddaraju has been promoted to Chief Technology Officer of Compute at OpenAI, a year after joining from Elon Musk's xAI. The Indian-origin tech leader will oversee the company's mission to build the world's largest compute footprint, supporting frontier models like GPT-5.6. His team focuses on scaling compute infrastructure across distributed systems, hardware, and data centers as the AI race intensifies.
Etched, founded by three Harvard dropouts in 2022, has closed a $300 million Series C funding round at a $10.3 billion valuation—doubling its worth in just seven months. Led by Sequoia with participation from SK Hynix and Andreessen Horowitz, the AI chip startup is building specialized hardware for inference to challenge Nvidia's dominance with custom chips designed for low-voltage operation and cluster-scale memory.
General Compute landed $400 million in debt financing from Upper90 to build an AI inference cloud using SambaNova and AMD chips instead of Nvidia GPUs. The deal marks the first time inference-specific chips have been used as loan collateral, signaling a shift toward cost-efficient alternatives as AI infrastructure costs come under scrutiny.
Apple briefly surpassed Nvidia on Friday with a market valuation of $4.88 trillion versus Nvidia's $4.86 trillion, reclaiming the top spot for the first time since April last year. The shift signals changing investor sentiment as Apple's lower capital expenditure approach to artificial intelligence gains favor over Nvidia's infrastructure-heavy model, despite the chipmaker's dominance in generative AI hardware.
South Korea's $4 trillion equity market has transformed from a peripheral investment destination into a critical bellwether for global AI sentiment. Fund managers in London, New York, and Tokyo now start their day by checking Korean stocks, as swings in Samsung Electronics and SK Hynix ripple through chip markets worldwide. But the influence comes with extreme volatility driven by leveraged trading.
Nvidia's Vera Rubin platform has reached full production, with CoreWeave reporting 10 times more tokens per watt compared to the previous GB200 NVL72 system. Major customers including OpenAI, Microsoft Azure, Google Cloud, and Meta are deploying the next-generation AI infrastructure, which combines 72 Rubin GPUs with custom Vera CPUs to handle agentic AI workloads more efficiently.
Chief Minister Yogi Adityanath unveiled plans to transform Uttar Pradesh into a global deep tech manufacturing hub with PRAGATI, India's first integrated robotics and advanced manufacturing cluster spanning 75 acres in Noida. The initiative targets over one lakh jobs and Rs 2,000 crore in value addition over five years while reducing import dependence.
China-based Montage Technology forecasts first-half net profit growth of up to 81% driven by strong AI demand for memory interface chips and interconnect chips. But the semiconductor company faces regulatory scrutiny after Korean prosecutors raided its Seoul office over suspected price-fixing, sending shares plummeting over 20% and prompting a share buyback plan to restore investor confidence.
Intel is expanding its long-running partnership with Google Cloud by deploying Gemini Enterprise across its global workforce to accelerate chip design and automate workflows. The chipmaker will use AI agents for engineering, supply chain, and marketing operations while leveraging Google Cloud's C4 and N4 instances to run multiple high-performance computing simulations concurrently and dramatically speed up semiconductor development timelines.
Don’t drown in AI news. We cut through the noise - filtering, ranking and summarizing the most important AI news, breakthroughs and research daily. Follow topics that matter to you and stay ahead.