South Korea's Free AI for All Rides on 512 B200 GPUs
South Korea picked SK Telecom, KT, and Kakao consortia to run a free, token-unlimited AI service for all citizens using just 512 Nvidia B200 GPUs.
69 verified stories covering AI infrastructure, product updates and industry developments.
South Korea picked SK Telecom, KT, and Kakao consortia to run a free, token-unlimited AI service for all citizens using just 512 Nvidia B200 GPUs.
Qualcomm says 6G's real shift is AI-native integration, not speed, and expects carriers to move from selling data to selling compute and inference tokens.
Five months after ordering 1 million Nvidia GPUs, AWS has doubled the order to 2 million, even as Amazon's own AI chips hit a $25 billion revenue run rate.
SemiAnalysis founder Dylan Patel projects Anthropic and OpenAI will control 80% of new AI compute by 2028, scaling to 54GW each on $5 trillion in debt.
Alibaba raised HK$80 billion in an hour-long share placement earmarked for AI infrastructure, but shares fell 9% and closed below the offer price.
NVIDIA's early tests show Vera Rubin NVL72 hitting 30x GB300 NVL72's per-megawatt throughput on agentic AI workloads, with independent review from SemiAnalysis still pending.
Meta detailed MetaRoCE, an in-house RDMA transport that skips PFC and still works at 10% packet loss; full spec ships at OCP in October.
Ant Group and SGLang keep quantized weights resident in GPU memory, cutting Ling-2.6-1T restart time from 8.8 minutes to about half a minute.
Google Cloud's updated ScaNN index for AlloyDB now scales past 10 billion vectors at 51ms P95 latency and 95% recall, as AI agents drive retrieval volumes far beyond traditional RAG.
Stripe is buying AI model router OpenRouter in a deal reported at more than $7 billion, folding model selection into payment settlement just three months after a $1.3 billion valuation.
OpenAI, Nvidia and SB Energy have set a 20-year structure for the PORTS-Pike campus. The important constraint is power, financing and construction—not another chip launch.
Tencent's second-quarter report shows AI moving from product demos into capex, cash flow and WeChat-scale deployment tests.
Nvidia is working with Apollo, BlackRock, Blackstone, Brookfield, Goldman Sachs and KKR on compute-financing platforms that could mobilize more than $500 billion for AI infrastructure.
Firebird has launched a NVIDIA DSX AI factory in Armenia, with a roadmap for 70,000 GPUs, 300MW in the country, and a wider 2GW frontier-market infrastructure plan.
Volta and Bitdeer disclosed a Norway AI factory deal that Bloomberg-linked reports identify as Anthropic's latest compute push.
Nvidia and Safe Superintelligence tied Ilya Sutskever’s closed research lab to Vera Rubin compute, turning a quiet superintelligence bet into an infrastructure story.
Alphabet says Gemini has reached 950 million monthly active users while it raises its 2026 capex outlook to $195 billion-$205 billion.
Microsoft is committing billions of dollars to Mistral's European AI infrastructure, turning the French model lab into a compute and sovereign-AI partner.
SenseTime and ADA Space plan four compute satellites this year, with a 2030 target of a thousand-satellite AI compute constellation.
Approaching.ai raises 1B yuan for a token factory. A concise localization of the verified facts and the industry signal behind the story.
Ollama raises $65 million for local models is reframed for global technology readers, with the key numbers, product claims and open questions kept intact.
Nvidia tests revenue sharing for AI cloud GPUs is reframed for global technology readers, preserving the key numbers, claims and open questions.
Subsea bandwidth, GPU clusters and AI data-center corridors.
AI power demand, grid queues and private gas-fired capacity.
Neocloud capacity, power access and hyperscale AI contracts.
Omen AI raised a $31 million Series A for inline sensors that track coolant contamination and wear as AI data centers move deeper into liquid cooling.
Oxmiq Labs wants to license OxCore, a GPU-style IP block that combines GPU, CPU and tensor engine pieces for custom AI silicon.
Aramco Ventures led the Series C, valuing the Nvidia-backed open-model infrastructure provider at $8.3 billion.
Valar Atomics powered an Nvidia Blackwell chip from its Ward250 reactor in Utah, framing the test as a first for nuclear-fed AI infrastructure.
Meta is exploring a business that rents out excess AI compute, turning a costly infrastructure buildout into a possible revenue line.
South Korea puts $518B behind four AI chip clusters
Baseten’s new Series F values the inference infrastructure company at up to $13 billion, reflecting investor demand for cheaper, faster and more reliable model serving.
The U.S. energy regulator ordered six major grid operators to justify or rewrite large-load interconnection rules as AI data centers strain power queues.
Anthropic reported elevated errors across Claude.ai, API, Claude Code and government services, with Opus, Sonnet and Haiku models recovering one by one.
Goldman Sachs estimates global AI infrastructure spending from 2026 to 2031 could reach $7.6 trillion, with compute, data centers and power as the major buckets.
The reported 10-gigawatt lease would shift OpenAI from owning more infrastructure to long-term capacity commitments backed by partners such as Nvidia.
The interruption showed how dependent users have become on hosted AI assistants and why AI infrastructure now faces cloud-service reliability expectations.
The data-center operator has moved from an $11 billion take-private deal to a potential $50 billion-plus valuation as AI demand turns power and land into scarce assets.
Google says it will replenish more water than its data centers consume by 2030, pairing public water reporting with projects across 97 watersheds.
SoftBank wants to build 5 GW of AI data center capacity in northern France, tying Masayoshi Son’s infrastructure wager to Europe’s nuclear power advantage.
DriveNets raised a $410 million Series D at an $8.5 billion valuation, with AMD joining a bet on Ethernet fabrics for large AI deployments.
DriveNets is valued at $8.5 billion as AMD backs its Ethernet fabric bet against Nvidia-style closed AI clusters.
Alphabet plans to raise $80 billion through stock sales to fund AI infrastructure, with Berkshire Hathaway contributing $10 billion.
OpenRouter, an AI model routing platform, has raised $113 million in Series B funding led by CapitalG, valuing the company at $1.3 billion.
Five companies including Broadcom and Meta jointly invest $125 million over five years to establish a semiconductor R&D center at UCLA, focusing on next-gen AI chips and workforce training.
SendCutSend, a metal fabrication startup, raised $110 million led by Paradigm and Sequoia, hitting a $1 billion valuation as AI data center demand drives orders.
Armada is selling deployable AI infrastructure for ships, energy sites and remote industrial operations.
Dell reports $64 billion in AI-related orders and $43 billion backlog, adding 1,000 new AI Factory customers in the latest quarter.
European industrial electricity costs are double those in the US and 50% higher than in China and India, threatening AI data center expansion.
The UK liquid-cooling company wants to scale immersion systems as next-generation AI racks push far beyond what air cooling can handle.
London startup Fractile raised a $220 million Series B led by Accel, Factorial Funds and Founders Fund to attack the inference bottleneck with a memory-centric chip design.
Cisco raised its full-year AI infrastructure order target to $9 billion while announcing a 5% workforce reduction and restructuring charge.
Google’s Project Suncatcher is reportedly moving from research toward launch talks with SpaceX as AI data centers run into power and siting constraints on Earth.
Lightspeed backed Judgment Labs twice in six months as the startup builds a feedback layer for agent failures.
ByteDance has raised its 2026 AI infrastructure plan by 25%, with memory inflation and a larger domestic-chip allocation reshaping the bill.
NVIDIA paired a five-year IREN share option with a $3.4 billion GPU cloud commitment, signaling that AI infrastructure now starts with power, land and data centers.
Nvidia is tying itself to Corning as AI data centers move from copper cabling to optical links, funding new U.S. plants and taking warrant-based upside.
DeepInfra's Series B signals that independent inference infrastructure for open models and agent workloads is becoming its own layer.
Peter Thiel led a $140 million round for Panthalassa, which wants wave-powered offshore nodes to run AI workloads from 2027.
Greg Brockman told a court OpenAI expects to spend $50 billion on compute this year, exposing the scale and circular financing behind its AGI push.
Thinking Machines Lab, founded by former OpenAI CTO Mira Murati, has secured a multibillion-dollar deal with Google Cloud for GB300 systems and reached a $12 billion valuation just 14 months after launch.
A major global outage on April 20 took ChatGPT offline for roughly three hours, affecting Codex, the API, and Projects. OpenAI provided minimal information and no root cause analysis, raising concerns about AI infrastructure reliability for businesses.
AI infrastructure startup Fluidstack, which builds custom data centers for Anthropic, Meta, and others, has more than doubled its valuation to $18 billion within four months, driven by surging demand for specialized AI compute.
Amazon unveiled the Nova 2 family of four models, alongside Nova Forge for custom enterprise training and Nova Act for browser automation, bundling AI models, training, and automation into a single ecosystem.
Japan is deploying physical AI at scale, driven by a severe labor shortage rather than a love of technology. The country faces a shrinking workforce, with robots now working in warehouses, construction sites, farms, and convenience stores.
Anthropic's annualized revenue surpassed $30 billion in March 2026, and the company signed a massive compute contract with Google and Broadcom for 3.5GW of TPU capacity starting in 2027.
Eli Lilly has launched LillyPod, an in-house AI supercomputing cluster built with 1,016 NVIDIA Blackwell Ultra GPUs, and announced a joint $1 billion AI co-innovation lab with NVIDIA in San Francisco.
Global data center electricity consumption is projected to exceed 1,000 TWh by the end of 2026, driven almost entirely by AI workloads, raising concerns about grid strain, rising residential bills, and the limits of energy supply.
Nvidia's B200 and GB200 NVL72 racks are sold out until mid-2026, with cloud giants holding over 3.6 million units in backlog. The article examines supply chain bottlenecks, the upcoming Blackwell Ultra (GB300), and what this means for AI companies.