
Mit új a mesterséges intelligencia infrastruktúrában és irányításban idén
SZERZŐ
Alex Barrett
FORRÁS
Cloud Blog
DATE
READ
4 perc olvasás
A Google AI ökoszisztémája magában foglal modelokat, eszközöket és infrastruktúrát, amelyek havonta frissülnek. Az utóbbi kiadások közé tartoznak a Managed Lustre (GA, akár 8 PB), a C4N virtuális gépek (400 Gbps, 25 …
At Google, AI is a comprehensive endeavor. We offer leading AI models like Gemini and Nano Banana. We integrate AI into your daily tools, such as Gmail, BigQuery, AlloyDB, Google Cloud Code, and Google Cloud Assist. We provide software frameworks like Gemini Enterprise Agent Platform, JAX, and MaxTest to help you build with AI. We also co-design the powerful infrastructure platform that supports all of this, including a wide range of standard compute, accelerators like TPUs and GPUs, optimized networks and storage, as well as orchestration software like GKE and Cluster Director. We package all of this into supercomputing platforms like AI Hypercomputer to power the industry-wide transformation to the AI and agentic future. This is critical in today’s agentic era, where AI is evolving from answering questions to reasoning and taking action. Companies that want to lead in this next phase of AI need computing infrastructure that’s designed and optimized for these new requirements, so they can innovate faster, deliver compelling user and customer experiences, and optimize for cost and energy efficiency — all at massive scale. To support this, we are making AI infrastructure and orchestration news at a furious pace. In this blog, we provide a monthly snapshot of the recent launches and milestones that you need to know about to keep up-to-date, deep dives on architecture and performance tuning, and discussions of specialized use cases, always with pointers to where you can learn more. Keep an eye out for updates to this blog every month. Google Cloud Managed Lustre is now generally available, offering throughput ranging from 125 MB/s to 1000 MB/s per TiB of capacity. C4N network and storage optimized VMs are now generally available. GKE Dataplane V2 up to 15K Nodes with Network Policies is now generally available. Co-operative time-slicing in llm-d is now available. Protecting sensitive data used with AI is a critical part of advanced and secure cloud infrastructure. Confidential Computing cryptographically protects data in use in hardware-based Trusted Execution Environments (TEEs) with verifiable data integrity, and is now available on the accelerator-optimized G4 machine series, featuring NVIDIA RTX PRO 6000 Blackwell Server Edition GPUs. The new TPU Developer Hub is the place to go for model builders, optimizers, and developers to learn to unlock the full performance of Google Cloud TPUs. Scale your AI workloads with the new OpenTelemetry-Based TPU AI Telemetry Collector Agent. GKE Agent Sandbox is now generally available. Agent Substrate is a new open-source project aimed at continuing to push the limits of agentic infrastructure density. Google AI Edge Portal, a solution for testing and benchmarking on-device machine learning (ML) at scale, now supports benchmarking and debugging on-device LLMs. Cloud Storage Rapid is a new family of high-performance storage offerings for AI workloads. Google Cloud Managed Lustre is now generally available, offering throughput ranging from 125 MB/s to 1000 MB/s per TiB of capacity. C4N network and storage optimized VMs are now generally available. GKE Dataplane V2 up to 15K Nodes with Network Policies is now generally available. Co-operative time-slicing in llm-d is now available. Protecting sensitive data used with AI is a critical part of advanced and secure cloud infrastructure. Confidential Computing cryptographically protects data in use in hardware-based Trusted Execution Environments (TEEs) with verifiable data integrity, and is now available on the accelerator-optimized G4 machine series, featuring NVIDIA RTX PRO 6000 Blackwell Server Edition GPUs. The new TPU Developer Hub is the place to go for model builders, optimizers, and developers to learn to unlock the full performance of Google Cloud TPUs. Scale your AI workloads with the new OpenTelemetry-Based TPU AI Telemetry Collector Agent. GKE Agent Sandbox is now generally available. Agent Substrate is a new open-source project aimed at continuing to push the limits of agentic infrastructure density. Google AI Edge Portal, a solution for testing and benchmarking on-device machine learning (ML) at scale, now supports benchmarking and debugging on-device LLMs. Cloud Storage Rapid is a new family of high-performance storage offerings for AI workloads.