Cloud Computing

Build with more flexibility: New open models arrive in the Vertex AI Model Garden

In our ongoing effort to provide businesses with the flexibility and choice needed to build innovative AI applications, we are expanding the catalog of open models available as Model-as-a-Service (MaaS) offerings in Vertex AI Model Garden. Following the addition of Llama 4 models earlier this year, we are announcing DeepSeek R1 is available for everyone

Build with more flexibility: New open models arrive in the Vertex AI Model Garden Read More »

How Renault Group is using Google’s software-defined vehicle industry solution

It’s funny to think of Renault Group, the massive European car manufacturer, as a software company, but in many ways, it is. Renault Group subsidiary Ampere Software Technology is dedicated to developing and integrating advanced software solutions for intelligent electric vehicles, aiming to create software-defined vehicles (SDVs) with enhanced customer experiences and new services.  Ampere

How Renault Group is using Google’s software-defined vehicle industry solution Read More »

How to integrate your Cloud SQL for MySQL database with Vertex AI & vector search

Search is a critical component of many modern applications – whether searching for products in an online storefront, finding solutions to your customers’ support cases, or building the perfect playlist. But traditional keyword searches often miss the deeper meaning of data. Vector embeddings, however, capture the complexities of your data, enabling highly accurate and powerful

How to integrate your Cloud SQL for MySQL database with Vertex AI & vector search Read More »

Implementing High-Performance LLM Serving on GKE: An Inference Gateway Walkthrough

The excitement around open Large Language Models like Gemma, Llama, Mistral, and Qwen is evident, but developers quickly hit a wall. How do you deploy them effectively at scale?  Traditional load balancing algorithms fall short, as they fail to account for GPU/TPU load status, leading to inefficient routing for computationally intensive AI inference with its

Implementing High-Performance LLM Serving on GKE: An Inference Gateway Walkthrough Read More »

How to enable real time semantic search and RAG applications with Dataflow ML

Embeddings are a cornerstone of modern semantic search and Retrieval Augmented Generation (RAG) applications. In short, they enable applications to understand and interact with information on a deeper, conceptual level. In this post, we’ll show you how to create and retrieve embeddings with a few lines of Dataflow ML code to enable both of these

How to enable real time semantic search and RAG applications with Dataflow ML Read More »

Engineering Deutsche Telekom’s sovereign data platform

Imagine transforming a sprawling, 20-year-old telecommunications data ecosystem, laden with sensitive customer information and bound by stringent European regulations, into a nimble, cloud-native powerhouse. That’s precisely the challenge Deutsche Telekom tackled head-on, explains Ashutosh Mishra. By using Google Cloud’s Sovereign Cloud offerings, they’ve built a groundbreaking “One Data Ecosystem.” When we decided to modernize our

Engineering Deutsche Telekom’s sovereign data platform Read More »