Back to all posts

#Cloud Run

14 articles filed under this tag

02/Google Cloud
13 min

Google Cloud Blueprint: Reimagining work: How Pythian’s internal AI playbook delivers customer — How Does It Work?

Architectural Thesis: When Pythian rolled out Google Cloud’s Gemini Enterprise across our 500-person company in 27 countries, the goal was simple: use our own company as a proving ground to discover how enterprise AI actually delivers ROI. What we found changed... Real-World Field Use Cases: 1. High-Throughput Enterprise...

02/Google Cloud
13 min

Google Cloud Blueprint: What’s new in AI infrastructure and orchestration in August — How Does It Work?

Architectural Thesis: Welcome back to What’s new in AI infrastructure and orchestration this month , a collection of product updates, how-tos, customer stories, research and other resources about all the AI compute, networks, storage, frameworks, and... Real-World Field Use Cases: 1. High-Throughput Enterprise Workloads...

02/Google Cloud
14 min

Getting started with Mantis, our open-source bug finding-and-fixing harness — How Does It Work in Production?

<div class="block-paragraph_advanced"><p><span style="vertical-align: baseline;">AI models have clearly proven their ability to discover and exploit vulnerabilities without much, if any, human assistance. To help defenders gain the advantage with AI, we built Real-World Field Use Cases: 1. High-Throughput Production...

02/Google Cloud
12 min

Agentic FinOps on Google Cloud: Automating Cloud Spend Accountability at Scale — How Does It Work in Production?

How enterprise engineering teams deploy Agentic FinOps architectures using Google ADK and Gemini on Google Cloud to automate BigQuery billing anomaly detection and cloud spend accountability at scale.

02/Google Cloud
14 min

Cloud CISO Perspectives: Autonomous AI Threat Detection and Enterprise Defense on GCP — How Does It Work?

Welcome to the first Cloud CISO Perspectives for September 2026. Sandra Joyce shares the latest details on Google visibility into how attackers use AI and how enterprise security teams use autonomous defenses to stop them. Real-World Field Use Cases: 1. High-Throughput Enterprise Workloads: Isolating P99 tail-latency and...

02/Google Cloud
13 min

Google Cloud Media Delivery at Scale: Serving Billions of Concurrent IPL Requests — How Does It Work in Production?

For the millions of fervent fans of the Indian Premiere League (IPL), being able to count on a flawless live streaming cricket experience is never up for debate. For Airtel, producing league TV broadcasts with some of the world's most massive concurrent viewer Real-World Field Use Cases: 1. High-Throughput Enterprise...

02/Google Cloud
12 min

Google Cloud Blueprint: Spanner Migrations and Automating Dual-Write with Antigravity CLI — How Does It Work?

Architectural Thesis: When Google's Finance Engineering team needed to modernize their legacy data layer, they chose Spanner , a globally distributed, strongly consistent, multi-model database with high availability capabilities. But migrating to Spanner... Real-World Field Use Cases: 1. High-Throughput Enterprise Workloads...

02/Google Cloud
13 min

Google Cloud Blueprint: Google is a Leader in the 2026 Gartner® Magic Quadrant™ for Enterprise — How Does It Work?

Architectural Thesis: We are excited to share that Gartner has named Google a Leader in its inaugural 2026 Magic Quadrant for Enterprise AI Assistants . In this comprehensive evaluation of top enterprise AI assistant vendors, Gartner placed Google in the... Real-World Field Use Cases: 1. High-Throughput Enterprise Workloads...

02/Google Cloud
14 min

Cloud SQL PostgreSQL 17 + pgvector: Hybrid HNSW Indexing at 10,000 QPS

Combining relational ACID guarantees with sub-10ms semantic vector retrieval over Unix domain sockets in Cloud Run.

02/Google Cloud
16 min

Cloud Storage FUSE Gen 2: Zero-Redeploy Content Lakes for Next.js 15

Eliminating daily Docker rebuilds by mounting GCS buckets directly into Cloud Run with gRPC metadata caching.

02/Google Cloud
16 min

BigQuery Physical Storage Billing: Achieving 8x Compression on Telemetry Tables

Step-by-step FinOps migration blueprint from logical active/long-term bytes to physical compressed storage with time-travel tuning.

02/Google Cloud
14 min

Vertex AI Context Caching: Cutting Enterprise LLM Inference Costs by 75%

Production patterns for caching massive system prompts, RAG corpora, and multi-turn conversation prefixes in Gemini.

02/Google Cloud
14 min

Google Cloud Enterprise AI: Production Multi-Agent Systems with ADK

Deep-dive into enterprise agent governance, Vertex AI grounding, and zero-cold-start Cloud Run serverless deployment.

01/News Flash
6 min

Google Cloud Run Introduces Native GPU Support for Serverless AI Microservices

Why lease a luxury penthouse year-round just to sleep there on weekends? Google Cloud Run now supports NVIDIA L4 GPUs with true scale-to-zero economics, eliminating the costly idle-GPU penalty for AI microservices.

Found this helpful?