his session explores how Canva leverages Karpenter to scale and optimize diverse workloads on Amazon EKS. Learn how Canva manages AI workloads using On-Demand Capacity Reservations (ODCRs) and EC2 Capacity Blocks for ML, while maximizing resource utilization by intelligently co-locating CPU and GPU workloads on GPU nodes. We will dive into NodePool management strategies for efficient scheduling of AI workloads and examine how Canva uses a range of Amazon EC2 instance types to operate a multi-tenant container orchestration platform for all workloads, optimizing for cost-effectiveness and resource efficiency. Ideal for platform engineers and Kubernetes operators looking to optimize their EKS clusters for both AI and general workloads at scale.
What this session is about
Playbook
Editorial commentary · what to actually do about this on Monday
Independent editorial perspective — not an official AWS or speaker statement. Designed for executives evaluating what to brief their teams on next.
Live updates related to this session LIVE
Sourced via Parallel AI Monitor — continuous web watch on 21 topical streams. Updated .
- cloud.google.com Scaling infra for agent workloads
AlloyDB's agentic database architecture
Google Cloud announced AlloyDB’s agentic database architecture, including an agent pool that can scale from zero to thousands of nodes for bursty agent activity and scale back down when demand falls. This directly addresses database capacity management for highly variable agent w
- aigrants.in Scaling infra for agent workloads
Agentic Runtime Framework: Architecture, Tools and Patterns
Google Cloud announced AlloyDB’s agentic database architecture, including an agent pool that can scale from zero to thousands of nodes for bursty agent activity and scale back down when demand falls. This directly addresses database capacity management for highly variable agent w
- cloud.google.com high confidence Scaling infra for agent workloads
Memorystore for Valkey 9.1: 3x QPS Caching
Google Cloud announced general availability of Memorystore for Valkey 9.1, which delivers up to 3× higher queries per second at microsecond latency and is designed for workloads scaling to millions of concurrent users, including AI applications. The release adds dynamic I/O-threa
- cloud.google.com Scaling infra for agent workloads
Announcing PostgreSQL for agents in AlloyDB
Google Cloud announced AlloyDB’s agentic database architecture, including an agent pool that can scale from zero to thousands of nodes for bursty agent activity and scale back down when demand falls. This directly addresses database capacity management for highly variable agent w
- sg.finance.yahoo.com Scaling infra for agent workloads
CoreWeave to Offer NVIDIA Vera, the First CPU Built for AI ...
CoreWeave announced that NVIDIA Vera CPU rack-scale systems, designed for demanding agentic AI workloads, would be available on its cloud. The system places 128 CPUs and 11,264 cores in a rack, a new compute-capacity option relevant to scaling agent-native workloads.
External links matched to this session via topic relevance. The KB does not endorse third-party content; verify before citing.