Nvidia and AWS Deepen Ties to Speed AI Inference and Vector Search

New Products

Summary

AWS and Nvidia have expanded their partnership to make AI inference and vector search cheaper and easier to run in production. AWS introduced the Blackwell-based EC2 G7 instance for mid-tier inference, graphics, spatial computing, and analytics workloads. AWS also made Nvidia cuVS the default for vector indexing in next-generation OpenSearch Serverless, which strengthens retrieval-augmented generation and semantic search use cases. The moves push Blackwell-class infrastructure deeper into everyday enterprise AI operations rather than only large training clusters. The announcement also signals possible future region expansion to Frankfurt and Tokyo.

Classifications

industries
Entertainment
applications
Anti Piracy

AskAI Classifications

Labels
AI Software Developer Tools MLOps

Linked Companies