SendTech Times
News
DEPLOYMENT WATCH:

Nvidia And AWS Add Blackwell G7 GPUs To Production AI Stack

Newsroom brief

AWS is adding EC2 G7 instances with Nvidia RTX PRO 4500 Blackwell GPUs, cuVS-backed OpenSearch vector indexing and GB300 Exemplar Cloud status for AI training workloads.

Verified against source materialEdited by SendTech Times Cloud & Infrastructure DeskSource: NVIDIA Blog
Nvidia And AWS Add Blackwell G7 GPUs To Production AI Stack
Image source: NVIDIA

AWS Adds Blackwell G7 Instances

Nvidia and Amazon Web Services are expanding the AWS AI infrastructure stack by putting Nvidia RTX PRO 4500 Blackwell Server Edition GPUs inside the EC2 G7 instance family.

The launch targets production workloads that need inference, graphics, spatial computing and GPU-accelerated analytics without customers managing their own GPU platform.

The hardware claim is specific.

At the largest configuration, a G7 instance can carry eight GPUs, 256GB of combined GPU memory, EFA networking at 700 Gbps and local NVMe SSD storage reaching 7.6TB.

AWS is offering one-, two-, four- and eight-GPU configurations, with bare metal coming soon.

Nvidia says the instances deliver up to 4.6x AI inference performance and up to 2.1x graphics performance compared with G6 instances.

The same platform is also positioned for Amazon EMR analytics workloads using the Nvidia cuDF library for Apache Spark.

Vector Search Moves Into OpenSearch

The update also changes the retrieval layer for AI applications.

Amazon OpenSearch Serverless now sets Nvidia cuVS GPU acceleration as the default path for vector indexing in vector collections.

For teams building retrieval-augmented generation, semantic search, recommendation systems and agentic AI applications, the managed OpenSearch path changes the deployment work.

Instead of treating GPU vector search as a separate optimization project, AWS is making it part of the managed OpenSearch Serverless path.

Nvidia says the customer impact is vector indexing that can run up to 10x faster while costing a quarter as much as CPU-only builds.

It also says billion-scale vector databases can be built in under an hour.

Those are vendor performance claims, but they identify the operating burden AWS is trying to reduce: moving raw enterprise data into searchable AI retrieval systems without running separate infrastructure.

The managed-service angle is as important as the speed claim.

Enterprises building AI retrieval systems often need vector search, serverless scaling and idle-time cost control in the same workflow.

AWS and Nvidia are packaging those pieces inside OpenSearch rather than asking each team to build a separate GPU indexing pipeline.

G7 also reaches more than one buyer group.

AI teams can use the instances for lower-latency inference, media teams can use the same family for high-resolution video and rendering, and data teams can apply the GPU memory, storage and networking to analytics pipelines.

That breadth is useful for procurement teams because one instance family can support several production workloads instead of a single AI pilot.

GB300 Status Targets Training Buyers

AWS has also achieved Nvidia Exemplar Cloud status on Nvidia GB300 for training workloads.

Nvidia describes the status as evidence that AWS meets the performance thresholds it uses to benchmark AI workloads against its reference architecture.

The designation is aimed at companies comparing cloud providers for large-scale training.

It does not name customer deployments or pricing, but it gives procurement and AI infrastructure teams another benchmark when they compare training performance, total cost of ownership and the move from pilots to production.

For AWS customers, the new stack now covers GPU instances, managed vector indexing and a GB300 training-performance benchmark.

The public record still lacks regional availability, customer adoption figures or pricing in the announcement, leaving buyers to test whether the claimed inference, search and training gains hold inside their own workloads.

Share this article
inXf

Related articles

More
Blackwell’s MLPerf Run Puts AI Training Bottlenecks At Rack Scale
Cloud & Data Centers

Blackwell’s MLPerf Run Puts AI Training Bottlenecks At Rack Scale

NVIDIA says Blackwell led MLPerf Training 6.0 across all seven benchmarks, with submissions scaling to 8,192 GPUs and GB300 NVL72 training up to 1.6x faster than GB200 NVL72 at the same scale.

Digi Power X Lands $19.6 Million Blackwell Capacity Deal
Cloud & Data Centers

Digi Power X Lands $19.6 Million Blackwell Capacity Deal

Digi Power X signed a $19.6 million, 24-month agreement with SubQ AI for bare metal Nvidia Blackwell GPU capacity, but did not identify which data center will host the deployment.

Armenia AI Plan Centres On Blackwell Compute, Not Chip Fabrication
Cloud & Data Centers

Armenia AI Plan Centres On Blackwell Compute, Not Chip Fabrication

AI News places Armenia’s Firebird plan around imported NVIDIA Blackwell infrastructure, US export approval and power capacity rather than domestic chip manufacturing.

Nvidia Japan AI Factory Plan Lists 140MW Capacity
Cloud & Data Centers

Nvidia Japan AI Factory Plan Lists 140MW Capacity

Tech Wire Asia reported that Nvidia’s Japan expansion includes a 140-megawatt AI factory using 13,750 Vera CPUs and 27,500 Rubin GPUs for the METI-backed FRONTia Project. The report did not name the site, capital cost, power supplier, construction timetable or first enterprise workloads.

Crusoe Adds Serverless Fine-Tuning To AI Infrastructure Platform
Cloud & Data Centers

Crusoe Adds Serverless Fine-Tuning To AI Infrastructure Platform

Crusoe added managed fine-tuning and inference services for open-weight models inside Intelligence Foundry. The launch moves its AI infrastructure pitch beyond rented GPU access, while prices, named customers and verified savings remain outside the public record.

Nvidia Names $500 Billion US AI Infrastructure Plan But Leaves Timing Open
Cloud & Data Centers

Nvidia Names $500 Billion US AI Infrastructure Plan But Leaves Timing Open

Nvidia says it and partners including TSMC, Foxconn, Wistron, Corning, Lumentum, Coherent and Amkor plan up to $500 billion of US AI infrastructure production. The account comes from Nvidia's own company blog; it names factories, suppliers and job figures, but gives no full production timetable for the programme.

AWS Adds Persistent Runtime Instances For Production AI Agents
Cloud & Data Centers

AWS Adds Persistent Runtime Instances For Production AI Agents

AWS announced runtime instances for Amazon Bedrock AgentCore Runtime, adding managed infrastructure for multi-agent workflows, shared sessions lasting up to 14 days and GPU-supported production agent deployments.

Cisco Adds Supermicro Rack Systems to Secure AI Factory Stack
Cloud & Data Centers

Cisco Adds Supermicro Rack Systems to Secure AI Factory Stack

ServeTheHome wrote that Cisco will add Supermicro liquid-cooled and air-cooled rack-scale systems to Secure AI Factory with NVIDIA, with October availability and Cisco validation, support and Cloud Control management around the hardware.

Keep Reading

More Stories

Latest
Ethereum Testnet Update Targets 200 Million-Gas BlocksCrypto/Web3Oct 6, 2026Ethereum Testnet Update Targets 200 Million-Gas BlocksEthereum developers released Prysm 7.2.1 so the Sepolia trial of Glamsterdam can test 200 million-gas blocks, more than three times the prior 60 million setting, before any main-network change.Kepler Targets 2027 Production for HBM Replacement MemoryCloud & Data CentersOct 6, 2026Kepler Targets 2027 Production for HBM Replacement MemoryEE Times reports that Kepler Computing is preparing 3D ferroelectric memory for 2027 production, promising higher capacity and bandwidth per watt while limiting reliance on advanced-node lithography.Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceCapital & PolicyOct 6, 2026Yokogawa Opens Singapore Hub For Industrial Cyber ResilienceYokogawa Engineering Asia has launched a Singapore center focused on OT cyber resilience, training, response planning and recovery coordination for Southeast Asia, Oceania and Taiwan.ClickFix Attack Uses Browser Cache To Hide Malware PayloadCybersecurityOct 6, 2026ClickFix Attack Uses Browser Cache To Hide Malware PayloadMicrosoft Threat Intelligence traced a ClickFix cache-smuggling method that preloads malware into browser caches, then uses file size checks and a pasted Run command to launch later credential-theft stages.VOA Tests Six-Month Startup Buildout Before Funding DecisionsFintech & Digital PaymentsOct 6, 2026VOA Tests Six-Month Startup Buildout Before Funding DecisionsTechCabal’s interview with VOA Venture Partners founder Victoria Olayide Adesanya describes a six-month build programme that lets the firm work inside African financial-infrastructure startups before deciding whether to invest.Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCrypto/Web3Oct 6, 2026Bitcoin Holds $86,000 As Dollar Index Hits 18-Month HighCoinDesk reported that bitcoin stayed near $86,000 while the U.S. Dollar Index reached about 102.5, with U.S. rate expectations and European political risks strengthening the dollar backdrop.Google Freezes OSS Bug Bounty Reports After AI Submission FloodCybersecurityOct 6, 2026Google Freezes OSS Bug Bounty Reports After AI Submission FloodGoogle has stopped accepting new product vulnerability reports in its OSS VRP after invalid automated submissions swamped reviewers, while older reports and some Cloud VRP routes remain open.Fleuret AI Raises €4M For Continuous AI Pentesting PlatformCybersecurityOct 6, 2026Fleuret AI Raises €4M For Continuous AI Pentesting PlatformTech.eu reported that French startup Fleuret AI raised €4 million in pre-seed funding to develop an agentic-AI platform that turns penetration testing into a continuous security process.GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%Fintech & Digital PaymentsOct 6, 2026GFT Analysis Says AI Documentation Can Cut Maintenance Work 30%A GFT Technologies analysis says AI-linked software documentation can cut maintenance effort and speed developer onboarding when knowledge assets stay synchronized with code changes.Schneider Electric Lines Up $22.6 Billion PTC DealAIOct 5, 2026Schneider Electric Lines Up $22.6 Billion PTC DealSchneider Electric plans to buy PTC in a cash transaction valuing the US engineering software provider’s equity at about $22.6 billion, adding product-lifecycle software to its industrial AI push.Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueCapital & PolicyOct 5, 2026Aggarwal Pledges Ola Electric Stake To Fund ₹1,000 Cr Rights IssueOla Electric founder Bhavish Aggarwal pledged 20 Cr shares to finance his participation in a rights issue that forms part of a larger ₹1,500 Cr fundraising plan.Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseAIOct 5, 2026Natrona Schools AI Review Puts Student Privacy Ahead Of Classroom Tool UseNatrona County trustees questioned whether teacher AI tools expose student data, even as existing district rules already ban unauthorized generative AI use by students.