Hardware Acceleration for Next-Generation Models
NVIDIA highlighted OpenAI's GPT-6 Astra Ultrafast running on Blackwell GPUs for faster token generation. The company also announced the DGX Spark 64GB for local deployment and partnered with CoreWeave to support agentic AI infrastructure.
Co-optimizing hardware and software is essential for reducing processing latency and making real-time agent deployments cost-effective.
Source posts · 4
NVIDIA DGX Spark 64GB Gives Developers More Ways to Build and Scale Local AI — Local AI is becoming more useful by the token. As AI agents move from experiments into everyday development, increasingly capable open models are shrinking to fit on more devices, giving builders more to run locally. Coming this month, NVIDIA DGX Spark will be available with 64GB of unified memory from top manufacturer partners — Acer, […]
NVIDIA announced the DGX Spark with 64GB of unified memory to help developers build and scale local AI models.
How NVIDIA GPUs Help Accelerate OpenAI’s GPT-6 Astra Ultrafast — GPT-6 Astra Ultrafast, running on NVIDIA Blackwell GPUs, is available now in the OpenAI API and to eligible ChatGPT Work and Codex users. Accelerated by inference optimizations through OpenAI’s models that tap into the capabilities of the NVIDIA Blackwell architecture, Ultrafast offers up to 8x faster token generation than the Astra Standard mode. For developers, […]
NVIDIA GPUs accelerate OpenAI's GPT-6 Astra Ultrafast, offering eight times faster token generation in the OpenAI API.
Productive, Durable, Fungible: How NVIDIA AI Factories Maximize Return on Investment — AI factories are built by the megawatt, even by the gigawatt. Each megawatt factory costs roughly $60 million, and AI factory operators will only commit capital on that scale with a clear view of the return on investment. Three key things shape AI factory returns: Earning capacity: What the factory could earn in a year […]
NVIDIA explains how AI factories maximize return on investment through earning capacity, productivity, durability, and fungibility.
From Training to Production, NVIDIA and CoreWeave Close the Loop on Agentic AI — Building on nearly a decade of co-engineering, CoreWeave has built NVIDIA compute, networking and software into a cloud purpose-built for AI that’s still returning on investment across multiple generations of deployment. Now, CoreWeave is bringing the next generation of NVIDIA infrastructure to production. At CoreWeave Fully Connected, running this week in San Francisco, CoreWeave announced […]
NVIDIA and CoreWeave are partnering to bring next-generation AI infrastructure and software to production for agentic AI workflows.