How NVIDIA GPUs Help Accelerate OpenAI’s GPT-6 Astra Ultrafast

NVIDIA Blog ·

A source describing OpenAI's GPT-6 Astra Ultrafast model deployment on NVIDIA Blackwell GPUs, its inference performance relative to Astra Standard mode, and OpenAI's use of its own models to refine inference software on NVIDIA hardware. Lisez 3 points de vue avec leurs éléments à l’appui et les liens vers les sources.

En un coup d’œil

  • GPT-6 Astra Ultrafast available via OpenAI API and to eligible ChatGPT Work/Codex users

    GPT-6 Astra Ultrafast, running on NVIDIA Blackwell GPUs, is available now in the OpenAI API and to eligible ChatGPT Work and Codex users.

    Lire le moment probant · Paragraphe 1
  • Ultrafast offers up to 8x faster token generation than Astra Standard mode

    Accelerated by inference optimizations through OpenAI’s models that tap into the capabilities of the NVIDIA Blackwell architecture, Ultrafast offers up to 8x faster token generation than the Astra Standard mode.

    Lire le moment probant · Paragraphe 2
  • OpenAI uses its own models to refine inference software on NVIDIA GPUs

    Performance gains don’t stop when a model is deployed. OpenAI is using its own models to help refine the inference software running on NVIDIA GPUs, taking advantage of the platform’s programmability to test and implement improvements.

    Lire le moment probant · Paragraphe 6

Passages clés3

Passages attribués et accompagnés du contexte nécessaire à leur vérification. Ouvrez le texte original pour vérifier la source.

infrastructure optimization

OpenAI uses its own models to refine inference software on NVIDIA GPUs

Extrait original

OpenAI is using its own models to help refine the inference software running on NVIDIA GPUs, taking advantage of the platform’s programmability to test and implement improvements.
Contexte

Performance gains don’t stop when a model is deployed. That ongoing work can make model responses faster and deployed infrastructure more productive over time.

model deployment

GPT-6 Astra Ultrafast available via OpenAI API and to eligible ChatGPT Work/Codex users

Extrait original

GPT-6 Astra Ultrafast, running on NVIDIA Blackwell GPUs , is available now in the OpenAI API and to eligible ChatGPT Work and Codex users.
inference performance

Ultrafast offers up to 8x faster token generation than Astra Standard mode

Extrait original

Ultrafast offers up to 8x faster token generation than the Astra Standard mode.
Contexte

Accelerated by inference optimizations through OpenAI’s models that tap into the capabilities of the NVIDIA Blackwell architecture, For developers, faster generation can shorten coding agents’ edit-test-debug cycles, reduce the time spent generating responses between tool calls and make interactive applications feel more responsive.

Source et méthodologie

Ces points de vue renvoient à leurs sources originales. Les reformulations sont signalées et ne sont pas des citations mot à mot.

Ouvrir la transcription ou les documents sources (s’ouvre dans un nouvel onglet)Signaler un problème