NVIDIA Developer

Vertically Integrated, Horizontally Open | AI Factory Insider Ep 5

Published 2026-09-22 · Duration 40:14

Summary

The discussion outlines the concept of 'vertically integrated, horizontally open' AI factories, emphasizing that a common, accelerated computing foundation is required to support specialized AI across diverse industries. NVIDIA's CUDA platform is presented as the core enabler, providing a full-stack architecture that allows for maximum performance optimization while remaining open to external ecosystem partners (ISVs, middleware, etc.). The evolution of the platform is driven by reusable libraries (CUDAX) and the emerging paradigm of Agentic AI, where these libraries will be exposed as 'skills' for autonomous agents.

Download summary

Key takeaways

  1. Understanding 'Vertically Integrated, Horizontally Open' 2:00

    Vertically integrated means designing the platform for maximum performance across the entire AI full-stack (hardware, software, libraries). Horizontally open means that every layer in this stack is open to the entire ecosystem, allowing partners to contribute their IP and workloads.

  2. The Role of CUDA and CUDAX Libraries 5:40

    CUDA is the foundational platform connecting applications to GPU hardware. CUDAX libraries are reusable, highly optimized Intellectual Property (IP) built on top of CUDA, providing domain-specific solutions (e.g., QDF for data processing, cuGraph for graphs) that accelerate development and time-to-market.

  3. Agentic AI and Skills 24:30

    The future of the platform involves exposing CUDAX libraries as 'skills' for autonomous agents. This allows agents to leverage specialized tools—such as a data processing pipeline or a visual profiler—to perform complex tasks, acting as a 'force multiplier' for developers.

  4. Hardware and Software Co-evolution 31:00

    NVIDIA maintains a tight coupling between hardware and software roadmaps (e.g., Hopper $ ightarrow$ Blackwell $ ightarrow$ Rubin). This co-evolution ensures that software advancements can often be applied to older hardware, maximizing the lifespan and utility of the GPU architecture.

Technical details

  • AI Factory Architecture 120s

    The model is described as 'vertically integrated' (optimized full-stack performance) and 'horizontally open' (allowing diverse partners and industries to build upon a common foundation).

  • CUDA Platform 340s

    CUDA is the fundamental platform that enables accelerated computing by connecting applications to GPU hardware. It forms a deep platform stack from frameworks down to hardware controllers.

  • Library Interoperability 1120s

    The platform relies on interoperability between core libraries (e.g., cuBLAS for linear algebra, cuDNN for AI) and specialized libraries, allowing new AI components to plug into established classical computing mechanisms.

  • Agentic Workflow Enhancement 1470s

    CUDAX libraries are being exposed as 'skills' for agents. This allows agents to autonomously identify and utilize the correct tool (e.g., QDF for data processing) to solve complex, multi-step problems.

Mentioned resources

  • CUDA (Platform/API)
  • CUDAX Libraries (Software/IP)
  • QDF (Library (Data Processing))
  • cuGraph (Library (Graph Processing))
  • Bioneo / Qaquarians (Library (Bio/Pharma))
  • Metropolis (Library (Smart Cities))
  • cuBLAS (Library (Linear Algebra))
  • cuDNN (Library (AI Core))
  • Hopper, Blackwell, Rubin (Hardware Architecture)

Channel & topics

Watch on YouTube · Back to latest

This independent, AI-assisted summary is provided for commentary and informational purposes. It may contain errors or omit important context. Please watch the original video for the creator's complete presentation. Video, thumbnail, and related copyrights belong to their respective owners.