AI infrastructure and developer tool updates #133

Today's Letter

  1. NVIDIA, CUDA Toolkit 13.4 released
  2. Qwen3.8-Omni-Flash Launches

NVIDIA, CUDA Toolkit 13.4 released

NVIDIA, CUDA Toolkit 13.4 released
  • CUDA Toolkit 13.4 adds Windows on Arm support, extending CUDA development beyond Linux on Arm
  • Preview support for NVIDIA Rubin GPUs introduces the compute capability 107 target for early porting
  • Multi-Process Service V3 adds a scriptable CLI, named server instances, TOML configuration, SM partitioning, and cgroup-integrated GPU memory limits
  • CUDA Compute Fabric Transport enables asynchronous data transfers across NVIDIA NVLink using named logical endpoints
  • CUDA Python adds texture and surface programming, NUMA-aware managed memory, and type stubs through cuda.core 1.1.0
  • SDK installers no longer bundle the NVIDIA driver, which must be installed separately

Source: developer.nvidia.com


Qwen3.8-Omni-Flash Launches

  • Qwen launched Qwen3.8-Omni-Flash as a native omnimodal model for agentic audio and video workflows
  • The model accepts text, image, audio, and video inputs with a 1M-token context window
  • Target workflows include video editing, music video creation, film commentary, audiovisual summarization, and real-time conversation
  • Qwen-MM-Plugins adds perception, tool use, and workflow execution for long-form audio and video
  • Qwen-Live Harness was open-sourced as a runtime for continuous, real-time omnimodal interaction

Source: qwen.ai
More: byteshape.com · techgenyz.com · gigazine.net


Jocoletter curates AI, software, and product trends for developers and builders.

#NVIDIA #Qwen

Subscribe to Jocoletter

Read more