Make Enterprise Agentic Inference Production-Ready with PyTorch and vLLM & the PyTorch Landscape

Make Enterprise Agentic Inference Production-Ready with PyTorch and vLLM & the PyTorch Landscape

Video by PyTorch via YouTube
Make Enterprise Agentic Inference Production-Ready with PyTorch and vLLM & the PyTorch Landscape

Joseph Groenenboom of Red will give two talks at PyTorch Conference North America:

Ecosystem Working Group: Unlocking Community Impact Through the PyTorch Landscape

Created by the PyTorch Foundation in early 2025, the PyTorch Ecosystem Working Group spotlights, through inclusion in the PyTorch Landscape, projects that demonstrate technical excellence and active community engagement in their respective domains. The Landscape includes projects that are both Foundation-hosted and community-hosted. With more than 70 active Landscape projects, including Helion, SGLang, and vLLM, membership creates opportunities for broader impact and recognition within the PyTorch community. In this session, led by the Ecosystem Working Group, you will learn:
– What the Ecosystem Landscape is and why it matters: how membership drives visibility and community engagement for independent projects
– How to apply: a lightweight, GitHub-based process that makes it straightforward for projects to apply for ecosystem status
– How lifecycle management works: what ongoing membership looks like and how the Working Group supports active projects

Attendees will leave with a clear understanding of the benefits of joining the Landscape, the minimum governance standards to qualify for inclusion, and will have the opportunity to meet working group members.

Making Enterprise Agentic Inference Production-Ready with PyTorch and vLLM

While serving AI models for research and pilot use cases is a well solved problem, moving it to 24/7 Enterprise ready systems is the next important phase in AI maturity. The reliability, observability, KV cache management, and concurrency requirements for Enterprise readiness are non-trivial problems. Rising to this challenge, PyTorch, vLLM, and the other foundation projects and broader ecosystem have begun adding Enterprise level features and enhancements.

In this session we will cover the various upstream work committed to helping PyTorch and its ecosystem support Enterprise workloads. This talk will cover a sampling of some of the project work; from core PyTorch project build infrastructure up to model serving improvements to account for tool calling support and long context multi-turn chat.

You will leave this session not only with a deeper understanding of what Enterprise ready” actually entails but also understanding how those changes can be created in a large, diverse, open source ecosystem. There will be code shown and tales shared from the engine room.

Join us in San Jose, October 20-21: https://hubs.la/Q04v4SL60

Source

How Non-Code Contributions Drive Open Source | CNCF Ambassador

How Non-Code Contributions Drive Open Source | CNCF Ambassador

Video by CNCF [Cloud Native Computing Foundation] via YouTube
How Non-Code Contributions Drive Open Source | CNCF Ambassador

Showing up, sharing knowledge, and connecting people is how open source grows. CNCF Ambassador Leon Nunes reflects on three years of building community across working groups and global events.

Every talk given and every connection made opens new pathways for builders everywhere.

Join the cloud native community in person at KubeCon + CloudNativeCon: https://events.linuxfoundation.org/kubecon-cloudnativecon-north-america/

#CloudNative #OpenSource #CNCFAmbassadors

Source

How Banks Keep Data Private With Open AI #EnterpriseAI #DataPrivacy

How Banks Keep Data Private With Open AI #EnterpriseAI #DataPrivacy

Video by FINOS via YouTube
How Banks Keep Data Private With Open AI #EnterpriseAI #DataPrivacy

Learn how banks leverage open foundation models to achieve proprietary precision.

Financial institutions are moving toward platform independence by using open AI models to secure data privacy and customize performance. This guide breaks down how post-training adjustments allow banks to maintain full control over their internal data and AI infrastructure.

Subscribe for more deep dives into the future of enterprise AI.

Source

30 Years of KDE: Inside Plasma 6.8, Wayland Switch, & more with Nate Graham & Aleix Pol

30 Years of KDE: Inside Plasma 6.8, Wayland Switch, & more with Nate Graham & Aleix Pol

Video by Michael Tunnell via YouTube
30 Years of KDE: Inside Plasma 6.8, Wayland Switch, & more with Nate Graham & Aleix Pol

KDE is celebrating its 30th anniversary, Plasma 6.8 is right around the corner, and Akademy is bringing the KDE community together once again.

I sat down with KDE’s Nate Graham and Aleix Pol to talk about Plasma 6.8, the move forward with Wayland, what happens behind the scenes at Akademy, how KDE has evolved over the past three decades, and what the future could look like for one of Linux’s biggest desktop projects.

Source

New Apple Watch Feature: Shazam!

New Apple Watch Feature: Shazam!

Video by TWiT Tech Podcast Network via YouTube
New Apple Watch Feature: Shazam!

Got the new Apple Watch Ultra, but where are all the promised features? Live Shazam is cool, but still waiting for the real magic. #AppleWatchUltra #TechReview #Unboxing #NewGadget #Shazam

Source

Open Source News: R Linting, Chrome Translator, Qubit Performance

Open Source News: R Linting, Chrome Translator, Qubit Performance

Introduction Welcome to our open source digest, where we bring you the latest updates from the community. From code linting in R to quantum qubit performance, and from browser translation tools to network testers, we’ve got you covered. Let’s dive in. Community and Collaboration Social Coworking and Office Hours – Code Linting in R: Join … Read more

Open Source Faces AI Integration Tensions

Open Source Faces AI Integration Tensions

Elastic AI and Open Source: A New Era of Flexibility The open-source world is buzzing with innovation, but also with growing pains. This week, we see major advancements in AI infrastructure and a fierce debate over how open-source projects should integrate AI tools. From NVIDIA’s work on elastic expert parallelism to KDE’s controversial AI policy, … Read more

Open Source News: Linting, Qubits, and Security Fixes

Open Source News: Linting, Qubits, and Security Fixes

Open Source Community and Tools Social Coworking and Office Hours: Join the community for a session on code linting in R, focusing on best practices and tools. Asgard v0.4.0 released: The latest version of Asgard now supports scheduling periodic tasks, enhancing automation capabilities. Bean Network Tester v0.7: A new release of this Windows network conditions … Read more

AI and Open Source Weekly: vLLM, Linux, and More

AI and Open Source Weekly: vLLM, Linux, and More

Elastic Expert Parallelism: A Game-Changer for Scalable AI In the rapidly evolving world of AI infrastructure, scalability and flexibility are paramount. NVIDIA’s recent presentation at PyTorch Conference 2026 introduced Elastic Expert Parallelism (EP) in vLLM, a technique that allows dynamic addition or removal of GPUs from a Mixture-of-Experts (MoE) deployment without significant downtime. This innovation … Read more

Elastic Expert Parallelism in vLLM

Elastic Expert Parallelism in vLLM

Video by PyTorch via YouTube
Elastic Expert Parallelism in vLLM

Elastic Expert Parallelism in vLLM lets you add or remove GPUs from an active Mixture-of-Experts deployment during traffic with minimal interruption to serving and minimal downtime.

At #PyTorchCon North America 2026, Itay Alroy of NVIDIA will present “Elastic Expert Parallelism in vLLM,” covering the architecture, key implementation details, open challenges, and future roadmap for Elastic EP.

Alroy will also examine what happens when EP size changes and how NIXL EP enables grow/shrink under live traffic.

Register for PyTorch Conference North America 2026: https://hubs.la/Q04v4SL60

Source