Ai infrastructure

Flexible Orchestration for AI & ML: Beyond Kubernetes Automation

Flexible Orchestration for AI & ML: Beyond Kubernetes Automation

Explore the concept of flexible workload orchestration as a unified solution to manage diverse application types, from traditional web services to complex AI/ML pipelines. This approach simplifies operations, breaks down tooling silos, and provides a future-proof infrastructure for evolving AI technologies.

Fully Connected keynote: Building tools for agents at Weights & Biases

Fully Connected keynote: Building tools for agents at Weights & Biases

A summary of the keynote by Lukas Biewald (Weights & Biases) and Camille Fournier (CoreWeave) at Fully Connected London 2025. They discuss recent product updates for W&B Models and Weave, the synergy behind the CoreWeave acquisition, and a deep dive into building and automating an autonomous software engineer agent.

The Power of AI Agents and Agentic AI Explained

The Power of AI Agents and Agentic AI Explained

AI agents represent a paradigm shift from traditional reactive AI models. This summary explores their proactive, goal-driven nature, detailing how they autonomously plan and execute complex workflows by interacting with a diverse ecosystem of models, APIs, hardware, and even other agents to solve real-world problems.

The CEO Behind the Fastest-Growing AI Inference Company | Tuhin Srivastava

The CEO Behind the Fastest-Growing AI Inference Company | Tuhin Srivastava

Tuhin Srivastava, CEO of Baseten, joins Gradient Dissent to discuss the core challenges of AI inference, from infrastructure and runtime bottlenecks to the practical differences between vLLM, TensorRT-LLM, and SGLang. He shares how Baseten navigated years of searching for a market before the explosion of large-scale models, emphasizing a company-building philosophy focused on avoiding premature scaling and "burning the boats" to chase the biggest opportunities.

Surviving the AI Workforce Shakeup

Surviving the AI Workforce Shakeup

Ben Lorica and Evangelos Simoudis analyze the nuances of AI-driven layoffs, categorizing them into upskilling gaps, automation, and strategic R&D shifts. They also explore the immense pressure for ROI on AI infrastructure investments, leading to the emergence of LLMOps as a form of financial management and the critical need for breaking down organizational silos.

Inside The Startup Launching AI Data Centers Into Space

Inside The Startup Launching AI Data Centers Into Space

Starcloud launched the first NVIDIA H100 GPU into orbit, a first step towards building massive AI data centers in space. Their vision is to leverage continuous solar power and zero-water cooling to overcome Earth's energy and resource limitations for large-scale compute.