InfoQ Homepage QCon San Francisco 2025 Content on InfoQ
-
Realtime and Batch Processing of GPU Workloads
Joseph Stein explains how to build a highly available private AI cloud. He shares blueprints on scaling vLLM on enterprise GPUs, implementing gateway guardrails, and optimizing batch workloads.
-
The Ironies of A^2 I^2
J. Paul Reed explains the "ironies of automation" and AI in incident response. He discusses how reliance on AI can erode manual skills and camouflage system failures during high-stakes outages.
-
Powering the Future: Building Your GenAI Infrastructure Stack
Merrin Kurian discusses Intuit’s GenOS, a generative AI operating system powering agents for 100M users. She explains the transition from chat assistants to "done-for-you" autonomous experiences.
-
Accelerating LLM-Driven Developer Productivity at Zoox
Amit Navindgi explains how Zoox built Cortex, an internal AI platform that streamlines the developer lifecycle by moving beyond the hype to deliver secure, agentic workflows and real-world impact.
-
Beyond Coding: How Senior ICs Grow Influence and Drive Impact
Kasia Trapszo, engineering leader at Netflix, explains why the hardest problems aren't systems, but people. She shares how to scale your impact through clarity, alignment, and building trust.
-
Engineering at AI Speed: Lessons from the First Agentically Accelerated Software Project
Adam Wolff shares how the Claude Code team uses their own agentic tools to ship 90% of their production code. He explains why implementation isn't the bottleneck and why feedback loops are everything.
-
How Netflix Shapes our Fleet for Efficiency and Reliability
Joseph Lynch and Argha C. discuss how Netflix balances hardware supply and software demand. They explain techniques like risk-adjusted net value, buffer management, and priority-based load shedding.
-
Stripe’s Docdb: How Zero-Downtime Data Movement Powers Trillion-Dollar Payment Processing
Jimmy Morzaria explains how Stripe scales its MongoDB infrastructure to process $1.4 trillion in payments. He shares the engineering behind their zero-downtime data movement platform and DocDB.
-
Week-Long Outage: Lifelong Lessons
Molly Struve shares a "murder mystery" outage story from a massive Elasticsearch upgrade. She explains why you need a rollback plan, how to check biases, and why leadership support is a stabilizer.
-
How to Build an Exchange: Sub Millisecond Response Times and 24/7 Uptimes in the Cloud
Frank Yu explains how Coinbase builds ultra-high-performance exchanges. He discusses using single-threaded, deterministic core logic and Raft consensus to achieve sub-millisecond P99s in the cloud.
-
Dynamic Moments: Weaving LLMs into Deep Personalization at DoorDash
Sudeep Das and Pradeep Muthukrishnan discuss how DoorDash combines LLMs with deep learning to move from "classic" collaborative filtering to "hyper-personalization" in real-time commerce.
-
From VR to Flat Screens: Bridging the Input and Immersion Gap
Dany Lepage explains how Lucky VR scaled "Vegas Infinite" from Meta Quest to PS5, PC, and mobile. He shares the technical hurdles of cross-play, dual avatar systems, and the "product fit" trap.