InfoQ Homepage Presentations
-
Turning Outward: Growing From Code to Influence
Brad Grantham explains how senior engineers and architects can transition from writing code to building influence by sharpening communication, mastering business alignment, and empowering teams.
-
From Thousands to One: Building LLM-Powered Selection Systems
Jendrik Jördening explains how to build reliable LLM architecture by enforcing strict schemas, separating AI text reasoning from deterministic code, and applying automated output validation.
-
From Models to Agents: Building Context-Aware Consumer AI at Scale at DoorDash
Sudeep Das explains how DoorDash transforms e-commerce recommendations using LLM-driven semantic consumer memory, hierarchical semantic IDs, grounded agentic search, and multi-tiered LLM ranking.
-
The Right 300 Tokens Beat 100k Noisy Ones: The Architecture of Context Engineering
Baruch Sadogursky and Patrick Debois explain context engineering antipatterns in coding agents, demonstrating how skills, right-tool retrieval, and external memory optimize developer workflows.
-
Migrating Uber Eats Feeds to Webview
Nick DiStefano discusses how Uber Eats migrated core mobile surfaces to a hybrid WebView stack, sharing insights on maintaining native UI quality and shipping features faster.
-
Adopting Memory-Safety and Fine-Grained Compartmentalisation with CHERI
David Chisnall explains how CHERI architecture unifies hardware capabilities and pointer metadata to deliver memory safety and fine-grained, efficient software compartmentalization.
-
Producing the World's Cheapest Tokens: A How-to Guide
Meryem Arik explains how to cut AI inference costs by up to 90% by optimizing batch sizes, hardware selection, and request scheduling for high-volume, non-real-time workloads.
-
Leveraging Adversary Emulation for GenAI Red Teaming
Kennedy Torkura explains how to apply GenAI red teaming to secure AWS Bedrock models and knowledge bases, leveraging MITRE ATLAS to discover cloud supply chain vulnerabilities.
-
Keeping ChatGPT Fast as AI Development Accelerates
Martin Spier shares how agentic coding accelerates release velocity at OpenAI and explains how always-on AI agents redefine performance engineering to keep ChatGPT fast at massive scale.
-
Rewriting All of Spotify's Code Base, All the Time
Aleksandar Mitic and Jo Kelly-Fenton discuss how Spotify uses "Honk," a background AI coding agent, to automate codebase migrations at scale and tackle the ongoing maintenance problem.
-
From ms to µs: OSS Valkey Architecture Patterns for Modern AI
Dumanshu Goyal explains how moving from proxied architectures to direct access in Redis/Valkey slashes tail latency to microseconds, lowers costs, and eliminates single points of failure.
-
Automatically Retrofitting JIT Compilers
Laurence Tratt explains yk, a meta-tracing JIT compiler technology that automatically speeds up C-based interpreters like Lua and Python with minimal, non-invasive code changes.