Dallas skyline illustration

CYC26 / AI

Gating Cognition: How Modern AI Architectures Minimize Runtime Reasoning

This talk explores the evolution of generative AI architectures and the design trend toward minimizing runtime LLM calls to improve latency, reliability, and cost. I'll walk through ReAct, ReWOO, and agent/skills-based approaches — covering implementation details, comparing their strengths and weaknesses, and demonstrating each with a live demo.

Session abstract

What you’ll learn

This talk explores the evolution of generative AI architectures and the design trend toward minimizing runtime LLM calls to improve latency, reliability, and cost. I'll walk through ReAct, ReWOO, and agent/skills-based approaches — covering implementation details, comparing their strengths and weaknesses, and demonstrating each with a live demo.