Skills, Memory, or Fine-Tuning? The Engineering Loop Behind Self-Improving Agents

QCon San Francisco 2026

Session

Skills, Memory, or Fine-Tuning? The Engineering Loop Behind Self-Improving Agents

Tuesday Nov 17 / 01:35PM PST, Ballroom A at Hyatt Regency, San Francisco

Register

$2,835, Conference (3 days)
Current pricing ends September 8th

Abstract

As agents become mainstream, everyone wants to improve theirs either by making fewer mistakes on existing tasks or by taking on harder ones. This usually happens once an agent is already deployed in production. So when teams try to make an existing agent system better, they see a bad trace and always end up asking the same question:

Should the team update the prompt, add a skill, change memory, fine-tune a model, rewrite a tool, add an eval, or rethink the architecture?

This talk presents practical heuristics for improving production agents safely and repeatedly and for building the kind of self-improving agents teams are now after. It focuses on two core questions: how do you know the agent actually got better, and what part of the agent should you update when something goes wrong? We'll cover failure attribution, scalable vs. one-off fixes, overfitting to individual traces, regression prevention, and how teams can build a manual improvement loop that turns agent failures into durable system improvements.

Main Takeaways

  1. How to tell whether an agent actually improved, rather than just performing better on a single failure case.

  2. How to identify which part of an agent system should change: prompt, skill, memory, fine-tune a model, tool, eval, workflow, or architecture.

  3. How to distinguish scalable improvements from brittle one-off patches that create future maintenance problems.

Register

$2,835, Conference (3 days). Current pricing ends September 8th. All pass options.

76% senior dev or higher
1:11 speaker ratio
60+ practitioners

QCon San Francisco 2026 is a three day conference for senior software engineers, architects and team leads. An international program committee of working engineers selects every session. Patterns and practices, not products and pitches.

Share

From the same track

Tuesday 17 November

10:35 Ballroom A Session Progressive Failure Modes of Modern AI Serving Systems Abi Aryan AI Infrastructure Engineer and Educator 11:45 Ballroom A Session The Revenge of the Data Scientist: Why Reliable AI Needs Evals, Traces, and Metrics Hamel Husain Machine Learning Engineer, 20+ Years in Applied AI, Machine Learning, and Data Science 13:35 Ballroom A Session Skills, Memory, or Fine-Tuning? The Engineering Loop Behind Self-Improving Agents Abhinav Sinha CEO @Lucidic AI, Previously @Stanford AI Lab, @Citadel and Susquehanna International Group, and @Apple 14:45 Ballroom A Session Lessons from Building a $100M Product in Six Weeks at OpenAI Brian Yang Member of Technical Staff @OpenAI 15:55 Ballroom A Session Performance Engineering in the Age of AI 17:05 Seacliff D Unconference Unconference: Engineering AI Systems

Current pricing ends September 8th
$2,835, Conference (3 days)

Register