Observe and evaluate LLM Coding Agents
- Time
- 2026-08-09 09:35 ~ 10:05
- Speaker
- Che Chia Chang
- Room
- AU
- Co-write
Abstract
LLM coding agents are changing software delivery, but most teams still lack a reliable way to measure quality and risk before deployment. This talk shows how to move from raw observability to practical decision-making by combining tracing, dataset extraction, LLM-as-a-judge, and regression evaluation, so model and framework upgrades become testable engineering decisions instead of guesswork.
Speaker
Che Chia Chang
Che-Chia Chang 是一名專注於後端開發、開發維運、容器化應用及 Kubernetes 開發與管理的技術專家,同時也是 Microsoft 最有價值專業人士(MVP)。
活躍於台灣技術社群,經常在 CNTUG、DevOps Taipei、GDG Taipei、Golang Taipei Meetup 等社群分享 DevOps、SRE、Kubernetes 及雲端運算相關技術。致力於推動開發與維運的最佳實踐,並熱衷於研究與應用最新的雲端與 AI 技術。
個人部落格:https://chechia.net
Che-Chia Chang is a technology expert specializing in backend development, DevOps, site reliability engineering (SRE), containerized applications, and Kubernetes development and management. He is also recognized as a Microsoft Most Valuable Professional (MVP).
Actively engaged in the Taiwanese tech community, he frequently shares insights on DevOps, SRE, Kubernetes, and cloud computing at CNTUG, DevOps Taipei, GDG Taipei, and Golang Taipei Meetup. Passionate about promoting best practices in development and operations, he continuously explores and applies the latest advancements in cloud and AI technologies.