Kiro12d agoAre AI coding agents actually getting better?Six months of diagnostics data say yes, with caveatsAIDevTools1 min
Kiro17d agoContinuous Prompt Evaluation: How We Use LLM Judges and Live Signals to Improve Kiro Agent QualityPrompt behavior is difficult to validate exhaustively. A system prompt operates across combinations of models, tools, codebases, tasks, and users that no test...AIDevTools1 min