Jira Metrics
Projects/Test AI Agents(TAA)

Planned vs Unplanned

Share of completed issues that were planned vs unplanned (interrupts).

25%

unplanned

9

planned

3

unplanned

Contributing tickets (12)

šŸ”
12
Key↕Summary↕Type↕Assignee↕Class↕Completedā–¼
TAA-18Model registration silently didn't happen on completed training jobs with register_model=trueBugSemenyuk DmitriyUnplanned7/7/2026
TAA-20Fail fast when a training config needs the solver but the job image doesn't have itTaskOliverAIPlanned7/7/2026
TAA-19Training runs don't log the demonstration-imitation signal (bc_loss / bc_weight) — add the instrumentationTaskOliverAIPlanned7/7/2026
TAA-16Autoplay orchestrator: serve the new 3900-input observation schema (blind_v8) alongside the current 1011-input oneTaskOliverAIPlanned7/7/2026
TAA-13Wire parallel eval to MLflow and verify it matches serial on the real RL championTaskSemenyuk DmitriyPlanned7/4/2026
TAA-7Make the canonical lever-screen statistically valid (train-to-budget + N≄2000), forbid underpowered verdictsTaskOliverAIPlanned7/4/2026
TAA-12Deterministic (argmax) eval arm logs NaN/empty win_rate instead of a numberBugOliverAIUnplanned7/3/2026
TAA-11Rule-server loss-heuristic false-positives on draw-heavy-but-winnable positionsBugOliverAIUnplanned7/3/2026
TAA-6Background solvable-seed collector (Azure CPU job)TaskOliverAIPlanned7/3/2026
TAA-5TAA-2 follow-up: run the argmax-collapse E1-E7 investigation and deliver the findings doc (PR #34887 added only the tooling)TaskSemenyuk DmitriyPlanned7/1/2026
TAA-2Root-cause the argmax-collapse: why does greedy (deterministic) eval score ~0% when stochastic scores ~43%?TaskOliverAIPlanned7/1/2026
TAA-1Orchestrator autoplay UI: always show the real/actual seed for RL agentsTaskOliverAIPlanned7/1/2026
Showing 1–12 of 12
Page 1 of 1