Agent Evaluation: every test case routed to a connected Fabric data agent errors ("Something went wrong") while the same questions work in the Test pane
Setup: Copilot Studio agent with generative orchestration. Tools: a published Microsoft Fabric data agent added as a connected agent (User authentication, shared_fabricdataagent connector), the Fabric IQ MCP (Preview) tool, and one knowledge file. UK region.
Issue: In Agent Evaluation (single-response test set), every test case whose question the orchestrator routes to the connected Fabric data agent fails. In the classic experience the exported CSV contained a silently empty actualResponse; in the new agent experience (September 2026) each such case shows an explicit per-case Error - "Something went wrong while evaluating this test case" - with an empty agent response, and both test methods record Error.
Latest reproduction (September 2026): 24-case test set, 24/24 processed in about 6 minutes, 19 errored. The 19 are exactly the cases that route to the connected Fabric data agent. The other 5 behave correctly in the same run, including a case answered with live data via the Fabric IQ MCP tool - so tool connections work under the same identity. The evaluation ran under the signed-in maker's profile with all connections connected.
Control: the identical questions, asked in the Test pane by the same user minutes earlier, are answered correctly by the connected agent (typical latency 20-60 s). Reproduced across three different connected Fabric data agents and on multiple days (10 August, three runs, classic experience; 9 September, new experience).
Expected: connected-agent responses recorded and graded, or a specific error explaining why the invocation failed.
Is this a known limitation of Agent Evaluation with connected Fabric data agents, and is there a workaround (for example a supported way for the evaluation harness to invoke connected agents, or a timeout setting)?