最新的Microsoft Designing and Implementing Multi-Agent AI Solutions - AI-500免費考試真題

問題1
You have a Microsoft Foundry multi-agent solution that includes a planning agent and an execution agent You are designing a fixed order process where executions can run for several hours and can be interrupted by worker restarts.
You need to implement orchestration for the process. The solution must meet the following requirements
* Pause at predefined points.
* Resume from the last saved state after an interruption.
Which mechanisms should you use? To answer, select the appropriate options in the answer area.
NOTE: Each correct selection is worth one point.
正確答案:

Explanation:
Process structure: A workflow with agent executors; Pause mechanism: A human-in-the-loop (HITL) gate; Interruption recovery: Checkpoints.
The process has a fixed execution order and can run for several hours, so an explicit workflow with agent executors is preferable to free-form agent delegation. The requirement to pause at predefined points is a human-in-the-loop requirement: the workflow must be able to suspend and emit a request for approval or operator input. The requirement to resume after worker restarts is addressed by checkpoints. Microsoft Agent Framework checkpoints persist workflow progress, executor state, pending messages, shared state, and outstanding requests so execution can be restored from the last saved point. A retry policy alone re-executes operations but does not reconstruct the complete workflow state after a process interruption. The three selected mechanisms therefore cover structure, controlled pause, and durable recovery as separate but complementary workflow concerns. In production, add telemetry and regression tests around this behavior so changes to prompts, models, tools, or orchestration do not silently alter the intended contract. The selected approach is the one that best matches the platform ' s native execution semantics.
Official Microsoft reference: Microsoft Agent Framework - Checkpoints and Human-in-the-loop
問題2
You are designing a multitenant software as a service (SaaS) platform that uses multiple agents. Users will send latency-sensitive inference requests to the platform by using a shared API.
Initially, there will be 20 tenants, and the platform will expand to 200 tenants.
You need to identify the compute component for a production agent runtime. The solution must meet the following requirements:
Isolate workloads for each tenant by using containerization.
Dynamically scale based on demand.
Minimize administrative effort.
What should you use?

正確答案: D
問題3
You are designing Microsoft Foundry multi-agent solution. The agents will use Agent-to-Agent (A2A) delegation and access separate Azure Storage containers within a resource group named RG1.
You need to recommend identity components for the design. The solution must meet the following requirements:
* Eliminate stored application secrets.
* Limit the impact of a compromised agent or deployment.
Solution: Use delegated user permissions for agent actions. Assign Azure roles by using a shared security group. Does this meet the goal?

正確答案: A
說明:(僅 PDFExamDumps 成員可見)
問題4
You have a Microsoft Foundry project that includes four independent analysis agents. Each agent invocation consumes 500 tokens per minute (TPM) from a Foundry deployment that has a TPM rate limit of 1,000.
After each agent completes, it writes 200 small records to Microsoft Dataverse. Running multiple agents simultaneously causes write bursts that result in HTTP 429 (Too Many Requests) responses.
You need to reduce the end-to-end task duration, while preventing provider and platform throttling. The solution must meet the following requirements:
* Keep as much agent parallelism as the TPM rate limit permits.
* Handle Dataverse throttling without sending premature retries.
How should you configure the orchestration? To answer, select the appropriate options in the answer area.
NOTE: Each correct selection is worth one point.
正確答案:

Explanation:
Maximum agent concurrency: Two agents; Dataverse retry behavior: Respect the Retry-After duration returned by Dataverse.
Each analysis agent consumes 500 TPM and the deployment permits 1,000 TPM, so at most two agents can run simultaneously without exceeding the stated model limit. Running only one wastes available parallelism; running three or four violates the limit. The separate Dataverse problem is HTTP 429 service-protection throttling. Microsoft Dataverse returns a `Retry-After` duration that tells clients how long to wait before retrying. Respecting that value prevents premature retries from worsening the overload condition. A fixed retry interval or immediate retry ignores the server ' s current capacity signal. The orchestration should therefore cap model-side concurrency at two and, after each agent finishes, pace Dataverse writes according to the returned retry guidance. This combination minimizes task duration while honoring both provider and platform throttling constraints. At implementation time, the same rule should be expressed through the framework or service configuration rather than left only as a natural-language convention. That makes the behavior repeatable across runs, easier to test, and less sensitive to model variability.
Official Microsoft reference: Dataverse service protection API limits
問題5
You have a Microsoft Foundry multi-agent solution. The solution includes an orchestrator agent that sends a mix of simple requests, reasoning-heavy tasks, and tool-calling workflows to a single premium model deployment.
Response quality is acceptable, but costs are too high, and latency varies significantly across requests.
You need to recommend a solution to reduce token utilization based on the complexity of the user input.
What should you include in the recommendation?

正確答案: C
說明:(僅 PDFExamDumps 成員可見)
問題6
You have a Microsoft Foundry multi-agent solution that routes requests from an intake agent to a retrieval agent, and then to a resolution agent. The retrieval agent uses a knowledge search tool.
Security testing reveals the following recurring issues:
* Some users submit jailbreak-style prompts at the start of a conversation.
* Some retrieved documents contain hidden instructions intended to manipulate the downstream agent The legal department at your company requires that final responses be flagged for review if they contain protected text. You need to configure guardrails to resolve the security issues and meet the legal requirements.
What should you configure?

正確答案: C
說明:(僅 PDFExamDumps 成員可見)
問題7
HOTSPOT -
You have the following persistence configuration for a Microsoft Foundry multitenant, multi-agent solution.

For each of the following statements, select Yes if the statement is true. Otherwise, select No.
NOTE: Each correct selection is worth one point.
正確答案:

Explanation:
No / Yes / Yes
The shared-team cache key uses `teamId` but no tenant identifier, so two tenants that use the same team value can collide; the first statement is therefore false. Microsoft multitenant cache guidance recommends including a tenant dimension in shared cache keys when data must remain isolated. The long-term semantic-memory tier is durable Cosmos DB storage partitioned by `tenantId` and is outside the service-managed Foundry Memory feature, making the second statement true, assuming application authorization enforces the same tenant boundary. The session-state tier uses durable workflow checkpoint storage, so it can restore execution after a process restart even if service-managed conversation history is disabled. Microsoft Agent Framework checkpoints are specifically designed for durable, cross-process workflow resumption. The correct sequence is No, Yes, Yes. The architecture should still be validated with representative end-to-end tests, but the selected component establishes the correct structural boundary first. Microsoft ' s AI-500 blueprint consistently favors explicit scopes, interfaces, and persistence or identity boundaries over prompt-only conventions.
Official Microsoft reference: Azure Architecture Center - multitenant cache and state patterns
問題8
You have a multi-agent solution in Microsoft Foundry. Every request begins with the same 1,800-token instruction block and Model Context Protocol (MCP) tool definitions, and then appends a unique user message.
Input-token costs and Time to First Token (TTFT) increase during peak hours.
You need to reduce the input-token costs and TTFT for requests that share common instructions and tool definitions. The solution must meet the following requirements:
Reuse cached work only when the common prefix matches exactly.
Generate a new completion for each user request.
Minimize application changes.
Which type of caching should you use?

正確答案: A
說明:(僅 PDFExamDumps 成員可見)

專業認證

PDFExamDumps模擬測試題具有最高的專業技術含量,只供具有相關專業知識的專家和學者學習和研究之用。

品質保證

該測試已取得試題持有者和第三方的授權,我們深信IT業的專業人員和經理人有能力保證被授權産品的質量。

輕松通過

如果妳使用PDFExamDumps題庫,您參加考試我們保證96%以上的通過率,壹次不過,退還購買費用!

免費試用

PDFExamDumps提供每種産品免費測試。在您決定購買之前,請試用DEMO,檢測可能存在的問題及試題質量和適用性。