An Azure virtual machine named VM1 runs Windows Server 2019 and holds 500 GB of data files. You will use Azure Data Factory to transform these files and then load them into Azure Data Lake Storage. What should you deploy on VM1 to support this design?
Choose an answer
Tap an option to check your answer.
Correct answer: the self-hosted integration runtime.
Why this is the answer
The self-hosted integration runtime (IR) is the correct choice because it's required for Azure Data Factory to access data sources located in private networks or on-premises, such as a virtual machine like VM1. The self-hosted IR acts as a secure bridge, allowing Data Factory to connect to VM1 and read its 500 GB of data files for transformation and loading into Azure Data Lake Storage. The On-premises data gateway is used for Power BI, Azure Logic Apps, and other Azure services, but not directly for Azure Data Factory's data movement. The Azure Pipelines agent is for CI/CD pipelines, not data integration. The Azure File Sync agent synchronizes file shares between Windows Servers and Azure Files, which is not the primary goal here.
Pass your exam — without the endless answer hunt
Get every verified question and explanation for this exam in one place, and save hours of prep. 1,000+ certifications · 20+ languages · free to start.
Pass your exam faster → No card needed