Skip to main content Explore View all products (200+) Microsoft Foundry Azure Copilot GitHub Copilot Azure Kubernetes Service (AKS) Azure Cosmos DB Azure Database for PostgreSQL Azure Arc Microsoft Fabric Linux virtual machines in Azure Foundry Models Foundry Agent Service Foundry IQ Foundry Tools Foundry Control Plane Observability in Foundry Control Plane Azure OpenAI in Foundry Models Azure Speech in Foundry Tools Azure Machine Learning View all databases Azure Cosmos DB Azure DocumentDB Azure SQL Azure Database for PostgreSQL Azure Managed Redis Microsoft Fabric Azure Databricks Linux virtual machines in Azure Windows Server on Azure Azure Functions Azure Virtual Machine Scale Sets Azure API Management Azure Container Apps Azure Kubernetes Service (AKS) Azure Kubernetes Fleet Manager Azure Container Registry Azure Red Hat OpenShift Azure Container Instances Azure Container Storage Azure Arc Azure Local Microsoft Defender for Cloud Azure Monitor Microsoft Sentinel Azure Migrate View all solutions (40+) Cloud solutions for small and medium businesses Cloud migration and modernization center Data analytics for AI Azure Databases AI apps and agents Microsoft Marketplace Microsoft Sovereign Cloud AI apps and agents Responsible AI with Azure AI Infrastructure Data analytics for AI Machine learning operations (MLOps) Low-code application development on Azure Integration Services Serverless computing DevOps Migration and modernization center .NET apps migration Databases on Azure Linux on Azure Oracle on Azure SAP on the Microsoft Cloud Adaptive cloud High-performance computing (HPC) Infrastructure as a service (IaaS) Resiliency Azure Essentials Frontier Accelerate for Azure FinOps on Azure Microsoft Marketplace Azure pricing overview Create an Azure account Free Azure services Flexible purchase options Pricing calculator FinOps on Azure Maximize ROI from AI Azure savings plans Azure reservations Azure Hybrid Benefit Virtual Machines Azure SQL Microsoft Foundry Microsoft Fabric Azure Kubernetes Service (AKS) Microsoft Defender for Cloud View more Software Development Companies Microsoft Marketplace Find a partner Resources for Azure partners Get started with Azure Customer stories Analyst reports, white papers, and e-books Videos Learn more about cloud computing Documentation Explore Azure portal Developer resources Quickstart templates Resources for startups Developer community Students Azure for partners Blog Events and Webinars Learn Support Contact Sales Get started with Azure Sign in

SQL ServerThis blog post has been co-authored by Bhanu Prakash, Principal Program Manager, Azure Databricks.

We are now announcing the general availability of the Apache Spark 3.0 compatible Apache Spark Connector for SQL Server and Azure SQL, accessible through Maven.

The Spark 3.0 compatible connector went into preview early this year. Since then, we have seen tremendous customer adoption and received helpful customer feedback. Over the last few months, after incorporating enhancements and bug fixes to the connector, we are now excited about the general availability of this connector so that customers can expand their usage for even more workloads.

The Apache Spark Connector for SQL Server is a high-performance connector that enables users to use transactional data in big data analytics and persist results for ad-hoc queries or reporting. It allows you to use SQL Server or Azure SQL as input data sources or output data sinks for Spark jobs. It provides bulk insert data into the database and can outperform row-by-row insertion with 10 to 20 times faster performance, as compared to just using Java Database Connectivity (JDBC). In addition, customers can use this connector to score machine learning models from SQL Server Machine Learning Services, or score results in SQL after doing machine learning in Spark.

Why use the Apache Spark Connector for SQL Server and Azure SQL

The Apache Spark Connector for SQL Server and Azure SQL is based on the Apache Spark DataSourceV1 API and SQL Server Bulk API and uses the same interface as the built-in JDBC Spark-SQL connector. This allows you to easily integrate the connector and migrate your existing Spark jobs by simply updating the format parameter.
Notable features and benefits of the connector:

  • Compatible with Apache Spark 3.0.
  • Support for all Apache Spark bindings (Scala, Python, R).
  • Basic authentication, Active Directory (AD) Key Tab, and Azure Active Directory support.

To learn more about the connector and how to use it, visit the GitHub page. To configure the compatible connector using Maven coordinates, reach out to the Apache Spark Connector for SQL Server and Azure SQL Maven page, links to specific builds are also provided on the GitHub page.

Get involved

The Apache Spark Connector for SQL Server and Azure SQL makes the interaction between SQL Server and Apache Spark flawless. The connector has a growing and engaged community, and has been installed thousands of times. We are continuously evolving and improving the connector, and we look forward to your feedback and contributions.

Want to contribute or have feedback or questions? Check out the project on GitHub and follow us on Twitter.

Note: The connector is community-supported and does not include Microsoft SLA support. Please file an issue on GitHub to engage the community for help.

WE ARE MICROSOFT

Explore Microsoft Foundry

The future of AI starts here. Envision your next great AI app with the latest technologies. Get started with Azure.