| Data Engineer with Databricks at Charlotte, North Carolina, USA |
| Email: [email protected] |
|
http://bit.ly/4ey8w48 https://jobs.nvoids.com/job_details.jsp?id=2664066&uid=bfb8ea66292c4988a86f15683332312c From: Navya, Fisec Global [email protected] Reply to: [email protected] Job Title: Data Engineer with Databricks Location: Charlotte, NC (Onsite) Looking for only local candidates Onsite Interview for last round Databricks certification is Mandatory No Job Description: Design and Develop Data Pipelines: Create robust ETL/ELT workflows in Azure Databricks (or AWS/GCP Databricks) for ingesting, processing, and transforming large datasets. Data Modeling: Implement efficient data models (star, snowflake, data vault) to support analytics and reporting needs. Performance Optimization: Tune Databricks jobs, Spark configurations, and storage formats (Delta Lake, Parquet, ORC) for optimal performance and cost efficiency. Data Quality & Governance: Implement validation rules, schema enforcement, and data profiling techniques. Integrate with Unity Catalog or other governance tools. Integration: Connect Databricks with various data sources such as relational databases, APIs, streaming services (Kafka, Event Hub), and data warehouses (Snowflake, Synapse, BigQuery, Redshift). Collaboration: Work with data scientists to prepare training datasets and deploy machine learning models in Databricks MLflow. Automation & CI/CD: Build reusable notebooks, deploy jobs with Databricks Repos and integrate with DevOps pipelines. Documentation: Maintain technical documentation for pipelines, architecture, and data dictionaries. Required Skills & Qualifications Bachelors degree in Computer Science, Information Technology, or related field (Masters preferred). 37+ years of experience as a Data Engineer or similar role. Strong hands-on experience with Databricks and Apache Spark (PySpark/Scala). Proficiency in SQL and at least one programming language (Python, Scala, or Java). Experience with Delta Lake and data lakehouse architecture. Knowledge of cloud platforms (Azure, AWS, or GCP) and their data services (Azure Data Lake Storage, S3, GCS). Familiarity with streaming data pipelines (Structured Streaming, Kafka, Kinesis, Event Hubs). Strong understanding of data warehousing concepts and BI integration. Experience with version control (Git) and CI/CD tools (Azure DevOps, Jenkins, GitHub Actions). Keywords: continuous integration continuous deployment business intelligence sthree card North Carolina Data Engineer with Databricks [email protected] http://bit.ly/4ey8w48 https://jobs.nvoids.com/job_details.jsp?id=2664066&uid=bfb8ea66292c4988a86f15683332312c |
| [email protected] View All |
| 11:57 PM 07-Aug-25 |