×
Register Here to Apply for Jobs or Post Jobs. X

Data Engineer; local to NY

Job in New York, New York County, New York, 10261, USA
Listing for: Quinnox
Full Time position
Listed on 2026-09-01
Job specializations:
  • Software Development
    Data Engineering, SQL Developer
Salary/Wage Range or Industry Benchmark: 120000 - 160000 USD Yearly USD 120000.00 160000.00 YEAR
Job Description & How to Apply Below
Position: Data Engineer (local to NY)
Location: New York

Role
- Data Engineer Location
- New York Job Description

We are seeking an experienced Data Engineer to join our team and build robust, scalable data pipelines. In this role, you will:

  • Design and implement scalable PySpark data pipelines for batch and streaming workloads
  • Optimize Spark jobs and queries for performance and cost efficiency
  • Build and maintain ETL/ELT processes following data engineering best practices
  • Troubleshoot and resolve complex data pipeline and processing issues
  • Collaborate with data teams to ensure data quality and reliability
Top Skills Databricks Platform Experience
  • Hands-on development experience with Databricks notebooks and workflows
  • Proficiency in Python and PySpark for data transformation and processing
  • Working knowledge of Unity Catalog for data discovery and lineage
  • Experience with cluster configuration and job scheduling
  • Delta Lake development and optimization techniques
  • Databricks SQL for data analysis and reporting
Data Engineering & Pipeline Development
  • Advanced ETL/ELT pipeline design and development
  • Delta Lake performance tuning (Z-ordering, data skipping, compaction, vacuuming)
  • Real-time streaming data pipelines using Structured Streaming and Delta Live Tables
  • Query performance optimization and debugging slow-running jobs
  • Data quality validation and testing frameworks
  • Incremental data processing patterns (CDC, SCD Type
    2)
Data Processing & Optimization
  • Spark optimization techniques (partitioning, bucketing, caching, broadcast joins)
  • Working with large-scale datasets (terabytes to petabytes)
  • Data pipeline orchestration and scheduling
  • Monitoring and alerting data pipelines
  • Implementing Bronze/Silver/Gold (Medallion) data layer patterns
Required Technical Skills
  • Databricks & Spark Proficiency: 3+ years of hands-on experience building data pipelines in Databricks; deep understanding of Spark fundamentals, transformations, actions, and performance optimization techniques including partitioning, caching, and resource management
  • Advanced PySpark and SQL

    Skills:

    Expert-level proficiency writing production-quality PySpark code and complex SQL queries for data transformation, aggregation, and analysis; experience with Data Frame API, Spark SQL, and UDFs; strong understanding of lazy evaluation and execution plans
  • Data Engineering & ETL/ELT: Proven experience building and maintaining production data pipelines; hands-on experience with incremental data loading, change data capture (CDC), and slowly changing dimensions; experience handling data quality issues and implementing data validation frameworks
  • Cloud & Big Data Technologies: Strong proficiency with AWS services (S3, EC2, IAM, Glue, Athena); experience working with large-scale distributed data processing; familiarity with data formats (Parquet, Delta, JSON, Avro) and compression techniques
  • Dev Ops & CI/CD: Experience with version control (Git) and CI/CD pipelines using Git Lab, Git Hub Actions, or similar tools; familiarity with testing data pipelines and deployment automation; experience with Databricks Repos and workspace-level integrations
  • Data Governance: Understanding of data lineage, cataloging, and metadata management; experience implementing data quality checks and monitoring; knowledge of data privacy and security best practices in cloud environments (nice to have)
Preferred Qualifications
  • Bachelor's degree in Computer Science, Engineering, or related field
  • Databricks Certified Data Engineer Associate or Professional certification
  • Experience with data orchestration tools (Apache Airflow, Databricks Workflows)
  • Strong debugging and problem-solving skills
  • Excellent communication skills for technical collaboration
#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary