Data Engineer
Location: Piscataway, NJ
Experience: 4–10 Years
Job Type: Contract – W2
Work Model: Hybrid / Onsite
Visa: USC
Job Overview
We are looking for an experienced Data Engineer to design, develop, and maintain scalable data pipelines and data integration solutions. The ideal candidate will have strong hands-on experience with Python, SQL, ETL/ELT, cloud platforms, and modern data engineering technologies.
The candidate will work closely with data scientists, analysts, application teams, and business stakeholders to build reliable data solutions and transform large and complex datasets into high-quality, analytics-ready data.
Key Responsibilities
- Design, develop, test, and maintain scalable ETL/ELT data pipelines.
- Develop data ingestion and transformation processes using Python and SQL.
- Integrate data from multiple sources including relational databases, APIs, files, and cloud platforms.
- Build and maintain data pipelines for both structured and semi-structured data.
- Develop optimized SQL queries, stored procedures, views, and data transformation logic.
- Design and implement data models for data warehouses and data lakes.
- Work with cloud-based data platforms such as AWS, Azure, or GCP.
- Develop data pipelines using technologies such as AWS Glue, S3, Lambda, Azure Data Factory, Databricks, or equivalent tools.
- Work with Snowflake, Databricks, Redshift, PostgreSQL, SQL Server, or other cloud databases.
- Implement workflow orchestration using Airflow or similar scheduling tools.
- Perform data validation, reconciliation, quality checks, and troubleshooting.
- Optimize pipeline performance, SQL queries, storage, and cloud resource utilization.
- Monitor production pipelines and resolve data-processing failures.
- Participate in code reviews and follow software engineering best practices.
- Use Git/GitHub for source control and collaborate within Agile development teams.
- Create technical documentation for pipelines, data models, and integration processes.
- Collaborate with business and technical stakeholders to understand requirements and deliver data solutions.
Required Skills
- 4–10 years of experience in Data Engineering, ETL, Data Integration, or related roles.
- Strong hands-on experience with Python.
- Advanced SQL skills with experience writing complex queries and optimizing query performance.
- Strong understanding of ETL/ELT concepts and data pipeline architecture.
- Experience working with AWS, Azure, or GCP.
- Experience with at least one cloud data platform such as Snowflake, Databricks, Redshift, BigQuery, or Synapse.
- Experience with data warehousing and data lake concepts.
- Hands-on experience with Airflow or another workflow orchestration tool.
- Strong understanding of relational databases and data modeling.
- Experience working with Git/GitHub and CI/CD processes.
- Strong analytical, troubleshooting, and problem-solving skills.
- Excellent communication and collaboration skills.
Preferred Skills
- PySpark / Apache Spark
- AWS S3, Glue, Lambda, Redshift, Athena
- Azure Data Factory, Azure Databricks, ADLS
- Snowflake and dbt
- Terraform or CloudFormation
- Docker/Kubernetes
- Kafka or other streaming technologies
- REST APIs and data integration
- Data quality and data governance
- CI/CD using GitHub Actions, Jenkins, Azure DevOps, or similar
- Experience preparing datasets for AI/ML applications
Education
Bachelor’s degree in Computer Science, Information Technology, Engineering, Data Science, or a related field.
Share your resume to akshitha@ashratech.com
Pay: $100,000.00 - $500,000.00 per year
Work Location: In person