NiteSun Technologies is seeking a talented Data Engineer to join our dynamic team in Johar Town, Lahore. This full-time, onsite role offers an exciting opportunity to design and implement scalable data pipelines, build robust data models, and develop cloud-based data solutions that support both client and internal product needs. The position requires 2 to 4 years of experience and involves working evening hours from 5:00 PM to 2:00 AM. The ideal candidate will be passionate about data engineering and eager to contribute to innovative projects using modern technologies.
Key Responsibilities
- Design, develop, and maintain scalable ETL/ELT data pipelines to support data ingestion and transformation from multiple sources.
- Build efficient data processing solutions using SQL and Python.
- Create and maintain data models tailored to analytical and business requirements.
- Utilize PySpark and other distributed data-processing frameworks to handle large datasets.
- Develop and manage workflows with tools such as Apache Airflow, dbt, or Azure Data Factory.
- Monitor pipeline performance, troubleshoot failures, and resolve data quality issues.
- Optimize queries and data-processing jobs for enhanced performance and scalability.
- Work extensively with cloud data platforms and storage services.
- Implement data validation, quality checks, and monitoring processes to ensure reliability.
- Collaborate closely with engineers, analysts, product teams, and stakeholders to gather and deliver data requirements.
- Maintain comprehensive technical documentation and adhere to Git-based development and deployment workflows.
Required Qualifications
- 2 to 4 years of professional experience in Data Engineering or a related discipline.
- Strong proficiency in SQL and Python programming.
- Hands-on experience with ETL/ELT processes, data pipelines, and data warehousing solutions.
- Demonstrated ability to design data pipelines integrating multiple data sources.
- Solid understanding of data modeling, transformation techniques, and data quality management.
- Practical experience with PySpark or similar distributed data-processing technologies.
- Familiarity with at least one cloud data platform such as Databricks, Snowflake, Azure Synapse/Azure Data Factory, or AWS Glue/Amazon Redshift.
- Experience using orchestration or transformation tools like Apache Airflow, dbt, or Azure Data Factory.
- Proficiency with Git and preferably Azure DevOps or GitHub Actions for version control and CI/CD.
- Strong analytical, troubleshooting, and problem-solving skills.
- Excellent communication skills with the ability to explain technical concepts to both technical and non-technical audiences.
- Bachelor’s degree in Computer Science, IT, Software Engineering, or a related field, or equivalent practical experience.
Preferred Qualifications and Benefits
- Experience with streaming platforms such as Kafka, Azure Event Hubs, or Amazon Kinesis.
- Knowledge of Infrastructure-as-Code tools like Terraform.
- Familiarity with data governance tools such as Unity Catalog, Microsoft Purview, or Collibra.
- Exposure to AI/ML data use cases and supporting data workflows.
- Experience in data quality monitoring, lineage, or metadata management.
- Relevant certifications such as Databricks Data Engineer, SnowPro, DP-203, or AWS Data Engineer are a plus.
What We Offer
- Competitive salary package aligned with industry standards.
- Opportunities for professional growth and continuous learning.
- Exposure to international clients and diverse projects.
- Collaborative and professional work environment.
- Access to cutting-edge cloud, data, and AI technologies.
- Career development through challenging, real-world engineering projects.
If you are passionate about data engineering and eager to work in a forward-thinking environment, we encourage you to apply and join the NiteSun Technologies team.