**Data Engineer** **Primary Skills** Python, Java, Kotlin - Proficient in writing clean, efficient, and maintainable code using modern programming languages. Apache Spark, Hadoop, Kafka - Experience with Spark for large-scale data processing and distributed data engineering environments. Apache Hive - Knowledge of Hive for building and maintaining data warehousing solutions. Apache Airflow - Skilled in orchestrating complex data workflows using Airflow. SQL Server - Experience with SQL Server for database management and querying. Software Testing Principles - Ability to design unit and integration tests (PyTest, JUnit) ensuring data accuracy and reliability across pipeline stages. Docker & Kubernetes - Hands-on experience deploying and managing data services in containerized environments. CI/CD Tools - Familiarity with CI/CD tools such as Jenkins, GitHub Actions, or GitLab CI. **Role Overview** We are seeking a highly skilled Data Engineer to design, build, and maintain scalable data infrastructure and pipelines. The ideal candidate has strong experience with distributed systems, modern data processing frameworks, and cloud-native deployment practices. **Key Responsibilities** Pipeline Development - Design, build, and maintain scalable data pipelines across distributed systems. Data Governance - Implement and manage frameworks ensuring data quality, security, and compliance. Performance Optimization - Optimize and troubleshoot data processing jobs for performance and reliability. Testing Data Workflows - Apply unit and integration testing to validate data workflows and transformations. Data Warehousing - Develop and maintain data warehousing solutions using Apache Hive. Workflow Orchestration - Orchestrate data workflows using Apache Airflow. Database Management - Manage and query databases using SQL Server. Cross-functional Collaboration - Work closely with data scientists and analysts to deliver data solutions. Containerized Deployment - Use Docker and Kubernetes to deploy and manage data services. CI/CD Automation - Implement CI/CD pipelines for automated testing, deployment, and monitoring. Data Quality - Ensure data integrity and reliability across all pipelines. **Qualifications** Data Engineering Experience - Proven experience as a Data Engineer or similar role. Programming Skills - Strong programming skills in Python, Kotlin, or Java. Distributed Systems - Hands-on experience with Apache Spark, Apache Hive, Apache Airflow, and SQL Server. Cloud Platforms - Familiarity with AWS, GCP, or Azure is a plus. Problem-solving - Excellent analytical skills and attention to detail. Teamwork - Strong communication and collaboration abilities. Cognizant is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to sex, gender identity, sexual orientation, race, color, religion, national origin, disability, protected Veteran status, age, or any other characteristic protected by law.



