Skill: Databricks + Pyspark
Experience: 6 to 9 years
Location: AIA Kochi
Job Summary
Join our multinational corporation as a Developer focusing on PySpark and Python to design and implement scalable data solutions using Databricks platforms. This role demands expertise in managing workflows and SQL in Databricks to support data analytics initiatives in a hybrid work environment during daytime shifts with no travel requirements.
Responsibilities
- Develop test and deploy robust PySpark scripts and Python code to process large-scale datasets efficiently in conjunction with cloud-based data lake infrastructure.
- Design and implement Databricks Workflows to automate data pipeline orchestration ensuring timely data availability for analytics and reporting purposes.
- Create optimized SQL queries using Databricks SQL tools to extract transform and load data while maintaining accuracy and performance.
- Utilize Databricks CLI to manage workspace resources automate routine tasks and enhance productivity throughout the development lifecycle.
- Collaborate closely with data engineers and data scientists to integrate new data sources and refine existing pipelines for meeting evolving business requirements.
- Troubleshoot and resolve performance bottlenecks within data processing workflows by analyzing runtime metrics and applying optimization techniques.
- Apply data validation and quality assurance checks within pipelines to uphold data integrity and compliance with organizational standards.
- Participate in code reviews and contribute to maintaining coding standards documentation and best practices for sustainable development.
- Continuously monitor system activity and usage to proactively identify risks or opportunities for improvements in data infrastructure.
- Support cross-functional teams by delivering data-driven insights and technical expertise to accelerate the achievement of strategic goals.
- Adhere to internal security protocols and data privacy regulations while handling sensitive information throughout the data lifecycle.
- Engage in knowledge sharing and collaborative learning sessions to foster an innovative and inclusive team environment.
- Communicate progress and technical information effectively to stakeholders ensuring transparency and alignment with project timelines.
Qualifications
- Possess a minimum of two years hands-on experience developing data pipelines predominantly using PySpark and Python programming languages.
- Demonstrate proficiency with Databricks Workflows to construct and automate scalable ETL processes within cloud ecosystems.
- Show competence in writing complex queries using Databricks SQL to support advanced analytics use cases and data visualization efforts.
- Have practical skills in utilizing Databricks CLI for environment configuration and administrative operations.
- Exhibit understanding of distributed computing frameworks data architecture principles and best practices in data engineering.
- Hold good-to-have knowledge of additional big data tools and frameworks facilitating seamless integration across platforms.
- Adapt comfortably to a hybrid work model blending onsite collaboration and remote productivity with self-driven accountability.
- Operate primarily in daytime shift hours following company schedule demands without requirements for business travel.
- Communicate fluently in English to interact with global team members and stakeholders effectively.
Certifications Required
Databricks Certified Associate Developer for Apache Spark or equivalent certification preferred.
关于高知特 (Cognizant)
高知特(Cognizant)(纳斯达克代码:CTSH)作为一家AI Builder和相关技术服务提供商,致力于通过打造全栈AI解决方案,帮助企业将人工智能投资转化为实际价值。公司凭借深厚的行业经验、流程优化和工程技术专长,将企业独特的业务场景融入科技系统,赋能组织释放人才潜能,推动切实成果,并帮助全球企业在瞬息万变的环境中保持领先。如需了解更多详情,敬请访问 cognizant.ai 或关注@cognizant。
补充雇佣信息
薪酬信息截至本职位发布之日为准。Cognizant 保留在适用法律允许的范围内随时修改该信息的权利。
申请人可能需要通过现场面试或视频会议的方式参加面试。此外,候选人在每次面试时可能需要出示其当前所在州或政府签发的有效身份证件。
Cognizant 是一家提供平等就业机会的雇主。在招聘过程中,您的申请和候选资格不会因种族、肤色、性别、宗教、信仰、性取向、性别认同、国籍、残疾、遗传信息、怀孕、退伍军人身份或任何其他受联邦、州或地方法律保护的特征而受到影响。







