Job Summary
Sr Developer role in a global media and entertainment environment focusing on advanced data engineering solutions using PySpark in a hybrid work model. The candidate will design optimize and support data pipelines that power content analytics audience insights and revenue reporting while collaborating with cross functional teams to drive measurable impact on business decisions and viewer experiences.
Responsibilities
- Design efficient PySpark based data pipelines that ingest transform and prepare large scale media and entertainment data sets to support analytics and reporting for content performance and audience engagement.
- Develop reusable PySpark frameworks and modules that standardize data processing for streaming metrics viewership trends recommendation inputs and campaign attribution across multiple media platforms.
- Optimize PySpark jobs by refining partitioning strategies caching resource configurations and query logic to improve reliability processing speed and cost efficiency in large data environments.
- Collaborate with product data science and business stakeholders in media and entertainment to translate analytical needs into scalable PySpark solutions that support audience insights and content valuation.
- Implement robust data quality checks validation rules and monitoring mechanisms in PySpark workflows to ensure accuracy of key performance indicators for ratings engagement and revenue metrics.
- Integrate data from diverse media sources including streaming logs ad servers customer profiles and content catalogs into unified PySpark based datasets that enable comprehensive analysis and reporting.
- Maintain clear technical documentation for PySpark pipelines data models and operational procedures to support knowledge sharing and smooth onboarding for other developers and analysts.
- Troubleshoot production issues in PySpark jobs and related data workflows by performing root cause analysis applying fixes and implementing preventive enhancements to improve system stability.
- Collaborate in a hybrid work model with onsite and remote team members by using effective communication practices regular checkpoints and shared tools to ensure alignment on deliverables.
- Ensure that PySpark solutions comply with data governance privacy and security standards relevant to media and entertainment data including viewer behavior and content metadata.
- Contribute to continuous improvement by evaluating new PySpark features big data tools and best practices that can enhance media analytics capabilities and operational efficiency.
- Provide guidance to peers on PySpark usage patterns coding standards and performance tuning techniques to promote consistency and maintain high quality data engineering outputs.
- Align day to day development activities with the broader purpose of delivering engaging and responsible media experiences by enabling accurate insights that inform content creation and audience personalization.
Qualifications
- Demonstrate strong proficiency in PySpark including data frame operations window functions performance tuning and integration with big data ecosystems for large media data sets.
- Show proven experience of at least eight years in data engineering or development roles with a focus on analytics solutions and complex data workflows in relevant industries.
- Apply domain knowledge of media and entertainment including concepts such as content catalogs ratings audience segmentation and advertising metrics to design meaningful data transformations.
- Utilize solid understanding of distributed data platforms such as Spark clusters and cloud based data services to deploy and maintain PySpark pipelines in production environments.
- Exhibit strong skills in writing clean modular and well tested code using Python and PySpark that supports maintainability and quick enhancements for evolving media analytics needs.
- Demonstrate experience working in hybrid work setups and day shifts by coordinating effectively with cross functional teams and managing time zones and collaboration expectations.
- Leverage familiarity with media analytics tools reporting environments or data visualization platforms to ensure that PySpark data outputs are well structured for downstream consumption.
- Display strong communication and problem solving abilities to convert business questions from media and entertainment stakeholders into clear technical requirements and deliverable solutions.
- Showcase experience in implementing data quality frameworks including validation rules anomaly detection and reconciliation processes to maintain trust in key media metrics.
- Apply knowledge of software development practices such as version control code reviews and continuous integration to maintain stability and reliability of PySpark projects.
- Maintain awareness of industry trends in streaming media digital advertising and audience measurement to keep data solutions aligned with evolving business models and regulatory expectations.
- Demonstrate capacity to mentor junior team members informally by sharing practical insights on PySpark techniques domain patterns and troubleshooting methods that uplift team capability.
关于高知特 (Cognizant)
高知特(Cognizant)(纳斯达克代码:CTSH)作为一家AI Builder和相关技术服务提供商,致力于通过打造全栈AI解决方案,帮助企业将人工智能投资转化为实际价值。公司凭借深厚的行业经验、流程优化和工程技术专长,将企业独特的业务场景融入科技系统,赋能组织释放人才潜能,推动切实成果,并帮助全球企业在瞬息万变的环境中保持领先。如需了解更多详情,敬请访问 cognizant.ai 或关注@cognizant。
补充雇佣信息
薪酬信息截至本职位发布之日为准。Cognizant 保留在适用法律允许的范围内随时修改该信息的权利。
申请人可能需要通过现场面试或视频会议的方式参加面试。此外,候选人在每次面试时可能需要出示其当前所在州或政府签发的有效身份证件。
Cognizant 是一家提供平等就业机会的雇主。在招聘过程中,您的申请和候选资格不会因种族、肤色、性别、宗教、信仰、性取向、性别认同、国籍、残疾、遗传信息、怀孕、退伍军人身份或任何其他受联邦、州或地方法律保护的特征而受到影响。







