Position
Data Engineer Analyst
Company
Optum (UnitedHealth Group)
Location
Hyderabad, India
Job type
Full-time
Job mode
Onsite
Job requisition id
2355738
Years of experience
0 to 3 Years
Company description
Optum is a highly dynamic and globally recognized organization that operates at the critical intersection of healthcare and cutting-edge technology, focusing on delivering exceptional care and improving the lives of millions of people worldwide through innovative data-driven solutions.
As a prominent and integral part of the UnitedHealth Group family of businesses, the organization is deeply committed to its foundational mission of helping people live healthier lives and making the health system work significantly better for everyone, regardless of their background or circumstances.
The overarching corporate culture is meticulously built upon core values that serve as the guiding principles for every team member: integrity, compassion, inclusion, relationships, innovation, and performance, ensuring that every project and initiative aligns with these deeply held beliefs.
The company places a massive emphasis on health equity, firmly believing that every single individual—irrespective of their race, gender, sexuality, age, geographical location, or income bracket—deserves an equal and unimpeded opportunity to achieve and maintain their healthiest possible life.
Recognizing the numerous systemic barriers to good health that disproportionately affect people of color, historically marginalized communities, and individuals with lower incomes, the organization actively works to mitigate these disparities through targeted technological interventions, equitable care delivery, and comprehensive data analysis.
Environmental sustainability and corporate social responsibility are also enterprise priorities, with the organization actively committing to reducing its environmental footprint and mitigating its impact on the planet while simultaneously advancing its core healthcare objectives.
Working at this organization means becoming part of a highly talented and deeply inclusive team of peers who are consistently encouraged to collaborate, share knowledge, and support one another in an environment that fosters both personal and professional growth.
Team members are provided with comprehensive and holistic benefits that go far beyond standard compensation, encompassing financial security, physical wellness, emotional support, and continuous educational opportunities to ensure they can thrive both inside and outside the workplace.
The organization leverages advanced analytics, pharmacy benefits management, and an extensive network of resources to connect individuals with the precise care and information they need to feel their absolute best, fundamentally transforming the way healthcare is accessed and delivered on a global scale.
Ultimately, joining this team means stepping into a role where your daily contributions have a direct, tangible, and profoundly positive impact on the communities served, allowing you to advance health optimization and start caring, connecting, and growing together with a global leader in healthcare innovation.
Profile overview
The Data Engineer Analyst role is a technically demanding and highly rewarding position that requires a passionate individual to dive deep into the architecture, development, enhancement, and maintenance of robust data pipelines that serve as the backbone for advanced analytics and enterprise-level reporting solutions.
In your day-to-day activities, you will be heavily involved in constructing, testing, and meticulously optimizing complex data processing workflows, utilizing powerful programming languages and frameworks such as Python, Apache Spark, and Scala to ensure maximum efficiency, scalability, and reliability.
You will strictly adhere to firmly established development standards, industry best practices, and organizational guidelines, ensuring that all code and data architectures are clean, maintainable, and aligned with the overarching strategic goals of the enterprise data ecosystem.
A critical component of this role involves supporting both traditional batch processing and modern real-time streaming data ingestion mechanisms, requiring a comprehensive understanding of diverse data movement strategies, including significant exposure to high-throughput, Kafka-based data pipelines.
The successful candidate will systematically apply core fundamentals of data warehousing and advanced data modeling techniques, seamlessly navigating the complexities of both highly structured relational databases and fluid, semi-structured data formats to derive meaningful insights.
Beyond development, you will actively participate in crucial operational activities, which include the continuous monitoring of automated data jobs, the rapid identification and resolution of system incidents, and the proactive implementation of measures to guarantee the day-to-day stability and high availability of the data platform.
Security is paramount in the healthcare sector, meaning you will consistently assist with the identification and remediation of security vulnerabilities, promptly implementing necessary fixes, and rigorously following stringent data security guidelines across all development, testing, and production environments to protect sensitive health information.
You will play a supportive yet vital role in the lifecycle management of essential tools and software, actively participating in system upgrades, rigorously testing configuration changes, and carefully validating the integrity and performance of data pipelines immediately following any deployment to production environments.
Under the valuable guidance and mentorship of senior engineering staff, you will be empowered to create, modify, thoroughly test, and maintain intricate Azure Data Factory (ADF) pipelines, contributing to the organization's broader cloud migration and modernization strategies.
Your responsibilities will also extend to the active support of varied Databricks workloads, which involves overseeing job execution, conducting in-depth performance troubleshooting, and implementing basic system optimizations to ensure that big data processes run as efficiently and cost-effectively as possible.
You will be tasked with the management, monitoring, and support of advanced cloud-based data storage solutions, facilitating seamless data movement and integration using enterprise-grade tools such as Azure Blob Storage and AZ Copy.
Collaboration is deeply embedded in the workflow; you will frequently engage with cross-functional teams comprising data scientists, product managers, and software engineers, ensuring that all technical solutions are comprehensively documented and participating actively in knowledge transfer sessions to elevate the collective expertise of the team.
Embracing the future of technology, you will systematically leverage enterprise-approved Artificial Intelligence tools to significantly enhance your daily productivity, drive technological innovation, streamline complex workflows, and automate repetitive tasks, all while continuously evaluating emerging industry trends to push the boundaries of strategic innovation within the company.
Qualifications
A solid educational foundation is required, typically demonstrated through a Bachelor's degree in Computer Science, Information Technology, Data Science, Engineering, or a closely related technical field, or an equivalent combination of education, training, and relevant professional experience.
The candidate must possess foundational, hands-on experience navigating and operating within UNIX and Linux operating systems, showcasing comfort with command-line interfaces and basic shell scripting for server interaction and task automation.
Practical exposure to traditional enterprise data integration and ETL tools, specifically IBM DataStage, is necessary to understand legacy systems and facilitate potential migration or integration with modern cloud architectures.
A foundational understanding of Teradata or similar enterprise-grade relational database management systems is required to effectively query, manipulate, and extract large volumes of data stored within traditional data warehouses.
Working knowledge and practical proficiency in Python programming is an absolute must, as this language will be heavily utilized for various data processing tasks, scripting, data manipulation, and the automation of repetitive engineering workflows.
Candidates should have meaningful exposure to Apache Spark, demonstrating a fundamental grasp of distributed computing principles, in-memory processing, and basic big data processing concepts required to handle massive datasets efficiently.
Given the organizational context, any prior exposure to, or a strong, demonstrable interest in working with complex healthcare data systems, medical claims data, or electronic health records is considered highly advantageous and critical to understanding the business domain.
Familiarity with Apache Airflow or similar modern workflow scheduling and orchestration tools is necessary for creating, scheduling, and monitoring complex Directed Acyclic Graphs (DAGs) that manage the dependencies of various data pipelines.
A basic but solid understanding of the Databricks unified analytics platform is required, encompassing familiarity with workspaces, clusters, and notebooks used for collaborative data science and engineering tasks.
Knowledge of the Snowflake cloud data platform is similarly important, requiring an understanding of its unique architecture, data sharing capabilities, and scalable compute features used for modern data warehousing.
Deep foundational knowledge of ETL (Extract, Transform, Load) and ELT (Extract, Load, Transform) methodologies is critical, along with a firm grasp of core data warehousing concepts, dimensional modeling, and schema design.
Exceptional SQL fundamentals are mandatory, requiring the ability to write complex, optimized queries, joins, aggregations, and window functions to interact with various relational database systems.
The candidate must possess proven, solid communication skills, both written and verbal, to effectively articulate technical concepts to non-technical stakeholders, write clear documentation, and collaborate seamlessly with diverse team members.
Strong analytical thinking and complex problem-solving abilities are essential, enabling the candidate to dissect intricate data issues, identify root causes of pipeline failures, and design elegant, efficient technical solutions.
Above all, an insatiable eagerness to learn, adapt, and grow is required, as the data engineering landscape is constantly evolving, necessitating a proactive approach to mastering new tools, technologies, and methodologies.
Additional info
This position is classified as an exempt role under overtime status regulations, implying a professional level of responsibility where compensation is based on completing the job rather than strictly tracking hourly work, though the standard schedule remains full-time.
Currently, the role requires no travel, allowing the successful candidate to remain focused entirely on their core engineering responsibilities within the primary office location without the disruption of business trips.
Employees are expected to strictly comply with all terms and conditions of their employment contract, as well as all comprehensive company policies, standard operating procedures, and management directives.
The nature of the enterprise environment means that employees must be adaptable to potential organizational changes, which could include transfers or re-assignments to different physical work locations based on evolving business needs.
Candidates must remain flexible regarding potential changes in their assigned teams, reporting structures, or specific work shifts, ensuring that the company can rapidly respond to operational demands and strategic shifts.
The company retains the absolute discretion to adopt, vary, or completely rescind policies regarding the flexibility of work benefits, work environments, and alternative work arrangements without any implied limitations, meaning employees must be comfortable in a dynamic, policy-driven corporate setting.
A preferred qualification includes tangible experience in supporting live, production-grade systems, demonstrating an understanding of the pressures and strict protocols required when participating in critical operational and on-call activities.
Candidates who have had exposure to Apache Kafka, event-driven architectures, or real-time streaming data concepts will find themselves at a distinct advantage, as these technologies are becoming increasingly central to modern data strategies.
Familiarity with modern version control systems, specifically GitHub, is highly preferred, along with an understanding of collaborative coding practices, pull requests, code reviews, and branching strategies.
Experience with GitHub Copilot or similar AI-assisted coding tools is viewed favorably, aligning perfectly with the company's directive to leverage enterprise-approved AI applications to streamline development workflows and boost engineering productivity.
Basic knowledge and practical experience with the broader ecosystem of Azure cloud services are strongly preferred, with a particular emphasis on mastering Azure Data Factory for orchestration and Azure Blob Storage for scalable, secure object storage.
The ideal candidate will have a proven, demonstrable ability to work effectively and harmoniously within a team-oriented, large-scale enterprise environment, navigating complex organizational structures and collaborating across multiple departments to achieve shared technological goals.
The company is deeply committed to being an Equal Employment Opportunity employer, ensuring that all qualified applicants receive fair and unbiased consideration for employment without any regard to race, national origin, religion, age, color, sex, sexual orientation, gender identity, disability, or protected veteran status.
The organization strictly maintains a drug-free workplace environment, meaning that all successful candidates will be required to pass a comprehensive drug test before they can officially begin their employment, ensuring safety and compliance across all facilities.
Accommodations are readily available and thoughtfully provided to individuals with physical and mental disabilities throughout the entire application and interviewing process, reflecting the company's commitment to accessibility, inclusion, and supporting all candidates in presenting their best selves.
Please click here to apply.

