Job Description - Lead System Engineer - Storage (260007H7)
Lead System Engineer - Storage - (260007H7)
Missions
MSA - L2 Storage
The person must possess self confidence in his / her technical abilities to perform his role on Storage Administrator & project Migration.
Resolve technical issues of projects and Explore alternate designs
Able to thrive in rapid environment and adopt new technology rapidly
Good Team handling skills and also a team player.
Should be open to work on 24x7 environment
Participate in business meetings with various stake holders
Provide recommendations to improve the SAN/NAS infrastructure , and address / mange critical issues and root cause analysis
Profile
Expertise in NetApp AFF storage technology and installation of new NetApp clusters to the Production for regular scaleup of Infrastructure.
Expertise in NetApp provisioning and snap mirror replication for both CIFS and NFS.
Hands on experience in ONTAP firmware update on NetApp clusters.
Experience in troubleshooting Level 2/3 issues on the NetApp filers in SAN and NAS environments
Exposure to Cisco IC Switches and intermediate troubleshooting skills.
Exposure to managing NetApp storage related errors and hardware failures and vendor management.
Experience in Preparation of Production referential like Runbook, Disaster recovery scenario management, managing operations on storage clusters.
Ability to plan , manage and coordinate new feature designing and development as well as run activities.
Exposure to Incident management, configuration management and capacity Management.
Exposure to Develop and maintain standard metrics for service performance analysis , component reports and service reports , based on Customer requirements
Applied knowledge in the administration of windows and Linux servers
Good knowledge on ITIL and experience on handling incident / change / problem on regular basis.
Good in Technical Documentation and design process and standards
Expert on Critical incident management and SLA management.
Proven track record in the implementation and execution of documented support processes and can direct others in same.
Intermediate level experience in Python scripting and automation.
Good to have knowledge on cloud storage basics and market trend for storage services.
Enthusiast who focuses on regular upskilling and share the knowledge with colleagues.
Mentoring new team members and work towards team's success.
Good to have Intermediate skills in DEVOPS and opensource tools.
Aware of cloud services and market trends on Cloud storage to contribute value to the automations.
As a Storage administrator, ensuring stable Production on Infrastructure, maintaining resiliency, and managing disruptions is your top priority.
As an Agile team member, embrace changing requirements, prioritize customer satisfaction through timely delivery, and utilize Scrum ceremonies for collaborative planning within team goals.
Own tasks, adapt to flexible working hours during launches, and actively participate in key forums for team goals.
Responsibilities
Salary : As per industry standard.
Industry :IT-Software / Software Services
Functional Area : IT Software - Application Programming , Maintenance
Role Category :Programming & Design
Role :Job Description - Lead System Engineer - Storage (260007H7)
We are seeking a Engineer (L2 Support) with 3-5 years of experience to join our technical support team on a Contract-to-Hire basis. The ideal candidate will be responsible for providing technical support, troubleshooting application issues, and assisting in maintaining optimal system performance in AWS cloud environments. This role requires good analytical skills, proficiency with monitoring and logging tools, and the ability to work collaboratively with cross-functional teams to resolve incidents and support service reliability.
Shift: 2PM to 11PM
JOB RESPONSIBILITY
· Incident Management and Resolution: Respond to and resolve L2 support tickets in a timely manner, ensuring adherence to SLA commitments.
· AWS Cloud Monitoring: Monitor cloud infrastructure and applications using AWS CloudWatch, Datadog, and other monitoring tools to identify and address performance issues.
· Log Analysis and Troubleshooting: Analyze application logs, system logs, and error traces to diagnose root causes of incidents and assist in implementing corrective actions.
· Application Support: Provide troubleshooting support for production applications, working with development teams to resolve technical issues.
· Linux System Administration: Execute Linux commands and scripts to investigate system behavior, manage processes, and perform routine maintenance tasks.
· Incident Ticketing and Documentation: Create, update, and manage incident tickets using ticketing systems, ensuring accurate documentation of issues and resolutions.
· Escalation Management: Escalate complex issues to senior engineers or L3 support teams when necessary, providing detailed context and initial analysis.
· Performance Monitoring: Assist in identifying performance bottlenecks and work with engineering teams to support optimization efforts.
· Cross-functional Collaboration: Collaborate with DevOps, development, and infrastructure teams to ensure effective incident resolution.
· Communication and Reporting: Provide clear and timely updates to stakeholders, including status reports and incident summaries.
· Proactive Monitoring: Support the setup and maintenance of alerts and dashboards to detect anomalies and prevent potential incidents.
· Knowledge Sharing: Contribute to internal documentation, runbooks, and knowledge base to support team efficiency.
· On-call Support: Participate in on-call rotation to provide support coverage for production systems.
· Continuous Learning: Stay updated with AWS services, monitoring tools, and support best practices to improve technical capabilities.
QUALIFICATION
B.E/B.Tech in Computer Sciences, IT, Engineering, or related field
EXPERIENCE
· 3 to 5 years of experience in L2 technical support, application support, or production support roles
SKILLS AND COMPETENCIES
Technical Skills:
· Proficiency in AWS Cloud services including EC2, S3, Lambda, RDS, CloudWatch, and other core AWS components.
· Working knowledge of monitoring and observability tools such as AWS CloudWatch, Datadog, Prometheus, Grafana, or similar platforms.
· Good log analysis skills using tools like CloudWatch Logs, Splunk, ELK Stack, or other log management solutions.
· Proficiency in Linux/Unix systems administration, including command-line operations, shell scripting, and basic system troubleshooting.
· Experience with incident ticketing systems such as ServiceNow, Jira Service Desk, or similar ITSM tools.
· Application troubleshooting skills across web applications, APIs, and microservices.
· Basic knowledge of networking concepts, including DNS, load balancing, VPC, and security groups.
· Familiarity with CI/CD pipelines and deployment processes.
· Understanding of database systems (SQL and NoSQL) and ability to perform basic query troubleshooting.
· Experience with scripting languages such as Bash, Python, or PowerShell for automation and troubleshooting.
· Exposure to containerization technologies (Docker, Kubernetes) is a plus.
· Understanding of ITIL processes and best practices for incident and problem management.
Domain Expertise:
· Incident Management: Understanding of incident lifecycle, severity classification, SLA management, and escalation procedures.
· Root Cause Analysis: Ability to assist in RCA activities and document findings for preventive measures.
· Cloud Architecture: Basic understanding of cloud-native architectures and distributed systems troubleshooting.
· Performance Monitoring: Knowledge of application and infrastructure performance monitoring techniques.
Professional Competencies:
· Good problem-solving and analytical skills with the ability to diagnose technical issues systematically.
· Effective communication skills, both written and verbal, capable of explaining technical issues clearly and providing status updates.
· Ability to work independently on assigned tickets while seeking guidance when needed for complex issues.
· Customer-focused approach with commitment to delivering quality support and ensuring user satisfaction.
· Sense of ownership and accountability, taking responsibility for assigned incidents through resolution.
· Ability to remain calm and focused during incidents and manage multiple tasks effectively.
· Willingness to learn and stay current with AWS services, monitoring tools, and industry best practices.
· Team collaboration skills, working effectively with cross-functional teams including developers and DevOps engineers.
· Time management skills with the ability to prioritize tasks based on business impact and urgency.
· Attention to detail in documentation, ensuring accurate incident records and knowledge articles.
· Flexibility to work in shifts or participate in on-call rotations as required for support coverage.
· Commitment to continuous improvement and contributing to enhanced support processes.
Responsibilities
We are seeking a Engineer (L2 Support) with 3-5 years of experience to join our technical support team on a Contract-to-Hire basis. The ideal candidate will be responsible for providing technical support, troubleshooting application issues, and assisting in maintaining optimal system performance in AWS cloud environments. This role requires good analytical skills, proficiency with monitoring and logging tools, and the ability to work collaboratively with cross-functional teams to resolve incidents and support service reliability.
Shift: 2PM to 11PM
JOB RESPONSIBILITY
· Incident Management and Resolution: Respond to and resolve L2 support tickets in a timely manner, ensuring adherence to SLA commitments.
· AWS Cloud Monitoring: Monitor cloud infrastructure and applications using AWS CloudWatch, Datadog, and other monitoring tools to identify and address performance issues.
· Log Analysis and Troubleshooting: Analyze application logs, system logs, and error traces to diagnose root causes of incidents and assist in implementing corrective actions.
· Application Support: Provide troubleshooting support for production applications, working with development teams to resolve technical issues.
· Linux System Administration: Execute Linux commands and scripts to investigate system behavior, manage processes, and perform routine maintenance tasks.
· Incident Ticketing and Documentation: Create, update, and manage incident tickets using ticketing systems, ensuring accurate documentation of issues and resolutions.
· Escalation Management: Escalate complex issues to senior engineers or L3 support teams when necessary, providing detailed context and initial analysis.
· Performance Monitoring: Assist in identifying performance bottlenecks and work with engineering teams to support optimization efforts.
· Cross-functional Collaboration: Collaborate with DevOps, development, and infrastructure teams to ensure effective incident resolution.
· Communication and Reporting: Provide clear and timely updates to stakeholders, including status reports and incident summaries.
· Proactive Monitoring: Support the setup and maintenance of alerts and dashboards to detect anomalies and prevent potential incidents.
· Knowledge Sharing: Contribute to internal documentation, runbooks, and knowledge base to support team efficiency.
· On-call Support: Participate in on-call rotation to provide support coverage for production systems.
· Continuous Learning: Stay updated with AWS services, monitoring tools, and support best practices to improve technical capabilities.
QUALIFICATION
B.E/B.Tech in Computer Sciences, IT, Engineering, or related field
EXPERIENCE
· 3 to 5 years of experience in L2 technical support, application support, or production support roles
SKILLS AND COMPETENCIES
Technical Skills:
· Proficiency in AWS Cloud services including EC2, S3, Lambda, RDS, CloudWatch, and other core AWS components.
· Working knowledge of monitoring and observability tools such as AWS CloudWatch, Datadog, Prometheus, Grafana, or similar platforms.
· Good log analysis skills using tools like CloudWatch Logs, Splunk, ELK Stack, or other log management solutions.
· Proficiency in Linux/Unix systems administration, including command-line operations, shell scripting, and basic system troubleshooting.
· Experience with incident ticketing systems such as ServiceNow, Jira Service Desk, or similar ITSM tools.
· Application troubleshooting skills across web applications, APIs, and microservices.
· Basic knowledge of networking concepts, including DNS, load balancing, VPC, and security groups.
· Familiarity with CI/CD pipelines and deployment processes.
· Understanding of database systems (SQL and NoSQL) and ability to perform basic query troubleshooting.
· Experience with scripting languages such as Bash, Python, or PowerShell for automation and troubleshooting.
· Exposure to containerization technologies (Docker, Kubernetes) is a plus.
· Understanding of ITIL processes and best practices for incident and problem management.
Domain Expertise:
· Incident Management: Understanding of incident lifecycle, severity classification, SLA management, and escalation procedures.
· Root Cause Analysis: Ability to assist in RCA activities and document findings for preventive measures.
· Cloud Architecture: Basic understanding of cloud-native architectures and distributed systems troubleshooting.
· Performance Monitoring: Knowledge of application and infrastructure performance monitoring techniques.
Professional Competencies:
· Good problem-solving and analytical skills with the ability to diagnose technical issues systematically.
· Effective communication skills, both written and verbal, capable of explaining technical issues clearly and providing status updates.
· Ability to work independently on assigned tickets while seeking guidance when needed for complex issues.
· Customer-focused approach with commitment to delivering quality support and ensuring user satisfaction.
· Sense of ownership and accountability, taking responsibility for assigned incidents through resolution.
· Ability to remain calm and focused during incidents and manage multiple tasks effectively.
· Willingness to learn and stay current with AWS services, monitoring tools, and industry best practices.
· Team collaboration skills, working effectively with cross-functional teams including developers and DevOps engineers.
· Time management skills with the ability to prioritize tasks based on business impact and urgency.
· Attention to detail in documentation, ensuring accurate incident records and knowledge articles.
· Flexibility to work in shifts or participate in on-call rotations as required for support coverage.
· Commitment to continuous improvement and contributing to enhanced support processes.
Salary : As per industry standard.
Industry :IT-Software / Software Services
Functional Area : IT Software - Application Programming , Maintenance
Minimum Qualifications
Bachelor’s Degree in computer science or engineering; or equivalent technical training and experience.
5+ years of experience building and deploying large-scale data processing pipelines.
5+ years of experience with AWS, Spark with Python/Scala, SQL with Redshift, Postgre SQL and Columnar Databases
5+ years of experience with Workflow & pipeline systems with Airflow.
Experience in Databricks
Experience in GIT
Preferred Qualifications
Master’s Degree in computer science, computer engineering, IT or another related field of study.
8+ years of experience in Data & Analytics
Experience with message-based, loosely coupled architectures (e.g. Kafka).
Experience developing systems intended for cloud deployments (AWS, EKS, lambda’s, etc.). DevOps experience. Experience with analytics platforms like Looker or Tableau.
Experience in DevOps orchestration tools
Databricks Certified Data Engineer Professional certification
The Senior Data Engineer works on different projects of data engineering to support the use cases, data ingestion pipeline and identify potential process or data quality issues. This role also partners with the data science team on special projects related to model development and deployment framework. The Sr Data Engineer team supports marketing analytic teams with analytical tools that enable our analytics and business communities to do their job easier, faster and smarter. The team brings together data from different internal & external partners and builds a curated Marketing analytics focused data & tools ecosystem. The Sr Data Engineer plays a crucial role in building this ecosystem depending on the Marketing analytics communities need.
Essential Job Functions
Builds data pipeline for various data sources, particularly from many external data sources. Works with other project team members to understand the use cases, familiar with data sources, provides recommendation on data ingestion processes, complete data ingestion, monitor and identify potential process or data quality issues. - (15%)
Collaborates with internal/external stakeholders to manage data logistics – including data specifications, transfers, structures, and rules. Collaborates with business users, business analysts and technical architects in transforming business requirements into analytical workbenches, tools and dashboards reflecting usability best practices and current design trends. - (15%)
Demonstrates analytical, interpersonal and professional communication skills. Learns quickly and works effectively individually and as part of a team. Accesses, extracts, and transforms Credit and Retail data from a variety of sources of all sizes (including client marketing databases, 2nd and 3rd party data) using Hadoop, Spark, SQL, Big data technologies etc. - (10%)
Provides automation help to analytical teams around data centric needs using orchestration tools, SQL and possibly other big data/cloud solutions for efficiency improvement. - (5%)
Supports Lead Data Engineer and Principal Data Engineer in new analytical proof of concepts and tool exploration projects. Effectively manages time and computing resources in order to deliver on time/correctly on concurrent projects. Involved in creating POCs to ingest and process streaming data using Spark and HDFS. - (10%)
Answers and trouble shoots questions about data sets and analytical tools; Develops, maintains and enhances new and existing analytics tools/Frameworks to support internal customers/consumers. Design, develop and implement data infrastructure and pipelines that collect, connect, centralize, and curate data from various internal and external data sources. - (10%)
Create automation systems and tools to configure, monitor, and orchestrate our data infrastructure and our data pipelines. Ingests data from different sources, processes it according to the requirement document in order to store data to Hive or NoSQL database or different warehousing solutions. Involved in HDFS maintenance, and loading of structured and unstructured data. - (10%)
Evaluate new technologies for continuous improvements in Data Engineering. Collaborate closely with the product team to build out new data features. Work with the data analysts and scientists to implement descriptive, forecasting, and predictive algorithms and models using the latest technologies. Make technology decisions for our data infrastructure. - (10%)
Consistently write production-ready code that is easily testable, easily understood by other developers, and accounts for edge cases and errors. Use systematic debugging to diagnose all issues located to a single service as well as cross service issues. Approach all engineering work with a security lens. Actively look for security vulnerabilities in the code and when providing peer reviews. - (10%)
Provide technical leadership and guidance to junior team members, fostering a collaborative environment for continuous learning and innovation. Translate complex technical subjects for diverse audiences, ensuring alignment and understanding across technical and non-technical stakeholders. - (5%)
Responsibilities
Minimum Qualifications
Bachelor’s Degree in computer science or engineering; or equivalent technical training and experience.
5+ years of experience building and deploying large-scale data processing pipelines.
5+ years of experience with AWS, Spark with Python/Scala, SQL with Redshift, Postgre SQL and Columnar Databases
5+ years of experience with Workflow & pipeline systems with Airflow.
Experience in Databricks
Experience in GIT
Preferred Qualifications
Master’s Degree in computer science, computer engineering, IT or another related field of study.
8+ years of experience in Data & Analytics
Experience with message-based, loosely coupled architectures (e.g. Kafka).
Experience developing systems intended for cloud deployments (AWS, EKS, lambda’s, etc.). DevOps experience. Experience with analytics platforms like Looker or Tableau.
Experience in DevOps orchestration tools
Databricks Certified Data Engineer Professional certification
The Senior Data Engineer works on different projects of data engineering to support the use cases, data ingestion pipeline and identify potential process or data quality issues. This role also partners with the data science team on special projects related to model development and deployment framework. The Sr Data Engineer team supports marketing analytic teams with analytical tools that enable our analytics and business communities to do their job easier, faster and smarter. The team brings together data from different internal & external partners and builds a curated Marketing analytics focused data & tools ecosystem. The Sr Data Engineer plays a crucial role in building this ecosystem depending on the Marketing analytics communities need.
Essential Job Functions
Builds data pipeline for various data sources, particularly from many external data sources. Works with other project team members to understand the use cases, familiar with data sources, provides recommendation on data ingestion processes, complete data ingestion, monitor and identify potential process or data quality issues. - (15%)
Collaborates with internal/external stakeholders to manage data logistics – including data specifications, transfers, structures, and rules. Collaborates with business users, business analysts and technical architects in transforming business requirements into analytical workbenches, tools and dashboards reflecting usability best practices and current design trends. - (15%)
Demonstrates analytical, interpersonal and professional communication skills. Learns quickly and works effectively individually and as part of a team. Accesses, extracts, and transforms Credit and Retail data from a variety of sources of all sizes (including client marketing databases, 2nd and 3rd party data) using Hadoop, Spark, SQL, Big data technologies etc. - (10%)
Provides automation help to analytical teams around data centric needs using orchestration tools, SQL and possibly other big data/cloud solutions for efficiency improvement. - (5%)
Supports Lead Data Engineer and Principal Data Engineer in new analytical proof of concepts and tool exploration projects. Effectively manages time and computing resources in order to deliver on time/correctly on concurrent projects. Involved in creating POCs to ingest and process streaming data using Spark and HDFS. - (10%)
Answers and trouble shoots questions about data sets and analytical tools; Develops, maintains and enhances new and existing analytics tools/Frameworks to support internal customers/consumers. Design, develop and implement data infrastructure and pipelines that collect, connect, centralize, and curate data from various internal and external data sources. - (10%)
Create automation systems and tools to configure, monitor, and orchestrate our data infrastructure and our data pipelines. Ingests data from different sources, processes it according to the requirement document in order to store data to Hive or NoSQL database or different warehousing solutions. Involved in HDFS maintenance, and loading of structured and unstructured data. - (10%)
Evaluate new technologies for continuous improvements in Data Engineering. Collaborate closely with the product team to build out new data features. Work with the data analysts and scientists to implement descriptive, forecasting, and predictive algorithms and models using the latest technologies. Make technology decisions for our data infrastructure. - (10%)
Consistently write production-ready code that is easily testable, easily understood by other developers, and accounts for edge cases and errors. Use systematic debugging to diagnose all issues located to a single service as well as cross service issues. Approach all engineering work with a security lens. Actively look for security vulnerabilities in the code and when providing peer reviews. - (10%)
Provide technical leadership and guidance to junior team members, fostering a collaborative environment for continuous learning and innovation. Translate complex technical subjects for diverse audiences, ensuring alignment and understanding across technical and non-technical stakeholders. - (5%)
Salary : As per industry standard.
Industry :IT-Software / Software Services
Functional Area : IT Software - Application Programming , Maintenance