Senior AWS Data Engineer
About Us:
Tech Mahindra offers technology consulting and digital solutions to global enterprises across industries, enabling transformative scale at unparalleled speed. With 149k+ professionals across 90+ countries helping 1100+ clients, TechM provides a full spectrum of services including consulting, information technology, enterprise applications, business process services, engineering services, network services, customer experience & design services, AI & analytics, and cloud & infrastructure services. It is the first Indian company in the world to have been awarded the Sustainable Markets Initiative’s Terra Carta Seal, in recognition of actively leading the charge to create a climate and nature-positive future. Tech Mahindra (NSE: TECHM) is part of the Mahindra Group, founded in 1945, one of the largest and most admired multinational federations of companies.
Job Details
- Should be able to Develop and write telecom network performance data parsers in for multiple file formats (XML, ASN.1) & convert into an acceptable Parquet columnar output. Knowledge & experience of Iceberg Table format is preferred.
- Implement Lambda functions triggered by S3 event notifications via SQS queues for high‐throughput parsing.
- Build & implement code for robust error handling with retries, DLQ, and SNS notifications
- Build a schema management in line with vendor mapping files & write code for updating the schema dynamically with addition of new counters to an extent of automating it.
- Support schema deployment, upgrades, and rollback processes not including user configurations for ignoring counters, grouping overrides, dimension preservation, and counter name mappings.
- Handle schema upserts and merges to simultaneously insert new records and update existing ones.
- Own the end‐to‐end lifecycle of schemas including versioning, updates, and retirement.
- Thorough knowledge of writing building audit logging and metadata in DynamoDB & forwarding the details via CloudWatch
- Should be able to understand the 15 / 60 PM counter granularity and able to write codes for aggregation
- Build and optimize distributed data processing jobs using PySpark on AWS EMR with deployment of these jobs on Data warehouse platforms like Amazon Redshift & ClickHouse
- Work with Amazon MSK (Kafka) for real‐time data ingestion and streaming.
- Automate pipeline orchestration with AWS Step Functions.
- Ensure schemas are aligned with vendor (OEM) specifications and business requirements.
How To Apply:
It's easy to apply online; you just need a copy of your up-to-date CV and to follow the step-by step process. Don't worry if you need to make changes - you'll have the opportunity to review and edit your work on the final page, or you can also share resume directly to provided email address. We look forward to receiving your application!
Tech Mahindra is an Equal Employment Opportunity employer. We promote and support a diverse workforce at all levels of the company. All qualified applicants will receive consideration for employment without regard to race, religion, color, sex, age, national origin or disability. All applicants will be evaluated solely on the basis of their ability, competence, and performance of the essential functions of their positions.