Data Engineer
Technology, Data & Digital · Data, AI & Analytics · Data Engineering
I korthet
Ericsson is seeking a Data Engineer in Kolkata, India, to build and operate large-scale data platforms for GenAI systems. The role involves developing complex solutions using Python, Java, Spark, and cloud technologies, with a focus on low-latency, high-throughput data processing and MLOps pipelines.
Ansvarsområden
- Build and operate large-scale, high-throughput data systems handling massive datasets
- Develop complex, scalable solutions using Python and Java
- Design and optimize distributed data pipelines using Apache Spark
- Engineer low-latency, high-performance data processing systems (batch+streaming)
- Work with Cassandra and OpenSearch/Elasticsearch for high availability and scale
- Develop scalable backend services and REST APIs (Spring Boot-based microservices)
- Experience to work with AWS (Kiro) / Microsoft Copilot stack
- Develop MCP-based applications and integrations with enterprise systems
- Build RAG pipelines across network, service, customer, and operational data
- Engineer data pipelines for embeddings, vector stores, and retrieval systems
- Implement end-to-end Data/MLOps pipelines using Docker, Kubernetes, Kubeflow, and CI/CD
- Ensure system performance, scalability, observability, and reliability
- Manage and mitigate FOSS (Free & Open Source Software) vulnerabilities using security scanning and patching practices
Krav
- Strong Python and Java expertise with experience building production-grade systems is mandatory
- Hands-on experience with Apache Spark (PySpark/Scala/Java) and distributed processing
- Proven experience in high-volume, low-latency system design and optimization
- Strong knowledge of Cassandra, OpenSearch/Elasticsearch, and NoSQL data modeling is mandatory
- Experience building scalable APIs and microservices (Spring Boot)
- Hands-on experience with cloud platforms (AWS or GCP)
- Strong working knowledge of Docker and Kubernetes
- Experience with vulnerability management and non functional features(alarm , logging , fault management etc)
- Good understanding of LLMs, embeddings, RAG, and GenAI data pipelines
- Exposure to developer copilots / AI-assisted coding tools
Önskade kvalifikationer
- Telecom OSS/BSS domain knowledge (business + data)
- Experience with Databricks, Snowflake, or similar platforms
- Understanding of AI observability, explainability, and responsible AI
Förmåner
- Insights from previous hires
- Previously worked at
- Previously worked as
#GenAI#AI agents#enterprise copilots#Telecom OSS/BSS#automation#intelligence#decisioning#high-volume#low-latency#data platforms#data pipelines#distributed processing#batch processing#streaming#backend services#REST APIs#microservices#cloud platforms#MLOps#RAG#vector stores#retrieval systems