We are seeking a Senior Data Engineer to build scalable data platforms, real-time streaming solutions, and graph-based data ecosystems that support analytics, AI, and agentic applications. The role involves working with modern data architectures, knowledge graphs, semantic technologies, and cloud platforms.
Key Responsibilities
Design and develop scalable batch and real-time data pipelines.
Build event-driven solutions using Apache Kafka.
Develop and maintain knowledge graphs, semantic data models, and SPARQL-based solutions.
Support AI/LLM initiatives through RAG, GraphRAG, MCP, and A2A integrations.
Implement cloud-native data solutions and data engineering best practices.
Collaborate with cross-functional teams to deliver high-quality data products.
Required Skills
8+ years of experience in Data Engineering.
Strong Python programming skills.
Hands-on experience with Apache Kafka and event-driven architectures.
Experience with graph technologies, knowledge graphs, SPARQL, RDF, and graph databases (Neo4j, Neptune, GraphDB, etc.)
Familiarity with MCP, A2A, RAG, and AI/LLM data architectures.
Strong SQL and data modeling skills.
Experience with Apache Spark and workflow orchestration tools.
Excellent communication and problem-solving skills.
Preferred (Nice to Have)
AWS or other cloud platform experience.
Databricks, Iceberg, Delta Lake, or Hudi.
Vector databases and GraphRAG implementations.
Docker, Kubernetes, and Terraform.
Experience with AI agent frameworks such as LangGraph or LangChain