Translate requirements, estimate effort, and address or escalate blockers as needed.
Write clean, maintainable code that adheres to best practices in readability, design patterns, reusability, and testing.
Own end-to-end projects, understanding and contributing to all aspects (infrastructure, application tiers, and data tiers).
Continuously monitor performance metrics and recommend improvements or refactors.
Our Tech Stack:
Scala, Apache Spark, Python
Airflow
ScyllaDB, PostgreSQL, ClickHouse, Hive, Redis, Kafka, Kafka Connect
Docker, Jenkins, Terraform, AWS
Required Qualifications:
5+ years of experience working in the Data Engineering space, and with building and maintaining big data pipelines
3+ years of experience working in agile environments (ideally Scrum), collaborating with cross-functional teams (engineering, design, product).
Proficient in Spark
Experience with strongly-typed languages (Java / Scala preferred)
Experience designing, building, and maintaining RESTful APIs and integrating with external services.
Participate in code reviews to ensure best practices, maintainability, and continuous improvement of the codebase.
Ability to write and maintain unit and integration tests based on acceptance criteria, ensuring code quality and reliability.
Proficiency with version control tools, particularly Git, for collaborative development and code management.
Preferred Qualifications:
Strong understanding of building scalable and high-performance back-end systems, optimizing for low-latency and high-throughput.
Worked with a variety of data (structured/unstructured), data formats (flat files, XML, JSON, relational, parquet).
Experience in functional programming languages (Scala preferred)
Experience in Cyber Security
Communication and Collaboration:
Strong attention to documentation and maintaining standards across projects.
Ability to present and defend technical decisions with confidence.