Software Engineer, Lakehouse (Data Platform Group)
Cato Networks
שכר לא צויןTel Aviv District, Israel, היברידיסניורמשרה מלאה
משרה חיצונית, ההגשה באתר החברהאושר שהמשרה פתוחה לפני 9 שעות
Build and maintain Cato Networks’ data platform, including cloud-based microservices and data pipelines that process large volumes of data at low latency. The Lakehouse team owns analytical storage, batch processing, and APIs used by customers and teams across the company.
מה תעשו
- Own the analytical storage layer, including data modeling, schema and table design, partitioning, retention, and query performance.
- Design and develop batch processing over the data lake using Spark on EMR, Java, and PySpark.
- Build and evolve Java/Spring Boot services that expose data through APIs.
- Optimize query performance, file layout, compaction, cluster sizing, and storage efficiency.
- Research lakehouse and analytical-storage technologies and adapt them for the product.
- Work with product, DevOps, and security teams.
דרישות
- 5+ years of hands-on experience designing and developing large-scale distributed data systems in production.
- Deep hands-on expertise at significant scale in a columnar/analytical database or Apache Spark.
- Experience with open table formats such as Iceberg, Delta Lake, or Hudi.
- Experience with data lake technologies including Parquet and S3, and SQL query engines such as Athena, Trino, or Presto.
- Strong Java skills; experience with Kubernetes microservices and AWS, particularly EMR, S3, and Glue.
תנאי סף
- 5+ years of hands-on experience designing and developing large-scale distributed data systems in production
- B.Sc. in Computer Science, Software Engineering, or a related field, or equivalent practical experience
JavaApache SparkPySparkClickHouseAmazon EMRAmazon S3KubernetesAnalytical data modeling