职位介绍
We are looking for a Trino Developer who can own query engine performance, lead migration of existing
pipelines and warehouses, and help establish a scalable, cost-efficient lakehouse architecture.
You will work at the intersection of query engineering, data modeling, and platform operations — translating
legacy Snowflake/Spark logic into performant Trino + Iceberg workloads while ensuring correctness, parity, and
reliability.
-
Hands-on production experience with Trino (or PrestoSQL/Presto).
-
Strong, deep SQL expertise — complex analytical queries, window functions, CTEs, query optimization.
-
Hands-on experience with Apache Iceberg (or comparable open table formats — Delta Lake, Hudi)
including schema/partition evolution and table maintenance. -
Practical experience with Snowflake and/or Apache Spark — enough to read, understand, and migrate
existing workloads. -
Understanding of distributed query execution: MPP architecture, join distribution, memory/spill behavior,
partition pruning, and predicate pushdown. -
Experience with cloud object storage and columnar file formats (Parquet, ORC).
-
Proficiency in at least one programming language (Python, Java, or Scala) for tooling, UDFs, and
automation. -
Version control (Git) and CI/CD for data pipelines.
-
Experience leading a Snowflake ® Trino or Spark ® Trino migration at scale.
-
Trino cluster administration and deployment on Kubernetes.
-
Workflow orchestration (Airflow, Dagster, or similar) and dbt.
-
Experience building data-validation / reconciliation frameworks for migration parity.
-
Knowledge of Trino internals or connector development (contributing custom connectors/UDFs).
-
Streaming/CDC ingestion into Iceberg (Kafka, Flink, Debezium).
-
Data governance, lineage, and cost-optimization tooling.
-
Strong analytical and problem-solving mindset for debugging correctness and performance issues.
-
Clear communication — able to document migration decisions and work with analytics, platform, and
business teams. -
Ownership mentality with attention to data correctness and reliability.
Education: Bachelor of Engineering
- Preferred skills: Technology->Big Data
- Data Processing->Spark->SparkSQL,Technology->Data on Cloud-Data Store->Snowflake
福利待遇
•Learning Budget
必备技能
Software engineering
System design
Troubleshooting
关于Infosys
BANGALORE
总部位置
