Experience with distributed computing tools like Apache Hudi, Spark, Kafka, Flink, Beam, Trino, DataBricks and other big data technologies.
Experience with distributed storage systems like ADLS, HDFS, S3, DLT etc.
Familiarity with Hadoop, Spark, Databricks or other distributed computing systems.
Understanding of data partitioning and sharding techniques.
Knowledge of distributed computing principles and how they apply to large-scale data processing.
Experience in writing clean code that performs well at scale using languages such as Python, Java etc.
Knowledge of relational databases (e.g. Microsoft SQL Server, MySQL).
Experience using system and performance monitoring tools (e.g. New Relic, DataDog).
Excellent organization, critical-thinking and personal leadership skills
Self-starter with the ability to deliver with minimal supervision.
Being okay with the uncomfortable feeling that comes from learning new things.
BSc/BA in Computer Science or a related degree .
Identify, prioritize and execute tasks in the software development life cycle.
Develop tools and applications by producing clean, efficient code.
Perform validation and verification testing in a test-driven manner
Review the work of others, and invite others to review your work.
Collaborate with internal teams and vendors to fix and improve products.
Work with distributed computing systems like Apache Hudi and Trino for big data processing.
Experience with distributed computing
Nice to have React, Selenium automation and cloud experience.
Knowledge of scripting languages such as Python, Bash or Groovy.
Generative AI Code Assistants - Use of Generative AI Code Assistants (e.g. Github Copilot) and knowledge of latest Generative AI Model capabilities would be an asset.
This posting was published by Pointclickcare on their own careers system and is shown here with a direct
link to apply there. Employers: for corrections or removal, contact jobs@veritahire.com.