Experience: 4+ years
We're building the data infrastructure behind some of the most demanding AI training workloads in the world, and we want sharp, curious people to help us do it. In this role, you'll build and maintain the high-performance data layer our Modeling teams rely on for training and evaluation jobs.
- Work directly on petabyte-scale storage infrastructure, and the networking and performance challenges that come with it.
- Collaborate daily with researchers and engineers who are some of the best in the world at what they do.
- 4+ years of experience working on data storage infrastructure
- Kubernetes experience, especially on the storage side (Persistent Volumes, CSI drivers, etc.)
- The ability to transform unstructured data into performant datasets across diverse storage backends including S3, GCS, and POSIX
- Experience with distributed data processing frameworks such as Apache Beam, Spark, or Flink
- [Nice-to-have] Familiarity with modern analytics tooling such as BigQuery, Airflow, or dbt
- Genuine excitement about AI. You follow the research, have opinions, and enjoy being in the weeds
- Comfort operating at the edge of what's known, with a desire to build something genuinely new rather than optimize what already exists
Search Software Engineer, Data Infrastructure jobs near New York (Remote) → Browse all live jobs
This posting was published by Cohere on their own careers system and is shown here with a direct link to apply there. Employers: for corrections or removal, contact jobs@veritahire.com.