Thinking Machines is seeking an engineer to join their data infrastructure team. This role involves architecting and scaling the core infrastructure for distributed training pipelines, multimodal data catalogs, and intelligent processing systems that handle petabytes of data. The engineer will work closely with researchers to accelerate experiments, develop new datasets, improve infrastructure efficiency, and enable key insights. The ideal candidate is excited by distributed systems, large-scale data mining, open-source tools like Spark, Kafka, Beam, Ray, and Delta Lake, and enjoys building from the ground up. This is an evergreen role, meaning applications are continuously reviewed for current and future opportunities.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Senior
Education Level
Associate degree