The mission of Thinking Machines is to build AI that extends human will and judgment. We are a team of full stack generalists with strong product instincts who work closely with researchers. We build systems that compound research and engineering velocity over time. We own the internal platform researchers use every day to manage and monitor training runs and evaluations, inspect and debug model trajectories, and compare results on shared leaderboards. You’ll own key parts of this platform, including evaluation and training libraries, experiment-tracking systems, and visualization tools. You’ll identify researchers’ most important bottlenecks and turn them into reliable, generalizable systems. Our team is still small—expect to participate in research meetings, build close relationships with researchers, and gather feedback frequently to develop conviction about where we should invest next. This role requires technical judgment, close cross-functional collaboration, and product intuition. Success means researchers trust your systems to work, rely on them every day, and find them genuinely delightful to use.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level