As a Research Engineer - Model Architectures, you will be a core contributor to Zyphra’s AI Architecture Research Team. This will involve designing and rigorously testing novel model architectures and training methodologies, with a focus on improving core modeling capabilities (e.g., loss per flop or loss per parameter) and addressing fundamental bottlenecks in contemporary models. You will also work extremely closely with our pre-training team, who will integrate your insights into our next-generation models.
Stand Out From the Crowd
Upload your resume and get instant feedback on how well it matches this job.
Job Type
Full-time
Career Level
Mid Level