- Company: Get A Job.ai
- Salary: Pay not listed
- Work type: Remote
Website Get A Job.ai
Represented by Get A Job.ai
About This Opportunity
We are representing a confidential cloud infrastructure organization that is building advanced AI platform capabilities at massive scale. Our client operates one of the world's largest GPU clouds and is seeking a Senior ML Engineer to join their inference and fine-tuning platform team.
This is a remote position where you'll work on cutting-edge foundation model optimization, training efficiency, and production deployment challenges across text, vision, audio, and multimodal architectures.
Responsibilities
You'll contribute to several high-impact technical directions:
- Advanced Fine-Tuning: Enhance fine-tuning methodologies (both LoRA-based and full-parameter) for state-of-the-art large language models, focusing on model quality and training efficiency improvements
- Inference Optimization: Identify and resolve LLM inference bottlenecks to drive production speedups by building model training and evaluation pipelines in JAX for speculative decoding, experimenting with different architectures (dense/MoE, auto-regressive/parallel), and deriving scaling laws to guide resource allocation
- Low Precision Training & Inference: Investigate and implement low-precision methodologies (FP8, NVFP4/MXFP4) for supervised fine-tuning and reinforcement learning, optimized for modern hardware across both inference and training workflows
What We're Looking For
Required qualifications:
- Profound understanding of theoretical foundations of machine learning and reinforcement learning
- Deep expertise in modern deep learning for language processing and generation
- Experience training large models on multiple computational nodes
- Solid understanding of performance aspects of large neural network training (sharding strategies, custom kernels, hardware features)
- Strong software engineering skills with Python
- Deep experience with modern deep learning frameworks (our client uses JAX)
- Proficiency in contemporary software engineering approaches including CI/CD, version control, and unit testing
- Strong communication and leadership abilities
Preferred experience:
- Previous experience working with language models or similar NLP technologies
- Familiarity with important concepts in the LLM space such as MHA, RoPE, ZeRO/FSDP, Flash Attention, and quantization
- Track record of building and delivering products in dynamic, fast-paced environments
- Experience developing large distributed systems or high-load web services
- Open-source projects that demonstrate your engineering capabilities
- Excellent English language skills with superior written and verbal communication abilities
How We Work With You
Get A Job.ai represents top talent to leading organizations in the AI and cloud infrastructure space. When you apply through our platform:
- Our recruiting team will review your application and schedule an initial screening conversation
- We'll discuss your background, technical expertise, and career goals in detail
- If there's a strong match, we'll submit your profile to our client for consideration
- We'll guide you through the interview process and provide coaching at each stage
- Please do not contact the client directly—all communication should flow through Get A Job.ai to ensure the best candidate experience
Pay
Compensation details will be discussed during the screening process and are competitive for senior ML engineering roles at scale-focused AI infrastructure organizations.
Equal Employment Opportunity: Get A Job.ai is committed to providing equal employment opportunities to all applicants regardless of race, color, religion, sex, national origin, age, disability, or any other characteristic protected by law.
Apply with Get A Job.ai
A recruiter will review your profile and submit you to the client. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you.
Local insights for this role are preparing — this section updates automatically in a few seconds (or refresh).
Listing facts
- Role Senior ML Engineer (Token Factory)
- Employer Get A Job.ai
- Location Remote
- Type Full Time
- Pay (from listing) Pay not listed
- Posted September 15, 2026
- Apply by October 16, 2026
- Overview Full job description on this page (485 words)
Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.
Limited public data for this employer
We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.
Explore related openings
Occupation family: ML Engineer
Keep exploring on Get A Job.ai
Not quite the right fit? Your next opportunity is a click away.
- Browse all jobs
- More jobs by category
- Remote jobs you can do from anywhere
- Research typical pay for this role
- Set a job alert so new matches reach you first
- Upload your resume to apply faster
Hiring instead? Post a job and reach candidates searching right now.
