- Company: Get A Job.ai
- Salary: Pay not listed
- Work type: Remote
Website Get A Job.ai
Represented by Get A Job.ai
About This Opportunity
We are representing a confidential AI cloud infrastructure organization that is building a next-generation inference and fine-tuning platform for foundation models. This remote Senior ML Engineer role focuses on advanced model optimization, training efficiency, and production-scale deployment of large language models across massive GPU clusters.
Our client operates at the cutting edge of AI infrastructure, tackling challenges in model performance, low-precision training, and inference acceleration for text, vision, audio, and multimodal architectures.
Responsibilities
You will contribute to one or more of the following technical directions:
- Advanced Fine-Tuning: Enhance fine-tuning methodologies including LoRA-based and full-parameter approaches for state-of-the-art LLMs, optimizing both model quality and training efficiency
- Inference Optimization: Identify and eliminate LLM inference bottlenecks to drive production speedups. Build model training and evaluation pipelines in JAX for speculative decoding, experiment with dense/MoE and auto-regressive/parallel architectures, and derive scaling laws to guide resource allocation
- Low-Precision Training & Inference: Investigate low-precision methodologies (FP8, NVFP4/MXFP4) for supervised fine-tuning and reinforcement learning, optimized for modern hardware across both inference and training workflows
- Train and deploy large models on multiple computational nodes at scale
- Develop and maintain model training pipelines with a focus on performance and reliability
What We're Looking For
Required:
- Profound understanding of theoretical foundations of machine learning and reinforcement learning
- Deep expertise in modern deep learning for language processing and generation
- Experience training large models on multiple computational nodes
- Strong understanding of performance aspects of large neural network training (sharding strategies, custom kernels, hardware features)
- Strong software engineering skills in Python
- Deep experience with modern deep learning frameworks (JAX preferred)
- Proficiency in contemporary software engineering approaches including CI/CD, version control, and unit testing
- Strong communication and leadership abilities
Nice to Have:
- Previous experience with language models or similar NLP technologies
- Familiarity with important concepts in the LLM space: MHA, RoPE, ZeRO/FSDP, Flash Attention, quantization
- Track record of building and delivering products in dynamic startup-like environments
- Experience developing large distributed systems or high-load web services
- Open-source projects demonstrating your engineering expertise
- Excellent command of English with superior writing and communication skills
How We Work With You
Get A Job.ai is partnering with our client to find exceptional ML engineering talent. Here's our process:
- Apply directly through the Get A Job.ai platform
- Our talent team will screen your application and conduct an initial conversation
- Qualified candidates are submitted to our client with our full support throughout the interview process
- Please do not contact the client directly; all communication flows through Get A Job.ai to ensure the best candidate experience
Pay
Compensation details will be discussed during the screening process based on experience and qualifications.
Equal Opportunity: Get A Job.ai is committed to inclusive hiring practices. We welcome applications from candidates of all backgrounds and provide reasonable accommodations throughout our recruitment process.
Apply with Get A Job.ai
A recruiter will review your profile and submit you to the client. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you. Do not contact the client directly.
Apply through Get A Job.ai. A recruiter will review your profile and submit you.
Local insights for this role are preparing — this section updates automatically in a few seconds (or refresh).
Listing facts
- Role Senior ML Engineer (Token Factory)
- Employer Get A Job.ai
- Location Remote
- Type Full Time
- Pay (from listing) Pay not listed
- Posted September 15, 2026
- Apply by October 15, 2026
- Overview Full job description on this page (468 words)
Facts above come from this job record on Get A Job.AI — not copied from third-party review sites.
Limited public data for this employer
We only show facts we can ground in public sources (Wikidata, O*NET, news/discussion links, or this listing). We do not invent Glassdoor-style ratings, salaries, or testimonials when data is thin. Use the listing facts, occupation context, and related openings below while we continue researching.
Explore related openings
Occupation family: ML Engineer
Keep exploring on Get A Job.ai
Not quite the right fit? Your next opportunity is a click away.
- Browse all jobs
- More jobs by category
- Remote jobs you can do from anywhere
- Research typical pay for this role
- Set a job alert so new matches reach you first
- Upload your resume to apply faster
Hiring instead? Post a job and reach candidates searching right now.
