Applied AI/ML Engineer

Boundless Networks, Inc.· Boundless Team
Apply Now ↗
🌍 Remote📍 TELECOMMUTE📍 United StatesFull time

About this role

Boundless is coordinating GPU compute at scale and building toward becoming a leader in AI. As an Applied AI/ML Engineer, you'll ship AI-powered products end-to-end on top of our growing GPU inference fleet — owning everything from serving low-latency inference to standing up reinforcement-learning post-training pipelines. This is a builder's role: you take an idea from prototype to production, tune it for throughput and cost on real GPUs, and iterate fast on customer and internal feedback.

You should be comfortable operating with a high degree of autonomy, navigating ambiguity, and defaulting to a strong bias for action.

What You'll Do

End-to-End AI Product Delivery: Own AI features and products from prototype through production — model selection, serving, evaluation, and iteration — shipping working software rather than research artifacts.

Inference Serving: Deploy and optimize LLM inference across the fleet using vLLM and SGLang. Tune continuous batching, KV-cache management, quantization, speculative decoding, and multi-model routing to maximize throughput and minimize latency and cost per token.

RL & Post-Training Harnesses: Build and operate reinforcement-learning and post-training pipelines using slime (Megatron-LM + SGLang) and Prime Intellect (prime-rl + the Environments Hub / verifiers). This includes reward and verifier design, rollout orchestration, weight synchronization, and keeping long-running training stable.

Evaluation & Iteration: Build eval harnesses and benchmarks that measure quality, throughput, and cost together, and use them to drive fast, data-informed iteration.

Work Across the Stack: Partner with Infrastructure on GPU scheduling and fleet utilization, and with Product on what to build next and why.

  • 3+ years shipping ML/AI systems to production
  • Hands-on experience serving LLM inference with vLLM, SGLang, or TensorRT-LLM
  • Experience with RL / post-training methods (GRPO, PPO, DPO, or SFT), or strong adjacent experience and a clear desire to go deep here
  • Strong Python and PyTorch
  • Working understanding of GPU execution: batching, memory, and basic CUDA concepts
  • Comfort operating in ambiguity with a strong bias for action

Nice to Have

  • Direct experience with slime, prime-rl, the verifiers library, or Megatron-LM
  • Distributed training experience (FSDP, TP/PP/DP parallelism)
  • Quantization (FP8/INT8), P/D disaggregation, or speculative decoding
  • Experience with verifiable inference or large-scale distributed systems
  • Kubernetes and container-based deployment
  • Familiarity with GPU fleet orchestration (Ray, SkyPilot, Slurm)

Additional Requirements

  • Candidates must include a public GitHub profile in their application.
  • The GitHub profile should demonstrate a minimum of 1 year of activity/history.
  • Applications that do not include a GitHub profile, or show insufficient activity, will not be considered.

At Boundless, we take care of our people, because building the future of AI compute starts with an empowered team. Here's what you can expect when you join us:

  • Competitive salary (proposed band b/t US$175k and $250k annually) + equity allocation
  • Health, dental, vision (for U.S. employees; region-adjusted globally)
  • Flexible PTO
  • Professional development and conference travel budget
  • Remote-first with regular off-sites and a high-trust, high-velocity team environment

We are a global team, and applicants from around the world are welcome to apply.

Frequently Asked Questions

Is the salary disclosed for the Applied AI/ML Engineer position at Boundless Networks, Inc.?
The salary for this Applied AI/ML Engineer role at Boundless Networks, Inc. is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Is the Applied AI/ML Engineer job at Boundless Networks, Inc. remote?
Yes, this Applied AI/ML Engineer position at Boundless Networks, Inc. is remote, with team members based in TELECOMMUTE, United States. You can work from home or anywhere in the supported regions.
Is the Applied AI/ML Engineer role at Boundless Networks, Inc. full-time or part-time?
This is listed as a Full time position. It is posted as a Applied AI/ML Engineer role in the Boundless Team department at Boundless Networks, Inc..
Which team or department does the Applied AI/ML Engineer at Boundless Networks, Inc. belong to?
This Applied AI/ML Engineer position is part of the Boundless Team department at Boundless Networks, Inc.. See the full job description for more information about the team structure and responsibilities.
How do I apply for the Applied AI/ML Engineer position at Boundless Networks, Inc.?
Click the "Apply Now" button on this page. You will be redirected to Boundless Networks, Inc.'s official application portal hosted on workable where you can submit your application directly.
When was the Applied AI/ML Engineer job at Boundless Networks, Inc. posted?
This Applied AI/ML Engineer position at Boundless Networks, Inc. was posted on Aug 3, 2026. Apply as soon as possible — early applications are often reviewed first.
Applied AI/ML Engineer
Boundless Networks, Inc.
Apply for this role ↗

You'll be redirected to Boundless Networks, Inc.'s official application page on workable.