Developer Advocate, MAX Inference & Serving

Modular
United States, Canada
Workplace: HybridFull timeUSD 150,200 - 225,400 annuallyFunction: Software EngineeringSkills: ["Communication","Writing","Technical communication","Community engagement","Growth mindset"]

Evangelize the MAX platform’s inference and serving capabilities by creating technical content (guides, blog posts, tutorials, and video) and engaging directly with developers and customers via GitHub, Discord, Twitter/X, and LinkedIn. Build and publish benchmarks comparing MAX to vLLM, Triton Inference Server, and TensorRT-LLM, collaborate with engineering and product teams to showcase CPU/GPU performance, and represent Modular at conferences through talks and workshops.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Modular
Modular
2 days ago

Developer Advocate, MAX Inference & Serving

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Evangelize the MAX platform’s inference and serving capabilities by creating technical content (guides, blog posts, tutorials, and video) and engaging directly with developers and customers via GitHub, Discord, Twitter/X, and LinkedIn. Build and publish benchmarks comparing MAX to vLLM, Triton Inference Server, and TensorRT-LLM, collaborate with engineering and product teams to showcase CPU/GPU performance, and represent Modular at conferences through talks and workshops.
Location: United States, Canada
Workplace: Hybrid
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Support customers evaluating MAX for inference and serving and answer technical questions.
  • •Establish, run, and publish benchmarks comparing MAX to vLLM, Triton Inference Server, and TensorRT-LLM.
  • •Create tutorials, video guides, and examples demonstrating how the MAX platform handles inference and serving workloads efficiently on CPUs and GPUs.
  • •Write blog posts and other educational content about MAX inference and serving performance and functionality.
  • •Engage with the developer community across GitHub, Discord, Twitter/X, and LinkedIn; represent Modular at conferences through presentations and workshops.
Travel: Medium travel

Pay and Benefits

Salary: USD 150,200 - 225,400 annually
Equity and Bonus:Equity
Perks:Health Insurance401kPaid Leave

Key Requirements

  • •Demonstrable experience creating technical content for developer audiences (portfolio requested: blogs, tutorials, videos, docs, or courses).
  • •Understanding of the ML inference stack, including model serving architectures and GPU acceleration, and how MAX compares to vLLM, Triton Inference Server, and TensorRT-LLM.
  • •Strong Python skills; systems programming experience (C++, Rust, or similar) is an advantage.
  • •Ability to learn new tools quickly and produce accurate content on a fast cadence.
  • •Ability to record, edit, and publish technical videos with clear audio, good pacing, and technical accuracy.
Experience:Developer toolsAI infrastructureMachine learning
Skills:CommunicationWritingTechnical communicationCommunity engagementGrowth mindset
Tech Stack:MAXVLLMTriton Inference ServerTensorRT-LLMPythonC++RustCUDAROCmMojoGitHubDiscordTwitter/XLinkedIn

Company Brief

Modular
Provides infrastructure and developer tools to build, deploy, and scale large AI models and foundation-model applications, including model hosting, orchestration, and SDKs to accelerate AI product development.
Industry: AI & Machine Learning
Website