Senior Member of Technical Staff, Multimodal AI

Cohere
San Francisco, New York, Toronto, Montreal, Paris
Workplace: RemoteFull timeFunction: Data Science & Machine LearningSkills: ["Communication","Problem-solving","Teamwork","Collaboration"]

Design and build cutting-edge multimodal AI systems, integrating text, speech, and vision; conduct research on advanced compute infrastructure; collaborate with world-class teams to push the boundaries of multimodal models and deliver scalable, robust solutions.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Cohere
Cohere
1 year ago

Senior Member of Technical Staff, Multimodal AI

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 8 hours agoStatus: Live

Job Summary

Design and build cutting-edge multimodal AI systems, integrating text, speech, and vision; conduct research on advanced compute infrastructure; collaborate with world-class teams to push the boundaries of multimodal models and deliver scalable, robust solutions.
Location: San Francisco, New York, Toronto, Montreal, Paris
Workplace: Remote
Employment Type: Full time
Job Function: Data Science & Machine Learning

Key Responsibilities

  • •Design and develop cutting-edge multimodal AI systems, integrating modalities such as text, speech, and vision.
  • •Conduct research and experiments on advanced compute infrastructure, exploring multimodal representation learning, transfer learning, and more.
  • •Collaborate closely with world-class teams, learning from and contributing to their expertise in the field.
  • •Tune, evaluate, and optimize large multimodal models, building evaluations to measure performance.
  • •Dive into complex ML codebases to identify issues and ensure smooth operation of our systems.

Pay and Benefits

Perks:Health InsuranceDentalParental LeaveRemote WorkMeal AllowanceCo-working StipendPaid Leave

Key Requirements

  • •Strong software engineering skills with a track record of building robust, scalable systems.
  • •Proficiency in Python and deep learning frameworks (JAX, PyTorch, TensorFlow) with an understanding of multimodal capabilities.
  • •Knowledge of distributed training strategies for large-scale multimodal models.
  • •Familiarity with autoregressive models for multimodal tasks such as image/video captioning and speech-to-text generation.
  • •Bonus: publications in top-tier venues; Bonus: experience writing efficient GPU kernels using CUDA to optimize multimodal tasks.
Experience:Multimodal AIAI researchDistributed trainingGPUs
Skills:CommunicationProblem-solvingTeamworkCollaboration
Languages:English
Tech Stack:PythonJAXPyTorchTensorFlowCUDAGPU

Company Brief

Cohere
Builds security-first foundation models and enterprise AI products (LLMs, retrieval, agent platforms) for regulated industries, enabling customizable, private deployments across cloud and on-premises for real-world business applications.
Industry: AI & Machine Learning
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series D
Headquarters: Toronto, Canada
Founded: 2019
Glassdoor
Glassdoor: 2.9
WebsiteLinkedInGlassdoor