AI Research Engineer (Multi-Modal & Vision) - 100% Remote Worldwide

Tether.io
Rome, Buenos Aires, Bogotá, São Paulo, Montevideo, Bengaluru, Chittagong, Karachi, Tokyo, Hanoi, Taipei, Bangkok, London, Barcelona, Dubai, Sofia, Prague, Belgrade, Copenhagen, Tallinn, Athens, Tbilisi, Dublin, Oslo, Amsterdam, Warsaw, Lisbon, Bucharest, Stockholm, Brussels, Tel Aviv, Nicosia, Argentina, Colombia, Brazil, Uruguay, India, Bangladesh, Pakistan, Vietnam, Taiwan, Thailand, United Kingdom, Switzerland, Spain, United Arab Emirates, Bulgaria, Czech Republic, Serbia, Denmark, Estonia, Greece, Georgia, Budapest, Ireland, Malta, Norway, Poland, Portugal, Romania, Sweden, Belgium, Israel, Cyprus
Workplace: RemoteFull timeFunction: Research & Scientific (R&D)Education: bachelorsSkills: ["Communication","Problem-solving","Collaboration"]

Join a small, high-caliber AI model team to drive end-to-end research and engineering on vision-language models. You will design and optimize training pipelines, manage multimodal datasets, and push model efficiency for deployment across distributed GPU infrastructure, with opportunities to publish and contribute to open-source tooling in a remote, globally distributed setting.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Tether.io
Tether.io
3 months ago

AI Research Engineer (Multi-Modal & Vision) - 100% Remote Worldwide

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 hours agoStatus: Live

Job Summary

Join a small, high-caliber AI model team to drive end-to-end research and engineering on vision-language models. You will design and optimize training pipelines, manage multimodal datasets, and push model efficiency for deployment across distributed GPU infrastructure, with opportunities to publish and contribute to open-source tooling in a remote, globally distributed setting.
Location: Rome, Buenos Aires, Bogotá, São Paulo, Montevideo, Bengaluru, Chittagong, Karachi, Tokyo, Hanoi, Taipei, Bangkok, London, Barcelona, Dubai, Sofia, Prague, Belgrade, Copenhagen, Tallinn, Athens, Tbilisi, Dublin, Oslo, Amsterdam, Warsaw, Lisbon, Bucharest, Stockholm, Brussels, Tel Aviv, Nicosia, Argentina, Colombia, Brazil, Uruguay, India, Bangladesh, Pakistan, Vietnam, Taiwan, Thailand, United Kingdom, Switzerland, Spain, United Arab Emirates, Bulgaria, Czech Republic, Serbia, Denmark, Estonia, Greece, Georgia, Budapest, Ireland, Malta, Norway, Poland, Portugal, Romania, Sweden, Belgium, Israel, Cyprus
Workplace: Remote
Employment Type: Full time
Job Function: Research & Scientific (R&D)

Key Responsibilities

  • •Conduct end-to-end research and engineering on vision-language models, covering training, evaluation, and optimization across the full model development lifecycle.
  • •Design and implement post-training pipelines including supervised fine-tuning, knowledge distillation, and reinforcement learning from human feedback.
  • •Develop and maintain high-quality multimodal datasets, including data curation, filtering, and balancing for domain-specific tasks.
  • •Drive model efficiency and deployability, adapting models for resource-constrained environments using compression and optimization techniques.
  • •Design and implement evaluation frameworks and benchmarks to measure model performance, robustness, and real-world task success.

Key Requirements

  • •Degree in Computer Science, Machine Learning, or a related field; MS/PhD preferred.
Experience:Multimodal AIVision-language models
Education:Bachelor's
Skills:CommunicationProblem-solvingCollaboration
Languages:English
Tech Stack:PythonPyTorchDistributed trainingReinforcement learningFine-tuningDistillation

Company Brief

Tether.io
Builds tools and services to help companies hire, onboard, and manage remote or distributed teams across borders, focusing on payroll, compliance, and global employment workflows.
Industry: HR Tech
Website