AI Research Engineer (Model Compression & Quantization)
Tether.io
Brazil
Workplace: RemoteFull timeFunction: Research & Scientific (R&D)Education: bachelorsSkills: ["Communication","Collaboration","Research"]Research and develop model compression techniques for multimodal AI (LLMs and VLMs) to reduce footprint and latency. Build robust compression pipelines using quantization, knowledge distillation, and pruning; evaluate trade-offs between size, speed, and accuracy; stay at the cutting edge with new techniques and publish findings. Remote role with a global, async team and edge deployment focus.

