GPU/AI Application Platform Architect - San Jose

TikTok
San Jose
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureEducation: mastersSkills: ["Self-motivated","Cross-functional collaboration","Systematic evaluation","Architecture exploration"]

Architect and design GPU/AI application platform and server/storage systems for high-performance, low-cost, and easy-to-operate infrastructure. Track and evaluate new GPU/AI LLM technologies from partners, then integrate them through platform customization and architecture explorations to improve performance and cost (Perf/TCO/TCO). Study and implement new solutions, evaluate performance on state-of-the-art LLM workloads, and collaborate with standards bodies, suppliers, and cross-functional teams, including international travel up to four times per year.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
TikTok
TikTok
1 month ago

GPU/AI Application Platform Architect - San Jose

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 23 hours agoStatus: Live

Job Summary

Architect and design GPU/AI application platform and server/storage systems for high-performance, low-cost, and easy-to-operate infrastructure. Track and evaluate new GPU/AI LLM technologies from partners, then integrate them through platform customization and architecture explorations to improve performance and cost (Perf/TCO/TCO). Study and implement new solutions, evaluate performance on state-of-the-art LLM workloads, and collaborate with standards bodies, suppliers, and cross-functional teams, including international travel up to four times per year.
Location: San Jose
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Track GPU/AI LLM technology from industry and partner vendors, evaluate/test new parts or technologies, and integrate them into the system.
  • •Drive GPU/AI LLM platform customization using application performance optimizations and architecture explorations to improve Perf/TCO and/or reduce TCO.
  • •Study and implement GPU/AI LLM new technology solutions.
  • •Evaluate GPU system performance on state-of-the-art LLM applications.
  • •Collaborate with industry consortiums and open standard committees; work with partners/suppliers to set up POCs or prototypes for evaluation and testing.
Travel: Low travel

Key Requirements

  • •Master’s degree or higher in Electrical Engineering, Computer Engineering, Computer Science, or related majors.
  • •Deep understanding of computer system architecture, especially GPU/AI SoC or platform architecture, interconnect fabric, and memory subsystem.
  • •Experience with GPU/AI system application performance optimization or software-hardware co-design.
  • •Understanding of LLM model architecture and training/inference requirements on accelerator, memory, and network.
  • •Understand GPU/AI virtualization technology, deep learning architecture, and distributed systems.
Experience:GPU/AILLM platform architectureSoftware-hardware co-designDistributed systems
Education:Master's in Electrical Engineering, Computer Engineering, Computer Science or related majors
Skills:Self-motivatedCross-functional collaborationSystematic evaluationArchitecture exploration
Tech Stack:GPUAILLMGPU/AI SoCPlatform architectureInterconnect fabricMemory subsystemGPU/AI virtualizationDeep learningDistributed system

Company Brief

TikTok
Short-form video platform that lets users create, share, and discover entertainment content through algorithmic recommendations. It also offers advertising and creator tools for brands, influencers, and businesses.
Industry: Digital Media
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Headquarters: Singapore, Singapore
Founded: 2016
WebsiteLinkedIn