ML Solution Architect (Early Talent)

at Nebius
USD 102-126 per hour
JUNIOR
✅ Remote

Tech Stack

AI @ 3 API AWS Azure Communication @ 6 DevOps Docker FastAPI Flask GCP GPU GenAI Generative AI @ 3 Git Kubernetes LLM @ 3 LangChain MLOps Machine Learning @ 3 Networking Prompt Engineering PyTorch @ 3 Python @ 6 SGLang @ 3 TensorRT @ 3 Vertex AI vLLM @ 3

Details

Nebius is building a full-stack AI cloud platform for the global AI economy, supporting developers and enterprises from data and model training through production deployment. The team works across GPU orchestration, inference optimization, compute, storage, networking, and applied AI.

This is a paid, three-month temporary contract for current university students, recent graduates, and early-career specialists. The role is remote from the United States, and applicants may work remotely from any time zone. Strong performers may be considered for a full-time Solutions Architect position at the end of the program.

Responsibilities

  • Help build and test LLM-based solutions and applications using Token Factory inference services, including multimodal text, vision, and audio models.
  • Assist senior Solutions Architects with prompt engineering, model selection, benchmarking, and inference optimization.
  • Run performance and quality experiments to support proof-of-concept work.
  • Contribute to internal tooling and automation that improves how the Solutions Architect team delivers.
  • Work closely with the backend team while learning how scalable AI applications are built and tuned on the platform.

Requirements

  • Currently pursuing or recently completed a BSc, MSc, or PhD in Computer Science, Machine Learning, or a related field.
  • Strong Python programming skills.
  • Hands-on generative AI experience, including experience with common machine learning frameworks such as PyTorch and Transformers.
  • Strong communication skills and a willingness to explain technical concepts to diverse audiences.

Nice-to-Haves

  • Experience deploying or serving LLMs with vLLM, SGLang, or TensorRT-LLM.
  • Familiarity with inference optimization techniques such as quantization, batching, caching, and routing.
  • Knowledge of model architectures and fine-tuning approaches.
  • Contributions to open-source ML or AI projects.

Preferred Technical Stack

  • Programming languages: Python
  • ML frameworks and libraries: vLLM, SGLang, TensorRT-LLM, Transformers, OpenAI SDKs, Anthropic SDKs
  • Agentic pipeline frameworks: LangChain, LangSmith, smolagents, or equivalent
  • API and web frameworks: FastAPI, Flask
  • MLOps and DevOps tools: Kubernetes, Docker, Git
  • Cloud platforms: AWS, including SageMaker and Bedrock; GCP, including Vertex AI; Azure, including Azure ML

Benefits

  • Competitive compensation and benefits packages.
  • Career growth and learning opportunities.
  • Flexibility and ownership.
  • Collaborative and innovative culture.
  • Opportunity to work on impactful AI projects.
  • International environment and talented teams.
  • For US employees: 100% company-paid medical, dental, and vision coverage for employees and families; a 401(k) plan with up to a 4% company match and immediate vesting; paid parental leave; remote work reimbursement of up to $85 per month for mobile and internet; and company-paid short-term, long-term, and life insurance coverage.

Nebius is an equal opportunity employer committed to an inclusive and diverse workplace. Applicants must be authorized to work in the country in which they apply and must provide proof of employment eligibility as a condition of hire.

More jobs at Nebius

Similar jobs