This is a community-generated profile. If you would like to claim it, please .
DiscoverStartupsAlibaba CloudJobs at Alibaba Cloud: Explore current Opportunities
Alibaba Cloud company logo
Alibaba Cloud
Actively Hiring
5000+ Employees

Jobs at Alibaba Cloud

Filter by
Team
Location
Type
Engineering

Site Reliability Engineer

NewPosted 3 days ago

Site Reliability Engineer

  • Developing stability standards and metrics
  • Driving major stability governance campaigns
Engineering

Infra Operations SRE

NewPosted 3 days ago

Infra Operations SRE

  • Perform daily monitoring, health checks, and capacity management of Kubernetes clusters; track node CPU/memory/disk utilization, Pod status, and database/middleware metrics; handle alerts, collect logs to identify risks, and execute scaling or disk cleanup as needed;
  • Respond to monitoring alerts and business feedback; troubleshoot failures of K8s nodes, middlew...
Engineering

Cloud Infrastructure – Site Reliability Engineer (SRE)-Sunnyvale

NewPosted 3 days ago

Cloud Infrastructure – Site Reliability Engineer (SRE)-Sunnyvale

Alibaba Cloud Native Message Middleware Team is responsible for message products, including RocketMQ and other messaging products. We are committed to creating a more stable, user-friendly, streaming, and large-scale messaging platform for the future.

Engineering

Staff SRE-Sunnyvale

NewPosted 6 days ago

Staff SRE-Sunnyvale

  • Build a service platform handling billion-scale computational unit scheduling.
  • Deliver an API system serving millions of developer requests per day.
Other

Infrastructure Technical Program Manager - IT Infrastructure Deployment -Sunnyvale

NewPosted 1 week ago

Infrastructure Technical Program Manager - IT Infrastructure Deployment -Sunnyvale

  • Own the deployment of Alibaba cloud overseas IT infrastructures in server rooms.
  • Partner with several stakeholders such as capacity planning, network, engineering, operation, procurement, supply chain management and finance team to ensure the successful delivery of IT infrastructures.
Engineering

Site Reliability Engineer-Bellevue

NewPosted 2 weeks ago

Site Reliability Engineer-Bellevue

  • Responsible for ensuring the stability of Alibaba Cloud Big Data and PAI products in the US region, including:
  • Service delivery and deployment
Engineering

Senior Security Engineer/Security Expert - Data Security and Business Risk Control-Sunnyvale

NewPosted 3 weeks ago

Senior Security Engineer/Security Expert - Data Security and Business Risk Control-Sunnyvale

  • Responsible for Alibaba Cloud's data compliance operations, including cross-border data transfer and privacy data compliance; handle third-party data compliance certifications and audits, and provide data compliance consulting to customers.
  • Responsible for data security policy management, risk operations, early-warning operations, eme...
Engineering

Senior Security Engineer - Network Security and Product Security-Sunnyvale

NewPosted 3 weeks ago

Senior Security Engineer - Network Security and Product Security-Sunnyvale

  • Responsible for emergency response and incident tracing for security incidents on the Alibaba Cloud platform side, covering both the office network and the production network.
  • Responsible for the detection, emergency response, and tracing of internal and external DDoS attack incidents against Alibaba Cloud; formulate and drive the implementation of mitigation plans; and govern botnet and DDoS issues on the cloud.
Engineering

GEN AI Solutions Architect

NewPosted 3 weeks ago

GEN AI Solutions Architect

Key Responsibilities

● As an AI Solutions Architect (SA) for international business, support the AI MaaS (Model-as-a-Service) sales team in expanding large AI model engagements with global customers, driving increased token consumption of Alibaba’s large models in overseas markets.

● Engage wit...

Engineering

Infra Operations SRE -Sunnyvale

NewPosted 3 weeks ago

Infra Operations SRE -Sunnyvale

  • Perform daily monitoring, health checks, and capacity management of Kubernetes clusters; track node CPU/memory/disk utilization, Pod status, and database/middleware metrics; handle alerts, collect logs to identify risks, and execute scaling or disk cleanup as needed;
  • Respond to monitoring alerts and business feedback; troubleshoot failures of K8s nodes, middlew...
Engineering

Site Reliability Engineer-Sunnyvale

NewPosted 3 weeks ago

Site Reliability Engineer-Sunnyvale

  • Developing stability standards and metrics
  • Covering robust architecture, R&D quality, release management, production environment operations, and more.
Engineering

Office Network Operations and Maintenance Engineer-Sunnyvale

NewPosted 3 weeks ago

Office Network Operations and Maintenance Engineer-Sunnyvale

  • Responsible for the architecture design, implementation, and day-to-day operations and maintenance of the office network, ensuring stable, secure, and efficient network operation.
  • Responsible for the planning, adjustment, and optimization of office network policies, including security policies, network access (admission) policies, and so on.
Engineering

ECS Site Reliability Engineer-Bellevue

NewPosted 3 weeks ago

ECS Site Reliability Engineer-Bellevue

  • Position Highlights
  • Drive the core operations of Alibaba Cloud's Elastic Compute product line, ensuring service stability for global users.
Engineering

Network SRE-Sunnyvale

NewPosted 3 weeks ago

Network SRE-Sunnyvale

  • More than 1 year of experience in network construction and operation and maintenance
  • Familiar with the deployment and operation and maintenance of equipment from mainstream network vendors such as Cisco, Juniper, h3c, and arista
Engineering

Site Reliability Engineering (SRE) Specialist -Bellevue

NewPosted 1 month ago

Site Reliability Engineering (SRE) Specialist -Bellevue

  • Our goal is not only to support enterprises in achieving elastic scalability but also to deeply empower infrastructure innovation in the New era . Our mission is to build an intelligent foundation of "Computing as a Service," enabling developers to focus on businesses to concentrate on breakthroughs, without worrying about the complex engineering implementations from chips to clusters .
Sales

Gen AI Business Development-Sunnyvale, CA

NewPosted 1 month ago

Gen AI Business Development-Sunnyvale, CA

  • Build and own relationships with AI-native companies founders, CTOs, engineers, and product leaders across the U.S.
  • Understand technical architectures and business models of AI-native companies—including LLM fine-tuning, RAG pipelines, AI agent orchestration, computer vision, and generative applications.
Sales

Business Development Manager (GEN AI focus)

NewPosted 1 month ago

Business Development Manager (GEN AI focus)

Alibaba Cloud Global team is looking for Business Development Managers based in Sunnyvale.

Key Responsibilities

● Build and own relationships with AI-native companies founders, CTOs, engineers, and product leaders across the U.

Engineering

RDMA Ops Engineer - Computing Infrastructure Networking-Sunnyvale

NewPosted 2 months ago

RDMA Ops Engineer - Computing Infrastructure Networking-Sunnyvale

  • Deploy, operate and maintain RDMA-based network architectures (RoCE/InfiniBand) for cluster with thousands of nodes
  • Optimize network performance for distributed collective communication workloads (NCCL, MPI, etc.)
Engineering

GenAI Solution Architect -International AI Native (Sunnyvale)

NewPosted 2 months ago

GenAI Solution Architect -International AI Native (Sunnyvale)

  • As an AI Solutions Architect (SA) for international business, support the AI MaaS (Model-as-a-Service) sales team in expanding large AI model engagements with global customers, driving increased token consumption of Alibaba’s large models in overseas markets.
  • Engage with global clients on both business and technical architecture discussions to understan...
Sales

Technical Account Manager-US (Sunnyvale)

NewPosted 2 months ago

Technical Account Manager-US (Sunnyvale)

  • Maintain the stability of the customer's cloud platform through risk governance, product changes and upgrades, and emergency response to faults;
  • Continuously optimize resources and performance to help customers improve cloud efficiency;