Top Tech Jobs & Startup Jobs in Los Angeles, CA

Reposted 21 Days AgoSaved
In-Office or Remote
USA
271K-425K Annually
Expert/Leader
271K-425K Annually
Expert/Leader
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
The Account CTO will be responsible for strategic technology leadership, defining technical strategies, and driving cloud transformation initiatives at an enterprise level while mentoring technical talent and influencing product strategy based on customer needs.
Top Skills: Ai/Ml TechnologiesDistributed File SystemsHigh-Performance Parallel File SystemsInfinibandLarge-Scale Gpu Cluster DeploymentsNvidia Dgx/Hgx/Mgx SystemsNvme-Based SolutionsObject StorageRoce Networking
Reposted 21 Days AgoSaved
In-Office or Remote
USA
125K-195K Annually
Senior level
125K-195K Annually
Senior level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
The Senior Incident Manager leads incident response for AI infrastructure, coordinating teams to resolve critical incidents, conducting post-incident analysis, and improving operational resilience across systems.
Top Skills: Cloud PlatformsDatadogGpu ClustersGrafanaJIRANetworkingPagerdutyPrometheusServicenow
23 Days AgoSaved
Remote
USA
214K-285K Annually
Mid level
214K-285K Annually
Mid level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Owns technical customer deployments from contract through production readiness. Validates GPU, cloud, compute, storage, networking, and connectivity configurations; troubleshoots issues; coordinates Infrastructure, Engineering, Product, and Data Center teams; tracks risks and dependencies; provides status updates; guides onboarding; documents reusable runbooks; and serves as the customer’s technical contact until the environment is stable.
Top Skills: Cloud PlatformsCompute InfrastructureGpu/Hpc InfrastructureKubernetesLinuxNetworkingStorage
One Month AgoSaved
Remote
USA
122K-162K Annually
Mid level
122K-162K Annually
Mid level
Artificial Intelligence • Cloud • Machine Learning • Infrastructure as a Service (IaaS)
Provide senior-level escalation and troubleshooting for GPU/HPC infrastructure, diagnosing hardware, driver, and kernel issues. Perform root-cause analysis across clusters, build automations and docs, mentor peers, collaborate with engineering for permanent fixes, and participate in on-call rotation and high-volume deployments.
Top Skills: Ci/CdCudaDatadogFirewallsGpudirect RdmaGrafanaInfinibandLinuxNcclNvlinkPrometheusRoceTcp/IpVpn
New

Cut your apply time in half.

Use ourAI Assistantto automatically fill your job applications.

Use For Free
Application Tracker Preview
All Filters
JobType
New Jobs
Job Category
Experience
Industry
Company Name
Company Size

Sign up now Access later

Create Free Account