Skip to main content
Remote Atlas
ashbyOfficial ATS boardOnsiteSeniorFullTime

Senior Software Engineer, AI Platform

Crusoe

San Francisco, CA - US

Job description

Crusoe is on a mission to accelerate the abundance of energy and intelligence. As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from electrons to tokens — to power the world's most ambitious AI workloads. When you join Crusoe, you join a team that is building the future, faster.

We're in the midst of the greatest industrial revolution of our time. The demand for AI compute is boundless, and power is a bottleneck. We're solving that — with an energy-first approach that makes AI infrastructure better for the world and faster for the people innovating with AI.

We're looking for problem-solving, opportunity-finding teammates with a sense of urgency, who believe in the scale of our ambition and thrive on a path not fully paved — people who want to grow their careers alongside a team of experts across energy, manufacturing, data center construction, and cloud services.

If you want to do the most meaningful work of your career, help our customers and partners advance their AI strategies, and be part of a high-performing team that believes in each other, come build with us at Crusoe.

About the Role:

The Crusoe Cloud Managed AI team seeks an ambitious and experienced Senior Software Engineer to join their team. You'll have a pivotal role in shaping the architecture and scalability of our next-generation managed AI services. You will lead the design and implementation of core systems for our AI offerings, including resilient fault-tolerant queues, model catalogs, and scheduling mechanisms optimized for cost and performance. This role offers the opportunity to build and scale critical infrastructure capable of handling millions of API requests per second across thousands of customers.

The team is working on building the next chapter of products to accelerate the process of adding intelligence into software.

This is an on-site role based in San Francisco, CA, or Sunnyvale, CA, requiring in-office presence..

What You'll Be Working On:

  • Design and Development:

    • Lead the design and implementation of core AI services, including:

    • Resilient fault-tolerant queues for efficient task distribution.

    • Model catalogs for managing and versioning AI models.

    • Scheduling mechanisms optimized for cost and performance.

    Scalability and Reliability:

    • Deep understanding of Kubernetes or other distributed cluster orchestration systems.

    • Architect and scale infrastructure to handle millions of API requests per second.

    • Implement robust monitoring and alerting to ensure system health and 24/7 availability.

    Collaboration and Innovation:

    • Collaborate closely with product management, business strategy, and other engineering teams to define the AI platform roadmap.

    • Influence the long-term vision and architectural decisions of the platform.

    • Contribute to open-source AI frameworks and actively participate in the AI community.

    • Prototype and rapidly iterate on emerging technologies and new features.

What You'll Bring to the Team:

  • Strong Engineering Fundamentals: Advanced degree in Computer Science/Engineering

  • 4-5+ years of industry experience with demonstrated history of consistent success leading a varied portfolio of initiatives across your function

  • Experience with distributed systems, cloud services (compute, storage, networking, database), and delivering early-stage projects quickly.

  • AI/ML Expertise: Experience with Generative AI (LLMs, Multimodal) and familiar with AI infrastructure (training, inference, ETL pipelines).

  • Software Engineering: Proficient with container runtimes (e.g., Kubernetes), microservices, REST APIs, gRPC, and the full software development lifecycle including CI/CD.

Preferred Qualifications:

  • Proficiency in Golang, Python or Rust for production services.

  • Contributions to open-source AI projects (e.g., VLLM).

  • Experience with performance optimizations on GPU systems and inference frameworks.

Personal Attributes:

  • Proactive, collaborative, and capable of working autonomously.

  • Excellent communication and interpersonal skills.

  • Passion for building cutting-edge AI products and solving complex technical problems

Benefits:

  • Competitive compensation and equity packages

  • Restricted Stock Units

  • Paid time off, paid holidays & leave of absence programs

  • Comprehensive health, dental & vision insurance

  • Employer contributions to HSA account

  • Paid parental leave

  • Paid life insurance, short-term and long-term disability

  • Professional development & tuition reimbursement

  • Mental health & wellness support

  • Commuter benefits (parking & transit)

  • Cell phone stipend

  • 401(k) Retirement plan with company match up to 4% of salary

  • Volunteer time off

  • Global travel insurance & emergency assistance

  • Daily meals allowance

  • Additional perks & programs specific to location

Compensation Range

Compensation will be paid in the range of up to $170,000 -$205,000 + Bonus. Restricted Stock Units are included in all offers. Compensation to be determined by the applicant's knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Apply kit

Sign in to copy a field card for the employer’s ATS. We never submit applications for you.

himalayasCurated job boardRemoteSenior

ML Ops Architect

Tiger Analytics

United States

Atlas fit 6

Tiger Analytics is an advanced analytics consulting firm. We are the trusted analytics partner for several Fortune 100 companies, enabling them to generate bus…

  • ai-ml-technical-architect
  • ai/ml
  • cloud
  • data-engineering
  • data-science

Sign in to track applications

Details
himalayasCurated job boardRemoteSenior

Senior Site Reliability Engineer, DGX Cloud

NVIDIA

United States

Atlas fit 5

NVIDIA is driving AI and high-performance computing forward. DGX Cloud aims to deliver a fully managed AI platform on major cloud providers, optimizing AI work…

  • ai/ml
  • cloud
  • cloud-computing
  • devops
  • devops-engineer

Sign in to track applications

Details
himalayasCurated job boardRemoteSenior

Senior Workday Developer

Bright Vision Technologies

United States

Atlas fit 4

Senior Workday Developer - Remote Bright Vision Technologies is a technology consulting and software development company delivering cloud, AI, data, and enterp…

  • ai/ml
  • cloud
  • erp-developer
  • workday-configuration-specialist
  • workday-developer

Sign in to track applications

Details
himalayasCurated job boardRemoteSenior

Senior Product Manager - US-Based

Toptal

United States

Atlas fit 3

About Toptal Toptal is a global network of top talent in business, design, and technology that enables companies to scale their teams, on-demand. With $200+ mi…

  • ai-product-management
  • ai/ml
  • product-development
  • product-management
  • remote-senior-product-manager

Sign in to track applications

Details