This job is no longer available

Crusoe

Senior Software Engineer, Managed AI

San Francisco, CA, US

HybridFull time roleSenior Level

8 months ago

About the Job

Crusoe is building the World’s Favorite AI-first Cloud infrastructure company. We’re pioneering vertically integrated,  purpose-built AI infrastructure solutions trusted by Fortune 500 companies to power their most advanced AI applications. Crusoe is redefining AI cloud infrastructure, with a mission to align the future of computing with the future of the climate. Our AI platform is recognized as the "gold standard" for reliability and performance. Our data centers are optimized for AI workloads and are powered by clean, renewable energy.

Be part of the AI revolution with sustainable technology at Crusoe. Here, you'll drive meaningful innovation, make a tangible impact, and join a team that’s setting the pace for responsible, transformative cloud infrastructure.

About This Role:

The Crusoe Cloud Managed AI team seeks an ambitious and experienced Senior Software Engineer to join their team. You'll have a pivotal role in shaping the architecture and scalability of our next-generation AI inference platform. You will lead the design and implementation of core systems for our AI services, including resilient fault-tolerant queues, model catalogs, and scheduling mechanisms optimized for cost and performance. This role gives you the opportunity to build and scale infrastructure capable of handling millions of API requests per second across thousands of customers.

From day one, you'll own critical subsystems for managed AI inference, helping to serve large language models (LLMs) to a global audience. As part of a dynamic, fast-growing team, you’ll collaborate cross-functionally, influence the long-term vision of the platform, and contribute to cutting-edge AI technologies. This is a unique opportunity to build a high-performance AI product that will be central to Crusoe's business growth.

What You’ll Be Working On:

  • Design and develop core infrastructure for the Managed AI platform:

    • Lead the design and implementation of resilient fault-tolerant queues, model catalogs, and scheduling mechanisms.

    • Build and optimize systems for high-volume AI inference, handling millions of API requests per second.

    • Develop and maintain core services for AI model deployment and management.

  • Collaborate with cross-functional teams:

    • Work closely with product managers, data scientists, and other engineers to define and deliver AI services.

    • Integrate with other Crusoe Cloud services to provide a seamless user experience.

    • Collaborate with research and development teams to explore and implement new AI technologies.

  • Ensure high availability and performance:

    • Design and implement systems with high availability, low latency, and fault tolerance.

    • Optimize performance across all stages of the AI inference pipeline, including model loading, execution, and response handling.

    • Continuously monitor and improve the performance and reliability of AI services.

What You’ll Bring to the Team:

  • Strong Engineering Foundations:

    • Advanced degree in Computer Science, Engineering, or a related field.

    • Proven experience in distributed systems design and implementation.

    • Expertise in using cloud-based services, such as elastic compute, object storage, virtual private networks, managed databases, etc.

    • Experience with container runtimes (e.g., Kubernetes) and microservices architectures.

    • Experience using REST APIs and common communication protocols, such as gRPC.

    • Demonstrated experience in the software development cycle and familiarity with CI/CD tools.

  • AI/ML Expertise:

    • Experience in Generative AI (Large Language Models, Multimodal).

  • Technical Skills:

    • Proficiency in Golang or Python for large-scale, production-level services. (Preferred)

    • Familiarity with AI infrastructure, including training, inference, and ETL pipelines. (Preferred)

    • Contributions to open-source AI projects such as VLLM or similar frameworks. (Preferred)

    • Performance optimizations on GPU systems and inference frameworks. (Preferred)

  • Soft Skills:

    • Proven track record of delivering early-stage projects under tight deadlines.

    • Strong communication and collaboration skills.

    • Proactive and results-oriented with the ability to work independently and as part of a team.

    • Passion for building high-quality, scalable, and impactful products.

Bonus Points:

  • Experience with AI/ML frameworks such as TensorFlow, PyTorch, or Hugging Face Transformers.

  • Knowledge of machine learning algorithms and data structures.

  • Experience with data streaming and real-time processing technologies.

  • Contributions to open-source projects outside of AI.

Benefits:

  • Hybrid work schedule

  • Industry competitive pay

  • Restricted Stock Units in a fast growing, well-funded technology company

  • Health insurance package options that include HDHP and PPO, vision, and dental for you and your dependents

  • Employer contributions to HSA accounts

  • Paid Parental Leave

  • Paid life insurance, short-term and long-term disability

  • Teladoc

  • 401(k) with a 100% match up to 4% of salary

  • Generous paid time off and holiday schedule

  • Cell phone reimbursement

  • Tuition reimbursement

  • Subscription to the Calm app

  • MetLife Legal

  • Company paid commuter benefit; $50 per pay period

Compensation:

Compensation will be paid in the range of $183,000 - $210,000 base salary. Restricted Stock Units are included in all offers. Compensation to be determined by the applicants knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

About the Company

Crusoe Logo

Crusoe

San Francisco, CA, USA

251-500

<div class="c-scrollbar__hider" role="presentation" data-qa="slack_kit_scrollbar"> <div class="c-scrollbar__child" role="presentation"> <div class="c-virtual_list__scroll_container" role="list" data-qa="slack_kit_list" aria-label="q (direct message, away)"> <div id="1685645130.304729" class="c-virtual_list__item c-virtual_list__item--initial-activeitem" tabindex="0" role="listitem" aria-setsize="-1" data-qa="virtual-list-item" data-item-key="1685645130.304729"> <div class="c-message_kit__background c-message_kit__background--hovered p-message_pane_message__message c-message_kit__message p-message_pane_message__message--last" role="presentation" data-qa="message_container" data-qa-unprocessed="false" data-qa-placeholder="false"> <div class="c-message_kit__hover c-message_kit__hover--hovered" role="document" aria-roledescription="message" data-qa-hover="true"> <div class="c-message_kit__actions c-message_kit__actions--default"> <div class="c-message_kit__gutter"> <div class="c-message_kit__gutter__right" role="presentation" data-qa="message_content"> <div class="c-message_kit__blocks c-message_kit__blocks--rich_text"> <div class="c-message__message_blocks c-message__message_blocks--rich_text" data-qa="message-text"> <div class="p-block_kit_renderer" data-qa="block-kit-renderer"> <div class="p-block_kit_renderer__block_wrapper p-block_kit_renderer__block_wrapper--first"> <div class="p-rich_text_block" dir="auto"> <div class="p-rich_text_section">Crusoe has pioneered infrastructure that taps into stranded energy &mdash; methane being flared or excess production from clean and renewable sources &mdash; to power the compute resources we need to drive our shared progress and reduce its environmental impact.</div> <div class="p-rich_text_section">&nbsp;</div> <div class="p-rich_text_section">In a few short years, data centers will consume more than 10% of the world&rsquo;s electricity. By powering the Crusoe Cloud&trade; platform and Crusoe Data Centers with clean, low-cost and emissions-reducing energy systems, we are fundamentally changing the economics and environmental impact of computation. We created the Crusoe Cloud&trade; platform and Crusoe Data Centers for users who need flexible, reliable solutions for their most demanding workloads, and for organizations looking to make measurable progress toward environmental goals.</div> <div class="p-rich_text_section"><br />There&rsquo;s never been a greater need to make traditional energy production more efficient, and bring more clean and renewable sources of energy online. Crusoe is a catalyst for doing both. While the complete transition to renewable energy will require new technologies, new infrastructure, and time, we&rsquo;re accelerating the process by creating alternative revenue streams and technical solutions for underutilized renewable and clean energy sources.</div> </div> </div> </div> </div> </div> </div> </div> </div> </div> </div> </div> </div> </div> </div>

Similar Jobs

Aidash Logo

Staff Machine Learning Engineer (LiDAR)

Staff Machine Learning Engineer (LiDAR)

  • Aidash
  • Palo Alto, CA, US
  • Hybrid, Remote
  • Full time role

Climate-resilient infrastructure with satellite-powered AI for sustainability and cost efficiency.

11 days ago

Crusoe Logo

Senior Staff Solutions Engineer

Senior Staff Solutions Engineer

  • Crusoe
  • San Francisco, CA, US
  • Hybrid
  • Full time role

Transforming stranded energy into eco-friendly power for data centers, reducing environmental impact significantly.

About 1 month ago

Crusoe Logo

Staff Solutions Engineer

Staff Solutions Engineer

  • Crusoe
  • San Francisco, CA, US
  • Hybrid
  • Full time role

Transforming stranded energy into eco-friendly power for data centers, reducing environmental impact significantly.

About 1 month ago

Crusoe Logo

Senior Solutions Engineer

Senior Solutions Engineer

  • Crusoe
  • San Francisco, CA, US
  • Hybrid
  • Full time role

Transforming stranded energy into eco-friendly power for data centers, reducing environmental impact significantly.

About 1 month ago

Crusoe Logo

Senior Solutions Engineer

Senior Solutions Engineer

  • Crusoe
  • Dublin, D, IE
  • Hybrid, Remote
  • Full time role

Transforming stranded energy into eco-friendly power for data centers, reducing environmental impact significantly.

26 days ago

Plus Logo

Senior/Staff Machine Learning Engineer, Planning

Senior/Staff Machine Learning Engineer, Planning

  • Plus
  • Santa Clara, CA, US
  • In-person, Hybrid
  • Full time role

AI-powered autonomous driving for a safer, greener world.

21 days ago

Zoox Logo

Principal Software Engineer, ML Infrastructure

Principal Software Engineer, ML Infrastructure

  • Zoox
  • Foster City, CA, US
  • Hybrid, Remote
  • Full time role

Pioneering electric autonomous vehicles for low-carbon, congestion-free urban transportation.

20 days ago

Iberdrola Logo

Senior Manager - Data & AI Solutions

Senior Manager - Data & AI Solutions

  • Iberdrola
  • Orange, CT, US
  • Remote
  • Full time role

Accelerating America's clean energy transition with sustainable power and industry-leading ethical practices.

18 days ago

Archer Logo

AI Research Engineer

AI Research Engineer

  • Archer
  • San Jose, CA, US
  • In-person
  • Full time role

Revolutionizing urban transport with climate-friendly electric vertical takeoff and landing aircraft.

13 days ago

Archer Logo

AI Systems Engineer

AI Systems Engineer

  • Archer
  • San Jose, CA, US
  • In-person
  • Full time role

Revolutionizing urban transport with climate-friendly electric vertical takeoff and landing aircraft.

13 days ago