Let’s get started
By clicking ‘Next’, I agree to the Terms of Service
and Privacy Policy
Jobs / Job page
Staff Software Engineer, Managed AI image - Rise Careers
Job details

Staff Software Engineer, Managed AI

Crusoe is building the World’s Favorite AI-first Cloud infrastructure company. We’re pioneering vertically integrated,  purpose-built AI infrastructure solutions trusted by Fortune 500 companies to power their most advanced AI applications. Crusoe is redefining AI cloud infrastructure, with a mission to align the future of computing with the future of the climate. Our AI platform is recognized as the "gold standard" for reliability and performance. Our data centers are optimized for AI workloads and are powered by clean, renewable energy.

Be part of the AI revolution with sustainable technology at Crusoe. Here, you'll drive meaningful innovation, make a tangible impact, and join a team that’s setting the pace for responsible, transformative cloud infrastructure.

About the Role

As a Staff Software Engineer on the Managed AI team at Crusoe, you'll have a pivotal role in shaping the architecture and scalability of our next-generation AI inference platform. You will lead the design and implementation of core systems for our AI services, including resilient fault-tolerant queues, model catalogs, and scheduling mechanisms optimized for cost and performance. This role gives you the opportunity to build and scale infrastructure capable of handling millions of API requests per second across thousands of customers.

From day one, you'll own critical subsystems for managed AI inference, helping to serve large language models (LLMs) to a global audience. As part of a dynamic, fast-growing team, you’ll collaborate cross-functionally, influence the long-term vision of the platform, and contribute to cutting-edge AI technologies. This is a unique opportunity to build a high-performance AI product that will be central to Crusoe's business growth.

A Day In the Life

As a Staff Software Engineer in the Managed AI team, you’ll play a crucial role in building the infrastructure to serve artificial neural networks and in the near term, large language models (LLMs) at scale. You’ll own the design and implementation of key subsystems for resiliency and quality of service. You will build model catalogs, billing systems, dynamic pricing models, and have the opportunity of going deep into the model deployment stack for cost-optimized scheduling. Each day, you’ll collaborate with a small but growing team of engineers to build scalable cloud-based solutions that can handle millions of requests per second.

You’ll work closely with cross-functional teams, including product management and business strategy, to develop a customer-facing API that serves real-world AI models. Every day will present an opportunity to influence the long-term vision and architectural decisions, from the first lines of code to full-scale implementation. You’ll also be prototyping rapidly, optimizing performance on GPUs, and ensuring high availability as part of the MVP development. Whether it’s contributing to open-source AI frameworks or diving into low-level performance optimizations, your contributions will directly impact both the company’s growth and the product’s success.

You Will Thrive In This Role If You Have:

  • You have a strong background in distributed systems design and implementation, with proven experience in early-stage projects and tight deadlines.

  • You are passionate about building scalable AI infrastructure and have experience with cloud-based services that can handle millions of requests.

  • You enjoy problem-solving around performance optimizations, particularly when it comes to AI inference on GPU-based systems.

  • You have a proactive and collaborative approach, with the ability to work autonomously while engaging with a rapidly growing team.

  • You have strong communication skills, both written and verbal, and can translate complex technical challenges into understandable terms for cross-functional teams.

  • You’re excited about working in a fast-paced environment, contributing to a new product category, and having a tangible influence on the long-term vision of the AI platform.

  • You are passionate about open-source contributions and AI inference frameworks like VLLM, with a desire to push the boundaries of performance and scalability.

  • You are keen on customer-facing product development, with a desire to build user-friendly APIs that integrate real-world feedback for continuous improvement.

Preferred Qualifications

  • Must-Have:

    • Advanced degree in Computer Science, Engineering, or a related field.

    • Demonstrable experience in distributed systems design and implementation.

    • Proven track record of delivering early-stage projects under tight deadlines.

    • Expertise in using cloud-based services, such as, elastic compute, object storage, virtual private networks, managed database, etc

    • Experience in Generative AI (Large Language Models, Multimodal).

    • Experience with container runtimes (e.g., Kubernetes) and microservices architectures.

    • Experience using REST APIs and common communication protocols, such as gRPC.

    • Demonstrated experience in the software development cycle and familiarity with CI/CD tools.

  • Nice-to-Have:

    • Proficiency in Golang or Python for large-scale, production-level services.

    • Familiarity with AI infrastructure, including training, inference, and ETL pipelines.

    • Contributions to open-source AI projects such as VLLM or similar frameworks.

    • Performance optimizations on GPU systems and inference frameworks.

Growth Opportunities

  • Shape the foundation of a cutting-edge, customer-facing AI inference platform.

  • Become a technical leader in performance optimization and AI infrastructure.

  • Collaborate with partners like Intel and NVIDIA on pushing the limits of AI performance.

  • Contribute to open-source AI frameworks and gain visibility in the AI community.

  • Take on leadership roles as the team scales, with opportunities to mentor junior engineers and influence the product roadmap.

Benefits

  • Hybrid work schedule

  • Industry competitive pay

  • Restricted Stock Units in a fast growing, well-funded technology company

  • Health insurance package options that include HDHP and PPO, vision, and dental for you and your dependents

  • Employer contributions to HSA accounts 

  • Paid Parental Leave 

  • Paid life insurance, short-term and long-term disability 

  • Teladoc 

  • 401(k) with a 100% match up to 4% of salary

  • Generous paid time off and holiday schedule

  • Cell phone reimbursement

  • Tuition reimbursement

  • Subscription to the Calm app

  • MetLife Legal

  • Company paid commuter benefit; $50 per pay period

Compensation Range

Compensation will be paid up to $250,000 base salary. Restricted Stock Units are included in all offers. Compensation to be determined by the applicants knowledge, education, and abilities, as well as internal equity and alignment with market data.

Crusoe is an Equal Opportunity Employer. Employment decisions are made without regard to race, color, religion, disability, genetic information, pregnancy, citizenship, marital status, sex/gender, sexual preference/ orientation, gender identity, age, veteran status, national origin, or any other status protected by law or regulation.

Crusoe Glassdoor Company Review
3.4 Glassdoor star iconGlassdoor star iconGlassdoor star icon Glassdoor star icon Glassdoor star icon
Crusoe DE&I Review
No rating Glassdoor star iconGlassdoor star iconGlassdoor star iconGlassdoor star iconGlassdoor star icon
CEO of Crusoe
Crusoe CEO photo
Chase Lochmiller
Approve of CEO

Average salary estimate

$225000 / YEARLY (est.)
min
max
$200000K
$250000K

If an employer mentions a salary or salary range on their job, we display it as an "Employer Estimate". If a job has no salary data, Rise displays an estimate if available.

What You Should Know About Staff Software Engineer, Managed AI, Crusoe

At Crusoe, we're not just an AI-first cloud infrastructure company; we're on a mission to revolutionize the way computing meets sustainability. As a Staff Software Engineer on our Managed AI team, you will play a crucial role in sculpting the architecture of our next-generation AI inference platform. Imagine designing resilient, fault-tolerant systems that power Fortune 500 companies' most sophisticated AI applications while ensuring that they are also eco-friendly. Your responsibilities will span the implementation of core services—think intricate model catalogs and efficient scheduling mechanisms that can handle millions of API requests per second. You’ll collaborate with a passionate group of engineers while influencing the platform's long-term vision. This isn't just a job; it's an opportunity to make a tangible impact through innovative technology. Your day-to-day will involve collaborating across teams—from product management to business strategy—to deliver user-friendly APIs, rapid prototyping, and autonomous development in a fast-paced environment. If you're excited about contributing to cutting-edge AI technologies and shaping a greener future, then this role is your chance to shine at Crusoe.

Frequently Asked Questions (FAQs) for Staff Software Engineer, Managed AI Role at Crusoe
What are the primary responsibilities of a Staff Software Engineer at Crusoe?

As a Staff Software Engineer at Crusoe, your main responsibilities will include designing and implementing core systems for our AI services. This encompasses building fault-tolerant queues, developing model catalogs, and optimizing scheduling mechanisms for cost and performance across our AI inference platform.

Join Rise to see the full answer
What qualifications are needed to become a Staff Software Engineer on the Managed AI team at Crusoe?

To become a Staff Software Engineer at Crusoe, you should have an advanced degree in Computer Science or a related field, along with proven experience in distributed systems design and implementation. Familiarity with cloud services, container runtimes, and experience in AI inference frameworks are essential.

Join Rise to see the full answer
What can I expect from the work culture as a Staff Software Engineer at Crusoe?

At Crusoe, the work culture is fast-paced and innovative. You will thrive in an environment that promotes collaboration across cross-functional teams, allowing you to influence the long-term vision of AI technologies while contributing to a stimulating and sustainable community.

Join Rise to see the full answer
How does Crusoe support professional growth for Staff Software Engineers?

Crusoe emphasizes growth opportunities through leadership roles, mentoring junior engineers, and collaboration with industry partners like Intel and NVIDIA, allowing you to shape the future of AI infrastructure while enhancing your technical expertise.

Join Rise to see the full answer
What unique challenges does a Staff Software Engineer face at Crusoe?

The unique challenges for a Staff Software Engineer at Crusoe involve building scalable infrastructure capable of handling vast amounts of API requests, optimizing AI inference performance, and ensuring high availability in a rapidly evolving technological landscape.

Join Rise to see the full answer
Common Interview Questions for Staff Software Engineer, Managed AI
Can you describe a recent project where you designed a distributed system?

In answering this question, you should focus on your role in the project, the challenges faced, and the technologies used. Highlight your problem-solving approach and how you ensured that the system was both scalable and fault-tolerant.

Join Rise to see the full answer
How do you approach performance optimization in AI systems?

Discuss your methods for performance optimization, including any specific tools and techniques you've used. Provide examples of how you've previously achieved significant performance improvements, particularly in GPU-based systems.

Join Rise to see the full answer
What experience do you have with cloud infrastructure services?

Share detailed examples of cloud services you've worked with and their specific applications. Mention any experience you have with elastic computing, object storage, or managed databases, emphasizing your hands-on experience and contributions.

Join Rise to see the full answer
How do you ensure the reliability of the services you build?

Talk about implementing rigorous testing protocols, monitoring systems, and failure recovery strategies. Provide examples of how you’ve created resilient systems and any challenges you overcame in achieving reliability.

Join Rise to see the full answer
What is your experience with microservices architectures?

Discuss your familiarity with microservices, how you have implemented them in past projects, and the benefits and challenges associated with microservices in building scalable applications.

Join Rise to see the full answer
How do you handle working in cross-functional teams?

Share your strategies for effective communication and collaboration with other teams, illustrating with examples where you successfully navigated these interactions to achieve project goals.

Join Rise to see the full answer
Can you give an example of how you've contributed to open-source AI projects?

Detail your involvement in open-source projects, your contributions, and the impact your work made on the community or industry. Highlight any specific frameworks you’ve contributed to.

Join Rise to see the full answer
How do you keep up with advancements in AI technology?

Discuss your methods for staying informed about AI advancements, including attending conferences, joining relevant forums, or working on personal projects related to new technologies. Share examples of how this knowledge has influenced your work.

Join Rise to see the full answer
Describe a time when you had to meet tight deadlines with an early-stage project.

Talk about a specific instance where you successfully prioritized tasks, collaborated with your team, and delivered quality work despite time constraints. Highlight the skills that helped you in that situation.

Join Rise to see the full answer
What are your expectations for the next stage of your career?

Be honest yet ambitious, outlining your goals for leadership, contributions to innovative projects, and your desire to influence product development within a company that aligns with your interests, like Crusoe.

Join Rise to see the full answer
Similar Jobs
Photo of the Rise User
Posted 7 days ago
Photo of the Rise User
Posted 7 days ago
Photo of the Rise User
Anduril Industries Hybrid Lexington, Massachusetts, United States
Posted 11 hours ago
Photo of the Rise User
Kaseya Careers Hybrid Miami, Florida, United States
Posted 13 days ago
Photo of the Rise User
Posted 12 days ago

We’re on a mission to align the future of computation with the future of the climate.

173 jobs
MATCH
Calculating your matching score...
FUNDING
SENIORITY LEVEL REQUIREMENT
TEAM SIZE
EMPLOYMENT TYPE
Full-time, hybrid
DATE POSTED
January 7, 2025

Subscribe to Rise newsletter

Risa star 🔮 Hi, I'm Risa! Your AI
Career Copilot
Want to see a list of jobs tailored to
you, just ask me below!