fokcus

Member of Technical Staff, Cloud Infrastructure, Singapore

fireworks · Singapore

по договорённости

Опубликовано 29 дн назад

Как распознать мошенников

Настоящий работодатель не просит войти через iCloud или Apple ID, прислать коды из SMS, установить программу для «тестового задания», перевести деньги или оплатить обучение. Если просят — это не работа.

middle
Свежая вакансия Появилась 29 дн назад

ABOUT US

At Fireworks, we’re building the future of generative AI infrastructure. Our platform delivers the highest-quality models with the fastest and most scalable inference in the industry. We’ve been independently benchmarked as the leader in LLM inference speed and are driving cutting-edge innovation through projects like our own function calling and multimodal models. Fireworks is a Series C company valued at $4 billion and backed by top investors including Benchmark, Sequoia, Lightspeed, Index, and Evantic. We’re an ambitious, collaborative team of builders, founded by veterans of Meta PyTorch and Google Vertex AI.

In the last few months alone we launched Fireworks Training, partnered with Microsoft Azure Foundry, and published research straight from our production systems. A few examples of what that looks like in practice:

  • Frontier RL is cheaper than the mega-cluster narrative suggests: we ran cross-region rollouts using 98% sparse weight deltas and published what we learned. (blog https://fireworks.ai/blog/frontier-rl-is-cheaper-than-you-think)
  • Open source agents with frontier advisors: matching frontier performance through training and harness engineering. (blog https://fireworks.ai/blog/open-source-agents-frontier-advisors)
  • The fine-tuning bottleneck is not the algorithm: integration friction and iteration speed are what actually stall teams; we documented the patterns across dozens of customer engagements. (blog) https://fireworks.ai/blog/fine-tuning-bottlenecks

ABOUT US

Here at Fireworks, we’re building the future of generative AI infrastructure. Fireworks offers the generative AI platform with the highest-quality models and the fastest, most scalable inference. We’ve been independently benchmarked to have the fastest LLM inference and have been getting great traction with innovative research projects, like our own function calling and multi-modal models. Fireworks is funded by top investors, like Benchmark and Sequoia, and we’re an ambitious, fun team composed primarily of veterans from Pytorch and Google Vertex AI.

THE ROLE

As a Backend Software Engineer, you’ll be responsible for designing and developing the core backend systems that power Fireworks AI’s high-performance generative AI platform. Your work will focus on ensuring efficiency, scalability, and stability in handling AI workloads.

KEY RESPONSIBILITIES

  • Design and build core backend software components, ensuring efficiency, scalability, and stability of system resources
  • Conduct design/code reviews and collaborate with cross-functional teams
  • Continuously analyze and optimize infrastructure efficiency for AI workloads (compute, storage, networking)

MINIMUM QUALIFICATIONS

  • Bachelor’s degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience
  • 3+ years experience working in ML infra (PyTorch, Vertex AI, Sagemaker, etc.)
  • Experience building, scaling, and optimizing enterprise-grade Machine Learning systems

PREFERRED QUALIFICATIONS

  • Experience in AI or large scale infrastructure
  • Master’s or PhD degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience

WHY FIREWORKS AI?

  • Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.
  • Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.
  • Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI—no bureaucracy, just results.
  • Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

WHY FIREWORKS AI?

  • Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.
  • Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.
  • Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI—no bureaucracy, just results.
  • Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.

Навыки

Показать как в источнике
Member of Technical Staff, Cloud Infrastructure, Singapore

ABOUT US:

At Fireworks, we’re building the future of generative AI infrastructure. Our platform delivers the highest-quality models with the fastest and most scalable inference in the industry. We’ve been independently benchmarked as the leader in LLM inference speed and are driving cutting-edge innovation through projects like our own function calling and multimodal models. Fireworks is a Series C company valued at $4 billion and backed by top investors including Benchmark, Sequoia, Lightspeed, Index, and Evantic. We’re an ambitious, collaborative team of builders, founded by veterans of Meta PyTorch and Google Vertex AI.

In the last few months alone we launched Fireworks Training, partnered with Microsoft Azure Foundry, and published research straight from our production systems. A few examples of what that looks like in practice:

 - Frontier RL is cheaper than the mega-cluster narrative suggests: we ran cross-region rollouts using 98% sparse weight deltas and published what we learned. (blog https://fireworks.ai/blog/frontier-rl-is-cheaper-than-you-think)

 - Open source agents with frontier advisors: matching frontier performance through training and harness engineering. (blog https://fireworks.ai/blog/open-source-agents-frontier-advisors)

 - The fine-tuning bottleneck is not the algorithm: integration friction and iteration speed are what actually stall teams; we documented the patterns across dozens of customer engagements. (blog) https://fireworks.ai/blog/fine-tuning-bottlenecks


ABOUT US:

Here at Fireworks, we’re building the future of generative AI infrastructure. Fireworks offers the generative AI platform with the highest-quality models and the fastest, most scalable inference. We’ve been independently benchmarked to have the fastest LLM inference and have been getting great traction with innovative research projects, like our own function calling and multi-modal models. Fireworks is funded by top investors, like Benchmark and Sequoia, and we’re an ambitious, fun team composed primarily of veterans from Pytorch and Google Vertex AI.


THE ROLE:

As a Backend Software Engineer, you’ll be responsible for designing and developing the core backend systems that power Fireworks AI’s high-performance generative AI platform. Your work will focus on ensuring efficiency, scalability, and stability in handling AI workloads.


KEY RESPONSIBILITIES:

 - Design and build core backend software components, ensuring efficiency, scalability, and stability of system resources

 - Conduct design/code reviews and collaborate with cross-functional teams

 - Continuously analyze and optimize infrastructure efficiency for AI workloads (compute, storage, networking)


MINIMUM QUALIFICATIONS:

 - Bachelor’s degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience

 - 3+ years experience working in ML infra (PyTorch, Vertex AI, Sagemaker, etc.)

 - Experience building, scaling, and optimizing enterprise-grade Machine Learning systems


PREFERRED QUALIFICATIONS:

 - Experience in AI or large scale infrastructure

 - Master’s or PhD degree in Computer Science, Computer Engineering, relevant technical field, or equivalent practical experience


WHY FIREWORKS AI?

 - Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.

 - Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.

 - Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI—no bureaucracy, just results.

 - Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

 

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.


WHY FIREWORKS AI?

 - Solve Hard Problems: Tackle challenges at the forefront of AI infrastructure, from low-latency inference to scalable model serving.

 - Build What’s Next: Work with bleeding-edge technology that impacts how businesses and developers harness AI globally.

 - Ownership & Impact: Join a fast-growing, passionate team where your work directly shapes the future of AI—no bureaucracy, just results.

 - Learn from the Best: Collaborate with world-class engineers and AI researchers who thrive on curiosity and innovation.

Fireworks AI is an equal-opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all innovators.