CareerCross uses cookies to enhance your experience on our websites. If you continue to view our sites without changing your browser settings, then it is assumed that we have your consent to collect and utilise your cookies. If you do not want to give us your consent, then please change the cookie settings on your browser. Please refer to our privacy policy for more information.
CareerCross uses cookies to enhance your experience on our websites. If you continue to view our sites without changing your browser settings, then it is assumed that we have your consent to collect and utilise your cookies. If you do not want to give us your consent, then please change the cookie settings on your browser. Please refer to our privacy policy for more information.
| Location | Singapore, Singapore |
| Job Type | Permanent Full-time |
| Salary | Negotiable, based on experience |
COMPANY OVERVIEW
The client is an innovative AI Fintech company on a mission to transform how people work by creating proactive digital assistants that do more than just respond to prompts. By combining AI reasoning, long-term context, and automation, they help users complete everyday tasks faster, smarter, and with less manual effort.
JOB RESPONSIBILITIES
Develop and maintain production-grade backend systems supporting AI-powered applications.
Design and optimize inference pipelines, model orchestration frameworks, and service architectures.
Own operational excellence across monitoring, logging, alerting, troubleshooting, and incident management.
Improve system performance through optimization of inference workloads, caching strategies, batching mechanisms, and streaming architectures.
Build robust, scalable APIs that enable seamless integration between frontend applications, machine learning systems, and external services.
Continuously enhance platform reliability, scalability, and observability based on production feedback and usage patterns.
JOB REQUIREMENTS
Strong backend engineering experience building and supporting production-grade applications.
Proven experience designing and operating high-throughput, low-latency distributed systems.
Familiarity with AI/ML serving architectures, including Large Language Models (LLMs), embeddings, and multimodal AI systems.
Experience debugging, monitoring, and optimizing complex distributed systems under production workloads.
Strong understanding of API design, system architecture, and cloud-native development practices.
Demonstrated ownership mindset with the ability to independently drive projects from design through deployment and optimization.
Comfortable working in fast-paced environments and making pragmatic engineering decisions.
TECHNICAL SKILLS
Python
Node.js
PyTorch
OpenAI, Anthropic, and open-source LLM ecosystems
SQL and NoSQL databases
Kubernetes
Docker
Working Location: Singapore
Syahirah Binte Ahmad Ja'afar (R22105331)
JAC Recruitment Pte. Ltd. (90C3026)
#LI-JACSG
Notice: By submitting an application for this position, you acknowledge and consent to the disclosure of your personal information to the Privacy Policy and Terms and Conditions, for the purpose of recruitment and candidate evaluation.
Privacy Policy Link: https://www.jac-recruitment.sg/privacy-policy
Terms and Conditions Link: https://www.jac-recruitment.sg/terms-of-use
| Minimum Experience Level | Over 3 years |
| Career Level | Mid Career |
| Minimum English Level | Native |
| Minimum Japanese Level | None |
| Minimum Education Level | Associate Degree/Diploma |
| Visa Status | No permission to work in Japan required |
| Job Type | Permanent Full-time |
| Salary | Negotiable, based on experience |
| Industry | IT Consulting |