Develop custom software solutions to design, code, and enhance components across systems or applications. Use modern frameworks and agile practices to deliver scalable, high-performing solutions tailored to specific business needs.
Summary:
We are seeking a motivated and technically proficient Custom Software Engineering Specialist to support, enhance, and operate enterprise AI-powered platforms and cloud-native applications. This role combines Python software development, AI agent deployment, cloud engineering, observability, and Level 3 application support responsibilities.
Roles & Responsibilities:
- Develop, deploy, and maintain AI agents and automation solutions using Python, LiteLLM, Vertex AI, and related technologies
- Build and enhance orchestration layers that improve platform security, resiliency, and operational efficiency.
- Create intelligent bots and automation capabilities to streamline incident triage and operational workflows.
- Design, develop, and enhance backend services, APIs, dashboards, and platform integrations supporting enterprise AI solutions
- Support cloud-native applications running primarily on Google Cloud Platform (GCP) and Kubernetes environments
- Ensure platform scalability, availability, and performance through effective container management, deployment strategies, and pod autoscaling
- Design and maintain monitoring dashboards, alerts, and observability solutions to proactively identify and resolve issues
- Serve as a Level 3 support engineer for complex application and platform incidents, including root cause analysis and remediation
- Improve operational excellence through continuous monitoring enhancements, automation, and reliability initiatives
- Collaborate directly with clients, architects, developers, and stakeholders to understand requirements and deliver technical solutions
- Communicate technical concepts effectively to both technical and non-technical audiences while driving successful project outcomes
Professional & Technical Skills:
- Strong proficiency in Python development, API integration, automation, and backend application development
- Experience with AI/LLM platforms such as LiteLLM, Vertex AI, or similar AI orchestration frameworks
- Understanding of AI agents, prompt engineering, model integration, and AI operations concepts
- Hands-on experience with Google Cloud Platform (GCP), including cloud-native application deployment and support
- Strong knowledge of Kubernetes administration, container orchestration, pod scaling, deployments, services, and troubleshooting
- Experience with Docker and modern cloud infrastructure practices
- Experience with observability and monitoring platforms such as Splunk, Splunk Observability Cloud, and Grafana
- Familiarity with ServiceNow for incident management and JIRA for Agile delivery and work management
- Understanding of production support, incident response, root cause analysis, and operational best practices
Additional Information:
- The candidate should have minimum 3 years of experience in Generative AI.
- This position is based at our Manila office.
- This is a night shift role
Minimum 3 year(s) of experience is required