NextPath Jobs

【SRE】Site Reliability Engineer

FunNow · Taipei City

Posted 2026-09-26 · Verified live 2026-09-26

See how your real experience scores against this role — our AI drafts an honest, verifiable resume tailor. Nothing invented, ever. You approve everything.

Get matched — join the waitlist Apply on company site ↗

About this role

<p>【Capsule】<br>At FunNow, we’re building joyful experiences, at the speed of now. As a Site Reliability Engineer, you’ll play a crucial role in ensuring our platform stays fast, resilient, and secure for millions of users booking spontaneous fun across Asia. But here’s the twist: we don’t just monitor uptime — we build with AI and automation. From Kubernetes tuning to auto-healing infrastructure, CI/CD pipelines to incident response, you'll be hands-on in evolving our DevOps culture. If you love scalable systems, believe in developer efficiency, and treat infrastructure as code, welcome aboard.<br><br></p>

<p>【Typical Accountability】</p>

<ul>

<li>Design robust architectures to comprehensively improve system availability, scalability, and service quality</li>

<li>Ensure stable service operation, monitor core service status, and quickly troubleshoot issues</li>

<li>Conduct in-depth analysis of system performance bottlenecks and propose and implement improvement solutions</li>

<li>Maintain and optimize Kubernetes clusters (EKS/GKE), effectively handling resource pressure, node anomalies, and other situations</li>

<li>Maintain and improve CI/CD pipelines and automated deployment systems (GitHub Actions / ArgoCD) to significantly enhance engineering team development efficiency</li>

<li>Establish and continuously optimize system monitoring and alerting mechanisms (Prometheus / Grafana / Alertmanager)</li>

<li>Assist with incident response and problem investigation</li>

<li>Regularly participate in system inspections and audits, proactively proposing and implementing improvements</li>

<li>Assist in maintaining and implementing fundamental security settings (e.g., IAM, resource permissions, encrypted storage)</li>

<li>Actively share your experience to collectively enhance the team's engineering culture<br></li>

</ul>

<p>【Essential Competencies】</p>

<ul>

<li>Familiarity with container technologies such as Docker or Kubernetes, and practical experience with Kubernetes operations (deployment, scheduling, resource management)</li>

<li>Familiarity with AWS services (e.g., ECS, EKS, S3, CloudFront, IAM, VPC, etc.), and practical experience maintaining AWS or GCP (we primarily use AWS)</li>

<li>Familiarity with at least one CI/CD tool (e.g., GitHub Actions, GitLab CI)</li>

<li>Proficiency in MySQL daily management and performance analysis</li>

<li>Familiarity with service-related log analysis and monitoring tools (e.g., CloudWatch, ELK/EFK, Grafana), and practical experience with Prometheus/Grafana</li>

<li>Experience maintaining Elasticsearch clusters</li>

<li>Familiarity with Git and basic Git flow operations</li>

<li>High degree of self-management, proactive and responsible work attitude, meticulousness, and excellent communication and teamwork skills<br><br></li>

</ul>

<p>【Desirable Competencies】</p>

<ul>

<li>Exposure to or familiarity with the Golang ecosystem</li>

<li>Familiarity with Infra-as-Code tools such as CDK, Terraform</li>

<li>Experience with IPO advisory or ISO audit</li>

<li>Security awareness<br><br></li>

</ul>

<p>【Who You Are】</p>

<ul>

<li>You enjoy solving real-world problems, are proactive in investigation, and act quickly</li>

<li>You value stability and data accuracy, and possess a high sense of responsibility</li>

<li>You are passionate about learning new tools and enjoy sharing improvement methods</li>

<li>You maintain clear communication and good documentation habits in team collaboration</li>

</ul>

One of thousands of fresh listings refreshed nightly, built for cleared & defense careers.

Browse all jobs