Auros

Senior DevOps Engineer

Hong Kong or Singapore; Remote - Asia, AustraliaFull timeSeniorPosted 20 days ago
Apply on Auros →

Sign into see who you know at Auros.

Auros - Senior DevOps Engineer   Location: Hong Kong OR Remote APAC Type: Full-Time On-Call Requirement: Yes (Rotational Operational Support)   About us: At Auros, we’re dedicated to advancing the cryptocurrency ecosystem through unparalleled liquidity and market-making services. We’re one of the largest participants in the market, trading across 10+ global locations, facilitating 3-4% of global daily volumes, and have connectivity to over 50 venues. What sets us apart, though, is our culture. We believe in hiring smart people and empowering them to do their best work. From day one you’ll have the autonomy and support to really excel. Our relentless focus on delivery drives us to continuously push boundaries and achieve exceptional results, all while offering abundant opportunities for personal and professional growth in the dynamic realm of digital assets.   About the role: We’re looking for a Senior DevOps Engineer to join our globally distributed team and help scale and operate the infrastructure that powers millions of trades daily across CeFi and DeFi venues. You’ll play a mission-critical role in ensuring the performance, reliability, and security of our systems in one of the most demanding environments in tech: digital asset trading. This isn’t just an infrastructure role. You’ll be a key player in building and operating systems where milliseconds matter and downtime is never an option. We need someone who can go beyond automation — someone who understands how low-level performance issues, network latency, and cloud inefficiencies can affect real-world trading outcomes. The position involves participating in a rotational on-call schedule collaborating with team members across Australia, Hong Kong, and Europe, and working proactively to ensure 24/7 uptime and operational excellence.   Key Responsibilities: Ensure the performance and reliability of mission-critical trading infrastructure by proactively identifying and resolving bottlenecks at all layers: compute, network, storage, and application. Support, operate, and continuously improve highly available, low-latency systems under global trading load, with a strong focus on automation and enabling self-service for internal teams. Participate in a day-time rotational on-call schedule, handling real-time operational incidents, root cause analysis, and long-term preventive improvements. Use metrics and observability tools (Prometheus, Grafana, etc.) to detect anomalies, monitor performance trends, and support incident response. Build, manage, and scale infrastructure-as-code using Terraform and configuration management tools like Ansible. Collaborate closely with trading, engineering, and security teams to align infrastructure improvements with application demands. Implement robust security practices, including IAM, threat detection, and secure CI/CD pipelines, to support high-trust production environments. Apply low-level optimization strategies to improve through...