Associate, Software Production Management & Reliability Engineering
Hong KongPosted Jul 6, 2026
Role Overview
The Real-Time Market Data (RTMD) team is responsible for operating and supporting the firm's mission-critical market data infrastructure that powers electronic trading, risk management, and front-office applications globally.
This position is primarily supports North American trading and 24-hour market data platforms, while providing secondary support for Asia-Pacific exchanges and market data services.
The successful candidate will be responsible for proactively monitoring real-time market data systems, managing production incidents, supporting business users, and driving operational excellence through AI-enabled workflows and automation. The ideal candidate possesses a strong ownership mentality, excellent communication skills, and the ability to identify and resolve issues before they impact users or trading activities.
Key Responsibilities
* Perform proactive real-time monitoring of mission-critical market data systems supporting global electronic trading.
* Detect potential issues before they become production incidents and take preventative actions to minimize business impact.
* Act as the primary operational interface for Trading, Technology, Infrastructure teams, and external vendors.
* Investigate, troubleshoot, and resolve production issues independently with minimal escalation to Engineering teams.
* Provide technical support for real-time market data services, applications, and distribution platforms.
* Manage production incidents from identification through resolution and post-incident review.
* Leverage AI tools and technologies to improve troubleshooting, operational efficiency, documentation, knowledge management, and incident response.
* Identify automation opportunities and drive operational improvement initiatives.
* Support exchange initiatives, market structure changes, weekend testing, production deployments, and disaster recovery exercises.
* Develop and maintain operational procedures, runbooks, monitoring standards, and knowledge repositories.
* Continuously improve platform reliability, operational controls, and service quality.
Skills Required
* Bachelor's degree in Computer Science, Engineering, Information Technology, or equivalent.
* 3+ years of experience in Production Management, Application Support, Technical Operations, Site Reliability Engineering (SRE), or related disciplines.
* Strong experience supporting mission-critical production environments.
* Demonstrated experience utilizing AI tools, AI-assisted workflows, Generative AI platforms, or operational automation solutions to improve productivity and operational effectiveness.
* Strong knowledge of Unix/Linux operating systems and troubleshooting.
Good understanding of networking fundamentals:TCP/IP,,UDPMulticast
* Network diagnostics and troubleshooting
* Experience supporting end users and managing production incidents in a time-sensitive environment.
* Excellent verbal and written communication skills.
* Strong analytical skills with a proactive and ownership-driven mindset.
* Ability to effectively manage multiple priorities in a fast-paced trading environment.
Skills Desired
* Perl and or Python scripting experience.
* Experience with market data systems, exchange connectivity, or electronic trading platforms.
* Knowledge of monitoring and observability platforms such as Splunk, Grafana, Geneos, or similar tools.
* Experience developing AI-powered operational tools, bots, agents, or workflow automation solutions.
WHAT YOU CAN EXPECT FROM MORGAN STANLEY:
At Morgan Stanley, we raise, manage and allocate capital for our clients – helping them reach their goals. We do it in a way that’s differentiated – and we’ve done that for 90 years. Our values - putting clients first, doing the right thing, leading with exceptional ideas, committing to diversity and inclusion, and giving back - aren’t just beliefs, they guide the decisions we make every day to do what's...