Job Description – Consultant Software Engineer - Python (13+ Years Experience)
We are seeking an experienced Consultant Software Engineer to lead complex engineering initiatives across mission‑critical banking and investment platforms. In this strategic and hands-on role, you will be responsible for designing and building scalable, high‑performance solutions across high‑availability, low‑latency, and risk‑sensitive environments, including pre‑trade, market data, and real‑time processing systems.
The ideal candidate brings 13+ years of strong software engineering experience, with solid expertise in object‑oriented programming (OOP) concepts using languages such as Python, along hands-on Python development experience. The role requires the ability to architect robust systems, contribute to modern engineering practices, and collaborate effectively with cross‑functional teams. A strong production mindset, technical leadership capability, and experience working in controlled enterprise environments with high expectations for stability, security, and performance are essential.
Key Responsibilities
- Design, implement, and scale end‑to‑end automation frameworks, including CI/CD modernization, automated infrastructure provisioning, and operational tooling across business‑critical systems.
- Drive infrastructure automation initiatives to improve reliability, consistency, and resilience across distributed and latency-sensitive platforms.
- Architect monitoring, alerting, and observability solutions for highly available, production‑grade environments, leveraging Python and industry-leading monitoring stacks.
- Define best practices, governance, and engineering standards for DevOps, automation, and operational excellence across global engineering teams.
- Act as a technical consultant to development, infrastructure, platform engineering, and production support teams, guiding them on automation and operational improvements.
- Partner with security, network, and infrastructure teams to ensure compliance with enterprise standards, risk controls, and regulatory requirements.
- Lead incident reviews, stabilize production platforms, and drive root cause analysis with a focus on long-term remediation and operational maturity.
- Oversee production readiness, deployment automation, environment consistency, and configuration management across Unix/Linux ecosystems.
- Manage obsolescence remediation and vulnerability management across infrastructure and application environments in line with enterprise risk and security standards.
- Coordinate and optimize server provisioning, decommissioning, patching cycles, and environment setup activities in line with enterprise hygiene standards.
- Troubleshoot and resolve complex production issues across applications, services, scripts, batch jobs, and market data scripts/pipelines.
- Support and validate firewall, and network flow requirements, ensuring secure and compliant connectivity across source‑to‑destination systems.
- Implement operational controls, audit traceability, and deployment discipline required for financial infrastructures.
- Lead disaster recovery (DR) design validation, failover readiness, and operational resilience exercises.
Required Qualifications
- Bachelor’s degree (BE / B.Tech.) in Computer Science, Information Technology, or a related engineering discipline.
- 10+ years of hands-on programming experience, maintenance and debugging of applications using any of the Object-Oriented programming languages
- 8+ years of experience in Python programming using oops concepts.
- Solid experience in DevOps engineering, infrastructure automation, production activities, and CI/CD in enterprise environments.
- Strong hands-on experience in:
- Unix/Linux systems – Unix/Linux systems, including Unix commands, Shell scripting, System utilities, Server-level troubleshooting, and administration.
- Containerization & Orchestration – Docker, Kubernetes.
- CI/CD Platforms – GitHub Actions, Jenkins, or equivalent.
- Observability & Monitoring – Prometheus, Grafana, Kibana, Elasticsearch, or similar.
- Proven experience with incident management, production issue analysis, and production reliability engineering.
- Strong understanding of production operations in a controlled enterprise environment, including release discipline, operational governance, and platform stability.
- Excellent communication, leadership, and stakeholder‑management skills, with the ability to influence cross‑functional teams across multiple geographies.
- Demonstrated ownership, accountability, and a strong production‑first mindset.
Good to Have
- Experience in banking, trading, investment platforms, or market data systems.
- Knowledge in weekend infrastructure checks, disaster recovery (DR) exercises, failover readiness, and platform resilience activities.
- Experience in cloud platforms (Azure, AWS, GCP) and hybrid cloud automation.