Staff Software Engineer
Brak informacji o wynagrodzeniu
SeniorFull-time
#444693·Dodano dziś·1
Źródło: CiscoTech Stack / Keywords
GoC++PythonBashKubernetesAWSGCPAzureTerraformGitLab
Firma i stanowisko
Cisco is a technology company innovating in data and infrastructure with a focus on security, visibility, and insights across digital footprints. The role is within the Foundations organization and relates to the Splunk Platform, which hosts enterprise software across multiple clouds and on-premises environments.
Wymagania
- 8+ years of professional experience designing and operating large-scale distributed systems or SaaS platform infrastructure.
- Strong proficiency in Go and/or C++ for production systems development, with scripting experience in Python or Bash.
- Proven expertise building Kubernetes-native software including Custom Controllers and Operators.
- Experience deploying resilient applications across major cloud providers (AWS, GCP, Azure).
- Deep understanding of concurrency, data replication, idempotency, partial failure modes, and eventual consistency.
- Proven experience leading technical design reviews, authoring RFCs/design documents, and mentoring teams.
- Excellent English communication skills, both verbal and written, with experience collaborating cross-functionally.
Preferred Qualifications:
- Hands-on experience scaling and running stateful workloads in Kubernetes such as PostgreSQL and distributed NoSQL databases.
- Experience with Infrastructure as Code using Terraform, CI/CD pipelines like GitLab or Jenkins, and service mesh/networking including load balancers, API gateways, DNS, TLS.
- Knowledge of SRE standards, distributed tracing, alerting frameworks, and chaos/recovery testing.
- Experience in fast-paced Agile environments (Scrum/Kanban) driving technical deliverables.
Obowiązki
- Design, develop, and maintain Splunk Platform components across AWS, Google Cloud Platform, Microsoft Azure, and on-premises infrastructure.
- Oversee and participate in designing, implementing, testing, and deploying distributed systems.
- Apply best practices in distributed systems programming including debugging, performance tuning, and reliability engineering.
- Define and improve service health indicators, observability, SLOs, alerting, runbooks, and recovery testing.
- Lead incident diagnosis and learning to reduce operational risk and customer impact.
- Mentor junior and mid-level engineers to promote knowledge sharing and growth.
- Convert roadmap items into clear designs, APIs, milestones, and engineering decisions.
- Lead design reviews and influence technical direction for projects and roadmap.
- Foster continuous learning and adaptability in an agile environment, encouraging innovation and new technology exploration.
Cisco
80 aktywnych ofert