Responsibilities
- Design scalable, fault-tolerant, and secure streaming architectures
- Monitor, troubleshoot, and performance-tune distributed systems; drive capacity planning
- Implement security controls, automation, and operational best practices
- Support integrations with data lakes, Snowflake, analytics platforms, and downstream systems
- Automate infrastructure provisioning using Terraform or similar IaC tools
- Leverage AI-assisted development tools (Claude or equivalent) to improve engineering productivity and platform automation
- Collaborate with engineering and analytics teams to enable real-time data processing and AI-driven data solutions
Requirements
- 6+ years of experience with Kafka administration and distributed streaming systems
- Strong AWS cloud services experience (MSK, Kinesis, and related services)
- Linux administration and automation/scripting skills
- Python development and platform reliability engineering
- Desire to learn and implement evolving and new technologies
Nice to Have
- Experience with Flink using Spring Boot
- Snowflake integration experience
- Kubernetes orchestration
- Administration knowledge in OpenSearch clusters
- Familiarity with AI-enabled engineering workflows and tools
Work Arrangement
Hybrid — Atlanta, Salt Lake City, Pune