About the Role
About the Role
The Platform Engineer / Distributed Systems Engineer role involves designing and developing a scalable observability platform that integrates AI agents for real-time data analysis. The position requires strong programming skills and experience in building distributed systems and analytics pipelines.
What You'll Work On
- Design and develop a next-generation scalable observability platform for modern cloud-native and hybrid infrastructures that works in tandem with AI agents.
- Create intelligent AI agents to analyze logs, traces, and metrics in real time, delivering automated insights and remediation.
- Build scalable and fault tolerant AI agent frameworks.
- Engineer and optimize large-scale analytics pipelines to process high-velocity telemetry data.
- Build resilient distributed systems with high reliability, performance, and fault tolerance.
- Implement and fine-tune LLMs for natural language querying and automated troubleshooting.
- Partner with ML engineers to streamline AI model deployment and management.
Who You Are
- Strong programming skills in C++ and Golang (experience with Rust is a plus).
- Track record of building distributed systems and large-scale analytics pipelines.
- Hands-on experience with cloud infrastructure (AWS, GCP, or Azure) and Kubernetes.
- Deep understanding of observability technologies (Prometheus, OpenTelemetry, Grafana, Elastic, etc.).
- Knowledge of LLMs, AI agents, agent frameworks like langchain, autogen is a plus.
- Experience with stream processing and real-time data processing frameworks.
- Proficiency in database technologies (SQL & NoSQL, Clickhouse, Time-Series DBs).
- 7+ years of relevant experience.
- Bachelor's degree in Computer Science, Engineering, or related field (Master's/PhD is a plus).
How to Apply
Step 1: Click On Apply! And Register or Login on our portal.
Step 2: Complete the Screening Form & Upload updated Resume.
Step 3: Increase your chances to get shortlisted & meet the client for the Interview!
About Uplers
Our goal is to make hiring reliable, simple, and fast. Our role will be to help all our talents find and apply for relevant contractual onsite opportunities and progress in their career. We will support any grievances or challenges you may face during the engagement.
So, if you are ready for a new challenge, a great work environment, and an opportunity to take your career to the next level, don't hesitate to apply today. We are waiting for you!
Requirements
Data Streaming
Experience with real-time data processing frameworks.
Distributed Systems
Proven track record of building distributed systems.
Programming Skills
Strong skills in C++ and Golang, with Rust as a plus.
Cloud Infrastructure
Hands-on experience with AWS, GCP, or Azure.
Observability Technologies
Deep understanding of tools like Prometheus and Grafana.
Nice to Have
Familiarity with Large Language Models and AI agents.
Experience with frameworks like langchain and autogen.