💼 Join Dialpad and Shape the Future of AI-Driven Business Communications
Are you passionate about Cloud Infrastructure, Observability Engineering, Kubernetes, Monitoring Platforms, and building highly scalable systems? Do you enjoy solving complex engineering challenges while working on cutting-edge technologies that power millions of business conversations worldwide?
Dialpad is actively hiring a Software Engineer – Observability for its Bengaluru office. This is an exciting opportunity to join one of the world’s fastest-growing AI-native business communications platforms and contribute to the future of cloud infrastructure, monitoring, logging, and reliability engineering.
Dialpad serves more than 70,000 organizations globally, including industry-leading companies such as WeWork, Asana, NASDAQ, Uber, AAA Insurance, COMPASS Realty, Randstad, and Tractor Supply. The company is transforming how businesses communicate through AI-powered calling, messaging, meetings, and contact center solutions.
With its groundbreaking DAART (Dialpad Agentic AI in Real Time) initiative, Dialpad is leading the next wave of intelligent automation by creating AI agents capable of understanding conversations, automating workflows, resolving customer issues, and driving business outcomes in real time.
If you’re excited about observability platforms, cloud-native infrastructure, distributed systems, and helping engineering teams build reliable services at scale, this opportunity could be the perfect next step in your career.
📌 Job Overview
🏢 Company: Dialpad
💼 Position: Software Engineer – Observability
📍 Location: Bengaluru, India
🎓 Qualification: Bachelor’s Degree in Computer Science or Engineering
🏢 Work Mode: On-Site
🕒 Employment Type: Full-Time
☁️ Department: Cloud & Infrastructure Operations
🚀 Industry: AI, Cloud Computing & Business Communications
🔍 About the Role
As a Software Engineer – Observability, you will play a critical role in building and maintaining Dialpad’s monitoring, logging, metrics, and tracing infrastructure.
You will work closely with infrastructure teams, software engineers, platform teams, and cloud specialists to ensure systems remain reliable, scalable, measurable, and highly available.
The role involves creating tools, libraries, automation solutions, and documentation that enable engineering teams to instrument applications effectively and gain deep visibility into system performance.
This position is ideal for engineers who enjoy cloud infrastructure, distributed systems, site reliability engineering, DevOps practices, and modern observability platforms.
🔧 Key Responsibilities
Monitoring & Observability
✔️ Develop monitoring instrumentation
✔️ Improve logging frameworks
✔️ Enhance metrics collection systems
✔️ Support distributed tracing initiatives
Platform Engineering
✔️ Maintain observability platforms
✔️ Optimize monitoring infrastructure
✔️ Improve system visibility
✔️ Ensure platform reliability
Engineering Enablement
✔️ Build tools and libraries
✔️ Support self-monitoring capabilities
✔️ Create reusable engineering solutions
✔️ Improve developer productivity
Best Practices & Standards
✔️ Define observability standards
✔️ Promote monitoring best practices
✔️ Improve service measurability
✔️ Drive reliability initiatives
Documentation
✔️ Create technical documentation
✔️ Maintain internal knowledge bases
✔️ Develop implementation guides
✔️ Support engineering onboarding
Innovation & Research
✔️ Explore emerging technologies
✔️ Evaluate monitoring tools
✔️ Research observability trends
✔️ Recommend platform improvements
Collaboration
✔️ Partner with engineering teams
✔️ Support infrastructure initiatives
✔️ Participate in architecture discussions
✔️ Contribute to platform strategy
💻 Required Skills
Dialpad is seeking professionals with strong expertise in modern cloud and observability technologies.
Programming Languages
🐍 Python
⚡ Go
🦀 Rust
💻 Software Development Fundamentals
Observability Platforms
📊 Metrics Collection
📝 Centralized Logging
🔍 Distributed Tracing
📈 Monitoring Systems
Monitoring Tools
📊 Grafana
📊 Prometheus
📊 Loki
📊 Cloud Monitoring Platforms
Cloud Technologies
☁️ Amazon Web Services (AWS)
☁️ Google Cloud Platform (GCP)
☁️ Cloud-Native Infrastructure
Infrastructure Automation
⚙️ Terraform
⚙️ Ansible
⚙️ Infrastructure as Code
⚙️ Automation Frameworks
Container Technologies
🐳 Docker
☸️ Kubernetes
☸️ GKE
☸️ EKS
Operating Systems
🐧 Linux Administration
⚡ Shell Operations
🔧 System Troubleshooting
📩 How to Apply
👉 Application Link: Click Here
