Posted today · be early
DevOps Engineer, India
Signalmash
IndiaPosted 1 day ago
Skill Required
FulltimeDevOps EngineerBackup and RecoveryShell ScriptingGitHub ActionsObservabilityPostgreSQLKubernetesPrometheusNode.jsPythonDockerArgoCDGitISO 27001AzureLinuxCI/CDAWSGCPDNSCryptography
Key highlights
- 4+ years of DevOps/SRE/platform engineering experience required
- Based in India with required availability until ~11:30 PM IST to overlap US management
- Strong Kubernetes expertise, including self‑managed/bare‑metal clusters
- Remote work possible initially, with potential relocation to Kochi office
- Hands‑on ownership of CI/CD, Kubernetes, PostgreSQL, observability, and security
- Practical experience with AI coding tools (Claude Code, Codex, Cursor, Copilot) is required
Role overview
Signalmash operates a cloud communications platform that enables businesses to manage and deliver RCS, SMS, and MMS messaging through carrier and technology‑provider integrations. The company is adding an India‑based DevOps Engineer to own the reliability, speed, and security of the infrastructure the platform runs on. The DevOps Engineer will own the build, deploy, and runtime infrastructure across self‑managed Kubernetes and cloud environments, making deployments faster and safer, surfacing failures before customers see them, and ensuring recovery procedures are tested.
Responsibilities
- Own CI/CD pipelines end to end: GitHub Actions (including self‑hosted runner infrastructure), build speed, caching, elimination of flaky jobs, and safe automated deploys.
- Operate and improve our Kubernetes environments (self‑managed K3s on bare metal) and cloud deployments (GCP), including environment parity from development through production.
- Manage PostgreSQL operations: schema migrations, backup and point‑in‑time recovery, and performance.
- Build and maintain observability: metrics, logs, and alerting that page the team before customers notice.
- Own backup and disaster recovery: automated offsite backups, periodic restore verification, and documented runbooks.
- Harden security: secrets management, network segmentation and overlay networking (Tailscale/WireGuard‑style), TLS and certificate management, least‑privilege access, and dependency hygiene.
- Coordinate production incidents and ensure root causes, corrective actions, and lessons are documented.
- Reduce infrastructure cost without reducing reliability.
- Introduce responsible AI‑assisted operations practices using tools such as Claude Code, Codex, Cursor, or Copilot.
- Work alongside internal engineers and external development partners, and connect with the US management team on priorities, risks, and incidents.
- Read application code and logs, trace failures across services, and own production issues from alert to documented root cause.
Requirements
- 4+ years in DevOps, SRE, or platform engineering with production ownership.
- Strong Kubernetes experience — deployment, networking, storage, and upgrades — ideally including self‑managed or bare‑metal clusters, not only managed cloud services.
- CI/CD depth with GitHub Actions or equivalent, including runner infrastructure and pipeline optimization.
- PostgreSQL administration, including migrations and tested backup/restore.
- Docker, Linux administration, shell scripting, and at least one of Python or Node.js.
- Cloud experience with GCP or transferable AWS/Azure experience.
- Monitoring and alerting stacks such as Prometheus/Grafana or equivalents.
- A track record of keeping a live, customer‑facing platform up through continuous change.
- Practical experience using AI coding tools in real repositories, including testing, review, and security controls.
- Strong spoken and written English.
- Based in India and able to work until approximately 11:30 PM IST to overlap with the US management team.
Nice to have
- Communications or CPaaS platform operations: messaging APIs, carrier integrations, high‑throughput webhooks and queues.
- Hybrid estates combining bare‑metal providers (such as Hetzner) with cloud.
- Cloudflare (DNS, R2 object storage), GitOps patterns (ArgoCD/Flux), and infrastructure as code.
- ORM‑driven migration workflows (Prisma or similar) and the failure modes that come with them.
- Quantifiable cost‑optimization wins.
- Security‑ or compliance‑sensitive environments (SOC 2 or similar).
Additional details
- IMPORTANT: PLEASE READ BEFORE APPLYING
- To be considered, you must follow the How to Apply instructions at the end of this job description and complete the Signalmash application form.
- Applications that do not include the completed form will not be reviewed. Automated or bulk applications that do not follow these instructions will be disregarded.
- About Signalmash Signalmash operates a cloud communications platform that enables businesses to manage and deliver RCS, SMS, and MMS messaging through carrier and technology‑provider integrations. The platform includes customer onboarding, messaging APIs, routing, compliance, vendor integrations, usage tracking, and billing.
- The Role This is a hands‑on role combining CI/CD ownership, Kubernetes operations, database reliability, observability, and security hardening. You will work alongside internal engineers and external development partners, and connect with the US management team on priorities, risks, and incidents. This is not a ticket‑taking position. The successful candidate must be comfortable reading application code and logs, tracing a failure across services, and owning a production issue from alert to documented root cause.
- Location The position is based in India and may initially be remote. Signalmash is evaluating an engineering office in Kochi during 2026‑2027. Candidates should be open to relocating to Kochi if an office is established.
- How to Apply Please complete the Signalmash application form: https://docs.google.com/forms/d/e/1FAIpQLSd0KZFpt1d1XpsI0_2kCC6v_XLLHEAYnvg7VLFXFe5kcWpm2A/viewform You can upload your resume directly through the form. You may also upload an optional two‑to‑three‑minute video or audio introduction. Applications that include an introduction will receive priority review. Recording quality, accent, appearance, and background are not evaluated. We care about clarity, authenticity, thoughtfulness, and communication. Applications without a completed Signalmash application form will not be reviewed.