PLATFORM ENGINEER
Engineering ยท Platform
๐Ÿ“ Noida ย  ๐Ÿ• Full-Time ย  ๐Ÿงญ 2-3 Years

THE MISSION
We aren't automating scripts โ€” we're deprecating the era of manual-heavy testing entirely. TestMu AI is building the world's first AI-native platform where Agentic Intelligence autonomously plans, authors, and self-heals the entire Quality Engineering lifecycle.

Drive reliability and scalability of TestMu AI's core infrastructure and service workflows.

THE PILLARS OF IMPACT

๐Ÿš€ 1. Platform Reliability (50%)
ย โ€“ Ensure robust cloud infrastructure and deployment pipelines.
ย โ€“ Improve observability tools and practices.
ย โ€“ Automate and streamline service workflows.
ย โ€“ Enhance platform systems for developer efficiency.

โš™๏ธ 2. Infrastructure Engineering (30%)
ย โ€“ Build and maintain cloud-based systems.
ย โ€“ Optimize performance and scalability.
ย โ€“ Implement infrastructure improvements based on feedback.

๐Ÿง  3. Incident Management (20%)
ย โ€“ Lead incident response with strong SRE principles.
ย โ€“ Analyze and resolve production issues swiftly.
ย โ€“ Develop strategies to prevent future incidents.

MUST-HAVES โ€” DO NOT APPLY UNLESS YOU HAVE THESE

ย โ€“ You must have strong cloud, scripting, and core DevOps fundamentals.
ย โ€“ You must have hands-on experience with observability tools like New Relic, Sumo Logic.
ย โ€“ You must demonstrate real coding ability.
ย โ€“ You must have a strong grasp of incident management and reliability/SRE thinking.
ย โ€“ You must explain service-layer architecture and reasoning effectively.

THE BAR โ€” WHAT YOU MUST PROVE

Cloud Expertise: Designed and optimized cloud infrastructure for scalability and reliability.
Coding Proficiency: Developed scripts and tools to automate deployment processes.
Incident Management: Led a team to resolve critical production incidents effectively.
Architecture Understanding: Explained complex service architectures to non-technical stakeholders.