Login Enter

MixMode

7 months ago

Senior Software Reliability Engineer for AI

Sign up free Log in

Sign up to save this job, get alerts, and apply with an optimized CV.

Company information

Company
MixMode
Location
United States
Posted
7 months ago
View all jobs at MixMode

Job description

MixMode is a leading provider of AI-powered cybersecurity solutions at scale, pioneering a patented third-wave, context-aware AI approach that automatically learns and adapts to dynamic environments. The MixMode platform delivers self-supervised, real-time threat detection for known and unknown threats across cloud, hybrid, and on-premises environments. Large organizations with big data workloads – including those in enterprise, critical infrastructure, US Department of War and US Intelligence Community – trust MixMode to defend their most important assets. Backed by PSG and Entrada Ventures, MixMode is headquartered in Santa Barbara, California. Learn more at www.mixmode.ai. Job Title: Senior Software Reliability Engineer for AI Location: Santa Barbara, CA or Remote Job Summary: We are looking for a Senior Software Engineer to improve the reliability, performance, and scalability of our production AI systems. This role focuses on understanding, refining, and strengthening existing distributed services across application, database, and Kubernetes layers. This individual will work closely with ML researchers to make our systems more robust, maintainable, flexible, and scalable. Responsibilities: • Own the reliability, performance, and operational health of production AI systems, focusing on improving complex, existing services. • Lead efforts to refactor and harden the AI codebase to improve observability, maintainability, and resilience. • Diagnose and resolve issues across distributed systems, including latency, throughput, data pipelines, and resource utilization. • Design and build monitoring, alerting, and debugging tools for high-availability services. • Partner with researchers and ML engineers to productionize models at scale. • Establish best practices for testing, deployment, capacity planning, and continuous integration/continuous delivery (CI/CD). • Collaborate with cross-functional teams to ensure seamless integration of AI-powered solutions into existing infrastructure. • Develop and maintain technical documentation, including code comments, API documentation, and release notes. • Participate in code reviews and contribute to the development of new features and improvements. • Stay up-to-date with industry trends, emerging technologies, and best practices in software reliability engineering, AI/ML, and cybersecurity. Requirements: • Bachelor's or Master's degree in Computer Science, Software Engineering, or related field. • 5+ years of experience in software development, with a focus on reliability, performance, and scalability. • Strong understanding of distributed systems, including application, database, and Kubernetes layers. • Experience with AI/ML frameworks such as TensorFlow, PyTorch, or Scikit-Learn. • Proficiency in programming languages such as Python, Java, C++, or Go. • Familiarity with containerization using Docker and orchestration using Kubernetes. • Knowledge of cloud-based services such as AWS, Azure, or Google Cloud Platform. • Experience with CI/CD pipelines and automated testing. • Strong problem-solving skills and attention to detail. • Excellent communication and collaboration skills. Benefits: • Competitive salary and benefits package. • Opportunity to work on cutting-edge AI-powered cybersecurity solutions. • Collaborative, dynamic work environment. • Professional development opportunities. • Flexible work arrangements, including remote work options.

Required skills

Interested in this position?

Create your free account and tailor your CV to match this job.