STAND 8 provides end to end IT solutions to enterprise partners across the United States and with offices in Los Angeles, New York, New Jersey, Atlanta, and more including internationally in Mexico and India.
Senior AI Platform Engineer will be responsible for designing, implementing, and managing scalable cloud infrastructure while enabling AI/ML platform capabilities. This role focuses on building resilient AWS and multi-cloud environments, developing CI/CD pipelines, and supporting AI-driven workloads. The engineer will collaborate with cross-functional teams to deliver secure, efficient, and scalable platform solutions. This position plays a critical role in advancing platform engineering standards and accelerating AI adoption across the organization.
Location & Work Type Location: Carrollton, TX Work Type: Onsite
Key Responsibilities
Design, implement, and manage scalable and resilient infrastructure on AWS
Architect and maintain Windows/Linux environments integrated with cloud platforms
Develop and maintain infrastructure-as-code using AWS CloudFormation/CDK and Terraform/OpenTofu
Implement configuration management for Windows and Linux servers using Chef
Build and optimize CI/CD pipelines using GitLab CI/CD for .NET applications
Integrate and support AI services including orchestration with AWS Bedrock and generative AI frameworks
Enable AI/ML workflows supporting large-scale model training, inference, and deployment across AWS and GCP
Automate model lifecycle management including training, deployment, and monitoring
Collaborate with AI and development teams to deliver scalable platform environments and APIs
Implement observability, security, data privacy, and cost optimization strategies for AI workloads
Enforce security best practices across infrastructure and deployment processes
Troubleshoot infrastructure and application deployment issues
Implement monitoring and logging solutions for system visibility and proactive detection
Contribute to platform engineering standards, documentation, and best practices
Stay current with cloud, AI, and platform engineering trends and technologies
Provide mentorship and guidance to junior team members
Qualifications Required:
Bachelor's degree in Computer Science, Engineering, or a related field (or equivalent experience)
5+ years of experience in a Platform Engineering, DevOps or Site Reliability Engineering (SRE) role
1+ year(s) of experience with AI services & LLMs
Extensive hands-on experience with Amazon Web Services (AWS)
Solid understanding of Windows/Linux Server administration and integration with cloud environments
Proven experience with infrastructure-as-code tools, specifically AWS CDK and Terraform
Strong experience designing and implementing CI/CD pipelines using GitLab CI/CD
Experience deploying and managing .NET applications in cloud environments
Deep understanding of security best practices and their implementation in cloud infrastructure and CI/CD pipelines
Solid understanding of networking principles (TCP/IP, DNS, load balancing, firewalls) in cloud environments
Experience with monitoring and logging tools (e.g., NewRelic, CloudWatch)