This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Lead AI Platform Engineer based in Brazil.
This role offers the opportunity to lead the engineering of a large-scale generative AI platform designed for enterprise environments. You will architect, build, and operate AI infrastructure that enables secure, scalable, and governed adoption of artificial intelligence. The position combines cloud engineering, platform development, automation, and AI expertise across multiple technology ecosystems. You will work with cutting-edge AI services and collaborate with global teams to create reliable solutions for complex business needs. Your contributions will directly impact how organizations consume, govern, and scale AI workloads across enterprise environments. This is an opportunity for a senior technology professional to shape the future of AI platforms through innovation, automation, and cloud-native engineering.
Accountabilities:
- Design, build, and operate a centralized enterprise AI Gateway using Azure API Management (APIM) to support scalable generative AI workloads.
- Develop advanced governance, authentication, routing, security, and observability policies for AI-powered applications.
- Integrate multiple AI providers and platforms, including Azure OpenAI, Azure AI Foundry, GCP Vertex AI, AWS Bedrock, and multimodal AI services.
- Implement FinOps strategies for AI workloads, including consumption monitoring, quota management, token budgeting, and cost attribution.
- Build and maintain infrastructure as code solutions using Terraform/OpenTofu to enable reliable and repeatable deployments.
- Develop and improve CI/CD pipelines with GitHub Actions, including secure authentication flows using OIDC.
- Create centralized observability solutions using tools such as Application Insights, KQL, Azure Workbooks, Datadog, and CloudWatch.
- Design and maintain API management policies involving streaming, request/response transformation, retries, fallback mechanisms, and backend routing.
- Strengthen platform security through identity management, JWT validation, WAF configuration, Front Door, and Key Vault integrations.
- Automate operational processes using scripting languages such as Bash, Python, and PowerShell.
- Create technical documentation, OpenAPI specifications, and communication materials for both technical and non-technical stakeholders.
- Support internal teams in adopting generative AI solutions securely and efficiently.
- Translate complex technical decisions into clear recommendations for different audiences.
- Work autonomously to identify improvements, prioritize initiatives, and continuously evolve the AI platform.
Requirements: