- Location
- San Juan, PR
- Workplace
- Hybrid
- Type
- Full-time
- Department
- IT
- Experience
- 5+ years
- Closing date
- Today
- Source
- Vincere
Description
Job Advertisement: AI Tooling and Infrastructure Engineer
Overview
We are seeking a talented AI Tooling and Infrastructure Engineer to join our team in a hybrid role, requiring an average of two days per week in the office. This position is a unique opportunity to work at the intersection of infrastructure engineering and applied AI, enabling the integration of AI technologies into tools and systems that support technical training and documentation. If you are passionate about leveraging AI to solve real-world business challenges and thrive in a collaborative, innovative environment, we encourage you to apply.
Responsibilities
- Design, build, and maintain infrastructure that integrates large language models (LLMs) into internal tools and customer-facing products.
- Evaluate, implement, and manage third-party AI tooling and platforms, including vendor relationship and integration.
- Develop custom tooling (e.g., pipelines, orchestration layers, retrieval systems, evaluation harnesses, monitoring) when off-the-shelf solutions are insufficient.
- Ensure the operational reliability of AI-powered systems, focusing on uptime, latency, cost, and observability.
- Implement guardrails for cost management, rate limits, prompt/version control, and failure handling for LLM-backed services.
- Collaborate with stakeholders to translate business use cases into scalable, maintainable systems.
- Set up monitoring, logging, and telemetry to assess the performance of AI systems in production.
- Contribute to internal standards and best practices for adopting and evolving AI tooling.
- Act as an AI advocate within the organization, sharing knowledge, demonstrating capabilities, and guiding teammates on AI applications.
Qualifications
- 5+ years of experience in software engineering, infrastructure, platform, or DevOps/SRE roles.
- Proficiency in scripting and programming languages such as Python or Java.
- Experience with prompt management, retrieval-augmented generation, agentic tooling, or similar AI-related technologies.
- Practical experience integrating third-party APIs into production systems (LLM APIs preferred).
- Hands-on experience with cloud infrastructure (AWS, GCP, or Azure) and infrastructure-as-code practices.
- Familiarity with containers and container tooling.
- Proficiency in operating within both Linux and Windows server environments.
- Experience in building and maintaining web services, including APIs and backend systems.
- Strong understanding of building reliable, observable, and cost-efficient backend systems.
- Pragmatic approach to balancing vendor-provided tools with custom-built solutions.
- A bias toward iterative development, delivering usable solutions quickly and enhancing them over time.
Day-to-Day
- Collaborate with cross-functional teams to identify and implement AI-driven solutions for technical training and documentation.
- Build and maintain infrastructure to support AI-powered tools and services.
- Monitor and optimize the performance, reliability, and cost of AI systems in production.
- Evaluate and integrate third-party AI tools and platforms, ensuring seamless functionality.
- Develop and refine custom tooling to address unique business needs.
- Provide technical guidance and share AI knowledge with team members and stakeholders.
Benefits
- Competitive salary and comprehensive benefits package.
- Flexible hybrid work environment, allowing for a balance between office and remote work.
- Opportunities for professional growth and career advancement.
- Collaborative and innovative work culture that values diverse perspectives.
- The chance to work on cutting-edge AI technologies and make a tangible impact on business processes.
If you are ready to take on the challenge of building and operating AI infrastructure that drives meaningful business outcomes, we look forward to hearing from you. Apply today to join our dynamic team!