Provectus Logo

Provectus

Senior AI/ML Engineer (GenAI, AWS)

Posted One Month Ago
Be an Early Applicant
In-Office or Remote
Hiring Remotely in Czechia
Senior level
In-Office or Remote
Hiring Remotely in Czechia
Senior level
Build and deploy production-grade generative AI and agentic systems for enterprise clients. Responsibilities include developing RAG applications, evaluation harnesses, backend services, data pipelines, and cloud-native AWS deployments; implementing LLMOps and AgentOps practices; monitoring quality, drift, cost, and latency; integrating APIs; improving reliability; contributing to architecture and reusable blueprints; and mentoring AI engineers in a forward-deployed consulting environment.
The summary above was generated by AI
  • Provectus is an AWS Premier Partner and an Anthropic Strategic Partner, working at the frontier of applied AI. We help enterprises turn Claude, agentic systems, and their own data into measurable business outcomes — through bespoke applications, managed services, and advisory engagements. With offices in North America, LATAM, and EMEA, we partner with clients worldwide.

  • Our work centers on two verticals — Financial Services & Insurance and Healthcare & Life Sciences — where we deploy five pre-built AI Blueprints: Submission Flow, Portfolio Lens, Asset Flow, Revenue Flow, and Evidence Lens. Each Blueprint rebuilds a critical business process front to back, shipped from working code and tuned to a client's specific book, regulators, and operating posture.

  • Our team holds 100+ AWS certifications, is Claude Code certified, and co-delivers Anthropic's Agentic SDLC program, Cowork Activation, and AI Blueprint engagements.

Where this role sits

You will work in a small, senior pod alongside an FDE and an FDX, 

  • Forward Deployed Executives (FDX) own the commercial relationship and the business outcome. Works alongside the client's leadership or C-suite level to move the client's KPIs.

  • Forward Deployed Engineers (FDE) embed with a client, map the client workflow, identify the business problem underneath it, design and build a working AI solution, present to the client, and transfer the knowledge to the client's team. Owns technical direction of the whole solution.

  • Senior AI Engineer. When an FDE comes back from the client with the business problem, you will help turn that into an agentic system that runs in production and will be responsible for evaluation, observability, and guardrails.  You'll have real ownership of components and of the technical decisions inside them.

Requirements:

    Mindset

  • Proactive and self-directed; you push for clarity rather than waiting for a ticket

  • Excellent communication and problem-solving skills

  • Comfort with ambiguity and ownership. 

  • B2+ English, comfortable collaborating across distributed, multicultural teams.

  • Technical depth

  • 5+ years in software or ML engineering, with production systems you were accountable for. 

  • Solid AI/ML foundations. You understand what the models do well enough to reason about failure modes.

  • Shipped to production LLM applications and agentic workflows — not demos, not POCs, not notebooks. 

  • Agentic orchestration: multi-step workflows, graph-based orchestration, tool use, state management, and recovery from partial failure 

  • Experience with LLM APIs (Anthropic, AWS Bedrock, or OpenAI) and agent frameworks.

  • Experience building and optimizing RAG systems in production.

  • Strong engineering fundamentals. Full-stack mindset, comfortable across AI, backend development, and cloud infrastructure. Python and/or TypeScript proficiency; depth matters more than stack. Dropped into an unfamiliar codebase, you're productive. 

  • Hands-on AWS in production: Bedrock, Bedrock AgentCore, Lambda, ECS, S3, SQS, ECR, or similar. GCP or Azure is a plus.

  • Cloud-native delivery: containers, ECS or Kubernetes, IaC, and CI/CD applied to AI pipelines.

  • You evaluate. You have built or owned an eval suite for a non-deterministic system, and you can explain what you measured, how you produced ground truth, and what gated a release.

  • Model and agent monitoring, drift detection. 

  • Cost and latency discipline: model tiering, caching, and the ability to say what a workload costs to run before it runs.

  • Hands-on production experience with the Claude ecosystem —  Claude Code, CLAUDE.md, hooks, skills files.  Spec-driven development — writing the intent, constraints, and acceptance criteria before you let an agent build — is a strong plus. 

  • MCP: you can say why an agent would prefer it to a REST integration. Having authored a server is a plus.

Nice to Have:

  • Experience in one of the industries: financial services, insurance, healthcare.

  • Consulting, professional services, or other embedded customer-facing delivery.

  • AWS and Claude Code Certifications

  • A2A: you can explain agent-to-agent interoperability 

  • CI/CD pipeline experience (GitHub Actions, GitLab CI)

  • Practical experience with one or more use cases from the following: NLP, LLMs, and Recommendation engines.

  • Experience in an additional language (Go, TypeScript, or Rust).

  • Experience with Apache Spark, Apache Airflow, Kafkа

Responsibilities:

  • Work in a pair with an FDE and an FDX. 

  • Build and ship production GenAI systems into the customer’s environment (cloud-native data, LLM-based, and agentic AI solutions). 

  • Build and optimize RAG systems for production use cases

  • Build the evaluation harness before you build the feature. 

  • Write production code across the stack — AI, backend services, data pipelines. We choose tools to fit the customer. 

  • Integrate AI components into backend services and RESTful APIs

  • Take systems to production on AWS (GCP or Azure where the customer requires it): containerised, CI/CD. Implement LLMOps and AgentOps practices: agent tracing, prompt and version management, cost and latency monitoring, regression testing, drift detection

  • Start from the blueprint, contribute to enablement and handover: clear documentation, runbooks, and pairing with the client engineers who will inherit the system. Feed reusable components and lessons back into the Provectus Blueprints

  • Participate in technical discussions and architectural decisions

  • Conduct model evaluation, improve failure modes you find, optimize model performance, efficiency, and reliability

  • Mentor junior and mid-level AI engineers, conduct code reviews and share knowledge across the team through documentation, presentations, and workshops.

What We Offer:

  • The chance to shape how leading enterprises across LATAM, Europe, and North America adopt AI, from strategy through first deployment

  • A forward-deployed model working in small, senior teams alongside FDE and FDX

  • A growing AI delivery practice where you help build the tooling and frameworks, not just use them

  • Remote-friendly culture

  • Internal training programs with full support for Claude, AWS, and other professional certifications, conference attendance

  • Career growth; we actively develop our engineers

  • Access to the latest AI tools and premium subscriptions

  • Long-term B2B collaboration

  • Private medical insurance or a budget for your medical needs

  • Paid sick leave, vacation, and public holidays

  • Equipment and all the tech you need for comfortable, productive work

How we hire:

    1. Intro conversation. The role, your background and aspirations, tech questions.

    2. Technical interview with live engineering sessions. Real problems, your own editor, you may use an LLM assistant 

    3. HR Interview. Soft skills and expectations

    4. HM interview. Tech questions; a live engineering session is also possible

Similar Jobs

Internship
Big Data • Food • Hardware • Machine Learning • Retail • Automation • Manufacturing
Supports global IT procurement through contract management, supplier performance analysis, negotiation preparation, sourcing strategy, KPI reporting, and AI-enabled process improvement. Collaborates with digital services, Finance, and Legal teams while using analytical tools to identify opportunities and support decisions. The six-month paid internship offers exposure to global IT sourcing and cross-functional business operations.
Top Skills: Artificial IntelligenceLarge Language Models (Llms)ExcelMS Office
Yesterday
In-Office or Remote
Senior level
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Provides technical leadership for Python-based AI backend platforms supporting LLM integrations, generative AI pipelines, and agent orchestration. Architects scalable AWS systems using FastAPI, Lambda, and EKS; establishes observability and engineering standards; drives performance, security, and reliability improvements; mentors engineers; and influences organization-wide technical strategy, platform capabilities, and best practices.
Top Skills: Amazon EksAnthropic ApiAWSAws LambdaCi/CdDistributed SystemsDockerEvent-Driven ArchitectureFastapiInfrastructure As CodeKubernetesLangchainLangfuseLitellmMicroservicesMlopsOpenai ApiPython
4 Days Ago
Remote or Hybrid
2 Locations
112K-207K Annually
Senior level
112K-207K Annually
Senior level
Artificial Intelligence • Healthtech • Machine Learning • Natural Language Processing • Biotech • Pharmaceutical
Leads Pfizer’s environmental sustainability strategy across climate, Scope 3 emissions, biodiversity, nature, product sustainability, and sustainable sourcing. Manages supplier decarbonization programs, sustainability metrics, regulatory horizon scanning, risk assessments, governance, training, and communications. Facilitates cross-functional teams, advises stakeholders, represents Pfizer in external initiatives, and serves as a sustainability subject matter expert across the enterprise.
Top Skills: Microsoft TeamsWebex

What you need to know about the NYC Tech Scene

As the undisputed financial capital of the world, New York City is an epicenter of startup funding activity. The city has a thriving fintech scene and is a major player in verticals ranging from AI to biotech, cybersecurity and digital media. It also has universities like NYU, Columbia and Cornell Tech attracting students and researchers from across the globe, providing the ecosystem with a constant influx of world-class talent. And its East Coast location and three international airports make it a perfect spot for European companies establishing a foothold in the United States.

Key Facts About NYC Tech

  • Number of Tech Workers: 549,200; 6% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Capgemini, Bloomberg, IBM, Spotify
  • Key Industries: Artificial intelligence, Fintech
  • Funding Landscape: $25.5 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Greycroft, Thrive Capital, Union Square Ventures, FirstMark Capital, Tiger Global Management, Tribeca Venture Partners, Insight Partners, Two Sigma Ventures
  • Research Centers and Universities: Columbia University, New York University, Fordham University, CUNY, AI Now Institute, Flatiron Institute, C.N. Yang Institute for Theoretical Physics, NASA Space Radiation Laboratory

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account