Remote Freedom

Data Engineer - Oracle to PostgreSQL Re-Platform ( 102-08SENG-02 )

OpsBrasil Serviços Cloud LTDA

Work from anywhereContract1w ago

See how well you match this role

Upload your résumé — we'll score your fit and check off the skills you already have. Free.

Requirements

  • Build and harden agent orchestration (LangGraph, LangChain, or equivalent)Must
  • 5+ years software engineering experienceMust
  • 2+ years at senior levelMust
  • Production Python proficiencyMust
  • Hands-on LLM application engineering: prompt design, tool calling, structured outputMust
  • Context management and token budgeting in production LLM systemsMust
  • AWS Bedrock and Bedrock AgentCore: model invocation, streaming, guardrailsMust
  • Managed LLM platforms in production (Bedrock, Vertex AI, or Azure AI Foundry)Must
  • Vector and hybrid search integration in productionMust
  • Observability: OpenTelemetry-based distributed tracingMust
  • Production AWS: IAM, networking, storage, observabilityMust
  • CI/CD pipelines and Infrastructure as Code (Terraform)Must
  • TypeScript or Go when required

What you'll do

  • End-to-end LLM agent delivery: requirement to production deployment
  • Defend architecture and trade-offs directly with clients
  • Agent orchestration frameworks (LangGraph, LangChain, or hand-rolled systems)
  • Build evaluation harnesses with golden sets and LLM-as-judge
  • Integrate tools via Model Context Protocol (MCP)
  • Token, cost, and latency attribution per session
  • Cost and latency optimization: model routing, prompt caching, payload pruning
  • Circuit breakers, fallbacks, and dead-letter handling
  • Design safe write-actions with least-privilege, human-in-the-loop approval

Keywords

Retrieval tools: BM25, hybrid, and vector searchAI-assisted engineering tools (Claude Code, Cursor, Copilot)

About the role

This is not a ticket execution role. You get the problem and the context, you propose the solution, you build it, you ship it, and you defend the technical decisions directly in front of the client. The mandate is to take a production-grade LLM agent from read-only insight toward supervised action — hardening it for scale and staging it up a capability ladder (Explains → Recommends → Orchestrates → Acts). AI-assisted engineering (Claude Code, Cursor, Copilot, or equivalent) is the baseline here, not a differentiator — but you sign the code, and "the AI wrote it" is never an answer when something breaks in production. What you will do • Own end-to-end delivery: take a production LLM agent from requirement to production deploy, and defend the architecture and trade-offs directly with the client. • Build and harden agent orchestration (LangGraph / LangChain or equivalent) — routing, tool-calling, planning, synthesis, and state management. • Integrate tools over MCP and keep a growing tool surface fast and correct, including BM25, hybrid, or vector retrieval as scale demands. • Run models on AWS Bedrock and Bedrock AgentCore — model selection/routing, guardrails, memory, and regional residency profiles. • Build the evaluation harness (golden sets, LLM-as-judge, quality gates wired into CI) and instrument the system with OpenTelemetry for per-session token, cost, and latency attribution. • Drive down cost and latency with real levers (model routing, prompt caching, payload pruning, parallelizing independent calls) behind a regression gate; build in circuit breakers, fallbacks, and dead-letter handling. • Design and stage safe write-actions with least-privilege permissions, human-in-the-loop approval, plan versioning, audit trail, and rollback — released behind feature flags to a small cohort first. Requirements Required • 5+ years in software engineering, with at least 2 genuinely at a senior level; strong Python in production, and comfort picking up TypeScript or Go when a project calls for it. • Hands-on LLM application engineering: prompt design, tool/function calling, structured output, context management, and token budgeting, in a system real users hit. • Agent orchestration experience: built or operated orchestration with a framework like LangGraph or LangChain, or hand-rolled, beyond single-prompt calls. • Managed LLM/agent platform in production: AWS Bedrock, Google Vertex AI, or Azure AI Foundry — model invocation, streaming, guardrails, and agent tooling. (We use Bedrock and Bedrock AgentCore; equivalent depth on Vertex AI or Azure transfers directly.) • Evaluation, retrieval & observability: eval harnesses and golden/reference sets, vector or hybrid search in production, and OpenTelemetry-based distributed tracing with token/cost/latency attribution. • Production AWS, CI/CD & IaC: real IAM, networking, storage, and observability experience; CI/CD pipelines versioned as code; Terraform in production. • Working English: comfortable defending system design and technical decisions directly on client calls. Highlights Tech stack AWS Bedrock, Bedrock AgentCore, LangGraph/LangChain, MCP, Terraform, GitHub Actions/GitLab CI, OpenTelemetry Originally posted on Himalayas

Sourced from Himalayas. Confirm the details and apply on the employer's site.