AI Engineer @Azumo
Artificial Intelligence
Salary unspecified
Remote Location
Employment Type full-time
Posted Today

[Hiring] AI Engineer @Azumo

Today - Azumo is hiring a remote AI Engineer. πŸ’Έ Salary: unspecified πŸ“Location: Latin America (LATAM)

Role Description

Azumo builds and operates production AI systems for companies ranging from seed-stage startups to Meta. We are hiring an AI Engineer to own what those systems do once they are live:

  • Retrieval pipelines
  • Tool-using agents
  • Evaluation harnesses
  • Guardrails that keep them dependable in front of real users

The role is fully remote across Latin America, aligned to your client's working day. You will not be building demos. Azumo has shipped production AI since 2016, and the work here starts where the prototype ends, making a system reliable, measurable, and affordable enough to put in front of customers.

Where this role sits

Azumo's engineering organization is built around four lanes:

  • The Data Engineer lane owns pipelines, storage, and the retrieval layer.
  • The Data Scientist lane owns the question and the method.
  • The Software Engineer lane owns AI-augmented product delivery.
  • This role is the AI Engineer lane, and it owns production behavior.

One question places the boundary: when the output is wrong, whose problem is it? "The method was inappropriate" is a Data Scientist question. "The system did the wrong thing with an appropriate method" is yours.

What you will build

  • Retrieval systems: Chunking and embedding pipelines, hybrid search, reranking, and evaluation of retrieval quality, built on pgvector, Pinecone, Qdrant, FAISS, or Azure AI Search.
  • Agentic workflows: Stateful multi-step execution, tool calling, MCP servers, structured output enforcement, context-window management, deterministic fallbacks, and human-in-the-loop gates for the decisions that need one.
  • Evaluation: Test sets that reflect the decision the system is actually making, model-as-judge scoring, regression tracking across prompt and model changes, and honest error analysis.
  • Reliability and safety: Prompt-injection defense, output validation, guardrails, PII handling, and graceful degradation when a model or tool call fails.
  • Production operation: Containerized deployment on Azure or AWS, CI/CD, observability, and explicit latency, cost, and token budgets that you own rather than discover after the invoice.
  • Work inside the client's environment: Their repositories, their standups, sometimes their customer calls.

Azumo is SOC 2 certified, client code stays in client repositories, and some engagements carry additional requirements such as HIPAA.

How we work

Our engineers build with AI every day. Claude Code, Codex, and similar tools are part of the standard toolchain here, not an experiment. We run an automated audit across the whole codebase on day one and every day after, grading security, cost, and architecture findings by severity with the exact file and line, so a small team can move quickly without quality drifting.

We stay vendor-neutral across OpenAI, Anthropic, and open-weight models, and we run Valkyrie, our own production layer, when a single interface to any model is the right call.

Qualifications

  • 4+ years building and shipping production software, with a modern backend language (e.g., Python) as your primary focus.
  • Demonstrable production experience with LLM-based systems: retrieval-augmented generation, function and tool calling, structured output enforcement, and prompt design.
  • Hands-on work with vector and retrieval infrastructure such as pgvector, Pinecone, Qdrant, FAISS, or Azure AI Search.
  • Experience with agent frameworks and tooling: LangGraph, LangChain, CrewAI, the Model Context Protocol (MCP), or native Python execution loops.
  • You have built an evaluation suite for an LLM system, including test-set design and regression tracking.
  • Cloud deployment experience, Azure preferred and AWS acceptable, with Docker, CI/CD pipelines, and infrastructure as code.
  • Working discipline around latency, token cost, and throughput.
  • Active use of AI-assisted coding tools such as Claude Code, Cursor, or GitHub Copilot.
  • Clear written and spoken English, C1 or above.
  • Bachelor's degree in Computer Science, Data Science, or a related field, or equivalent professional experience.

Preferred Qualifications

  • Fine-tuning and adaptation of open-weight models: LoRA, QLoRA, PEFT.
  • Self-hosted or open-weight inference and serving.
  • Multimodal systems covering vision, speech, or document understanding alongside text.
  • Security work specific to LLM systems: prompt-injection testing, red-teaming, and output sanitization.
  • Delivery under a compliance regime such as SOC 2 or HIPAA.
  • Streaming, real-time, or high-throughput inference workloads.
  • Contributions to open-source AI libraries, published technical writing, or active participation in the AI engineering community.

Benefits

  • 00% remote-first culture (work anywhere in Latin America)
  • Paid time off (PTO)
  • U.S. Holidays
  • AI Training and certifications
  • Mentored career development
  • Profit sharing
  • $US remuneration

Company Description

Azumo is a San Francisco based software development company that has been building intelligent applications since 2016. We provide nearshore AI engineering teams to organizations that need production AI faster than they can hire for it:

  • As an embedded engineering team
  • As AI staff augmentation alongside an existing team
  • As a full project build

Our engineers work from Latin America, aligned to United States time zones, and have delivered for Twitter, Meta, Discovery Channel, Omnicom, UnitedHealth, and CENTEGIX.

We hire for seniority and test for it before anyone joins a client team. We support engineers in going deep on the modern AI stack, and we give time back to open-source work, community teaching, and philanthropy.

Apply at https://azumo.com/join-our-team or write to us at [email protected] .

Before You Apply
️
remote Be aware of the location restriction for this remote position: Latin America (LATAM)
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
AI Engineer @Azumo
Artificial Intelligence
Salary unspecified
Remote Location
Employment Type full-time
Posted Today
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 130,000+ Remote Jobs
️
remote Be aware of the location restriction for this remote position: Latin America (LATAM)
β€Ό Beware of scams! When applying for jobs, you should NEVER have to pay anything. Learn more.
Apply for this position
Did not apply βœ“
Applied βœ“
Sent Follow-Up βœ“
Interview Scheduled βœ“
Interview Completed βœ“
Offer Accepted βœ“
Offer Declined βœ“
Application Denied βœ“
Unlock 130,000+ Remote Jobs
Γ—
Apply to the best remote jobs
before everyone else

Access 130,000+ vetted remote jobs and get daily alerts.

4.9 β˜…β˜…β˜…β˜…β˜… from 500+ reviews

⚑ 131,382+ remote jobs, refreshed hourly

πŸ”” Real-time alerts: Apply first

πŸ›‘οΈ Vetted companies, no scams, true remote only

Unlock All Jobs Now

Maybe later