LLM Systems / AI Agent Engineer
Job Description
Key Skills Required
Master these to land this role
Want to know if you're a match for this job?
As an LLM Systems / AI Agent Engineer, you will focus on building and evolving production AI agents on foundation models, covering orchestration, context engineering, evaluation pipelines, and production observability and monitoring. This is an LLM systems and AI agent engineering position rather than a traditional ML model-training role. You will be the first dedicated AI-agenting hire, working directly alongside the engineer who currently leads this function. The product currently uses a custom orchestration layer, and your knowledge of agent architecture patterns will help inform whether to continue with this approach or adopt a production framework such as LangGraph or LangChain. Roadmap timings are somewhat tentative, but you will be assigned measurable deliverables immediately upon onboarding.
Responsibilities
- Build and evolve production AI agents on foundation models, currently AWS Bedrock.
- Develop and evolve orchestration for production AI agents.
- Apply context engineering to production LLM and agent systems.
- Build and maintain evaluation datasets and pipelines, including tool-selection, trajectory and LLM-as-judge evaluations.
- Build and maintain production observability and monitoring for LLM and agent systems.
- Implement and work with tracing and instrumentation for production LLM systems.
- Work directly alongside the engineer currently leading the AI-agenting function as the first dedicated hire in this area.
- Apply strong agent-architecture fundamentals to help inform whether the existing custom orchestration layer should be retained or a production framework adopted.
- Work as part of a new three-person product team alongside the AI function lead and a Full Stack Engineer.
- Operate as an individual contributor working alongside the AI function lead rather than managing others.
- Take ownership of measurable deliverables immediately upon onboarding.
- Deliver similar AI/agent engineering work against roadmap timelines.
Requirements
- 2+ years of experience building production LLM agents, including tool-calling agent loops, streaming, context management, structured outputs and orchestration frameworks.
- Production experience with an agentic framework such as LangGraph, LangChain or custom orchestration. There is no fixed orchestration framework requirement; strong agent-architecture fundamentals and production experience with any agentic framework are in scope.
- 1.5+ years of experience with evaluation-driven development, including building and maintaining evaluation datasets and pipelines covering tool selection, trajectory evaluation and LLM-as-judge.
- 1+ year of experience with LLM observability, tracing and instrumentation using Langfuse, OpenTelemetry or similar tooling.
- Genuine production agent-observability exposure. Direct, hands-on Langfuse experience is strongly preferred because this is a confirmed skill gap within the team; OpenTelemetry or other tracing tools are acceptable only as a secondary signal alongside real agent-observability exposure.
- 1+ year of experience with LLM cost optimisation, including prompt caching, model selection and routing, and LLM FinOps.
- 5+ years of backend engineering proficiency, including TypeScript/Node, Postgres and serverless AWS.
- Proven experience delivering similar work on similar timelines.
- Experience shipping agentic AI systems to production, with the ability to speak to concrete failure modes and mitigations and operate end-to-end across the AI stack.
Nice to Have
- 1+ year of AWS Bedrock experience. Equivalent production experience with other foundation-model providers, including OpenAI, Anthropic API, Azure OpenAI or Vertex AI, is fully transferable.
- 1+ year of experience with AI safety and guardrails, including prompt-injection screening, output validation and handling untrusted input.
- 6+ months of familiarity with MCP and multi-agent patterns.
- Familiarity with geospatial data.
Benefits
- Fixed Shifts: 12:00 PM - 9:30 PM IST (Summer) | 1:00 PM - 10:30 PM IST (Winter)
- No Weekend Work: Real work-life balance, not just words
- Day 1 Benefits: Laptop and full medical insurance provided
- Support That Matters: Mentorship, community, and forums where ideas are shared
- True Belonging: A long-term career where your contributions are valued
How would you rate this job post?
See what other professionals think about this role.
Similar Opportunities
Senior AI Interaction Evaluator (Codex / Claude Code)
G2i
Calgary
BostonSenior AI Interaction Evaluator (Codex / Claude Code)
G2i
Calgary
BostonSenior AI Interaction Evaluator (Codex / Claude Code)
G2i
Calgary
BostonSenior AI Interaction Evaluator (Codex / Claude Code)
G2i
Calgary
BostonMore Openings at Smartworking
Explore Top Companies in this Space
Globaldev
IT Outsourcing / Software Engineering / AI Development
Dev Partners
IT Staff Augmentation / Software Development / Outsourcing
CaptivateIQ
Sales / Enterprise Software / Technology
ArenaNet
Computer Games / MMOs / MMORPGs / RPGs
Smartworking
View Company ProfileSmart Working is a specialized IT staffing and software development outsourcing platform that helps businesses build high-performing, remote engineering teams. They provide access to the top 1% of rigorously vetted global tech talent, enabling companies to hire elite Full-Stack, Front-End, Back-End, Mobile, and AI developers in as little as 10 days. Unlike traditional freelance marketplaces, Smart Working provides dedicated developers who are fully integrated into the client's team and work the same business hours, while the company handles all HR, payroll, and compliance overhead. This model allows businesses to save up to 50% on annual hiring costs compared to local hires. Additionally, the company features an in-house "AI Academy" that provides continuous, professional AI training to their developers at no extra cost, ensuring that teams remain future-proof, highly productive, and equipped with the latest technological skills.
Safety First
- Never pay for a job application.
- Do not share sensitive bank info.
- Verify the client before starting work.
