NextRaiseNextRaiseFind jobs
Sign inSign up free
Jobs / AI Engineer in United States of America
8 days ago
Apply with autofill
Apply with autofill
Artisan·8 days ago
8 days ago

Staff AI Engineer - Agent Architecture & Behavior

San Francisco, United States of AmericaFull-timeOn-siteSenior · 8-12 yearsAI Engineer

Sign up free to see how well your resume matches this role.

Boost your chances at artisan

How you compare FREE

?
Your scoreYour score: not yet known
→
70
Top 10%Top 10%: 70 out of 100

Top 10% of NextRaise users matched against AI Engineer roles in United States.

Must-have skills for this role

  • python
  • typescript
  • llm
  • agent architecture

PDF or DOCX · no account needed

Apply faster with autofill FREEartisan uses Ashby - autofill it instead of retyping.careers.example.com/applyAutofillingFull namePriya SharmaEmailpriya.sharma@example.comPhone+49 30 1234567LocationBerlGet the extension

What you'll do

  • Agent architecture and behavior. Design and implement agent execution loops, planning strategies, tool interfaces, and verification. Turn ambiguous technical requirements into clear system boundaries and working code.
  • Multi-agent systems. Build delegation, coordination, context sharing, and result synthesis. Handle concurrent work, conflicting updates, cancellation, and stale results. Establish when a multi-agent approach improves on a simpler baseline.
  • Reliable execution. Make complex, stateful workflows resilient to interruptions and partial failures. Build checkpoints, recovery strategies, and appropriate human intervention into the architecture.
  • Context, memory, and reusable methods. Improve retrieval, context construction, persistent state, and skill representation. Investigate how systems can use feedback and prior experience to perform better without introducing regressions.
  • Evaluation and experimentation. Build realistic evaluations, analyze task trajectories, and turn observed failures into measurable improvements. Compare approaches using quality, reliability, latency, and cost.
  • Model and tooling decisions. Evaluate models and emerging techniques, prototype promising approaches, and make informed build-versus-buy decisions. Choose tools because they solve the problem, and be willing to replace them when the evidence changes.
  • Technical leadership. Set engineering standards, review important design decisions, and help the team implement a coherent AI system. Stay close to the product and accountable for what ships.

What they're looking for

  • You have personally built and shipped a substantial agentic system. Production use or rigorous, reproducible open-source work matters more than the name of a framework or employer.
  • You have deep practical experience with LLM tool use, planning, context engineering, and evaluations. You have implemented multi-agent coordination or substantial parallel agent/tool execution and can explain its failure modes.
  • You have hands-on experience with browser or computer automation in an agentic system, including observing state, verifying effects, and recovering when an interface or execution path fails.
  • You are an excellent software engineer in Python, TypeScript, or a comparable language. You are comfortable with asynchronous services, state machines, persistence, concurrency, retries, and cancellation.
  • You know which decisions belong to a model and which guarantees must be enforced in code. You can reason carefully about permissions, untrusted inputs, uncertain external outcomes, and human approvals.
  • You can design meaningful experiments, debug real system behavior, and explain what improved, why it improved, and where the evidence is still weak.
  • You can take technical ownership of an unclear problem, work effectively with other engineers, and ship with urgency and care.

Nice to have

  • Depth in agent memory and retrieval, skill acquisition, reinforcement learning or post-training, trajectory datasets, sandboxed execution, distributed systems, inference optimization, or multimodal and voice models would be valuable.

Summarised by NextRaise from the employer’s description, which follows in full below.

Full description from employer

Artisan · Full-time · In person in San Francisco

Base salary: $250,000–$325,000 USD annually. Equity: 0.15%–0.30%. US visa sponsorship available.

Build something new at the frontier of applied AI

At Artisan, we're working on a new, ambitious project that will push the boundaries of what agentic AI can do. We're keeping the product details private ahead of launch, but we can tell you this: the technical problems are substantial, the scope for invention is real, and this hire will shape the core technology.

We're looking for a hands-on technical lead to design and build the underlying AI architecture. You'll work across agent behavior, complex multi-agent systems, tool use, context, and evaluation, taking promising ideas through to dependable production software.

This is an individual-contributor role with broad technical ownership. You'll make consequential architecture decisions, write the hardest parts of the system, and work closely with our existing engineers and leadership. You should enjoy both exploring an uncertain problem and doing the detailed engineering required to make a solution work.

What you'll own

  • Agent architecture and behavior. Design and implement agent execution loops, planning strategies, tool interfaces, and verification. Turn ambiguous technical requirements into clear system boundaries and working code.

  • Multi-agent systems. Build delegation, coordination, context sharing, and result synthesis. Handle concurrent work, conflicting updates, cancellation, and stale results. Establish when a multi-agent approach improves on a simpler baseline.

  • Reliable execution. Make complex, stateful workflows resilient to interruptions and partial failures. Build checkpoints, recovery strategies, and appropriate human intervention into the architecture.

  • Context, memory, and reusable methods. Improve retrieval, context construction, persistent state, and skill representation. Investigate how systems can use feedback and prior experience to perform better without introducing regressions.

  • Evaluation and experimentation. Build realistic evaluations, analyze task trajectories, and turn observed failures into measurable improvements. Compare approaches using quality, reliability, latency, and cost.

  • Model and tooling decisions. Evaluate models and emerging techniques, prototype promising approaches, and make informed build-versus-buy decisions. Choose tools because they solve the problem, and be willing to replace them when the evidence changes.

  • Technical leadership. Set engineering standards, review important design decisions, and help the team implement a coherent AI system. Stay close to the product and accountable for what ships.

You'll partner with product and infrastructure engineers on production services, integrations, secure execution, and observability. You'll own the AI architecture and its effectiveness, with implementation shared across the team.

What we're looking for

  • You have personally built and shipped a substantial agentic system. Production use or rigorous, reproducible open-source work matters more than the name of a framework or employer.

  • You have deep practical experience with LLM tool use, planning, context engineering, and evaluations. You have implemented multi-agent coordination or substantial parallel agent/tool execution and can explain its failure modes.

  • You have hands-on experience with browser or computer automation in an agentic system, including observing state, verifying effects, and recovering when an interface or execution path fails.

  • You are an excellent software engineer in Python, TypeScript, or a comparable language. You are comfortable with asynchronous services, state machines, persistence, concurrency, retries, and cancellation.

  • You know which decisions belong to a model and which guarantees must be enforced in code. You can reason carefully about permissions, untrusted inputs, uncertain external outcomes, and human approvals.

  • You can design meaningful experiments, debug real system behavior, and explain what improved, why it improved, and where the evidence is still weak.

  • You can take technical ownership of an unclear problem, work effectively with other engineers, and ship with urgency and care.

Useful additional experience

Depth in agent memory and retrieval, skill acquisition, reinforcement learning or post-training, trajectory datasets, sandboxed execution, distributed systems, inference optimization, or multimodal and voice models would be valuable. We expect strong foundations and particular depth in a few areas, rather than prior specialization in every one.

There is no required degree, publication record, previous employer, or agent framework. We're hiring for demonstrated engineering ability, judgment, and ownership.

How we work

This role is based in our San Francisco office. Expect a small team, short feedback cycles, direct communication, and high standards. We value people who move quickly, take responsibility for the result, surface problems early, and change their minds when the evidence calls for it.

You'll have substantial freedom to explore ambitious technical ideas and the responsibility to turn the best ones into software that works. We'll discuss the project in more detail during the interview process.

Interview process

Our conversations will center on systems you've built, a practical agent architecture and debugging exercise, and how you work with a team.

When applying, tell us about one agentic system you personally owned: what you built, the hardest failure you fixed, and how you measured the improvement. A project link, technical write-up, or open-source contribution is welcome. You can discuss confidential work without sharing proprietary material.

The base salary range is $250,000–$325,000 USD annually, plus equity. The final offer will reflect relevant experience, demonstrated skills, and the scope of responsibility.

About Artisan

Artisan builds AI employees that take on real work. Our first three are Ava, our outbound AI BDR; Aaron, our inbound AI SDR; and Aria, our AI account executive. Our broader mission is to build AI employees that can take responsibility for work across many roles and industries.

Company

Artisan
San Francisco, United States of America

Company facts come from this company's own listings. We only show what the postings themselves carry.

Sourced from Artisan's careers site·first seen 13 Sept 2026·last verified 13 Sept 2026·How we source jobs

Similar jobs

  • Senior Founding AI Engineer at cleraSan Francisco, United States of America–match not yet calculated
  • Sr. AI Engineer, CX at gametimeunitedUnited States - Remote–match not yet calculated
  • AI Engineer (On-Site, Indiana Only) at alliedsolutionsCarmel, United States of America–match not yet calculated
  • GTM AI Engineer -Deal Desk at gomotiveUnited States - Remote–match not yet calculated
  • Sr. Engineer - GenAi at skechersManhattan Beach, United States of America–match not yet calculated

Browse more jobs

  • AI Engineer jobs in United States
  • Machine Learning Engineer jobs in United States
  • AI / ML Researcher jobs in United States
  • Computer Vision Engineer jobs in United States
  • AI Engineer jobs in India
  • AI Engineer jobs in Germany