Software Engineer, Agents (Foundation Agents)
You find the fit. Your agent handles the form.
Choose a role or send your matches to the agent. It uses your original résumé and saved details, applies in the cloud, and keeps every result in one place.
What you'll need to apply
Fields this application requires
Company-specific questions
- In one or two sentences, what's one AI product feature you personally built or shipped?essay
- In 2-3 sentences, describe the most complex bug or failure you debugged in an AI/agent system and how you diagnosed it.essay
- What's a specific eval or test you've built to catch a failure mode in an AI system before it reached production? What did it check for?essay
About this role
Employer-provided description, formatted for easier reading.
About Us
Fieldguide is establishing a new state of trust for global commerce and capital markets by automating and streamlining the work of assurance and audit practitioners—specifically in cybersecurity, privacy, and financial audits. We build software for the people who enable trust between businesses.
We're based in San Francisco, CA, and backed by Goldman Sachs Alternatives, Bessemer Venture Partners, 8VC, Floodgate, Y Combinator, and more. Over 50 of the top 100 accounting and consulting firms trust Fieldguide to power mission-critical work.
About the Role
The Foundation Agents team stewards the long-horizon agents powering the Fieldguide AI platform. We work at the frontier of AI product development: agent knowledge, evaluations, and improving quality and reliability at scale. If you're excited by long-horizon agents that do real-world work, this is the team for you.
What You’ll
Own
- Build and maintain agent knowledge and evaluation infrastructure for Fieldguide's long-horizon agents
- Build our next generation of agents at Fieldguide, solving increasingly complex customer use cases
- Perform error analysis on agent behavior and help turn findings into concrete quality improvements
- Build backend systems that support agent execution, evaluation, and monitoring
- Partner with senior engineers on the team to execute the platform's reliability and quality roadmap
Who You Are
- You've built AI products, not just called an LLM API
- You're curious about evals and error analysis. You want to understand why an agent failed, not just that it did
- You're comfortable in the backend, with a bias toward reliable systems
- You're motivated by long-horizon agents that do real work in production
Experience
Must-have:
- Hands-on experience building AI products
- Working knowledge of evals and error analysis
- Backend engineering experience
Nice-to-have:
- Platform engineering skills
- Frontend experience
- Distributed systems experience
What Should Excite You
- Long-horizon agents: working on agents that do real, sustained work and not one-shot demos
- Evaluation as a craft: Evals and error analysis are core to how this team improves quality
- Platform-level impact: Your work shapes the reliability and quality of every agent built on top of it
- Frontier problems: You're working on open problems in agent reliability that don't yet have established playbooks
Benefits
- Competitive compensation with equity
- Comprehensive health and wellness benefits
- Flexible time off and work schedules
- Technology reimbursements
- 401(k) plan
- Twice-yearly in-person offsites across the U.S.
- Wellness benefits starting on your first day
Our Values
- Fearless — Inspire and break down seemingly impossible walls
- Fast — Launch fast with excellence; iterate to perfection
- Lovable — Deliver happiness and 11-star experiences
- Owners — Execute and run the business with ownership
- Win-win — Create mutual value and earn trust for life
- Inclusive — Scale the best ideas with inclusive teams