Fieldguide is the agentic AI platform for audit and advisory firms. Field Agents run planning, testing, and review—trusted by 50% of the top 100 firms. Backed by Bessemer.
About the role
The Foundation Agents team stewards the long-horizon agents powering the Fieldguide AI platform. We work at the frontier of AI product development: agent knowledge, evaluations, and improving quality and reliability at scale. As a Senior Software Engineer, Agents, you'll take ownership of how the team measures and improves agent quality, and help drive the platform forward.
What they're looking for
- You've built AI products end-to-end, with real ownership over agent quality outcomes
- You think in evals and error analysis as a discipline, you dig into why an agent failed and fix the systemic cause
- You're strong in the backend and comfortable owning platform-level systems
- You're motivated by long-horizon agents that do real work in production
- You multiply the people around you, not just your own output
- 1+ years working specifically on agents
More about this role
Fieldguide is establishing a new state of trust for global commerce and capital markets by automating and streamlining the work of assurance and audit practitioners—specifically in cybersecurity, privacy, and financial audits. We build software for the people who enable trust between businesses.
We're based in San Francisco, CA, and backed by Goldman Sachs Alternatives, Bessemer Venture Partners, 8VC, Floodgate, Y Combinator, and more. Over 50 of the top 100 accounting and consulting firms trust Fieldguide to power mission-critical work.
The Foundation Agents team stewards the long-horizon agents powering the Fieldguide AI platform. We work at the frontier of AI product development: agent knowledge, evaluations, and improving quality and reliability at scale. As a Senior Software Engineer, Agents, you'll take ownership of how the team measures and improves agent quality, and help drive the platform forward.
Evals strategy and error-analysis practice, shaping how the team measures and improves agent quality
Design and build agent knowledge and evaluation infrastructure for Fieldguide's long-horizon agents
Lead error analysis on agent behavior, turning findings into concrete...
Browse similar: AI jobs · AI startup jobs · Startup jobs · Remote jobs · San Francisco Bay Area