
Netskope
Actively Hiring
Netskope is redefining cloud, network, and data security
- B2B
- Scale StageRapidly increasing operations
- Top InvestorsThis company has received a significant amount of investment from top investors
- +4
Job Location
Visa Sponsorship
Not Available
RelocationNot Allowed
Hiring contact
Brendan Lynch
Talent Management • 11 years
San Francisco
About the job
As a Senior AI Quality & Red Team Engineer at Netskope, you will lead the charge in testing, stress-testing, and breaking our AI agents before they ever reach production. From automating multi-turn prompt injections to tracking fleet-wide drift in CI/CD, you will own the automated harness that ensures our AI systems are secure, resilient, and compliant. If you love the idea of being the person who proves an agent isn't ready yet, welcome home.
Skills and competencies:
- Build and grow the automated evaluation suite every agent runs against before it's approved for production, designed to run unattended and scale across a growing agent fleet, not something that needs a person babysitting each run.
- Design adversarial test scenarios — prompt injection attempts, sycophancy checks where an agent has to correctly push back on a false premise, multi-attempt attacks rather than single-shot ones — and automate them so they run on every relevant change, not just before a big release.
- Own the "break it on purpose" pass for every new agent: attempt to extract data it shouldn't expose, get it to act outside its registered tool boundaries, or get it to treat a synthetic test probe as real. As the fleet grows, build this into a repeatable, scriptable process rather than a manual exercise redone from scratch each time.
- Partner with the Data Steward on data sensitivity classification for the systems agents touch, so your test scenarios reflect what's actually at stake, not a generic checklist.
- Decide, for each agent capability, what "pass" actually means, and build that judgment into automated thresholds wherever possible so evaluation keeps up as the number of agents climbs into the hundreds.
- Maintain the guardrail and negative-test catalog (fail-closed vs. fail-graceful behavior) across the platform, and add new cases as new failure modes get discovered in the wild.
- Produce clear, audit-ready evidence for every agent's evaluation results, generated automatically as part of the pipeline rather than assembled by hand for each review.
- Track drift over time across the whole fleet, not agent by agent, so a slow-moving problem in one corner doesn't go unnoticed just because no one's looking at that specific agent that week. Must-Have:
- At least 4 years in software quality, security testing, or a related discipline, with 1–2 years specifically evaluating or red-teaming LLM-based systems — not just running unit tests against traditional code.
- Strong Python skills, since the evaluation harness, adversarial test scripts, and automated pipelines will mostly be built in it. Comfortable writing production-quality code, not just glue scripts.
- Working knowledge of REST APIs and webhook/event-driven patterns, enough to build test harnesses that call an agent's tools directly and validate its inputs and outputs, not just its final chat response.
- Real experience building automated test infrastructure and integrating it into CI/CD, not just manually running test cases — someone who thinks in pipelines and repeatability by default.
- Hands-on familiarity with at least one adversarial testing or LLM eval tool (DeepTeam, Garak, PyRIT, Promptfoo, or similar), and an understanding of how these map to standards like the OWASP LLM Top 10 or NIST's AI risk framework.
- Practical understanding of prompt injection, jailbreaking, and sycophancy failure modes — able to design new test cases for these, not just run ones someone else wrote.
- Comfort making a hard call: willing to block a release when an agent doesn't meet its bar, even under schedule pressure. Strong Advantage:
- Enough understanding of data sensitivity and compliance classification to design tests that reflect real risk, even though the Data Steward owns the classification system itself.
- Experience in a regulated environment where evaluation results had to hold up to an external audit, not just an internal review.
- Familiarity with multi-attempt or persistent-attack testing methodology, rather than only single-shot adversarial prompts.
- Some exposure to how agents are actually built (prompting, tool schemas, orchestration) — not required, but it makes it much easier to design tests that target real failure modes instead of generic ones.
#LI-CV1
About the company
1001-5000
Startup
SaaS
Enterprise Security
B2B · SaaS · Mobile · Artificial Intelligence / Machine Learning
- B2B
- Scale StageRapidly increasing operations
- Top InvestorsThis company has received a significant amount of investment from top investors
- Valuation $1B+This company has a valuation of $1B or more
- 4.2Highly ratedNetskope is highly rated on Glassdoor, with 4.2 out of 5 stars
- 4.1Work / Life BalanceEmployees rate Netskope 4.1/5 on Glassdoor for work / life balance
- 4.1Strong LeadershipEmployees rate Netskope 4.1/5 on Glassdoor for faith in leadership
Perks
Insurance, Health & Wellness
● Medical (UHC-HDHP, PPO, & EPO; CA Kaiser- HMO & HDHP)
● Dental
● Vision
● Equitable Life & AD&D Insurance
● Short & Long Term Disability
● Company HSA Contributions
● Employee Assistance Program (EAP)
401(k) Retirement Savings Plan
401(k)/ROTH offering through Newport Group ($20,500 Annual Max Contribution / $6,500 Annual Max Catch-Up). Eligible to start
deferring after the 1st paycheck.
Voluntary Life Insurance
Employees have the option to enroll in supplemental life insurance up to a max of $250,000. Netskope is pleased to provide spouse and dependent life insurance offerings upon an employee’s insurance election.
Paid Parental Leave
12 weeks Birth Parent Paid Parental Leave
8 weeks Non-Birth Parent Parental Leave
Commuter Benefits
Employees can contribute up to $280/month to a pre-tax account for mass transit and/or parking.
Additional Netskope Perks
● 13+ Company Observed Holidays
● Quarterly Global Wellness Days
● Unlimited Paid Time Off
● Discount program for popular brands, 30,000 national/local offers, and devices
● Meditation Hours
● Family Planning Assistance
● Travel Assistance
