OpenAI · Sep 2025
OpenAI said SafetyKit uses its latest models to expand multimodal risk agents for fraud and prohibited-activity detection across text, images, transactions, and product listings.
Public company, workplace, funding, and market signals
Updated Jul 30, 2026
SafetyKit is a San Francisco-based YC S23 software company that builds AI agents and a network foundation model to automate trust & safety, fraud prevention, content moderation, merchant investigations, and compliance for marketplaces, payments, and fintech platforms.
Primary product
AI agents for trust & safety, fraud prevention, and policy enforcement automation
Founded
2023
Headquarters
San Francisco, California, United States
Team size
20-30 employees
Work style
Onsite
Industry
Software Development
Sub-industry
Trust & Safety automation / fraud and compliance AI
Offices
Business model
Leadership
CEO / Founder
CTO / Founder
Head of GTM Strategy & Business Operations
Head of Risk and Policy
Chief of Staff to the CEO
Marketing Lead
Head of Strategic Accounts
Founders
Founder
Investors
High-autonomy, engineering-led, high-speed culture; small full-stack team; strong ownership over infrastructure and customer deployments; in-person San Francisco, five days a week; high trust and high expectations.
Work style
Onsite
Compensation
No public salary bands were found; a 2026 GTM hiring post mentioned a $1,000 referral bonus for successful introductions.
Differentiators
Technology
Customers
Competitors
Estimated revenue
≈$1M estimated ARR
Estimated monthly visits
9K
Traffic estimate as of Jul 2026
OpenAI · Sep 2025
OpenAI said SafetyKit uses its latest models to expand multimodal risk agents for fraud and prohibited-activity detection across text, images, transactions, and product listings.
SafetyKit · Jun 2025
SafetyKit announced a $27M raise and named Ribbit, First Round Capital, Y Combinator, and several angel investors; it highlighted customers such as Patreon, Eventbrite, Upwork, and Character.ai.
SafetyKit
SafetyKit described collaborating with OpenAI, ROOST, and Discord on gpt-oss-safeguard safety models for classifying online harms.
SafetyKit
SafetyKit says Upwork sends 100% of posts for review and reports $1M+ in savings, >95% accuracy, and large reductions in manual review volume.
SafetyKit
SafetyKit says Eventbrite adopted its policy enforcement in under two weeks and used it to review events, media, and external links at scale.