As AI systems grow more capable, ensuring they remain controllable and aligned with human intent becomes a central challenge. The organizations below treat safety not as an afterthought but as a core research discipline.
As AI systems become more powerful, ensuring they remain safe and aligned with human values is crucial. These organizations are dedicated to AI safety research and implementation.
Cloudflare
San Francisco, United States
Cloudflare provides a comprehensive connectivity cloud platform delivering security, performance, and reliability for websites, applications, and APIs. Their core AI offering, Workers AI, enables developers to deploy and run machine learning models directly on Cloudflare’s globally distributed edge network. This uniquely positions Cloudflare to serve organizations requiring low-latency AI inference and robust protection against evolving cyber threats, particularly for AI-powered applications and workloads.
Roblox AI
San Mateo, United States
Roblox AI develops and deploys artificial intelligence solutions to power the Roblox platform, a user-generated massive multiplayer online game. Their core technology focuses on generative AI tools for content creation alongside AI-driven systems for content moderation and platform safety. Targeting the vast Roblox developer community and its millions of daily users, Roblox AI aims to scale content creation and maintain a safe, engaging virtual environment through automated solutions.
OpenAI
San Francisco, United States
OpenAI is an AI research and deployment company whose stated mission is to ensure artificial general intelligence benefits all of humanity. It develops the GPT family of large language models and products including ChatGPT, the DALL-E and Sora generative models, and widely used developer APIs.
Anthropic
San Francisco, United States
Anthropic is an AI safety and research company that builds Claude, a family of large language models focused on being reliable, interpretable, and steerable. Founded by former OpenAI researchers, it pioneered Constitutional AI and delivers Claude through consumer apps, a developer API, and major cloud partnerships.
Klarna AI
Stockholm, Sweden
Klarna is a Swedish fintech company that provides flexible payment solutions for online and in-store purchases. Their core offering is “Buy Now, Pay Later” services – including options for payment in 3 installments, within 30 days, or longer-term financing – powered by AI-driven credit risk assessment and fraud prevention. Klarna targets consumers seeking convenient and adaptable payment methods, and partners with e-commerce merchants to increase sales through enhanced payment flexibility.
Thinking Machines Lab
San Francisco, United States
Thinking Machines Lab, founded by former OpenAI CTO Mira Murati, builds more understandable and customizable AI systems. Raised a record $2B seed round.
Wiz
New York, United States
Wiz is a cloud security platform that provides vulnerability detection and security posture management for cloud environments. Utilizing a knowledge graph-based approach, Wiz maps cloud assets and identifies misconfigurations, vulnerabilities, and compliance risks across multi-cloud deployments. The platform targets security, development, and DevOps teams seeking to integrate security into the cloud development lifecycle and improve collaboration.
Toss
Seoul, South Korea
Toss is a South Korean fintech super app consolidating personal financial management into a single platform. Utilizing AI-powered fraud detection during transactions and data aggregation, Toss provides users with a unified view of their bank accounts, loans, investments, and spending habits. Targeting individual consumers, Toss simplifies financial oversight and enhances security through proactive risk assessment.
Netskope
Santa Clara, United States
Netskope provides cloud-native security and networking solutions focused on securing data and applications across cloud environments, SaaS, and increasingly, generative AI. Their core technology is a context-aware Zero Trust Engine that analyzes user and entity activity to prevent data loss and mitigate threats. Netskope targets enterprises seeking to consolidate security and networking infrastructure into a single platform for improved visibility, reduced complexity, and consistent policy enforcement across all digital channels.
Fireblocks
New York, United States
Fireblocks provides a platform for the secure custody, transfer, and settlement of digital assets for financial institutions and fintech companies. Their core technology is a multi-layer, AI-powered security infrastructure that mitigates risks associated with blockchain transactions and wallet management. Fireblocks targets businesses requiring scalable and compliant digital asset operations, offering solutions ranging from treasury management to embedded crypto wallets for consumer applications.
Frequently asked questions
What is AI alignment?
AI alignment is the field of ensuring that AI systems pursue the goals their designers and users actually intend, behaving safely and predictably even as they become more capable.
Why does AI safety matter?
Powerful AI deployed at scale can amplify harms — from misinformation to unsafe autonomous actions — so safety research aims to make these systems reliable, interpretable, and controllable.
What do AI safety companies actually do?
They research interpretability, robustness, evaluation, and oversight techniques, and build tooling and benchmarks to test and constrain model behavior.
Is AI safety only about future risks?
No. Much of the work addresses present-day issues like bias, jailbreaks, and reliability, while also preparing for the risks of more autonomous future systems.