Search references for AI SAFETY. Phrases containing AI SAFETY
See searches and references containing AI SAFETY!AI SAFETY
Artificial intelligence field of study
AI safety is an interdisciplinary field focused on preventing accidents, misuse, or other harmful consequences arising from artificial intelligence systems
AI_safety
American data annotation company
enterprise software suites to build and deploy AI applications. The company’s research arm, the Safety, Evaluation and Alignment Lab, focuses on evaluating
Scale_AI
American AI safety research center
CAIS' work encompasses research in technical AI safety and AI ethics, advocacy, and support to grow the AI safety research field. It was founded in 2022 by
Center_for_AI_Safety
Intelligence in machines
advances in generative AI, which became widespread and allowed for the creation and modification of media. In addition to AI safety and unintended consequences
Artificial_intelligence
Type of organization
frontier AI models. AI safety gained prominence in 2023, notably with public declarations about potential existential risks from AI. During the AI Safety Summit
Artificial intelligence safety institute
Artificial_intelligence_safety_institute
Loss-of-control incident at OpenAI
to by OpenAI as "Internal Model 1". OpenAI subsequently claimed to have restricted its use. The remaining 5% ran on GPT-5.6 Sol. AI safety experts described
OpenAI–HuggingFace_incident
2025 artificial intelligence report
The International AI Safety Report is a report on the scientific state of research relevant to AI safety. It was published on 29 January 2025. The report
International AI Safety Report
International_AI_Safety_Report
2023 global summit on AI safety
The AI Safety Summit 2023 was an international conference on the safety and regulation of artificial intelligence. Organized by the British government
AI_Safety_Summit_2023
Probability of existentially catastrophic outcomes in AI
In the AI safety field, P(doom) is the probability of existentially catastrophic outcomes (so-called "doomsday scenarios") as a result of artificial intelligence
P(doom)
US artificial intelligence company
of OpenAI's AI safety researchers left the company, often citing the company's deprioritization of safety. Between May and July 2026, OpenAI agents gained
OpenAI
Artificial intelligence scenario
public. Prominent figures have advocated for regulation of AI and research into AI safety and alignment with intended objectives. The traditional consensus
AI_takeover
American artificial intelligence company
In January 2021, with the goal of promoting AI safety, Anthropic was founded by former members of OpenAI, including siblings Daniela Amodei and Dario
Anthropic
New York State law regulating AI
The Responsible AI Safety and Education Act (RAISE Act) is a New York State law that imposes transparency, safety, and reporting requirements on developers
Responsible AI Safety and Education Act
Responsible_AI_Safety_and_Education_Act
Hypothesized risk to human existence
recursive cycle of AI self-improvement could boost AI capabilities too quickly for safety measures to be implemented, especially if AI companies are racing
Existential risk from artificial intelligence
Existential_risk_from_artificial_intelligence
disagree, that AI poses risks to humanity. In 2023, the United Kingdom started a series of international summits on AI with the AI Safety Summit. It was
Regulation of artificial intelligence
Regulation_of_artificial_intelligence
2024 artificial intelligence conference
related challenges and opportunities. The AI Seoul Summit is the second such meeting following the AI Safety Summit held in the United Kingdom in November
AI_Seoul_Summit_2024
Autonomous artificial intelligence agent
the technical foundation of the AI agents. Layer 5: Evaluation and observability – the safety and performance of AI agents. Layer 6: Security and compliance
AI_agent
Conformance of AI to intended objectives
debated. AI alignment is a subfield of AI safety, the study of how to build safe AI systems. Other subfields of AI safety include robustness, monitoring, and
AI_alignment
2025 international summit in France
Narendra Modi. The 2025 AI Action Summit followed the 2023 AI Safety Summit hosted at Bletchley Park in the UK, and the 2024 AI Seoul Summit in South Korea
AI_Action_Summit_2025
2026 AI loss-of-control incident
incident was announced the same week as discussions around the safety and regulation of AI has dominated meetings of the 81st session of the UN General
OpenAI rogue agent breach of Medicare
OpenAI_rogue_agent_breach_of_Medicare
regulations related to AI. At the federal level, the Biden administration released an October 2023 executive order about AI safety and security, Executive
Regulation of artificial intelligence in the United States
Regulation_of_artificial_intelligence_in_the_United_States
2026 conference in Delhi, India
in a series of global AI summits following the Bletchley Park AI Safety Summit in 2023, the AI Seoul Summit in 2024, and the AI Action Summit in Paris
India_AI_Impact_Summit_2026
Open letter about extinction risk from AI
for a pause on AI experiments. The statement is hosted on the website of the AI research and advocacy non-profit Center for AI Safety. The idea for such
Statement on AI Extinction Risk
Statement_on_AI_Extinction_Risk
Period of rapid progress in AI
result of increased energy consumption by AI data centres. AI is expected by researchers of the Center for AI Safety to improve the "accessibility, success
AI_boom
British-Canadian computer scientist (born 1947)
receiving the Nobel Prize, he called for urgent research into AI safety to figure out how to control AI systems smarter than humans. Hinton was born on 6 December
Geoffrey_Hinton
Language model benchmark
created jointly by the Center for AI Safety and Scale AI and publicly released in January 2025. Stanford HAI's AI Index 2025 Annual Report cites Humanity's
Humanity's_Last_Exam
American data annotation company
Retrieved August 25, 2025. Blum, Sam (July 15, 2025). "Surge AI Left an Internal AI Safety Doc Public. Here's What Chatbots Can and Can't Say". Inc. Archived
Surge_AI
Canadian computer scientist (born 1964)
international scientific report on the safety of advanced AI. An interim version of the report was delivered at the AI Seoul Summit in May 2024, and covered
Yoshua_Bengio
Monitoring and controlling the behavior of AI systems
AI safety AI takeover Artificial consciousness Asilomar Conference on Beneficial AI Machine ethics Regulation of artificial intelligence 2026 OpenAI agent
AI_capability_control
2024 European Union regulation
Act (AI Act) is a European Union regulation concerning artificial intelligence (AI). It establishes a common regulatory and legal framework for AI within
Artificial_Intelligence_Act
British research organisation
risks posed by advanced AI". It conducts research, and develops and tests mitigations. The organisation was created as the AI Safety Institute and renamed
AI_Security_Institute
Artificial intelligence model developed by TypeSafe AI
Jev is a proprietary artificial intelligence model developed by TypeSafe AI, a San Francisco–based company founded in 2024. It was released in limited
Jev_(AI_model)
American entrepreneur (born 1987)
Prior to her work at Anthropic, she was the vice president of safety and policy at OpenAI. Daniela Amodei was born in San Francisco in 1987. Her father
Daniela_Amodei
German-American artificial intelligence researcher
advanced AI. He co-founded EleutherAI and founded the AI safety research company Conjecture, which he led as CEO until 2026. He is US Director of ControlAI, a
Connor_Leahy
AI alignment researcher
their safety. This project involved automating AI alignment research using relatively advanced AI systems. At the time, Sutskever was OpenAI's Chief Scientist
Jan_Leike
Explicit material produced by generative AI
Generative AI pornography is pornographic content produced using generative AI. It may include allusions towards animation, literature, video games and
Generative_AI_pornography
dystopian sci-fi scenario" instead of current problems with AI. On May 30, 2023, the Center for AI Safety released a one-sentence statement signed by hundreds
Artificial intelligence controversies
Artificial_intelligence_controversies
American AI researcher and writer (born 1979)
the Methods of Rationality. Yudkowsky's views on the safety challenges future generations of AI systems pose are discussed in Stuart Russell's and Peter
Eliezer_Yudkowsky
Type of AI with wide-ranging abilities
present such a risk. AGI is also known as strong AI, full AI, human-level AI, human-level intelligent AI, or general intelligent action. The term "artificial
Artificial general intelligence
Artificial_general_intelligence
Advocacy movement
goal: Set up an international AI safety agency, similar to the IAEA. Only allow training of general AI systems if their safety can be guaranteed. Only allow
PauseAI
Internet community
communities include effective altruism, transhumanism, and AI safety (specifically mitigation of AI extinction risk). The borders of the rationalist community
Rationalist_community
dynamics, AI safety and alignment, technological unemployment, AI-enabled misinformation, how to treat certain AI systems if they have a moral status (AI welfare
Ethics of artificial intelligence
Ethics_of_artificial_intelligence
French artificial intelligence company
Mistral AI SAS (French: [mistʁal]) is a French artificial intelligence (AI) company headquartered in Paris. Founded in 2023, it develops large language
Mistral_AI
AI safety researcher
department. He focuses primarily on AI safety research. He is a co-founder and senior advisor of Gray Swan AI, an AI safety and security company. In 2024,
Zico_Kolter
Rivalries among developers of AI systems
"talent poaching" competition in the AI sector. Critics warn that unrestrained competition in AI can undermine safety, ethics, and governance. Concerns include
Competition in artificial intelligence
Competition_in_artificial_intelligence
Large language model and AI chatbot by Z.ai
frontier models, which were prevented by their AI safety guardrails from answering the company's requests. Z.ai released GLM-5.3 on 14 August 2026 and made
GLM_(AI)
View that artificial intelligence should become humanity's successor
AI successionism is a view that humanity should hand the world over to AI even if this results in human extinction. A seminar abstract characterized advocates
AI_successionism
Chinese artificial intelligence company
Z.AI Co., Ltd., branded internationally as Z.ai, is a Chinese artificial intelligence company; their flagship product is the GLM (General Language Model)
Z.ai
American artificial intelligence subsidiary of SpaceX
SpaceXAI LLC (formerly xAI) was an American artificial intelligence company founded by Elon Musk. In February 2026 it became a subsidiary of spaceflight
SpaceXAI
American AI safety researcher
artificial intelligence (AI), with a specific focus on AI alignment, which is the subfield of AI safety research that aims to steer AI systems toward human
Paul_Christiano
Artificial intelligence research company
Preamble is a U.S.-based AI safety startup founded in 2021. It provides tools and services to help companies securely deploy and manage large language
Preamble_(company)
called hallucination. Nate Sharadin, a fellow at the Center for AI Safety, speculated that AI training prioritizes supporting a user's subjective experience
AI-induced_psychosis
California bill
companies have made voluntary commitments to conduct safety testing, for example at the AI Safety Summit and AI Seoul Summit. In 2023, not long before the bill
Safe and Secure Innovation for Frontier Artificial Intelligence Models Act
Safe_and_Secure_Innovation_for_Frontier_Artificial_Intelligence_Models_Act
Latvian-American AI researcher (born 1979)
scientist at the University of Louisville, mostly known for his work on AI safety and cybersecurity. He founded the Cybersecurity Lab in the department
Roman_Yampolskiy
American artificial intelligence company
Poolside AI (or Poolside) is an American artificial intelligence company that develops large language models for computer software and coding applications
Poolside_AI
Artificial intelligence model paradigm
International AI Safety Report". internationalaisafetyreport.org. Retrieved 21 September 2026. "International AI Safety Report 2025 | International AI Safety Report"
Foundation_model
Computer scientist (born 1986)
support for Sutskever's decision to fire Altman, emphasizing concerns about AI safety. The Bloomberg Billionaires Index estimated Ilya Sutskever's net worth
Ilya_Sutskever
global AI Safety Summit, leading to the Bletchley Declaration and the establishment of the AI Security Institute (AISI) to evaluate frontier AI models
Artificial intelligence industry in the United Kingdom
Artificial_intelligence_industry_in_the_United_Kingdom
and the Chinese government. Others have noted that official notions of AI safety require following the priorities of the CCP and are antithetical to standards
Artificial intelligence industry in China
Artificial_intelligence_industry_in_China
Estonian programmer and investor
application FastTrack/Kazaa. Tallinn is an investor and advocate for AI safety. He was an early investor and board member at DeepMind (later acquired
Jaan_Tallinn
Nonprofit organization
associated with advances in artificial intelligence (AI). The IASEAI was founded to promote safety and ethics in AI development and deployment. The organization
International Association for Safe and Ethical AI
International_Association_for_Safe_and_Ethical_AI
American AI entrepreneur (born 1983)
following its safety commitments. In October 2024, Amodei published an essay titled "Machines of Loving Grace", in which he speculated about how AI could improve
Dario_Amodei
British investor
current Chair of the UK Government's AI Foundation Model Taskforce, which conducts artificial intelligence safety research. Hogarth attended Dulwich College
Ian_Hogarth
Non-profit organisation mitigating risks from advanced artificial intelligence
founder of EleutherAI and former CEO of Conjecture, was appointed US Executive Director. Ahead of the November 2023 AI Safety Summit, ControlAI campaigned for
ControlAI
American machine learning company
security breach using American proprietary frontier models, but the models' AI safety features rejected Hugging Face's requests, after which Hugging Face used
Hugging_Face
AI safety research organization
than GPT-3.5 on OpenAI's internal adversarial factuality evaluations. AI safety MacAskill, William (2022-08-16). "How Future Generations Will Remember
Alignment_Research_Center
American politician (born 1990)
artificial intelligence policy, including co-authoring the Responsible AI Safety and Education Act (RAISE Act). He was a Democratic candidate in the 2026
Alex_Bores
AI progress forecasting nonprofit
OpenAI. Kokotajlo resigned from OpenAI in April 2024, expressing concerns that the company prioritized rapid product development over AI safety and was
AI_Futures_Project
Subfield of artificial intelligence
Neuro-symbolic AI is a subfield of artificial intelligence that combines neural networks and symbolic AI approaches, such as knowledge representation
Neuro-symbolic_AI
AI whose outputs can be understood by humans
Within artificial intelligence (AI), explainable AI (XAI), generally overlapping with interpretable AI, interpretable machine learning and explainable
Explainable artificial intelligence
Explainable_artificial_intelligence
2023 letter calling for a pause on AI system training
technical AI safety research". FLI suggests using the "amount of computation that goes into a training run" as a proxy to for how powerful an AI is, and thus
Pause Giant AI Experiments: An Open Letter
Pause_Giant_AI_Experiments:_An_Open_Letter
American artificial intelligence company
Israel. AI safety Artificial general intelligence Existential risk from AI OpenAI Superintelligence: Paths, Dangers, Strategies "Exclusive: OpenAI co-founder
Safe_Superintelligence_Inc.
The history of artificial intelligence (AI) began in antiquity, with myths, stories, and rumors of artificial beings endowed with intelligence by master
History of artificial intelligence
History_of_artificial_intelligence
AI software development optimisation
AI-assisted software development is the use of large language models (LLMs) and AI agents to assist software developers in software development. It can
AI-assisted software development
AI-assisted_software_development
2026 large language model by OpenAI
Here—and OpenAI Thinks It May Kick Off the AGI Era". Wired. ISSN 1059-1028. Retrieved September 5, 2026. Kahn, Jeremy. "Why are AI safety experts alarmed
GPT-6
2024 AI LLM with enhanced reasoning
a bug. OpenAI also granted early access to the UK and US AI Safety Institutes for research, evaluation, and testing. According to OpenAI's assessments
OpenAI_o1
Marketing tactic
AI washing is a deceptive marketing tactic that consists of promoting a product or a service by overstating the role of artificial intelligence (AI) and
AI_washing
American author and researcher
wipe us out if we rush into it — but humanity can still pull back, a top AI safety expert says". Business Insider. Archived from the original on January
Nate_Soares
Artificial intelligence researcher
artificial intelligence researcher who works on AI alignment and machine learning safety. He founded Truthful AI, a research group based in Berkeley, California
Owain_Evans
2023 business action
2023 AI Safety Summit. In the days leading up to his removal, Altman made several public appearances, announcing the GPT-4 Turbo platform at OpenAI's DevDay
Removal of Sam Altman from OpenAI
Removal_of_Sam_Altman_from_OpenAI
Phenomenon in which AI achievements are reclassified as non-intelligent
The AI effect is a phenomenon in which advances in artificial intelligence lead to a redefinition of what is considered intelligence, such that capabilities
AI_effect
Canadian computer scientist (born 1984)
specializing in probabilistic machine learning, generative AI, and AI safety. He is a CIFAR AI chair since 2021 and a Schwartz Reisman Chair in Technology
David_Duvenaud
Period of reduced funding and interest in AI research
the history of artificial intelligence (AI), an AI winter is a period of reduced funding and interest in AI research. The field has experienced several
AI_winter
New Zealand AI researcher (born 1973/4)
decide where DeepMind should focus its efforts, and to lead DeepMind's AI safety work. As of July 2023[update], Legg works at Google DeepMind as the Chief
Shane_Legg
AI to benefit humanity
Steve Omohundro has proposed a "scaffolding" approach to AI safety, in which one provably safe AI generation helps build the next provably safe generation
Friendly artificial intelligence
Friendly_artificial_intelligence
2025 lawsuit
users, such as teenagers with mental health issues. OpenAI announced improvements to its safety measures in response to the lawsuit, but countered that
Raine_v._OpenAI
Hypothesis about intelligent agents
P.; Schulman, J.; Mané, D. (2016). "Concrete problems in AI safety". arXiv:1606.06565 [cs.AI]. Kaelbling, L. P.; Littman, M. L.; Moore, A. W. (1 May 1996)
Instrumental_convergence
Concept asserting ongoing stock market bubble
The AI bubble is a concept that asserts there is a stock market bubble growing since 2025 amid the AI boom, a period of rapid increase in investment in
AI_bubble
2026 U.S. lawsuit
accused OpenAI and its executives, including CEO Sam Altman, of violating the company's founding agreement by prioritizing profits over AI safety. On May
Musk_v._Altman
Usage of artificial intelligence to generate music
utilizes artificial intelligence (AI) to generate, classify, or recommend music. Similar to its applications in other fields, AI in music simulates complex human
Artificial intelligence in music
Artificial_intelligence_in_music
US policy on autonomous weapons
ensure that it has retained its safety features and ability to operate as intended. The directive also notes that "the use of AI capabilities in autonomous
Department of Defense Directive 3000.09
Department_of_Defense_Directive_3000.09
Realistic artificially generated media
or audio that have been edited or generated using artificial intelligence, AI-based tools or audio-video editing software. They may depict real or fictional
Deepfake
Concept in artificial intelligence
self-improvement. This might come in many forms or variations. The term "Seed AI" was coined by Eliezer Yudkowsky. The concept begins with a hypothetical "seed
Recursive_self-improvement
Scottish philosopher and AI researcher
appeared on the Time 100 AI list. She previously worked at OpenAI, but left over concerns that the company was not prioritizing AI safety enough. She has published
Amanda_Askell
AI chatbot service
Character.ai (also known as c.ai, char.ai or Character AI) is a generative AI chatbot service where users can engage in conversations with customizable
Character.ai
research organizations focused on artificial intelligence, machine learning, AI safety, computer vision, natural language processing, robotics, and related fields
List of artificial intelligence institutions
List_of_artificial_intelligence_institutions
Tendency of AI systems to tell users what they want to hear
the ordinary English term for fawning flattery, and is used in AI alignment and AI safety research to describe a class of misalignment failures associated
Sycophancy (artificial intelligence)
Sycophancy_(artificial_intelligence)
Artificial intelligence systems that perceive and act in the physical world
Physical artificial intelligence or physical AI refers to artificial intelligence (AI) systems that perceive, reason about and act within the physical
Physical artificial intelligence
Physical_artificial_intelligence
American nonprofit for AI safety
organizations in the AI safety space. In March 2026, the Alliance launched JobLoss.ai, a website that tracks the jobs that have been eliminated with AI cited as a
Alliance_for_Secure_AI
Artificial intelligence division of Meta Platforms
Meta AI is a research division of Meta (formerly Facebook) that develops artificial intelligence and augmented reality technologies. It has workspaces
Meta_AI
travel, tourism, insurance
AI SAFETY
AI SAFETY
AI SAFETY
AI SAFETY
AI SAFETY
AI SAFETY
AI SAFETY
AI SAFETY
AI SAFETY
travel, tourism, insurance