Conceptio
›
ai-safety
Topic
ai-safety
Knowledge-graph topic
· documents ABOUT ai-safety across the archive
25
Documents about ai-safety
Documents about ai-safety
The Design Paradox: Why Conversational AI Safety Architecture Needs Longitudinal Monitoring
#681831
Zenodo (CERN)
STATIC and AI Safety
#17583
PhilArchive
AI for science with considerations for AI safety
#237410
UR Research
Evaluating whether AI models would sabotage AI safety research
#138968
arXiv CS
Levels of Self-Improvement in AI and their Implications for AI Safety
#774043
PhilArchive
Improving Trust in Safety-Critical AI Systems: Explainable AI and Anomaly Detection Frameworks for human safety in Smart Industries
#274754
IRIS
The Unfireable Safety Kernel: Execution-Time AI Alignment for AI Agents and Other Escapable AI Systems
#306943
arXiv CS
Deception-Enabling Cognitive Capabilities in AI Systems
#376411
OSF
(AI Rights 2): AI Safety Through Economic Integration: Why Markets Outperform Control
#92270
PhilArchive
Automating Inequality? AI, Women's Online Safety, and the Case for Proactive Governance
#638192
Zenodo (CERN)
The Adaptation Dilemma: Cultural Fit Does Not Guarantee Safety in Mental-Health LLMs
#976744
PsyArXiv
On AI Safety and Security Technical Debt in Engineering AI-Enabled Systems
#405576
arXiv CS
Harmonizing AI Safety Thresholds
#381770
arXiv CS
AI Rights for Human Safety
#704618
PhilArchive
Understanding Misalignment in AI Agents
#634949
eScholarship
When Human-in-the-Loop Fails: An Answerability Test for Deployed AI Systems
#28982
DataCite
When Human-in-the-Loop Fails: An Answerability Test for Deployed AI Systems
#28983
DataCite
AI identity and self-concern: A new theory for AI rights and safety
#119206
PhilArchive
The Möbius Strip — Why AI welfare, AI safety, and human welfare are the same problem
#691577
PhilArchive
Constraint Architecture of Physical AI Deployment: A Coupled Interaction Model
#181272
Zenodo (CERN)
The Cosmological Premise of AI Civilization Safety
#101247
PhilArchive
AI Safety: A Climb To Armageddon?
#723468
PhilArchive
(AI Rights 1): Beyond Control: AI Rights as a Safety Framework for Sentient Artificial Intelligence
#108635
PhilArchive
201 - the logical invalidation of rational thought according to AI - and incredibles consequences for AI SAFETY
#689943
PhilArchive
The Adaptation Dilemma: Cultural Fit Does Not Guarantee Safety during Mental-Health Interactions with LLMs
#1001906
PsyArXiv
Industry Self-Flagging and the Insufficiency Critique of Alignment
#304802
IRIS
Human Natural Structure (HNS): A Structural Operating System for Human Understanding in AI — Full Manuscript
#31488
DataCite
Human Natural Structure (HNS): A Structural Operating System for Human Understanding in AI — Full Manuscript
#31489
DataCite
Delusions by design? How everyday AIs might be fuelling psychosis (and what can be done about it)
#486111
PsyArXiv
Generative AI in healthcare education: How AI literacy gaps could compromise learning and patient safety
#724081
PhilArchive
AI Applications in Food Safety and Quality Control
#88228
PhilArchive
Position: AI Safety Requires Effective Controllability
#229554
arXiv CS
Adversarial Prompting Framework for AI Safety Assessment
#370283
arXiv CS
Item Response Theory for AI Safety
#429395
arXiv CS
AI Safety: Not Optional, Not Later
#673600
arXiv CS
Why Safety-First AI Governance Risks Producing Unsafe Systems
#692739
PhilArchive
A Generative AI-Based Construction Safety Assistant Using Retrieval-Augmented Generation
#675156
Digital Commons @ Michigan Tech
Unified Cognitive Dynamics v3.0: A Non-Markovian Framework for Cognitive Safety in High-Resonance Conversational AI
#142359
Zenodo (CERN)
Artificial intelligence (AI) psychosis: mechanisms, clinical risks and safety considerations in generative AI chatbots.
#295170
NCBI PubMed Central
Functional Safety in Industrial Automation: Integrating Programmable Logic Controllers, Safety PLCs, AI/ML/DL, Control Theory, Safety Invariants, and Uncertainty Quantification
#30177
DataCite
Functional Safety in Industrial Automation: Integrating Programmable Logic Controllers, Safety PLCs, AI/ML/DL, Control Theory, Safety Invariants, and Uncertainty Quantification
#30178
DataCite
AI Coach: A Prototype of Artificially Intelligent Coach for Psychological Safety and Mental Wellbeing in Healthcare
#196202
OSF
Water Efficient Suppression Material Fire Behaviour and AI Driven Safety in Future Fire Protection Research
#635124
HAL (France)
Safety as Understanding: Interpretive Braking, Comprehension-Based Restraint, and the Limits of Compliance-Based AI Safety
#714114
PhilArchive
From Optional Safety to Architectural Responsibility: AI Governance after Models
#716492
PhilArchive
Global Solutions vs. Local Solutions for the AI Safety Problem
#773979
PhilArchive
Acceleration AI Ethics, the Debate between Innovation and Safety, and Stability AI’s Diffusion versus OpenAI’s Dall-E
#74458
PhilArchive
allenai/wildjailbreak
#786974
Hugging Face Datasets
Unsafe at any AUC: Unlearned Lessons From Sociotechnical Disasters for Responsible AI
#682882
Calhoun
Agentic Microphysics: A Manifesto for Generative AI Safety
#19044
arXiv CS
Page 1 of 5
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.