Conceptio
›
ai-safety
Topic
ai-safety
Knowledge-graph topic
· documents ABOUT ai-safety across the archive
25
Documents about ai-safety
Documents about ai-safety
Recursive Closure in AI Systems: A Reflection Pattern Account of Stabilization, Permeability, and Safety
#728566
PhilArchive
When Patients Cut In: Extending Clinical Conversational AI Safety to Interruptions
#790033
arXiv (OAI Expanded)
When Patients Cut In: Extending Clinical Conversational AI Safety to Interruptions
#794052
arXiv (All)
Defining Operational Conditions for Safety-Critical AI-Based Systems from Data
#797334
arXiv (OAI Expanded)
Defining Operational Conditions for Safety-Critical AI-Based Systems from Data
#798748
arXiv (All)
Hacking the bomb? What Claude Mythos AI reveals about the gamble of nuclear deterrence
#205859
HAL (France)
The Inalienability of Human Assignment Sovereignty and the Necessity of Physical Governance for AI: An AI Safety Governance Framework Grounded in the Analysis of the Assignment Symbol “=”
#691026
PhilArchive
Governed MCP: Kernel-Level Tool Governance for AI Agents via Logit-Based Safety Primitives
#120451
arXiv CS
Explainable AI (XAI) for transparent resource allocation in public safety communications networks.
#157214
NCBI PubMed Central
Maturity, Safety, and Equity of AI-Enabled Systems and Triage in Integrated Primary Care.
#189981
NCBI PubMed Central
The Illusion of Safety: Multi-Tier Verification of AI vs. Human C++ Code
#329169
arXiv CS
AI-Driven Safety and Security for UAVs: From Machine Learning to Large Language Models
#339529
Semantic Scholar
MedFailBench: A Clinician-Built Open-Source Benchmark for Medical AI Safety Boundary Inspection
#373428
arXiv CS
When Robots Mishear Us: Mapping the Safety Risks of Voice-Controlled Embodied AI
#618595
arXiv CS
AI-optimized BPNN model for port safety risk prediction and management.
#636601
NCBI PubMed Central
Understanding as an Explicit and Assessable Component of Frontier AI Safety Decisions
#656777
arXiv (OAI Expanded)
Understanding as an Explicit and Assessable Component of Frontier AI Safety Decisions
#658913
arXiv (All)
Coverage-Driven Verification for Safety-by-Design in AI-Based Collision Avoidance Systems
#661007
arXiv (OAI Expanded)
Coverage-Driven Verification for Safety-by-Design in AI-Based Collision Avoidance Systems
#662057
arXiv (All)
Silent Revision: Measuring Undisclosed Change in the Safety Frameworks of Frontier AI Developers
#668093
arXiv CS
From Structural Dynamics to Robot Safety Runtime Assurance, Hard Constraints, and Persistent Correction in Embodied AI
#695093
PhilArchive
Legally mandated pre-deployment evaluations: promoting AI safety in clinical medicine across Africa
#956010
SU+ Digital Repository
Hub Architecture: A Layered Approach to Evidence-Based AI Reasoning
#201446
Zenodo (CERN)
Safety for Whom? Boundary-Aware Self-Distillation for Controlled LLM Safety Refusal
#927768
arXiv (All)
Benchmarking Relational Safety in Empathic AI: Perceived Empathy and Relational Appropriateness Can Diverge Across Large Language Models
#404635
OSF
Beyond AI Psychosis and Sycophancy: Structural Drift as a System-Level Safety Failure
#28630
bioRxiv / medRxiv
Range Control as the Foundational Unification of AI Alignment and Safety: Why Probabilistic Frameworks Necessarily Separate Them
#170353
PhilArchive
AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security
#238508
arXiv CS
Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety
#349658
arXiv CS
What Current AI Benchmarks Leave Unmeasured: Modality, Search, Citations, and Implications (for Safety Evaluations)
#432570
arXiv CS
Pre-implementation safety evaluation of an AI decision-support system for surgical antimicrobial prophylaxis.
#449951
NCBI PubMed Central
Operational Hallucination and Safety Drift in AI Agents
#633566
arXiv (OAI Expanded)
Operational Hallucination and Safety Drift in AI Agents
#635915
arXiv (All)
From Pocket God to Digital Jonestown: A Risk Taxonomy and Evaluation Framework for Spiritual AI Safety
#691590
PhilArchive
The Category Error in Contemporary AI Safety Discourse and Why Non-Sentient Systems Cannot Be Moral Machines
#719047
PhilArchive
OpenAgentFlow: Enabling System-Wide Safety Boundaries for Heterogeneous AI Agent Fleets
#822027
arXiv (All)
A Translational Note on AI Safety Evaluation
#974685
arXiv (All)
CARES: A Conversational AI System for Regulation-Grounded Safety Reporting in Construction Education
#1014023
arXiv (All)
GiriMedicolegal™ CDSS
#1028113
Zenodo (CERN)
The Credibility Illusion: When AI Writing Fluency Undermines Human Fact-Checking
#253459
Figshare
The Adaptation Dilemma: Cultural Fit Does Not Guarantee Safety during Mental-Health Interactions with LLMs
#1001815
OSF
PKU-Alignment/PKU-SafeRLHF
#786813
Hugging Face Datasets
Ahimsa AI Framework: A Multi-Layer Approach to Implementing Non-Violence Principles in Large Language Model Safety
#74876
PhilArchive
Transformative Integration of Agentic Generative AI in Food Safety Systems: Policy Framework, Implementation Guidelines, and Economic Impact Analysis
#98237
PhilArchive
Safety Evaluation of a Generative AI Agent for Anxiety and Depression Symptoms
#146577
OSF
New Frontiers in AI-Nano Converged Platforms for Intelligent Diagnostics, Therapeutics, and Safety Evaluation.
#323204
NCBI PubMed Central
ARTIFICIAL INTELLIGENCE (AI) IN PHARMACEUTICAL SCIENCES: REDEFINING DISCOVERY, DEVELOPMENT, AND PATIENT SAFETY
#628844
DOAJ
Using ambient AI in clinical consultations: reframing policy around clinical audit and patient safety.
#630891
NCBI PubMed Central
“If we are good friends, AI doesn't spy so much”: Children’s knowledge and misconceptions of AI safety
#672364
PsyArXiv
Biosecure-LLM Framework: Protecting LLMs from Cyberbiosecurity Threats and the Case for Independent AI Safety Governance
#292587
ODU Digital Commons
← Previous
Page 3 of 5
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.