Conceptio
›
ai-safety
Topic
ai-safety
Knowledge-graph topic
· documents ABOUT ai-safety across the archive
25
Documents about ai-safety
Documents about ai-safety
Illusions Of Control: Why Today’s AI‑Safety Plans Court Catastrophe
#78419
PhilArchive
Illusions Of Control: Why Today’s AI‑Safety Plans Court Catastrophe
#85658
PhilArchive
Containment Verification: AI Safety Guarantees Independent of Alignment
#175347
arXiv CS
The Alignment Tax: How Safety Training Creates Operationally Lethal AI Systems
#248522
PhilArchive
Safety, Security, and Cognitive Risks in Neuro-Symbolic AI
#282736
arXiv CS
Beyond Her: Safety Dynamics in Role-play AI Companions
#321776
arXiv CS
Tanyuan: A Hardware-Grounded Mission Alignment Paradigm for Endogenous AI Safety
#692906
PhilArchive
AI Safety: Not Optional, Not Later
#999906
arXiv (All)
Steering Generative AI Toward Developmentally Supportive Learning: The SCAFFOLD Framework and a Pilot in a School Setting
#989430
PsyArXiv
Safety Overreach in the AI Era: An LMT Analysis of Guardrail Hypersensitivity and the Structural Cost of Excessive Safety Optimization
#695813
PhilArchive
Guidelines to Preventing Artificial Intelligence Hallucinations in Microsoft Fabric
#122155
DataCite
Guidelines to Preventing Artificial Intelligence Hallucinations in Microsoft Fabric
#122156
DataCite
AI-DRIVEN TRAFFIC MANAGEMENT SYSTEM FORZERO VIOLATION AND ENHANCED ROAD SAFETY
#85788
PhilArchive
Data Flow Control: Data Safety Policies for AI Agents
#259587
arXiv CS
Reconceiving Safety Regulation for AI and ML Medical Software.
#439069
NCBI PubMed Central
Rules or Character? Scaling Laws for AI Safety Design
#451343
arXiv CS
INTERPRETIVE SOVEREIGNTY FAILURE: An Interaction-Level Safety Risk in Human–AI Systems
#708708
PhilArchive
The safety failures we are not instrumenting: a perspective on hidden safety-critical challenges in modern AI systems
#386907
arXiv CS
AI for Computational Design Science: A Responsible Human-AI Framework and Case Study on Short-Form Video Safety Surveillance
#660859
arXiv CS
AI for Computational Design Science: A Responsible Human-AI Framework and Case Study on Short-Form Video Safety Surveillance
#972206
arXiv (All)
Steering Generative AI Toward Developmentally Supportive Learning: The SCAFFOLD Framework and a Pilot in a School Setting
#976670
PsyArXiv
Questionnaire Responses Do not Capture the Safety of AI Agents
#95244
PhilArchive
How Did I Get Here: From Cowboy Dreams to AI Safety Research
#98245
PhilArchive
What is AI safety? What do we want it to be?
#103442
PhilArchive
Xenoreproduction: Exploration and Recovery of Collapsible Modes as Core AI Safety Objective
#107603
PhilArchive
Towards China-initiated actions on AI safety and governance.
#150290
NCBI PubMed Central
The Containment Paradox: Intelligence-Asymmetry and the Limits of Unamplified Supervisory AI Safety
#167256
PhilArchive
Physician-Reported Safety Outcomes of AI-Generated Hospital Course Summaries.
#170490
NCBI PubMed Central
Wearable AI to enhance patient safety and clinical decision-making
#240958
STORRE
Multi-Agent AI Safety as an Institutional Design Problem
#441790
arXiv CS
Why AI Safety Requires Uncertainty, Incomplete Preferences, and Non-Archimedean Utilities
#709617
PhilArchive
VVUQ Physical AI Oncology Trial Bill
#264277
Zenodo (CERN)
From doubt to trust: how AI safety voice bridges the gap between employees and AI teammates through perceived capability and benevolence.
#334636
NCBI PubMed Central
Let Me Check on You: Job Quality Under AI and Human Oversight
#370184
EconStor
Shared Answerability as a Condition of AI Safety: Beyond Alignment Theater and Behavioral Adequacy
#109453
PhilArchive
HaloGuard 1.0: An Open Weights Constitutional Classifier for Multilingual AI Safety
#332458
arXiv CS
Tool Specifications Matter: Uncovering and Mitigating Safety Risks in AI Agents
#422304
arXiv CS
The 2026 Singapore Consensus on Global AI Safety Research Priorities
#609049
arXiv (OAI Expanded)
Regret Dominates Surprise: Design-Time Requirements Engineering for Agentic-AI Safety
#972743
arXiv (All)
Automated Safety Testing and Reporting Application for Conversational Safety Monitoring of Generative AI Tools for Mental Health: Development and Validation Study.
#258025
NCBI PubMed Central
Safety Mechanisms and Risk Mitigation in Generative AI Mental Health Chatbots: A Systematic Scoping Review
#16421
OSF
OP9 — Operational Deployment and Regulatory Framework
#260945
Zenodo (CERN)
SSAIL: A Design Framework for Safe and Sound AI for Learning
#613638
OSF
Therapeutic Efficacy and Safety Evaluation of AI in the Management of Diabetes: A RCT Trial
#141168
ClinicalTrials.gov
AI Alignment Strategies from a Risk Perspective: Independent Safety Mechanisms or Shared Failures?
#96598
PhilArchive
Alignment of Policy, Practice, and Patient Safety for Trustworthy AI in Radiology.
#358343
NCBI PubMed Central
From Informational Complexity to Paradigmatic Safety: Theory, Diagnostics, and Nurturing of Agency in AI Systems
#688576
PhilArchive
Risk of What? Defining Harm in the Context of AI Safety
#688794
PhilArchive
Authority Before Action: A Conditional Failure-Chain Framework for Tool-Using AI Safety
#723250
PhilArchive
Subliminal Learning and Radiant Transmission in LLM Entrainment: Rethinking AI Safety with Quantitative Symbolic Dynamics
#723934
PhilArchive
← Previous
Page 2 of 5
Next →
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.