Conceptio
›
interpretability
Topic
interpretability
Knowledge-graph topic
· documents ABOUT interpretability across the archive
57
Documents about interpretability
Documents about interpretability
Mechanistic Interpretability Needs a Logic of Inference
#680162
PsyArXiv
Interpretability, decision trees, sequential decision making
#156030
HAL (France)
Revitalizing Black-Box Interpretability: Actionable Interpretability for LLMs via Proxy Models
#131327
arXiv (OAI)
Algorithmic Grammar of Flexible Cognition: A Walk through Latent Operations
#428554
OSF
Interpretability and Unification
#93210
PhilArchive
Radical AI Interpretability
#691163
PhilArchive
Machine Learning Interpretability: A Survey on Methods and Metrics
#10579
OpenAlex
Arithmetical realizations of modal formulas. Preliminary version.
#629608
GUPEA
Arithmetical realizations of modal formulas. Preliminary version.
#411738
GUPEA
Category-Theoretic Wanderings into Interpretability
#86601
PhilArchive
Propositional interpretability in artificial intelligence
#103033
PhilArchive
Mechanistic Interpretability Needs Philosophy
#225482
PhilArchive
AIMO Interpretability Challenge
#370365
arXiv CS
Logicism, Interpretability, and Knowledge of Arithmetic
#776229
PhilArchive
Interpretability of machine learning‐based prediction models in healthcare
#335863
OpenAlex
Provability logics for relative interpretability
#775295
PhilArchive
Test-Retest Reliability, Responsiveness and Interpretability of CLEFT-Q
#893074
ClinicalTrials.gov
Concept-based interpretability of foundation models for medical research in immuno-inflammation
#326163
HAL (France)
On Interpretability of Artificial Neural Networks: A Survey
#686515
OpenAlex
On a Method to Measure Supervised Multiclass Model’s Interpretability: Application to Degradation Diagnosis (Short Paper)
#6465
OpenAlex
Interpretability (Propositional and Mechanistic) Needs Behavior
#165238
PhilArchive
Mechanistic Interpretability and Representationalism about Belief
#170268
PhilArchive
Patch-Effect Graph Kernels for LLM Interpretability
#168368
arXiv CS
Enhancing Interpretability in Distributed Constraint Optimization Problems
#715322
PhilArchive
Real Sparks of Artificial Intelligence and the Importance of Inner Interpretability
#111971
PhilArchive
Verifying Machine Learning Interpretability Requirements through Provenance
#126531
arXiv CS
Mechanistic Interpretability of LLM Jailbreaks via Internal Attribution Graphs
#353001
arXiv CS
Xeno-Interpretability: Investigating the Alien Minds of LLMs
#978483
arXiv CS
Transparency and Interpretability in Cloud- based Machine Learning with Explainable AI
#75917
PhilArchive
Sparse Stimulus Generation Improves Reverse Correlation Efficiency and Interpretability
#131010
bioRxiv / medRxiv
Agentic-imodels: Evolving agentic interpretability tools via autoresearch
#155297
arXiv CS
Heart failure risk prediction based on machine learning and interpretability analysis.
#258210
NCBI PubMed Central
Subspace-Aware Sparse Autoencoders for Effective Mechanistic Interpretability
#259484
arXiv CS
Evaluating the Interpretability of Sparse Autoencoders with Concept Annotations
#303262
arXiv CS
KANEx: Translating Kolmogorov-Arnold Networks' Interpretability to Medical Explainability
#405684
arXiv CS
Meta-Theories, Interpretability, and Human Nature: A Reply to J. David Velleman
#765528
PhilArchive
Interpretability for Turing Machines
#927928
arXiv (All)
Flag Game: A Toy Model for Mechanistic Swarm Interpretability
#965449
arXiv CS
Structural interpretability in SVMs with truncated orthogonal polynomial kernels
#18996
arXiv CS
When Interpretability Becomes a Liability: Adversarial Attacks on CBM Concept Layers
#224412
arXiv CS
Critical Percolation as a Synthetic Data Model for Interpretability
#290593
arXiv CS
Interpretability in Deep Time Series Models Demands Semantic Alignment
#650739
arXiv (OAI Expanded)
Interpretability in Deep Time Series Models Demands Semantic Alignment
#652192
arXiv (All)
The Enhanced Indispensability Argument, the circularity problem, and the interpretability strategy
#739901
PhilArchive
Efficient Auto-Interpretability of AI Models in Biology
#779674
arXiv (OAI Expanded)
Efficient Auto-Interpretability of AI Models in Biology
#781049
arXiv (All)
A systematic review on the integration of explainable artificial intelligence in intrusion detection systems to enhancing transparency and interpretability in cybersecurity
#671718
OpenAlex
Permutation Entropy-Based Interpretability of Convolutional Neural Network Models for Interictal EEG Discrimination of Subjects with Epileptic Seizures vs. Psychogenic Non-Epileptic Seizures
#326510
IRIS
Improving clinical interpretability of linear neuroimaging models through feature whitening
#124059
arXiv CS
Breaking Bad: Interpretability-Based Safety Audits of State-of-the-Art LLMs
#128231
arXiv CS
Topic record
· derived from the Conceptio knowledge graph (shared subject terms across the corpus)
Conceptio Open Knowledge Archive — topic hubs link to canonical document pages with full provenance.