Conceptio
›
Archive
›
arXiv (All)
arXiv (All)
open access
Truthful AI Advisors: A Pre-Specified Benchmark for Large Language Model Honesty Under Preference Misalignment
Balyani, Hamidreza Hasani et al.
arXiv (All) · Papers · License: Open Access
Open Source ↗
Direct PDF ↓
computationandlanguage
computerscienceandgametheory
machine learning, computation and language, computer science and game theory
This document is indexed with metadata only — full text is not available in the archive for this record.
Open the official source ↗
Related documents
A Generalized Tangent Approximation based Variational Inference Framework for Strongly Super-Gaussian Likelihoods
#1000498
Reinforcement Learning from Human Feedback
#1000499
GLaMoR: Consistency Checking of OWL Ontologies using Graph Language Models
#1000502
Record
· ID 793704
Retrieved via
Conceptio
— every document is proof-bundled with source, license, and retrieval metadata.