Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
Paper Citation Record · LEDGER
As of 22 July 2026, this Paper Citation Record lists 15 of 15 outbound references and 48 inbound Pith citation observations for arXiv:2209.14375.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-07-20T06:30:07.809122+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-07-13T02:24:25.366633Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-07-10T06:15:00.866473Z
15 of 15 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d98b8c90-efa2-42a7-8bce-143ed9b0172d · outbound
Improving alignment of dialogue agents via targeted human judgements Supervising strong learners by amplifying weak experts
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation af213c98-4d24-491a-b09f-c987fb864e9c · outbound
Improving alignment of dialogue agents via targeted human judgements doi: 10.18653/v1/2021.emnlp-main.444
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation a5c68135-9860-48a8-9d32-47311cfb4db1 · outbound
Improving alignment of dialogue agents via targeted human judgements Adam: A Method for Stochastic Optimization
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 3afcb4b2-66bb-4530-be4b-2fb6908e4590 · outbound
Improving alignment of dialogue agents via targeted human judgements Procaccia
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 93835384-2bc5-459a-830b-87c943bf3349 · outbound
Improving alignment of dialogue agents via targeted human judgements WebGPT: Browser-assisted question-answering with human feedback
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 4debb08a-1c94-4462-82f0-c91fa6b6f80d · outbound
Improving alignment of dialogue agents via targeted human judgements V-MPO: On-Policy Maximum a Posteriori Policy Optimization for Discrete and Continuous Control
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 386f3ecc-a97c-47df-a4fd-0f8972cb10e7 · outbound
Improving alignment of dialogue agents via targeted human judgements LaMDA: Language Models for Dialog Applications
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation b5563820-f415-45ca-a8c1-9db788aba8a7 · outbound
Improving alignment of dialogue agents via targeted human judgements emnlp-main.308/
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 385683d5-b16f-4ce7-8f36-0c1a1fef9131 · outbound
Improving alignment of dialogue agents via targeted human judgements Conversational Information Seeking
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation ccc3b299-5485-4eff-bd8e-b8c437fcd67f · outbound
Improving alignment of dialogue agents via targeted human judgements Sparrow
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation e35bae23-e256-4b88-9636-d95f6660b7f7 · outbound
Improving alignment of dialogue agents via targeted human judgements Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 27850ea4-cc54-4c69-a177-cffd201e0bfc · outbound
Improving alignment of dialogue agents via targeted human judgements Search Query
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 40f41549-f49e-4e46-85db-c0682d42e3b5 · outbound
Improving alignment of dialogue agents via targeted human judgements Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 80346297-26f4-4f2c-abd1-8c83a5fe845b · outbound
Improving alignment of dialogue agents via targeted human judgements Sparrow
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 04d0b30b-d6a9-4603-99c2-81e90b9eee06 · outbound
Improving alignment of dialogue agents via targeted human judgements Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 1b11db04-25bd-472b-8a8a-ce02b4b07332 · inbound
The Flan Collection: Designing Data and Methods for Effective Instruction Tuning Improving alignment of dialogue agents via targeted human judgements
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 7d1a0ea5-566c-4df5-b07e-4a56bc74b0ec · inbound
PaLM-E: An Embodied Multimodal Language Model Improving alignment of dialogue agents via targeted human judgements
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation f5ba5e05-0800-4eb0-a25d-92dfe94671ba · inbound
Language Models can Solve Computer Tasks Improving alignment of dialogue agents via targeted human judgements
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 7e140afe-c497-49a5-b7fa-cf2b63fafc98 · inbound
BloombergGPT: A Large Language Model for Finance Improving alignment of dialogue agents via targeted human judgements
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 7bc91dd3-2385-4219-951f-cf7388e588af · inbound
CAMEL: Communicative Agents for "Mind" Exploration of Large Language Model Society Improving alignment of dialogue agents via targeted human judgements
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 9d49cffe-b716-4b73-bf07-82b6d5da00c9 · inbound
A Survey of Large Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 118
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation c20379ac-a2e3-438f-b4e4-1b5377f7b2f8 · inbound
RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment Improving alignment of dialogue agents via targeted human judgements
Reference 119
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 7d6d9625-e4ed-40c1-a1fd-606bc9cb6f16 · inbound
PaLM 2 Technical Report Improving alignment of dialogue agents via targeted human judgements
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 74a55406-f7e1-4278-b320-97902e6ee926 · inbound
QLoRA: Efficient Finetuning of Quantized LLMs Improving alignment of dialogue agents via targeted human judgements
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation de41d0a1-bb5b-48d2-b205-947e9475bae3 · inbound
A Comprehensive Overview of Large Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 167
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation d5dc8f3f-5cae-44a5-9567-4dfa8e5f5e6e · inbound
Universal and Transferable Adversarial Attacks on Aligned Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 9ca26385-f46c-49a6-bdb8-22a94774854d · inbound
XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 6ad3b56f-7190-45cc-b033-98d975fea961 · inbound
Simple synthetic data reduces sycophancy in large language models Improving alignment of dialogue agents via targeted human judgements
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 0297edf2-e3f1-4a01-98ac-521b63b696ff · inbound
Reinforced Self-Training (ReST) for Language Modeling Improving alignment of dialogue agents via targeted human judgements
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 648e8c07-7655-41b0-83f4-50a6a2465ecd · inbound
RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback Improving alignment of dialogue agents via targeted human judgements
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 8d7b35c0-c9a4-4a13-9a34-890f92b07e00 · inbound
Baseline Defenses for Adversarial Attacks Against Aligned Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 269a3f2c-6422-4d1c-a73d-754330a26a63 · inbound
Directly Fine-Tuning Diffusion Models on Differentiable Rewards Improving alignment of dialogue agents via targeted human judgements
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 0dfeb371-c573-4464-adc4-33a2426f89b2 · inbound
SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks Improving alignment of dialogue agents via targeted human judgements
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation c9ff2a56-a7c7-4879-975d-e00af6b1da31 · inbound
Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation Improving alignment of dialogue agents via targeted human judgements
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 3e3d8e77-0260-44de-b1b4-e21c32b4252c · inbound
Jailbreaking Black Box Large Language Models in Twenty Queries Improving alignment of dialogue agents via targeted human judgements
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 585d14de-d147-4067-9ef8-6c8578009894 · inbound
Towards Understanding Sycophancy in Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 8a1e8d6a-1c87-4171-aaf8-871695978d5c · inbound
Gemini: A Family of Highly Capable Multimodal Models Improving alignment of dialogue agents via targeted human judgements
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 7b856ae0-c26f-49b2-93e3-0039f8becdca · inbound
A Roadmap to Pluralistic Alignment Improving alignment of dialogue agents via targeted human judgements
Reference 215
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation e00f81c8-386b-480e-981e-65ba62dbdece · inbound
Large Language Models: A Survey Improving alignment of dialogue agents via targeted human judgements
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation d9c85f74-9275-46d0-83ca-9d327947aa60 · inbound
A Survey on Knowledge Distillation of Large Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 120
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation cc7489fc-2aa1-44da-acba-c1adcab5df13 · inbound
Yi: Open Foundation Models by 01.AI Improving alignment of dialogue agents via targeted human judgements
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation e36220b4-be3a-4dd6-a780-738a7987c6b7 · inbound
Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 129
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation b7a88c6c-f535-442c-843c-2e9cd9400dac · inbound
Training Language Models to Self-Correct via Reinforcement Learning Improving alignment of dialogue agents via targeted human judgements
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 5c50f8ee-303c-4fe1-9d6b-7472aac1f356 · inbound
Reinforcement Learning from Human Feedback Improving alignment of dialogue agents via targeted human judgements
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 080c4370-1ed8-48eb-af16-64edfc81021d · inbound
RewardBench 2: Advancing Reward Model Evaluation Improving alignment of dialogue agents via targeted human judgements
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 8ae01485-9409-4105-a344-7c086ba52a17 · inbound
Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback Improving alignment of dialogue agents via targeted human judgements
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 07cfc8f9-d13e-442e-a87c-976d63c92fbe · inbound
Don't Make Models Guess Security and Safety: Symbolic Guardrails for Domain-Specific AI Agents Improving alignment of dialogue agents via targeted human judgements
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 1470e8e3-8ad7-4e42-9225-99e9157e0c60 · inbound
Dynamics of Cognitive Heterogeneity: Investigating Behavioral Biases in Multi-Stage Supply Chains with LLM-Based Simulation Improving alignment of dialogue agents via targeted human judgements
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 2df8972c-2b47-4515-a319-d0a6a36482da · inbound
Large Language Models Outperform Humans in Fraud Detection and Resistance to Motivated Investor Pressure Improving alignment of dialogue agents via targeted human judgements
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation b4c9b0b7-9e65-44aa-8083-6a87537f1dbe · inbound
Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation e0e8e369-9000-42aa-8dab-f8e1c227278d · inbound
Three Models of RLHF Annotation: Extension, Evidence, and Authority Improving alignment of dialogue agents via targeted human judgements
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 939608e0-bca5-4f24-a158-87813973a08d · inbound
A Meta Reinforcement Learning Approach to Goals-Based Wealth Management Improving alignment of dialogue agents via targeted human judgements
Reference 294
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation c512bae5-55c5-49e5-b0c2-c50b5211c9b5 · inbound
Curated Synthetic Data Doesn't Have to Collapse: A Theoretical Study of Generative Retraining with Pluralistic Preferences Improving alignment of dialogue agents via targeted human judgements
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation ab3e76e7-7a9e-4f45-9d4e-3000530f44c7 · inbound
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Improving alignment of dialogue agents via targeted human judgements
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation ebe52ba7-0bec-4feb-bcd3-7c2c0a78f1d7 · inbound
TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Improving alignment of dialogue agents via targeted human judgements
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 47320623-2e93-49ee-9ee5-2039b641ae39 · inbound
TPMM-DPO: Trajectory-aware Preference-guided Model Merging for Iterative Direct Preference Optimization Improving alignment of dialogue agents via targeted human judgements
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 18f9985f-cd61-4bec-b599-35d08030f1d8 · inbound
Plans for Evaluating Structured Generative Search Summaries Improving alignment of dialogue agents via targeted human judgements
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation fe2b6c21-64bb-4e63-8515-8d80a6b5cc50 · inbound
Toward Agentic Governance: What Shapes LLM-Agent Intervention in Public Forums? Improving alignment of dialogue agents via targeted human judgements
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 58efed1a-af7d-4cc9-aee4-997dceeeac54 · inbound
SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models Improving alignment of dialogue agents via targeted human judgements
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation d2743a1b-ab1a-41a4-b490-b29184a08436 · inbound
Caring Without Feeling: Affective Dynamics as the Control Layer of Human-AI Agent Collaboration Improving alignment of dialogue agents via targeted human judgements
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 743815f2-937d-4135-b6fb-8c6b40a94d44 · inbound
Investigating The Security of Modern AI and Cloud Infrastructure Improving alignment of dialogue agents via targeted human judgements
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.
Observation 9d5b41cf-7330-4772-8969-f59b338a1c54 · inbound
Safe Inference-Time Alignment via Lagrangian Reward Augmentation Improving alignment of dialogue agents via targeted human judgements
Reference 91
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65ce5cd9-756d-4ecd-ac7b-684f3a9df1e8 · inbound
Balancing Usefulness and Naturalness: An LLM-based Curation Pipeline for Code Review Comments Improving alignment of dialogue agents via targeted human judgements
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.