Pith. sign in

Paper Citation Record · LEDGER

Improving alignment of dialogue agents via targeted human judgements

As of 22 July 2026, this Paper Citation Record lists 15 of 15 outbound references and 48 inbound Pith citation observations for arXiv:2209.14375.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2209.14375 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T11:50:26.030339Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-07-20T06:30:07.809122+00:00

measured 48 of 48 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-13T02:24:25.366633Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-10T06:15:00.866473Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact5
  • verified fuzzy2
  • unresolved3
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d98b8c90-efa2-42a7-8bce-143ed9b0172d · outbound

This paper cites Supervising strong learners by amplifying weak experts.

Improving alignment of dialogue agents via targeted human judgements Supervising strong learners by amplifying weak experts

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T17:54:02.244819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:2946e8a10ed47f22983d5778e69782f4bd9a5a043e2eecbc428dc75733b3f46d

Observation af213c98-4d24-491a-b09f-c987fb864e9c · outbound

This paper cites doi: 10.18653/v1/2021.emnlp-main.444.

Improving alignment of dialogue agents via targeted human judgements doi: 10.18653/v1/2021.emnlp-main.444

Reference 2

Resolution
verified exact
doi, observed 2026-05-14T17:54:02.249506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:b1c7e4198a961c9f4f0b0712d4d3a945777560b0f40540cd0edc260dc7041ebb

Observation a5c68135-9860-48a8-9d32-47311cfb4db1 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Improving alignment of dialogue agents via targeted human judgements Adam: A Method for Stochastic Optimization

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-05-14T17:54:02.287082Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:13c84a12b9f133a014e202139460e4452f6b1292a5e9490b0185dadd1c11a427

Observation 3afcb4b2-66bb-4530-be4b-2fb6908e4590 · outbound

This paper cites Procaccia.

Improving alignment of dialogue agents via targeted human judgements Procaccia

Reference 4

Resolution
verified exact
doi, observed 2026-05-14T17:54:02.254226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:560eee519dfd58bea63beb668ca2b06123b38f72fbca689ac5c4cd0a30678bb9

Observation 93835384-2bc5-459a-830b-87c943bf3349 · outbound

This paper cites WebGPT: Browser-assisted question-answering with human feedback.

Improving alignment of dialogue agents via targeted human judgements WebGPT: Browser-assisted question-answering with human feedback

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-05-14T17:54:02.260722Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:65c0de7e44bd4449ca27db37ff2fdd0a4faa3eac7cebf9d15258461db6d63968

Observation 4debb08a-1c94-4462-82f0-c91fa6b6f80d · outbound

This paper cites V-MPO: On-Policy Maximum a Posteriori Policy Optimization for Discrete and Continuous Control.

Improving alignment of dialogue agents via targeted human judgements V-MPO: On-Policy Maximum a Posteriori Policy Optimization for Discrete and Continuous Control

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.295634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:9be5aef57e02a260480132a55b96219021d2e06255cfe2e85ec79ac787e14bc0

Observation 386f3ecc-a97c-47df-a4fd-0f8972cb10e7 · outbound

This paper cites LaMDA: Language Models for Dialog Applications.

Improving alignment of dialogue agents via targeted human judgements LaMDA: Language Models for Dialog Applications

Reference 7

Resolution
malformed identifier
local_arxiv, observed 2026-05-14T17:54:02.267347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-07-11T11:50:26.030339Z digest=sha256:bfcd3da0c821bebebd06a9056d8d6785e4ebe04bc5e613550189c7de6f3a20e4

Observation b5563820-f415-45ca-a8c1-9db788aba8a7 · outbound

This paper cites emnlp-main.308/.

Improving alignment of dialogue agents via targeted human judgements emnlp-main.308/

Reference 8

Resolution
metadata mismatch
doi, observed 2026-05-14T17:54:02.272260Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:4a58cbee4d363f2a7319e262766a9dc8b374fb4a8ee837a4fed38a724e95f880

Observation 385683d5-b16f-4ce7-8f36-0c1a1fef9131 · outbound

This paper cites Conversational Information Seeking.

Improving alignment of dialogue agents via targeted human judgements Conversational Information Seeking

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.280056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:45781a8b77de230c30bb404426e6a0b7acdc711615ecda18f002cfcd9131fb38

Observation ccc3b299-5485-4eff-bd8e-b8c437fcd67f · outbound

This paper cites Sparrow.

Improving alignment of dialogue agents via targeted human judgements Sparrow

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T17:54:02.322557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:0b4a5f3ffd882ffb65159e60423c0c1fa03224cca6ba0e494acc5616e0aa378c

Observation e35bae23-e256-4b88-9636-d95f6660b7f7 · outbound

This paper cites an unresolved cited work.

Improving alignment of dialogue agents via targeted human judgements Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-05-14T17:54:02.313946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:d1d19b1464f5235c53da6eaad835c7b264d5bf18cc28862bb05aea1c82793f11

Observation 27850ea4-cc54-4c69-a177-cffd201e0bfc · outbound

This paper cites Search Query.

Improving alignment of dialogue agents via targeted human judgements Search Query

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-14T17:54:02.305203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:7b8f6f1ecf734a715009590a8bad61b7b6179c269a788d84ceba80982a257a0e

Observation 40f41549-f49e-4e46-85db-c0682d42e3b5 · outbound

This paper cites an unresolved cited work.

Improving alignment of dialogue agents via targeted human judgements Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-05-14T17:54:02.309212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:00a342486a50088992310254bbe16d73fc35e4c247a53018f68b9200572e31b1

Observation 80346297-26f4-4f2c-abd1-8c83a5fe845b · outbound

This paper cites Sparrow.

Improving alignment of dialogue agents via targeted human judgements Sparrow

Reference 14

Resolution
malformed identifier
raw_fallback, observed 2026-05-14T17:54:02.300685Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:a70cf64d7c9d0408286157a5edad4f81ad77e9e4497fa6c3bcf06a5488c0c76d

Observation 04d0b30b-d6a9-4603-99c2-81e90b9eee06 · outbound

This paper cites an unresolved cited work.

Improving alignment of dialogue agents via targeted human judgements Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-05-14T17:54:02.318171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:54:02.217086Z digest=sha256:69dd83d37938e44409e8e96aac173ef6cd8e5d573f449ec17291ecd9006b33f0

Pith citing papers

Observation 1b11db04-25bd-472b-8a8a-ce02b4b07332 · inbound

The Flan Collection: Designing Data and Methods for Effective Instruction Tuning cites this paper.

The Flan Collection: Designing Data and Methods for Effective Instruction Tuning Improving alignment of dialogue agents via targeted human judgements

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-05-24T09:14:16.363902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-24T09:13:30.054153Z digest=sha256:2312ab3945ceec01ef70f48193d2c477fad897f30fa9c6e2317c24c5680f8c36

Observation 7d1a0ea5-566c-4df5-b07e-4a56bc74b0ec · inbound

PaLM-E: An Embodied Multimodal Language Model cites this paper.

PaLM-E: An Embodied Multimodal Language Model Improving alignment of dialogue agents via targeted human judgements

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-10T22:29:29.631351Z digest=sha256:8c679986e4325c42cc8a2eedddc5749da6b8a9372f309418635ea8220c23fd6d

Observation f5ba5e05-0800-4eb0-a25d-92dfe94671ba · inbound

Language Models can Solve Computer Tasks cites this paper.

Language Models can Solve Computer Tasks Improving alignment of dialogue agents via targeted human judgements

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-17T12:17:26.817521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-17T12:17:26.602361Z digest=sha256:69aafe9ac6ce2ae0b40ea3f5e9cc4c9dd6427f49cc5981fa200eb815f5a9e772

Observation 7e140afe-c497-49a5-b7fa-cf2b63fafc98 · inbound

BloombergGPT: A Large Language Model for Finance cites this paper.

BloombergGPT: A Large Language Model for Finance Improving alignment of dialogue agents via targeted human judgements

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-13T23:19:46.231145Z digest=sha256:ab250b1f8f731dd2ec08fa0aed2aded73346e4e2a5cb51094180b2fe19848b9b

Observation 7bc91dd3-2385-4219-951f-cf7388e588af · inbound

CAMEL: Communicative Agents for "Mind" Exploration of Large Language Model Society cites this paper.

CAMEL: Communicative Agents for "Mind" Exploration of Large Language Model Society Improving alignment of dialogue agents via targeted human judgements

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T01:40:53.351795Z digest=sha256:63cbadcc8e18534ecb601ce18ab8542422c11bb47c2b2b9ebcaf07f7da9c2f44

Observation 9d49cffe-b716-4b73-bf07-82b6d5da00c9 · inbound

A Survey of Large Language Models cites this paper.

A Survey of Large Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 118

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-10T22:46:39.268353Z digest=sha256:a3136f1b884ba051e9a26d968200be0059f1285e3ff8e2c9305a99dae30609b6

Observation c20379ac-a2e3-438f-b4e4-1b5377f7b2f8 · inbound

RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment cites this paper.

RAFT: Reward rAnked FineTuning for Generative Foundation Model Alignment Improving alignment of dialogue agents via targeted human judgements

Reference 119

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T00:46:56.795939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-18T00:46:56.664582Z digest=sha256:0bad517b0f17e08716e2629e021fd871c3dee71c13fb13a89427956c76625a88

Observation 7d6d9625-e4ed-40c1-a1fd-606bc9cb6f16 · inbound

PaLM 2 Technical Report cites this paper.

PaLM 2 Technical Report Improving alignment of dialogue agents via targeted human judgements

Reference 53

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-12T11:59:25.813128Z digest=sha256:9597bcf28ee0a4a76db1c340c121ccb20882abbdc760d882490e6e49351c50e6

Observation 74a55406-f7e1-4278-b320-97902e6ee926 · inbound

QLoRA: Efficient Finetuning of Quantized LLMs cites this paper.

QLoRA: Efficient Finetuning of Quantized LLMs Improving alignment of dialogue agents via targeted human judgements

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-11T13:29:53.345251Z digest=sha256:0d9d9b438de92e92728643ec6fd6a58458f0c65f70fc32a7703da76ec55aca42

Observation de41d0a1-bb5b-48d2-b205-947e9475bae3 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 167

Resolution
verified exact
local_arxiv, observed 2026-05-19T20:28:39.538003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:29e820dfcbb61df90efb813b85f9a06e5ced5f6318847fae0d8b826bc7cc1e5c

Observation d5dc8f3f-5cae-44a5-9567-4dfa8e5f5e6e · inbound

Universal and Transferable Adversarial Attacks on Aligned Language Models cites this paper.

Universal and Transferable Adversarial Attacks on Aligned Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-24T07:44:08.427372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-24T07:42:09.112946Z digest=sha256:beb8b47a683fe7c2b14b2681238abea1c931fa138bb6b6cf65ca9dbab30f9a51

Observation 9ca26385-f46c-49a6-bdb8-22a94774854d · inbound

XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models cites this paper.

XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T06:51:50.830854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-15T06:51:50.771676Z digest=sha256:4fcc85cd66fd0fe4b9c0513c264d003e813cf56ac5819ca65139669e923b0703

Observation 6ad3b56f-7190-45cc-b033-98d975fea961 · inbound

Simple synthetic data reduces sycophancy in large language models cites this paper.

Simple synthetic data reduces sycophancy in large language models Improving alignment of dialogue agents via targeted human judgements

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-05-16T14:48:08.573326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-16T14:48:08.508109Z digest=sha256:2e1091c84b45aa82da1d19f6ae9e63ed8eca1f6dac2878e789dc4e114c74aa9a

Observation 0297edf2-e3f1-4a01-98ac-521b63b696ff · inbound

Reinforced Self-Training (ReST) for Language Modeling cites this paper.

Reinforced Self-Training (ReST) for Language Modeling Improving alignment of dialogue agents via targeted human judgements

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-13T07:59:55.849296Z digest=sha256:204bb101b3225e14b22c184ac719faafa170c7245c5112966730360efacd6d03

Observation 648e8c07-7655-41b0-83f4-50a6a2465ecd · inbound

RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback cites this paper.

RLAIF vs. RLHF: Scaling Reinforcement Learning from Human Feedback with AI Feedback Improving alignment of dialogue agents via targeted human judgements

Reference 79

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T21:32:28.041672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-15T21:32:27.806494Z digest=sha256:290516eb88e34d6c5f400bbdc4a8de964ae5bbc116046d57ab11793c514a2e59

Observation 8d7b35c0-c9a4-4a13-9a34-890f92b07e00 · inbound

Baseline Defenses for Adversarial Attacks Against Aligned Language Models cites this paper.

Baseline Defenses for Adversarial Attacks Against Aligned Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-13T23:24:39.835347Z digest=sha256:28a8bc31de32ec58f1364ad9c744a2cd13d93b8e92ea7c7cea00c2a1cea2c9ab

Observation 269a3f2c-6422-4d1c-a73d-754330a26a63 · inbound

Directly Fine-Tuning Diffusion Models on Differentiable Rewards cites this paper.

Directly Fine-Tuning Diffusion Models on Differentiable Rewards Improving alignment of dialogue agents via targeted human judgements

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-16T09:11:32.078538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-16T09:11:32.018262Z digest=sha256:c90694578e8c63b8fa1d02ca63ff24b5f9ef4b42ca621ea6f2851a02dee95947

Observation 0dfeb371-c573-4464-adc4-33a2426f89b2 · inbound

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks cites this paper.

SmoothLLM: Defending Large Language Models Against Jailbreaking Attacks Improving alignment of dialogue agents via targeted human judgements

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T17:11:00.639293Z digest=sha256:6ddc8bbe3792d7ce99cedc18412de3d4459cbc7c97b6209201a32ea93b7b1125

Observation c9ff2a56-a7c7-4879-975d-e00af6b1da31 · inbound

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation cites this paper.

Catastrophic Jailbreak of Open-source LLMs via Exploiting Generation Improving alignment of dialogue agents via targeted human judgements

Reference 9

Resolution
verified exact
local_arxiv, observed 2026-05-16T22:00:51.541408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-16T22:00:51.487120Z digest=sha256:89c31a58ec7d62ec0f4ac1b41da16b3e007a14158a3f0a170426ba81076a2867

Observation 3e3d8e77-0260-44de-b1b4-e21c32b4252c · inbound

Jailbreaking Black Box Large Language Models in Twenty Queries cites this paper.

Jailbreaking Black Box Large Language Models in Twenty Queries Improving alignment of dialogue agents via targeted human judgements

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-12T09:48:31.721745Z digest=sha256:a9bfa2f632595cb7f4fc7140a1c5c83235bb623282f2a2aa3996a68d319a888e

Observation 585d14de-d147-4067-9ef8-6c8578009894 · inbound

Towards Understanding Sycophancy in Language Models cites this paper.

Towards Understanding Sycophancy in Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-11T06:26:29.196349Z digest=sha256:a5b2ca9c58c61f9c85359bee1e3b27fc15e30a00d1571f037955e6c23f6785e4

Observation 8a1e8d6a-1c87-4171-aaf8-871695978d5c · inbound

Gemini: A Family of Highly Capable Multimodal Models cites this paper.

Gemini: A Family of Highly Capable Multimodal Models Improving alignment of dialogue agents via targeted human judgements

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-24T05:03:55.298300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-24T05:00:28.453838Z digest=sha256:859718af9f6011ff831c44238e09bf60d5fab32d0cd34f38c0639a8aa9cbd11c

Observation 7b856ae0-c26f-49b2-93e3-0039f8becdca · inbound

A Roadmap to Pluralistic Alignment cites this paper.

A Roadmap to Pluralistic Alignment Improving alignment of dialogue agents via targeted human judgements

Reference 215

Resolution
metadata mismatch
local_arxiv, observed 2026-05-16T14:37:53.548393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-16T14:37:53.279275Z digest=sha256:d3cf232005e9acf818c746cebf970cb47cab3bcc9acb2091fd06e4afe69faa34

Observation e00f81c8-386b-480e-981e-65ba62dbdece · inbound

Large Language Models: A Survey cites this paper.

Large Language Models: A Survey Improving alignment of dialogue agents via targeted human judgements

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-11T15:22:54.023279Z digest=sha256:1d87974deaf21d08ec6fd15fc993f470c0d5d50d683e6558194917abd954803b

Observation d9c85f74-9275-46d0-83ca-9d327947aa60 · inbound

A Survey on Knowledge Distillation of Large Language Models cites this paper.

A Survey on Knowledge Distillation of Large Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 120

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T23:31:11.648394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-17T23:31:11.213552Z digest=sha256:de484c833802293a65a2b13e32ca2818e881df705dde2e98b96bb05512d0cb66

Observation cc7489fc-2aa1-44da-acba-c1adcab5df13 · inbound

Yi: Open Foundation Models by 01.AI cites this paper.

Yi: Open Foundation Models by 01.AI Improving alignment of dialogue agents via targeted human judgements

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-13T05:47:27.775529Z digest=sha256:ec99734f2727c237f689030cd01c394a66ec409615a3b004373f6d8a6257472e

Observation e36220b4-be3a-4dd6-a780-738a7987c6b7 · inbound

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models cites this paper.

Inference Scaling Laws: An Empirical Analysis of Compute-Optimal Inference for Problem-Solving with Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 129

Resolution
metadata mismatch
local_arxiv, observed 2026-05-18T06:38:36.863441Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-18T06:38:36.517935Z digest=sha256:2f05e1a044cb395ac87bcb44c12ff520ecfc09c5553e0d48e0d82f628549dd89

Observation b7a88c6c-f535-442c-843c-2e9cd9400dac · inbound

Training Language Models to Self-Correct via Reinforcement Learning cites this paper.

Training Language Models to Self-Correct via Reinforcement Learning Improving alignment of dialogue agents via targeted human judgements

Reference 82

Resolution
metadata mismatch
local_arxiv, observed 2026-05-17T12:04:10.545237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-17T12:04:10.210508Z digest=sha256:da06f890005d623956c3101e33872fb54b619ccbd9a6d60e070ed16dce8484bf

Observation 5c50f8ee-303c-4fe1-9d6b-7472aac1f356 · inbound

Reinforcement Learning from Human Feedback cites this paper.

Reinforcement Learning from Human Feedback Improving alignment of dialogue agents via targeted human judgements

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-05-22T19:32:01.019179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-22T19:27:40.991325Z digest=sha256:e68979709c80efcf9beb179165bbe4200b9e163719c87f09082b807c67bc616f

Observation 080c4370-1ed8-48eb-af16-64edfc81021d · inbound

RewardBench 2: Advancing Reward Model Evaluation cites this paper.

RewardBench 2: Advancing Reward Model Evaluation Improving alignment of dialogue agents via targeted human judgements

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-19T11:22:16.854063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-19T11:18:03.965711Z digest=sha256:ed81fee3cd7d29b9d4cd8d2392d5b1631aa2efe46a6279fae68916fdfd3a72c3

Observation 8ae01485-9409-4105-a344-7c086ba52a17 · inbound

Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback cites this paper.

Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback Improving alignment of dialogue agents via targeted human judgements

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-14T21:19:29.052677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-14T21:03:48.813600Z digest=sha256:133df7fa74ff833b12c3dbfa748bbd68330928119658a1e7e8254acf38d4ff41

Observation 07cfc8f9-d13e-442e-a87c-976d63c92fbe · inbound

Don't Make Models Guess Security and Safety: Symbolic Guardrails for Domain-Specific AI Agents cites this paper.

Don't Make Models Guess Security and Safety: Symbolic Guardrails for Domain-Specific AI Agents Improving alignment of dialogue agents via targeted human judgements

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-10T10:15:27.926105Z digest=sha256:437b30d43c3f6779655d4036a47f4d3b4a0b84b8b24bce82e705a87bbdb5d915

Observation 1470e8e3-8ad7-4e42-9225-99e9157e0c60 · inbound

Dynamics of Cognitive Heterogeneity: Investigating Behavioral Biases in Multi-Stage Supply Chains with LLM-Based Simulation cites this paper.

Dynamics of Cognitive Heterogeneity: Investigating Behavioral Biases in Multi-Stage Supply Chains with LLM-Based Simulation Improving alignment of dialogue agents via targeted human judgements

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-10T06:06:48.422940Z digest=sha256:7403bc9cdd8eed16e1054eb36b7477087598b48d185d109587376d4b7ac3d44e

Observation 2df8972c-2b47-4515-a319-d0a6a36482da · inbound

Large Language Models Outperform Humans in Fraud Detection and Resistance to Motivated Investor Pressure cites this paper.

Large Language Models Outperform Humans in Fraud Detection and Resistance to Motivated Investor Pressure Improving alignment of dialogue agents via targeted human judgements

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-09T23:40:00.572491Z digest=sha256:a4bad6a62ebc62aaed7c19a7980df8b4ec6e0b10c26729e22ca0888a0b4ccbdd

Observation b4c9b0b7-9e65-44aa-8083-6a87537f1dbe · inbound

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models cites this paper.

Transient Turn Injection: Exposing Stateless Multi-Turn Vulnerabilities in Large Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-09T21:12:39.195573Z digest=sha256:81cb9ba77af06b63339b54c5a4b41f0d1ab58fe8d5c90325ae7c4fdd2a0eddb4

Observation e0e8e369-9000-42aa-8dab-f8e1c227278d · inbound

Three Models of RLHF Annotation: Extension, Evidence, and Authority cites this paper.

Three Models of RLHF Annotation: Extension, Evidence, and Authority Improving alignment of dialogue agents via targeted human judgements

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-07T14:28:31.451466Z digest=sha256:a77da213ea50b9430b5adff34824252cdea7f4a3cd7dff67e032535dfe35e534

Observation 939608e0-bca5-4f24-a158-87813973a08d · inbound

A Meta Reinforcement Learning Approach to Goals-Based Wealth Management cites this paper.

A Meta Reinforcement Learning Approach to Goals-Based Wealth Management Improving alignment of dialogue agents via targeted human judgements

Reference 294

Resolution
verified exact
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-08T18:42:50.962120Z digest=sha256:9c4cb9ae56c1d28747eeb669fe0428d5c868b43fded2f505ca86f60493b43a69

Observation c512bae5-55c5-49e5-b0c2-c50b5211c9b5 · inbound

Curated Synthetic Data Doesn't Have to Collapse: A Theoretical Study of Generative Retraining with Pluralistic Preferences cites this paper.

Curated Synthetic Data Doesn't Have to Collapse: A Theoretical Study of Generative Retraining with Pluralistic Preferences Improving alignment of dialogue agents via targeted human judgements

Reference 70

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-11T02:30:14.693348Z digest=sha256:67a73a0768aa68cf0d4d084554c4f5f453519ae3a3f36b674bfcaf6f8ae85483

Observation ab3e76e7-7a9e-4f45-9d4e-3000530f44c7 · inbound

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching cites this paper.

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Improving alignment of dialogue agents via targeted human judgements

Reference 104

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T17:54:02.323766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-13T04:55:55.013900Z digest=sha256:0cb9005b61757f1eec8927b87d0d7dce4397d4d22e229bd7f63307f5768c818a

Observation ebe52ba7-0bec-4feb-bcd3-7c2c0a78f1d7 · inbound

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching cites this paper.

TokenRatio: Principled Token-Level Preference Optimization via Ratio Matching Improving alignment of dialogue agents via targeted human judgements

Reference 104

Resolution
metadata mismatch
local_arxiv, observed 2026-05-15T05:45:06.480296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-05-15T05:41:10.714594Z digest=sha256:9ed447e0a5ca0214031d245fa7f556829be993637019f0d784caa8d847b82ca8

Observation 47320623-2e93-49ee-9ee5-2039b641ae39 · inbound

TPMM-DPO: Trajectory-aware Preference-guided Model Merging for Iterative Direct Preference Optimization cites this paper.

TPMM-DPO: Trajectory-aware Preference-guided Model Merging for Iterative Direct Preference Optimization Improving alignment of dialogue agents via targeted human judgements

Reference 27

Resolution
verified exact
local_arxiv, observed 2026-05-25T03:45:17.564323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-05-25T03:41:52.859647Z digest=sha256:9c2e81b544b096646c7665d23f758169fc57c346c0e202a52daac546aeb0ccef

Observation 18f9985f-cd61-4bec-b599-35d08030f1d8 · inbound

Plans for Evaluating Structured Generative Search Summaries cites this paper.

Plans for Evaluating Structured Generative Search Summaries Improving alignment of dialogue agents via targeted human judgements

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-06-29T16:33:39.106155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-06-29T16:26:58.658036Z digest=sha256:ee24ced09c64809c4ce7e1ace869758e30bcb3b6452a65ec125b3d5a08f8794e

Observation fe2b6c21-64bb-4e63-8515-8d80a6b5cc50 · inbound

Toward Agentic Governance: What Shapes LLM-Agent Intervention in Public Forums? cites this paper.

Toward Agentic Governance: What Shapes LLM-Agent Intervention in Public Forums? Improving alignment of dialogue agents via targeted human judgements

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-06-28T20:42:37.857551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=arxiv_source observed=2026-06-28T18:21:31.715046Z digest=sha256:adb0d07dfa4f573561f264b92ddf4e71a4c1bcb15f0f8081e2acbbd0bf7d0ee8

Observation 58efed1a-af7d-4cc9-aee4-997dceeeac54 · inbound

SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models cites this paper.

SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models Improving alignment of dialogue agents via targeted human judgements

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-07-02T16:27:09.453365Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-06-27T22:39:36.238242Z digest=sha256:456d4183aa8ea2869e4ad8ebe1bf3350c7f82bf396baebab7e63add5a875d1cf

Observation d2743a1b-ab1a-41a4-b490-b29184a08436 · inbound

Caring Without Feeling: Affective Dynamics as the Control Layer of Human-AI Agent Collaboration cites this paper.

Caring Without Feeling: Affective Dynamics as the Control Layer of Human-AI Agent Collaboration Improving alignment of dialogue agents via targeted human judgements

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-06-30T23:35:07.571667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-06-30T23:29:42.265595Z digest=sha256:ee0e83319a7b0493de41a97d079f851afc6de7b717af435e9486ab770f367c28

Observation 743815f2-937d-4135-b6fb-8c6b40a94d44 · inbound

Investigating The Security of Modern AI and Cloud Infrastructure cites this paper.

Investigating The Security of Modern AI and Cloud Infrastructure Improving alignment of dialogue agents via targeted human judgements

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-07-04T08:29:42.309868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-07-20T06:30:07.809122+00:00.

source=pdf_text observed=2026-06-26T11:31:39.910784Z digest=sha256:e38e147e85513f75b60eb5e9613c9bf3666ee8658e9f18872030990a62a39611

Observation 9d5b41cf-7330-4772-8969-f59b338a1c54 · inbound

Safe Inference-Time Alignment via Lagrangian Reward Augmentation cites this paper.

Safe Inference-Time Alignment via Lagrangian Reward Augmentation Improving alignment of dialogue agents via targeted human judgements

Reference 91

Resolution
unresolved
no resolver link, observed 2026-07-12T07:05:47.150308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T07:05:47.150308Z digest=sha256:55905144dca80e1fe46e406bb765beac1d76bb35c42e69aa4bdcecf6d581d385

Observation 65ce5cd9-756d-4ecd-ac7b-684f3a9df1e8 · inbound

Balancing Usefulness and Naturalness: An LLM-based Curation Pipeline for Code Review Comments cites this paper.

Balancing Usefulness and Naturalness: An LLM-based Curation Pipeline for Code Review Comments Improving alignment of dialogue agents via targeted human judgements

Reference 77

Resolution
unresolved
no resolver link, observed 2026-07-13T02:24:25.366633Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-13T02:24:25.366633Z digest=sha256:98f172788bc009a62e3a83df49898d24a42d5771c537f1dba16694d8cf54fee2