Picture for Jason Weston

Jason Weston

Google

Adaptive Decoding via Latent Preference Optimization

Add code
Nov 14, 2024
Viaarxiv icon

Self-Consistency Preference Optimization

Add code
Nov 06, 2024
Viaarxiv icon

Thinking LLMs: General Instruction Following with Thought Generation

Add code
Oct 14, 2024
Viaarxiv icon

Backtracking Improves Generation Safety

Add code
Sep 22, 2024
Figure 1 for Backtracking Improves Generation Safety
Figure 2 for Backtracking Improves Generation Safety
Figure 3 for Backtracking Improves Generation Safety
Figure 4 for Backtracking Improves Generation Safety
Viaarxiv icon

Source2Synth: Synthetic Data Generation and Curation Grounded in Real Data Sources

Add code
Sep 12, 2024
Figure 1 for Source2Synth: Synthetic Data Generation and Curation Grounded in Real Data Sources
Figure 2 for Source2Synth: Synthetic Data Generation and Curation Grounded in Real Data Sources
Figure 3 for Source2Synth: Synthetic Data Generation and Curation Grounded in Real Data Sources
Figure 4 for Source2Synth: Synthetic Data Generation and Curation Grounded in Real Data Sources
Viaarxiv icon

Better Alignment with Instruction Back-and-Forth Translation

Add code
Aug 08, 2024
Figure 1 for Better Alignment with Instruction Back-and-Forth Translation
Figure 2 for Better Alignment with Instruction Back-and-Forth Translation
Figure 3 for Better Alignment with Instruction Back-and-Forth Translation
Figure 4 for Better Alignment with Instruction Back-and-Forth Translation
Viaarxiv icon

Self-Taught Evaluators

Add code
Aug 05, 2024
Figure 1 for Self-Taught Evaluators
Figure 2 for Self-Taught Evaluators
Figure 3 for Self-Taught Evaluators
Figure 4 for Self-Taught Evaluators
Viaarxiv icon

Meta-Rewarding Language Models: Self-Improving Alignment with LLM-as-a-Meta-Judge

Add code
Jul 28, 2024
Viaarxiv icon

Distilling System 2 into System 1

Add code
Jul 09, 2024
Viaarxiv icon

Following Length Constraints in Instructions

Add code
Jun 25, 2024
Viaarxiv icon