algo
now
.net
new ·
18
pairs
atlas
problems
fields
listen
quant
AI
philosophy
algonow
/
algorithms
/
Constitutional AI
Constitutional AI
Pairings in the atlas
Self-critique and revision
standard
Preference alignment
llm-training-alignment
Rivals: other methods for the same problems
Direct preference optimization
GRPO
IPO
KTO
ORPO
Process reward modeling
RLHF
Rejection sampling fine-tuning
Reward model training
Where it sits
llm-training-alignment
·
Machine Learning & AI