ENGINEERING LLM POST-TRAINING: RLHF, Reward Modeling, DPO, GRPO, and Preference Optimization for Aligned Language Models

Prices from
30.14

Featured

COMPARE ALL WEBSHOPS (2)

Description

ENGINEERING LLM POST-TRAINING: RLHF, Reward Modeling, DPO, GRPO, and Preference Optimization for Aligned Language Models

Compare webshops (2)

Shop
Price
£ 30.14
£ 30.14
Description (0)

ENGINEERING LLM POST-TRAINING: RLHF, Reward Modeling, DPO, GRPO, and Preference Optimization for Aligned Language Models


Product specifications

Brand Independently Published
EAN
  • 9798193902928

Featured Choice
£ 30.14
To Shop