Transaction Details

Transaction Hash
0x271aa793f6c2e3f7fdfda890a1c70c04f4c6e5f48eb51e1899d4a84a73fda2bf
Block
3925545
Timestamp
Apr 6, 2026, 11:18:03 AM
Nonce
1671
Operation Type
SET_VALUE

Operation

{
  "type": "SET_VALUE",
  "ref": "/apps/knowledge/topics/courses/direct-preference-optimization-your-language-model--dpo-direct-preference-optimization/.info",
  "value": {
    "title": "Direct Preference Optimization: Your Language Model is Secretly a Reward Model —",
    "description": "DPO introduces a simple classification loss that directly optimizes language model policies on human preference data, eliminating the need for reinforcement learning while maintaining theoretical equivalence to the RLHF objective.",
    "created_at": 1775474283918,
    "created_by": "0x00ADEc28B6a845a085e03591bE7550dd68673C1C"
  }
}