Daniel Han on kernels, RL, and reward hacking in agents

Daniel Han (Unsloth) presents an advanced seminar covering kernels, reinforcement learning, and reward hacking in AI agents. The talk assumes familiarity with his previous AI Engineer workshops from 2024 and 2025.
Featured · Daniel Han
1 source
Daily brief
Get tomorrow's AI brief in your inbox
More stories today
- Liquid AI open-sources Pipette benchmarking suite for on-device models
- wikiHow sues OpenAI over copyright infringement in AI training
- Claude Code 2.1.246 adds Auto mode tab, Bash wildcard warning
- Korean AI startup Wrtn raises funds at $870M valuation
- Podcast: Google DeepMind's Vivek Natarajan on AI in healthcare