000 00847nam a22002537a 4500
001 57898
005 20260713151909.0
008 260713s2024 bg |||||||| |||| 00| 0 eng d
020 _a9781835882702
040 _aBD-DhIUB
_cBD-DhIUB
_dBD-DhIUB
082 _222
_a006.31
_bL299d
100 _aLapan, Maxim
_97239
242 _aDeep reinforcement learning hands- on
245 _aDeep reinforcement learning hands- on :
_ba practical and easy - to- follow guide to RL from Q- learning and DQNs to PPO and RLHF/
_cMaxim Lapan
250 _a3rd ed.
260 _aBirmingham ;
_bPackt Publishing Ltd.
_c2024
300 _axxviii, 684p:
_c26cm
526 _aCSE
_bps
_lREF
541 _aOmni concept
650 0 _aDeep learning
_97100
650 0 _aReinforcement learning
_97175
650 0 _aArtificial intelligence
_96057
942 _2ddc
_cBK
999 _c57898
_d58072