Elektrine lite

← Feed

@Reshirams_Rad_Slam@mastodo.neoliber.al

Post #935110

2026-04-05 04:16 UTC

I've definitely built a deterministic ai that never stores probabilities in its weights but I basically just skipped straight through dithering or noise building this from the get-go. I wonder if more purely deterministic exploration methods would've been almost as good if it's strictly on-policy, for example noisy nets only on the critic/dueling value head or rnd without any gaussian noise or epsilon-greedy or softmax.

Replies (0)

No replies.